Corpus Corporum: Difference between revisions

From The Digital Classicist Wiki
Jump to navigation Jump to search
(structure and fleshed out)
 
Line 1: Line 1:
==Available==
* https://mlat.uzh.ch/
==Director==
* Philipp Roelli
== Description ==
== Description ==


The Corpus Corporum is a repository of Latin texts being developed at the University of Zurich and led by Philipp Roelli. The project database is still under development. Its goals include:
The '''Corpus Corporum''' is a repository of Latin texts being developed at the University of Zurich and led by Philipp Roelli. The project database is still under development. Its goals include:


<ref>
<blockquote>
 
To provide a platform into which standardised (TEI) xml-files of Latin texts can be loaded (if you would like to share your texts, please contact us) and downloaded (unless copyrights or the texts' providers restrict this).
"To provide a platform into which standardised (TEI) xml-files of Latin texts can be loaded (if you would like to share your texts, please contact us) and downloaded (unless copyrights or the texts' providers restrict this).
<br>
<br>


To make these texts searchable in complex manners (including proximity
To make these texts searchable in complex manners (including proximity
search and lemmatised search). Search results, wordlists and concordances can be generated for the current text level at the right hand side of the page (we use the open-source software Sphinx).
search and lemmatised search). Search results, wordlists and concordances can be generated for the current text level at the right hand side of the page (we use the open-source software Sphinx).
<br>
<br>


To provide a platform to publish Latin texts online (cf. the Richard Rufus Project's corpus).
To provide a platform to publish Latin texts online (cf. the Richard Rufus Project's corpus).
Texts may be downloaded as TEI xml for non-commercial use and can thus be reused by other researchers.
Texts may be downloaded as TEI xml for non-commercial use and can thus be reused by other researchers.
Note that the XML files come from various sources and, although they should approach the TEI xml standard, some of them do not validate as TEI because of minor differences from the rigid TEI guide-lines." (Corpus Corporum). </ref>
Note that the XML files come from various sources and, although they should approach the TEI xml standard, some of them do not validate as TEI because of minor differences from the rigid TEI guide-lines." (Corpus Corporum).
 
</blockquote>
Source: mlat.uzh.ch/home


More information can be found on its website: [http://mlat.uzh.ch mlat.uzh.ch]
[[category:projects]]
[[category:corpora]]
[[category:linguistics]]

Latest revision as of 17:09, 13 August 2026

Available

Director

  • Philipp Roelli

Description

The Corpus Corporum is a repository of Latin texts being developed at the University of Zurich and led by Philipp Roelli. The project database is still under development. Its goals include:

To provide a platform into which standardised (TEI) xml-files of Latin texts can be loaded (if you would like to share your texts, please contact us) and downloaded (unless copyrights or the texts' providers restrict this).

To make these texts searchable in complex manners (including proximity search and lemmatised search). Search results, wordlists and concordances can be generated for the current text level at the right hand side of the page (we use the open-source software Sphinx).

To provide a platform to publish Latin texts online (cf. the Richard Rufus Project's corpus). Texts may be downloaded as TEI xml for non-commercial use and can thus be reused by other researchers. Note that the XML files come from various sources and, although they should approach the TEI xml standard, some of them do not validate as TEI because of minor differences from the rigid TEI guide-lines." (Corpus Corporum).