<?xml version='1.0' encoding='UTF-8'?>
<OAI-PMH xmlns="http://www.openarchives.org/OAI/2.0/" xmlns:xsi="http://www.w3.org/2001/XMLSchema-instance" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/ http://www.openarchives.org/OAI/2.0/OAI-PMH.xsd">
  <responseDate>2026-09-12T23:13:02Z</responseDate>
  <request set="user-hzsk" metadataPrefix="oai_dc" verb="ListRecords">https://www.fdr.uni-hamburg.de/oai2d</request>
  <ListRecords>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:1442</identifier>
        <datestamp>2022-06-28T11:26:16Z</datestamp>
        <setSpec>user-hzsk</setSpec>
        <setSpec>user-uhh</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:creator>Meyer, Bernd</dc:creator>
          <dc:date>2006-08-24</dc:date>
          <dc:description>Audio and video recordings of three lectures in Portuguese, one simultaneously and two consecutively professionally interpreted into German. For the simultaneouly interpreted lecture there are different recordings and transcriptions for the participants.

CoSi is a corpus of consecutive and simultaneous interpreting, created for the research project "Coherence in interpreter-mediated discourse" at the Research Center on Multilingualism, University of Hamburg.

 

CLARIN Metadata summary for Consecutive and Simultaneous Interpreting (CoSi)Konsekutives und Simultanes Dolmetschen (CoSi) (CMDI-based)

Title: Consecutive and Simultaneous Interpreting (CoSi)
Title: Konsekutives und Simultanes Dolmetschen (CoSi)
Description: Audio and video recordings of three lectures in Portuguese, one simultaneously and two consecutively professionally interpreted into German. For the simultaneouly interpreted lecture there are different recordings and transcriptions for the participants.
Publication date: 2010-02-26
Data owner:  Prof. Dr. Bernd Meyer, Arbeitsbereich Interkulturelle Kommunikation, Johannes Gutenberg-Universität Mainz, meyerb@uni-mainz.de
Contributors:  Prof. Dr. Bernd Meyer, Arbeitsbereich Interkulturelle Kommunikation, Johannes Gutenberg-Universität Mainz, meyerb@uni-mainz.de, Moritz Essmann (compiler)
Project:  K6 "Coherence in interpreter-mediated discourse", German Research Foundation (DFG)
Keywords:  professional interpreting, consecutive interpreting, simultaneous interpreting, expert-laymen communication, EXMARaLDA
Languages:  German (deu), Portuguese (por)
Size:  8 speakers (6 female, 2 male), 3 communications, 5 recordings, 345 minutes, 5 transcriptions, 35432 words
Temporal Coverage:  2006
Spatial Coverage:  DE
Genre:  discourse
Modality:  spoken

 </dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/1442</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.1442</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:1442</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.1441</dc:relation>
          <dc:rights>info:eu-repo/semantics/restrictedAccess</dc:rights>
          <dc:subject>German</dc:subject>
          <dc:subject>Portuguese</dc:subject>
          <dc:subject>professional interpreting</dc:subject>
          <dc:subject>consecutive interpreting</dc:subject>
          <dc:subject>simultaneous interpreting</dc:subject>
          <dc:subject>expert-laymen communication</dc:subject>
          <dc:subject>EXMARaLDA</dc:subject>
          <dc:subject>linguistics</dc:subject>
          <dc:title>Consecutive and Simultaneous Interpreting (CoSi)Konsekutives und Simultanes Dolmetschen (CoSi)</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:1496</identifier>
        <datestamp>2021-05-27T11:10:23Z</datestamp>
        <setSpec>user-hzsk</setSpec>
        <setSpec>user-uhh</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:creator>Lleó, Conxita</dc:creator>
          <dc:date>2020-09-02</dc:date>
          <dc:description>Audio recordings of five German and six Spanish speaking monolingual children. For the German children there are 176 recordings (interviewer/child interaction), on an average starting at 9 months and ending at 3 years; for the Spanish children there are 79 recordings, on average starting at 9 months and ending at 2;6 years.

PAIDUS is a phonetically and orthographically transcribed corpus of German and Spanish child language, finalized in the project Prosodic constraints on phonological and morphological development in bilingual first language acquisition at the Research Center on Multilingualism, University of Hamburg.

Lleó, Conxita. 2012. "Monolingual and Bilingual Phonoprosodic Corpora of Child German and Child Spanish." In Multilingual Corpora and Multilingual Corpus Analysis, edited by Thomas Schmidt and Kai Wörner, 14:pp. 107–22. Hamburg Studies in Multilingualism. John Benjamins.

 

CLARIN Metadata summary for Parameterfixierung im Deutschen und Spanischen (PAIDUS) (CMDI-based)

Title: Parameterfixierung im Deutschen und Spanischen (PAIDUS)
Description: Audio recordings of five German and six Spanish speaking monolingual children. For the German children there are 176 recordings (interviewer/child interaction), on an average starting at 9 months and ending at 3 years; for the Spanish children there are 79 recordings, on average starting at 9 months and ending at 2;6 years.
Publication date: 2010-09-30
Data owner:  Conxita Lleó, Institut für Romanistik / Von-Melle-Park 6 / D-20146 Hamburg, lleo@uni-hamburg.de
Contributors:  Conxita Lleó, Institut für Romanistik / Von-Melle-Park 6 / D-20146 Hamburg, lleo@uni-hamburg.de (depositor), Conxita Lleó, Institut für Romanistik / Von-Melle-Park 6 / D-20146 Hamburg, lleo@uni-hamburg.de (compiler)
Project:  E3 "Prosodic Constraints on Phonological and Morphological development in Bilingual First Language Acquisition", German Research Foundation (DFG)
Keywords:  child language acquisition, longitudinal data, monolingual data, EXMARaLDA
Languages:  German (deu), Spanish (spa)
Size:  67 speakers (44 female, 23 male), 255 communications, 138.47 hours, 8308 minutes, 255 recordings, 253 transcriptions, 166976 words
Genre:  discourse
Modality:  spoken
References:  Lleó, Conxita. 2012. "Monolingual and Bilingual Phonoprosodic Corpora of Child German and Child Spanish." In Multilingual Corpora and Multilingual Corpus Analysis, edited by Thomas Schmidt and Kai Wörner, 14:pp. 107–22. Hamburg Studies in Multilingualism. John Benjamins.

 </dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/1496</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.1496</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:1496</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.1452</dc:relation>
          <dc:relation>doi:10.25592/uhhfdm.1484</dc:relation>
          <dc:rights>info:eu-repo/semantics/restrictedAccess</dc:rights>
          <dc:subject>child language acquisition</dc:subject>
          <dc:subject>longitudinal data</dc:subject>
          <dc:subject>monolingual data</dc:subject>
          <dc:subject>EXMARaLDA</dc:subject>
          <dc:subject>linguistics</dc:subject>
          <dc:subject>German</dc:subject>
          <dc:subject>Spanish</dc:subject>
          <dc:title>Parameterfixierung im Deutschen und Spanischen (PAIDUS)</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:1544</identifier>
        <datestamp>2021-05-27T11:10:23Z</datestamp>
        <setSpec>user-hzsk</setSpec>
        <setSpec>user-uhh</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:creator>Lleó, Conxita</dc:creator>
          <dc:date>2020-09-07</dc:date>
          <dc:description>Audio recordings of four German/Spanish simultaneous bilingual children starting at approx. 1 year and ending between the ages 2;4 and 3 years. There are 144 recording sessions (interviewer/child interaction), half of them conducted in a German and half in a Spanish speaking environment.

Phonologie-Erwerb Deutsch-Spanisch als Erste Sprachen (PEDSES) is a phonetically and orthographically transcribed corpus of German and Spanish, finalized in the project Prosodic constraints on phonological and morphological development in bilingual first language acquisition at the Research Center on Multilingualism, University of Hamburg.

Lleó, Conxita. 2012. "Monolingual and Bilingual Phonoprosodic Corpora of Child German and Child Spanish." In Multilingual Corpora and Multilingual Corpus Analysis, edited by Thomas Schmidt and Kai Wörner, 14:pp. 107–22. Hamburg Studies in Multilingualism. John Benjamins.

 

CLARIN Metadata summary for Phonologie-Erwerb Deutsch-Spanisch als Erste Sprachen (PEDSES) (CMDI-based)

Title:  Phonologie-Erwerb Deutsch-Spanisch als Erste Sprachen (PEDSES)
Description: Audio recordings of four German/Spanish simultaneous bilingual children starting at approx. 1 year and ending between the ages 2;4 and 3 years. There are 144 recording sessions (interviewer/child interaction), half of them conducted in a German and half in a Spanish speaking environment.
Data owner:  Conxita Lleó, Institut für Romanistik / Von-Melle-Park 6 / D-20146 Hamburg, lleo@uni-hamburg.de
Contributors:  Conxita Lleó, Institut für Romanistik / Von-Melle-Park 6 / D-20146 Hamburg, lleo@uni-hamburg.de (depositor), Conxita Lleó, Institut für Romanistik / Von-Melle-Park 6 / D-20146 Hamburg, lleo@uni-hamburg.de (compiler)
Project:  E3 "Prosodic Constraints on Phonological and Morphological development in Bilingual First Language Acquisition", German Research Foundation (DFG)
Keywords:  child language acquisition, child bilingualism, longitudinal data, simultaneous bilingualism, L2 data, L1 data, L1-Daten, L2-Daten, EXMARaLDA
Languages:  German (deu), Spanish (spa)
Size:  22 speakers (15 female, 7 male), 146 communications, 137.79 hours, 4550 minutes, 144 recordings, 127 transcriptions, 101292 words
Genre:  discourse
Modality:  spoken
Modality:  Lleó, Conxita. 2012. "Monolingual and Bilingual Phonoprosodic Corpora of Child German and Child Spanish." In Multilingual Corpora and Multilingual Corpus Analysis, edited by Thomas Schmidt and Kai Wörner, 14:pp. 107–22. Hamburg Studies in Multilingualism. John Benjamins.</dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/1544</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.1544</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:1544</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.1540</dc:relation>
          <dc:rights>info:eu-repo/semantics/restrictedAccess</dc:rights>
          <dc:subject>child language acquisition</dc:subject>
          <dc:subject>child bilingualism</dc:subject>
          <dc:subject>longitudinal data</dc:subject>
          <dc:subject>simultaneous bilingualism</dc:subject>
          <dc:subject>L2 data</dc:subject>
          <dc:subject>L1 data</dc:subject>
          <dc:subject>L1-Daten</dc:subject>
          <dc:subject>L2-Daten</dc:subject>
          <dc:subject>EXMARaLDA</dc:subject>
          <dc:subject>linguistics</dc:subject>
          <dc:subject>German</dc:subject>
          <dc:subject>Spanish</dc:subject>
          <dc:title>Phonologie-Erwerb Deutsch-Spanisch als Erste Sprachen (PEDSES)</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:1601</identifier>
        <datestamp>2020-09-14T20:21:17Z</datestamp>
        <setSpec>user-hzsk</setSpec>
        <setSpec>user-uhh</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:creator>Fiedler, Ines</dc:creator>
          <dc:date>2015-09-14</dc:date>
          <dc:description>The corpus comprises out of a collection of texts from discussion forums in the web, randomly chosen for their near-standard like orthography and language, and treating different topics. The texts are translated manually by a mother tongue speaker and automatically tagged by a part-of-speech tagger. No further annotation is provided.

 

CLARIN Metadata summary for B7 Wolof (web) (CMDI-based)

Title: B7 Wolof (web)
Description: The corpus comprises out of a collection of texts from discussion forums in the web, randomly chosen for their near-standard like orthography and language, and treating different topics. The texts are translated manually by a mother tongue speaker and automatically tagged by a part-of-speech tagger. No further annotation is provided.
Publication date: 2015
Data owner:  Dr. phil. Ines Fiedler
Contributors:  Tom Güldemann (editor), Ines Fiedler (researcher), Peggy Jacob (researcher), Yokiko Morimoto (researcher), Anne Schwarz (researcher), Andreas Wetter (researcher)
Project:  Special Research Centre 632 Information structure, German Research Foundation
Keywords:  predicate-centered focus types, focus
Language:  Wolof (wol)
Size:  15335 Token
Segmentation units:  other
Genre:  discourse
Modality:  written

 </dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/1601</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.1601</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:1601</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.1600</dc:relation>
          <dc:rights>info:eu-repo/semantics/openAccess</dc:rights>
          <dc:rights>https://creativecommons.org/licenses/by/3.0/legalcode</dc:rights>
          <dc:subject>predicate-centered focus types</dc:subject>
          <dc:subject>focus</dc:subject>
          <dc:subject>linguistics</dc:subject>
          <dc:subject>information structure</dc:subject>
          <dc:subject>Wolof</dc:subject>
          <dc:title>B7 Wolof (web)</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:1062</identifier>
        <datestamp>2020-11-10T20:34:18Z</datestamp>
        <setSpec>user-hzsk</setSpec>
        <setSpec>user-uhh</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:creator>Meisel, Jürgen M.</dc:creator>
          <dc:date>2020-06-22</dc:date>
          <dc:description>Additional data from the DUFDE project.</dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/1062</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.1062</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:1062</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.1446</dc:relation>
          <dc:relation>doi:10.25592/uhhfdm.1061</dc:relation>
          <dc:rights>info:eu-repo/semantics/closedAccess</dc:rights>
          <dc:subject>German</dc:subject>
          <dc:subject>French</dc:subject>
          <dc:subject>Language Acquisition</dc:subject>
          <dc:subject>Bilingualism</dc:subject>
          <dc:subject>Linguistics</dc:subject>
          <dc:title>DUFDE Media</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>video</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:1084</identifier>
        <datestamp>2020-12-03T10:43:00Z</datestamp>
        <setSpec>user-hzsk</setSpec>
        <setSpec>user-uhh</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:creator>Meisel, Jürgen M.</dc:creator>
          <dc:date>2020-06-25</dc:date>
          <dc:description>Data from the BUSDE Project</dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/1084</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.1084</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:1084</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.1083</dc:relation>
          <dc:rights>info:eu-repo/semantics/closedAccess</dc:rights>
          <dc:subject>Basque</dc:subject>
          <dc:subject>Spanish</dc:subject>
          <dc:subject>Bilingualism</dc:subject>
          <dc:subject>Language Acquisition</dc:subject>
          <dc:subject>Linguistics</dc:subject>
          <dc:title>BUSDE</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:1735</identifier>
        <datestamp>2020-09-29T21:13:27Z</datestamp>
        <setSpec>user-hzsk</setSpec>
        <setSpec>user-uhh</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:creator>House, Juliane</dc:creator>
          <dc:date>2011-01-01</dc:date>
          <dc:description>Translation corpora of original texts with translations and comparable texts from the genre external business communication.

Übersetzungs- und Vergleichskorpus mit authentischen Texten aus der Wirtschaftskommunikation. Geschäftsberichte, Aktionärsbriefe und Selbstdarstellungen von deutschsprachigen und englischsprachigen Firmen jeweils im Original und der jeweiligen Übersetzung aus dem Internet aus der Zeit von September 2005 bis März 2006.</dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/1735</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.1735</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:1735</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.1734</dc:relation>
          <dc:rights>info:eu-repo/semantics/restrictedAccess</dc:rights>
          <dc:subject>translated texts</dc:subject>
          <dc:subject>business communication</dc:subject>
          <dc:subject>parallel corpus</dc:subject>
          <dc:subject>comparable corpus</dc:subject>
          <dc:subject>übersetzte Texte</dc:subject>
          <dc:subject>Populärwissenschaftstexte</dc:subject>
          <dc:subject>Parallelkorpus</dc:subject>
          <dc:subject>Vergleichskorpus</dc:subject>
          <dc:subject>Wirtschaftskommunikation</dc:subject>
          <dc:subject>linguistics</dc:subject>
          <dc:subject>German</dc:subject>
          <dc:subject>English</dc:subject>
          <dc:title>Covert translation: Business Communication (new)</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:1709</identifier>
        <datestamp>2020-09-29T19:05:23Z</datestamp>
        <setSpec>user-hzsk</setSpec>
        <setSpec>user-uhh</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:contributor>Hartmann, Katharina</dc:contributor>
          <dc:contributor>Jacob, Peggy</dc:contributor>
          <dc:creator>Hartmann, Katharina</dc:creator>
          <dc:date>2015-09-29</dc:date>
          <dc:description>Tangale sample: sample, status: final, manually transcribed, glossed and translated to English, annotated wrt. morphology, parts of speech, syntax, gramm. function, sem. roles, focus and focus position (e.g. ex situ) in EXMARaLDA.CLARIN Metadata summary for B2 Tangale (CMDI-based)    	    	    	    		Title: B2 Tangale    	        	Description: Tangale sample: sample, status: final, manually transcribed, glossed and translated to English, annotated wrt. morphology, parts of speech, syntax, gramm. function, sem. roles, focus and focus position (e.g. ex situ) in EXMARaLDA.					Publication date: 2015					Data owner: 			Univ.-Prof. Dr. Katharina Hartmann											                		Contributors:                 	Katharina Hartmann (editor), Peggy Jacob (researcher)                			                		            	Project:             	Special Research Centre 632 Information structure, German Research Foundation            				            			            	Keywords:             	focus, Fokus            				            			    	Language:     	Tangale (tan)    						    					            	Size:             	678 Token            					            				                		Segmentation units:                 	other                			                		            	Genre:             	discourse            				            			            	Modality:             	spoken            				            					</dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/1709</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.1709</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:1709</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.1708</dc:relation>
          <dc:rights>info:eu-repo/semantics/openAccess</dc:rights>
          <dc:rights>https://creativecommons.org/licenses/by-nc/3.0/legalcode</dc:rights>
          <dc:subject>focus</dc:subject>
          <dc:subject>Fokus</dc:subject>
          <dc:subject>linguistics</dc:subject>
          <dc:subject>information structure</dc:subject>
          <dc:subject>Tangale</dc:subject>
          <dc:title>B2 Tangale</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:1732</identifier>
        <datestamp>2020-09-29T21:12:41Z</datestamp>
        <setSpec>user-hzsk</setSpec>
        <setSpec>user-uhh</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:creator>House, Juliane</dc:creator>
          <dc:date>2009-01-01</dc:date>
          <dc:description>Translation corpora of original texts with translations and comparable texts from the genre external business communication.</dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/1732</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.1732</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:1732</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.1731</dc:relation>
          <dc:rights>info:eu-repo/semantics/restrictedAccess</dc:rights>
          <dc:subject>translated texts</dc:subject>
          <dc:subject>business communication</dc:subject>
          <dc:subject>parallel corpus</dc:subject>
          <dc:subject>comparable corpus</dc:subject>
          <dc:subject>linguistics</dc:subject>
          <dc:subject>German</dc:subject>
          <dc:subject>English</dc:subject>
          <dc:title>Covert translation: Business Communication (old)</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:8326</identifier>
        <datestamp>2021-05-13T10:26:34Z</datestamp>
        <setSpec>user-uhh</setSpec>
        <setSpec>user-hzsk</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:contributor>Andresen, Melanie</dc:contributor>
          <dc:contributor>Knorr, Dagmar</dc:contributor>
          <dc:creator>Knorr, Dagmar</dc:creator>
          <dc:creator>Andresen, Melanie</dc:creator>
          <dc:date>2017-11-18</dc:date>
          <dc:description>Authentic texts written by students of the University of Hamburg as part of their studies, the students have various L1 languages and study various subjects, all of the texts were subject of a writing counseling at the Writing Center Multilingualism (Schreibwerkstatt Mehrsprachigkeit), for some of the texts comments by peer tutors and several versions are available.

Authentische Texte Studierender der Universität Hamburg unterschiedlicher Herkunftssprachen und Studiengänge, die im Rahmen ihres Studiums als Prüfungsleistung entstanden sind und in einer individuellen Schreibberatung an der Schreibwerkstatt Mehrsprachigkeit besprochen wurden, zum Teil mit Kommentaren von Peer Tutoren und mehreren Versionen aus dem Schreibprozess.

Im Institut für Interkulturelle Bildung der Fakultät für Erziehungswissenschaft an der Universität Hamburg wird seit mehreren Jahrzehnten der Zusammenhang zwischen Bildungserfolg und Bildungssprache erforscht. Es wurde beobachtet, dass Studierende mit Migrationshintergrund Probleme beim Verfassen akademischer Texte haben. Die – naheliegende – Hypothese war, dass dies an den sprachlichen Fähigkeiten der Studierenden liegt. Allerdings ist eine Überprüfung der Hypothese ohne entsprechendes empirisches Datenmaterial schwierig. Die Fachliteratur zum deutschsprachigen akademischen Schreiben weist hier eine Lücke auf: Zugängliche Korpora von authentischen, deutschsprachigen akademischen Texten von Studierenden, die möglichst auch noch longitudinal untersucht werden können, existierten nicht. Mit dem Korpus KoLaS möchte die Schreibwerkstatt Mehrsprachigkeit einen Beitrag zur Forschung leisten, indem authentische Texte von Studierenden für Forschungszwecke zur Verfügung gestellt werden.

Andresen, Melanie; Knorr, Dagmar (2017) KoLaS: Kommentiertes Lernendenkorpus akademisches Schreiben. &lt;https://www.korpuslab.uni-hamburg.de/projekte/kolas/korpusdoku-2.pdf&gt;, (13.05.2021)

 

CLARIN Metadata summary for Commented Learner Corpus Academic WritingKommentiertes Lernendenkorpus akademisches Schreiben (CMDI-based)

Title: Commented Learner Corpus Academic Writing
Title: Kommentiertes Lernendenkorpus akademisches Schreiben
Description: Authentic texts written by students of the University of Hamburg as part of their studies, the students have various L1 languages and study various subjects, all of the texts were subject of a writing counseling at the Writing Center Multilingualism (Schreibwerkstatt Mehrsprachigkeit), for some of the texts comments by peer tutors and several versions are available.
Description: Authentische Texte Studierender der Universität Hamburg unterschiedlicher Herkunftssprachen und Studiengänge, die im Rahmen ihres Studiums als Prüfungsleistung entstanden sind und in einer individuellen Schreibberatung an der Schreibwerkstatt Mehrsprachigkeit besprochen wurden, zum Teil mit Kommentaren von Peer Tutoren und mehreren Versionen aus dem Schreibprozess.
Publication date: 2017
Data owner:  Dagmar Knorr, Leuphana Universität Lüneburg / Schreibzentrum/Writing Center / Universitätsallee 1 / 21335 Lüneburg, dagmar.knorr@leuphana.de
Contributors:  Dagmar Knorr (compiler), Melanie Andresen (compiler)
Project:  Schreibwerkstatt Mehrsprachigkeit, ZEIT-Stiftung Ebelin und Gerd Bucerius (2011–2014), Schreibwerkstatt Mehrsprachigkeit, Federal Ministry of Education and Research (2013–2016)
Keywords:  academic writing, student texts, written assignments, L2 data, L1 data, aquisition of academic writing, text comments, writing counselling, EXMARaLDA, akademisches Schreiben, wissenschaftlichen Schreiben, Studierendentexte, studentische Hausarbeiten, L2-Daten, L1-Daten, Erwerb der Wissenschaftssprache, Textkommentare, Schreibberatung, EXMARaLDA
Language:  German (deu)
Size:  853 texts
Segmentation units:  text
Temporal Coverage:  2011-01-01/2016-12-31
Spatial Coverage:  Hamburg, DE
Genre:  academic writing, akademisches Schreiben
Modality:  written

 </dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/8326</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.8326</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:8326</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.8985</dc:relation>
          <dc:relation>doi:10.25592/uhhfdm.8322</dc:relation>
          <dc:rights>info:eu-repo/semantics/restrictedAccess</dc:rights>
          <dc:subject>academic writing</dc:subject>
          <dc:subject>student texts</dc:subject>
          <dc:subject>written assignments</dc:subject>
          <dc:subject>L2 data</dc:subject>
          <dc:subject>L1 data</dc:subject>
          <dc:subject>aquisition of academic writing</dc:subject>
          <dc:subject>text comments</dc:subject>
          <dc:subject>writing counselling</dc:subject>
          <dc:subject>EXMARaLDA</dc:subject>
          <dc:subject>akademisches Schreiben</dc:subject>
          <dc:subject>wissenschaftliches Schreiben</dc:subject>
          <dc:subject>Studierendentexte</dc:subject>
          <dc:subject>studentische Hausarbeiten</dc:subject>
          <dc:subject>L2-Daten</dc:subject>
          <dc:subject>L1-Daten</dc:subject>
          <dc:subject>Erwerb der Wissenschaftssprache</dc:subject>
          <dc:subject>Textkommentare</dc:subject>
          <dc:subject>Schreibberatung</dc:subject>
          <dc:subject>EXMARaLDA</dc:subject>
          <dc:subject>linguistics</dc:subject>
          <dc:subject>German</dc:subject>
          <dc:title>Commented Learner Corpus Academic Writing; Kommentiertes Lernendenkorpus akademisches Schreiben</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:8308</identifier>
        <datestamp>2022-12-07T10:12:05Z</datestamp>
        <setSpec>user-uhh</setSpec>
        <setSpec>user-hzsk</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:contributor>Najhla</dc:contributor>
          <dc:contributor>Erkan</dc:contributor>
          <dc:contributor>Bührig,, Kristin</dc:contributor>
          <dc:contributor>Süreyya</dc:contributor>
          <dc:contributor>Ayzin, Alper</dc:contributor>
          <dc:contributor>Bernd</dc:contributor>
          <dc:contributor>Meyer,, Bernd</dc:contributor>
          <dc:contributor>Carla</dc:contributor>
          <dc:contributor>João</dc:contributor>
          <dc:contributor>Pawlack, Birte</dc:contributor>
          <dc:contributor>Schmidt,, Thomas</dc:contributor>
          <dc:contributor>Forschungsgemeinschaft, Deutsche</dc:contributor>
          <dc:contributor>Isabel</dc:contributor>
          <dc:contributor>Çakmak</dc:contributor>
          <dc:contributor>Maria</dc:contributor>
          <dc:contributor>Demet</dc:contributor>
          <dc:contributor>Latif</dc:contributor>
          <dc:contributor>Anna</dc:contributor>
          <dc:creator>Bührig, Kristin</dc:creator>
          <dc:creator>Meyer, Bernd</dc:creator>
          <dc:date>2009-01-05</dc:date>
          <dc:description>Transcription of audio recordings of various kinds of doctor-patient communication in hospitals. There are both monolingual conversations in German, Portuguese and Turkish, recorded in the respective country, and interpreted conversations recorded in Germany (i.e. in German-Turkish, German-Portuguese, and German-Portuguese/Spanish), about 15-20 recordings of each kind. The persons interpreting are bilingual hospital employees or relatives of the patients, who are all adults living in Germany but with varying knowledge of German.

The corpus Dolmetschen im Krankenhaus (DiK - Interpreting in Hospitals) was compiled between July 1999 and June 2005 in the research project "Interpreting in Hospitals" (K2, principal investigator: Kristin Bührig), part of the Collaborative Research Centre "Multilingualism", funded by the German Research Foundation (Deutsche Forschungsgemeinschaft, DFG) and hosted by the University of Hamburg. The corpus is based on doctor-patient-communications between German doctors or nursing staff and patients with Turkish or Portuguese as their mother tongue. These conversations were translated by laypersons (nursing staff, relatives). Furthermore, the corpus comprises monolingual doctor-patient-communications from Germany, Portugal and Turkey as comparative data.

Bührig, Kristin; Kliche, Ortrun; Meyer, Bernd and Pawlack, Birte. 2012. "The Corpus 'Interpreting in Hospitals'. Possible Applications for Research and Communication Training." In Multilingual Corpora and Multilingual Corpus Analysis, edited by Thomas Schmidt and Kai Wörner, pp. 305–15. Hamburg Studies in Multilingualism. John Benjamins.

 

CLARIN Metadata summary for Dolmetschen im Krankenhaus (DiK) (CMDI-based)

Title: Dolmetschen im Krankenhaus (DiK)
Description: Audio recordings of various kinds of doctor-patient communication in hospitals. There are both monolingual conversations in German, Portuguese and Turkish, recorded in the respective country, and interpreted conversations recorded in Germany (i.e. in German-Turkish, German-Portuguese, and German-Portuguese/Spanish), about 15-20 recordings of each kind. The persons interpreting are bilingual hospital employees or relatives of the patients, who are all adults living in Germany but with varying knowledge of German.
Publication date: 2020-09-30
Data owner:  Kristin Bührig, Institut für Germanistik I / Von-Melle-Park 6 / D-20146 Hamburg, kristin.buehrig@uni-hamburg.de; Bernd Meyer, Arbeitsbereich Interkulturelle Kommunikation / Fachbereich 06: Translations-, Sprach- und Kulturwissenschaft / Johannes Gutenberg-Universität Mainz / An der Hochschule 2 / D-76726 Germersheim, meyerb@uni-mainz.de
Contributors:  Kristin Bührig, Institut für Germanistik I / Von-Melle-Park 6 / D-20146 Hamburg, kristin.buehrig@uni-hamburg.de (depositor), Bernd Meyer, Arbeitsbereich Interkulturelle Kommunikation / Fachbereich 06: Translations-, Sprach- und Kulturwissenschaft / Johannes Gutenberg-Universität Mainz / An der Hochschule 2 / D-76726 Germersheim, meyerb@uni-mainz.de (depositor), Kristin Bührig, Institut für Germanistik I / Von-Melle-Park 6 / D-20146 Hamburg, kristin.buehrig@uni-hamburg.de (compiler), Bernd Meyer, Arbeitsbereich Interkulturelle Kommunikation / Fachbereich 06: Translations-, Sprach- und Kulturwissenschaft / Johannes Gutenberg-Universität Mainz / An der Hochschule 2 / D-76726 Germersheim, meyerb@uni-mainz.de (compiler), Kristin Bührig, Institut für Germanistik I / Von-Melle-Park 6 / D-20146 Hamburg, kristin.buehrig@uni-hamburg.de (compiler), Bernd Meyer, Arbeitsbereich Interkulturelle Kommunikation / Fachbereich 06: Translations-, Sprach- und Kulturwissenschaft / Johannes Gutenberg-Universität Mainz / An der Hochschule 2 / D-76726 Germersheim, meyerb@uni-mainz.de (compiler), Alper Ayzin (data_inputter), Anna (data_inputter), Bernd (data_inputter), Birte Pawlack (data_inputter), Çakmak (data_inputter), Carla (data_inputter), Demet (data_inputter), Erkan (data_inputter), Isabel (data_inputter), João (data_inputter), Latif (data_inputter), Maria (data_inputter), Najhla (data_inputter), Süreyya (data_inputter), Thomas Schmidt, thomas.schmidt@uni-hamburg.de (developer), Kristin Bührig, Institut für Germanistik I / Von-Melle-Park 6 / D-20146 Hamburg, kristin.buehrig@uni-hamburg.de (researcher), Bernd Meyer, Arbeitsbereich Interkulturelle Kommunikation / Fachbereich 06: Translations-, Sprach- und Kulturwissenschaft / Johannes Gutenberg-Universität Mainz / An der Hochschule 2 / D-76726 Germersheim, meyerb@uni-mainz.de (researcher), Deutsche Forschungsgemeinschaft (DFG) (sponsor)
Project:  K2 "Interpreting in Hospitals", German Research Foundation (DFG)
Keywords:  community interpreting, consecutive interpreting, interpreted communication, doctor-patient communication, communication in institutions, EXMARaLDA
Languages:  German (deu), Portuguese (por), Spanish (spa), Turkish (tur)
Size:  187 speakers (98 female, 89 male), 91 communications, 0.0 hours, 0 minutes, 0 recordings, 92 transcriptions, 170925 words
Annotation types:  transcription (manual): HIAT, akz: accentuation/stress, k: free comment, sup: suprasegmental information, de: German translation, en: English translation, mt: morphological transliteration
Temporal Coverage:  1997-07-11/2005-01-12
Spatial Coverage:  Hamburg, DE; Vila do Conde, PT; Viana do Castelo, PT; Ankara, TR
Genre:  discourse
Modality:  spoken
References:  Bührig, Kristin; Kliche, Ortrun; Meyer, Bernd; Pawlack, Birte (2012) The corpus "Interpreting in Hospitals": Possible applications for research and communication training. In: Schmidt, Thomas and Wörner, Kai (eds.): Multilingual Corpora and Multilingual Corpus Analysis. Hamburg Studies in Multilingualism (14), 305-315. Amsterdam: John Benjamins.

 </dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/8308</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.8308</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:8308</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.1439</dc:relation>
          <dc:rights>info:eu-repo/semantics/restrictedAccess</dc:rights>
          <dc:subject>community interpreting</dc:subject>
          <dc:subject>consecutive interpreting</dc:subject>
          <dc:subject>interpreted communication</dc:subject>
          <dc:subject>doctor-patient communication</dc:subject>
          <dc:subject>communication in institutions</dc:subject>
          <dc:subject>EXMARaLDA</dc:subject>
          <dc:subject>linguistics</dc:subject>
          <dc:subject>German</dc:subject>
          <dc:subject>Portuguese</dc:subject>
          <dc:subject>Spanish</dc:subject>
          <dc:subject>Turkish</dc:subject>
          <dc:title>Dolmetschen im Krankenhaus (DiK)</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:9195</identifier>
        <datestamp>2023-09-22T19:15:16Z</datestamp>
        <setSpec>user-uhh</setSpec>
        <setSpec>user-hzsk</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:contributor>Birr, Katharina</dc:contributor>
          <dc:contributor>Nagel, Norbert</dc:contributor>
          <dc:contributor>Mirahmadi, Hamasa</dc:contributor>
          <dc:contributor>Ihden, Sarah</dc:contributor>
          <dc:contributor>Nasielski, Vanessa</dc:contributor>
          <dc:contributor>Dücker, Lisa</dc:contributor>
          <dc:contributor>Tews, Sebastian</dc:contributor>
          <dc:contributor>Wallmeier, Nadine</dc:contributor>
          <dc:contributor>Hübener, Carlotta</dc:contributor>
          <dc:contributor>Recker, Anabel</dc:contributor>
          <dc:contributor>Peterberns, Hellen</dc:contributor>
          <dc:contributor>Dreessen, Katharina</dc:contributor>
          <dc:contributor>Bischoff, Annemarie</dc:contributor>
          <dc:contributor>Wollenschläger, Anna</dc:contributor>
          <dc:contributor>Schmitt, Eleonore</dc:contributor>
          <dc:contributor>Liebl, Mattias</dc:contributor>
          <dc:contributor>Schilling, Elmar</dc:contributor>
          <dc:contributor>Schroeder, Meile-Andrea</dc:contributor>
          <dc:contributor>Lehmberg, Timm</dc:contributor>
          <dc:contributor>Kahre, Paul</dc:contributor>
          <dc:contributor>Kleymann, Verena</dc:contributor>
          <dc:contributor>Sandmann, Lena</dc:contributor>
          <dc:contributor>Sturm, John</dc:contributor>
          <dc:contributor>Matthies, Sarah</dc:contributor>
          <dc:contributor>Sluyter-Gäthje, Henny</dc:contributor>
          <dc:contributor>Tran, Ilka</dc:contributor>
          <dc:contributor>Hütter, Julia</dc:contributor>
          <dc:contributor>Barteld, Fabian</dc:contributor>
          <dc:contributor>Schröder, Katharina</dc:contributor>
          <dc:contributor>Eichhorn-Hartmeyer, Christina</dc:contributor>
          <dc:contributor>Tiedemann, Meike</dc:contributor>
          <dc:contributor>Schnee, Lena</dc:contributor>
          <dc:contributor>Hübener, Louisa</dc:contributor>
          <dc:creator>ReN-Team</dc:creator>
          <dc:date>2021-01-06</dc:date>
          <dc:description>Search the corpus in ANNIS | Korpussuche in ANNIS

The project "Reference Corpus Middle Low German/Low Rhenish (1200–1650)", abbreviated as "ReN", is part of the "Corpus of Historical German Texts", which includes the projects "Old German Reference Corpus (750–1050)", the "Reference Corpus Middle High German (1050–1350)" and the "Reference Corpus EarlyModern New High German (1350–1650)". The project "Reference Corpus Middle Low German/Low Rhenish (1200–1650)" deals with a structured selection of Middle Low German and Low Rhenish monuments of speech from 1200 to 1650. This selection is based on the parameters "space", "time" and "field of writing".

The reference corpus of Middle Low German and Low Rhenish texts is based on manuscripts, prints and inscriptions. It is intended to provide an insight into the culture of speech and writing in Middle Low German and Low Rhenish regions. This spectrum of texttypes can be used to trace the linguistic development on the base of diatopic and diacronic subcategorisation. The aim of the project is the publication of diplomatic transcribed, lemmatised and grammatically annotated texts. The processed data – especially on the grammatical level – enables a linguistic analysis of the Middle Low German and Low Rhenish language, which goes far beyond what was possible until now.

Das „Referenzkorpus Mittelniederdeutsch/Niederrheinisch (1200–1650)“, kurz „ReN“, ist Teil des „Korpus historischer Texte des Deutschen“, zu welchem außerdem die Referenzkorpora Altdeutsch, Mittelhochdeutsch und Frühneuhochdeutsch zählen. Das ReN umfasst mittelniederdeutsche und niederrheinische Sprachdenkmäler von 1200 bis 1650 in einer strukturierten Auswahl. Diese ergibt sich aus den Parametern „Raum“, „Zeit“ und „Feld der Schriftlichkeit“.

Das Referenzkorpus mittelniederdeutscher und niederrheinischer Texte basiert auf Handschriften, Drucken und Inschriften. Es soll einen Einblick in die Sprach- und Textkultur des niederdeutschen und niederrheinischen Raums geben. Die historische Sprachentwicklung soll in ihrer diatopischen und diachronischen Untergliederung anhand des Textsortenspektrums nachgezeichnet werden. Das Ziel des Projektes besteht in der Veröffentlichung diplomatisch transkribierter, lemmatisierter und grammatisch annotierter Texte. Die so bearbeiteten Daten ermöglichen - insbesondere auf grammatischer Ebene - sprachwissenschaftliche Analysen des Mittelniederdeutschen und Niederrheinischen, die weit über das bisher Mögliche hinausgehen.

Ingrid Schröder. 2014. "Das Referenzkorpus: Neue Perspektiven für die mittelniederdeutsche Grammatikographie." Jahrbuch für Germanistische Sprachgeschichte, 5(1), pp. 150-164. http://www.degruyter.com/view/j/jbgsg.2014.5.issue-1/jbgsg-2014-0011/jbgsg-2014-0011.xml .

Robert Peters, Norbert Nagel. 2014. "Das digitale 'Referenzkorpus Mittelniederdeutsch / Niederrheinisch (ReN)'." Jahrbuch für Germanistische Sprachgeschichte, 5(1), pp. 165-175. https://www.degruyter.com/view/j/jbgsg.2014.5.issue-1/jbgsg-2014-0012/jbgsg-2014-0012.xml .

 

CLARIN Metadata summary for Referenzkorpus Mittelniederdeutsch/Niederrheinisch (1200-1650)Reference Corpus Middle Low German/Low Rhenish (1200-1650) (CMDI-based)

Title: Referenzkorpus Mittelniederdeutsch/Niederrheinisch (1200-1650)
Title: Reference Corpus Middle Low German/Low Rhenish (1200-1650)
Description: The reference corpus of Middle Low German and Low Rhenish texts is based on manuscripts, prints and inscriptions. It is intended to provide an insight into the culture of speech and writing in Middle Low German and Low Rhenish regions. This spectrum of texttypes can be used to trace the linguistic development on the base of diatopic and diacronic subcategorisation. The aim of the project is the publication of diplomatic transcribed, lemmatised and grammatically annotated texts. The processed data – especially on the grammatical level – enables a linguistic analysis of the Middle Low German and Low Rhenish language, which goes far beyond what was possible until now.
Description: Das Referenzkorpus mittelniederdeutscher und niederrheinischer Texte basiert auf Handschriften, Drucken und Inschriften. Es soll einen Einblick in die Sprach- und Textkultur des niederdeutschen und niederrheinischen Raums geben. Die historische Sprachentwicklung soll in ihrer diatopischen und diachronischen Untergliederung anhand des Textsortenspektrums nachgezeichnet werden. Das Ziel des Projektes besteht in der Veröffentlichung diplomatisch transkribierter, lemmatisierter und grammatisch annotierter Texte. Die so bearbeiteten Daten ermöglichen - insbesondere auf grammatischer Ebene - sprachwissenschaftliche Analysen des Mittelniederdeutschen und Niederrheinischen, die weit über das bisher Mögliche hinausgehen.
Data owner:  ReN-Team
Contributors:  Barteld, Fabian (developer, researcher), Birr, Katharina (annotator), Bischoff, Annemarie (annotator, researcher), Dücker, Lisa (annotator), Dreessen, Katharina (transcriber, annotator, researcher), Eichhorn-Hartmeyer, Christina (annotator), Hübener, Carlotta (transcriber, annotator), Hübener, Louisa (developer), Hütter, Julia (annotator), Ihden, Sarah (transcriber, annotator, researcher), Kahre, Paul (annotator), Kleymann, Verena (transcriber, annotator, researcher), Lehmberg, Timm (developer, researcher), Liebl, Mattias (transcriber), Matthies, Sarah (transcriber), Meike,, Tiedemann, (transcriber, annotator, researcher), Mirahmadi, Hamasa (transcriber), Nagel, Norbert (transcriber, researcher), Nasielski, Vanessa (annotator), Peterberns, Hellen (annotator), Recker, Anabel (annotator), Sandmann, Lena (developer), Schilling, Elmar (annotator, researcher), Schmitt, Eleonore (annotator), Schnee, Lena (annotator), Schröder, Katharina (transcriber), Schröder, Ingrid (rights holder, researcher), Schroeder, Meile-Andrea (annotator), Peters, Robert (rights holder, researcher), Sluyter-Gäthje, Henny (developer), Sturm, John (annotator), Tews, Sebastian (annotator), Tran, Ilka (transcriber), Wallmeier, Nadine (annotator, researcher), Wollenschläger, Anna (transcriber)
Project:  Referenzkorpus Mittelniederdeutsch/ Niederrheinisch (1200-1650), Deutsche Forschungsgemeinschaft (DFG)
Languages:  Middle Low German (gml), Low Rhenish (mis)
Size:  161 Text (annotated), 1485963 tok_anno (annotated), 1521346 tok_dipl (annotated), 74 Text (transcribed), 838400 tok_anno (transcribed), 854094 tok_dipl (transcribed)
Annotation types:  Transcription (manual), Part-of-speech (semi-automatic), Lemma (semi-automatic), Morphology (semi-automatic)
Temporal Coverage:  1200/1650
References:  Ingrid Schröder (2014) Das Referenzkorpus: Neue Perspektiven für die mittelniederdeutsche Grammatikographie References:  Robert Peters; Norbert Nagel (2014) Das digitale 'Referenzkorpus Mittelniederdeutsch / Niederrheinisch (ReN)'

 </dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/9195</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.9195</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:9195</dc:identifier>
          <dc:relation>url:https://dock.fdm.uni-hamburg.de/renannis/</dc:relation>
          <dc:relation>doi:10.25592/uhhfdm.1668</dc:relation>
          <dc:rights>info:eu-repo/semantics/openAccess</dc:rights>
          <dc:rights>https://creativecommons.org/licenses/by/4.0/legalcode</dc:rights>
          <dc:subject>Middle Low German</dc:subject>
          <dc:subject>Low Rhenish</dc:subject>
          <dc:subject>ANNIS</dc:subject>
          <dc:subject>linguistics</dc:subject>
          <dc:subject>diacronic</dc:subject>
          <dc:subject>historical</dc:subject>
          <dc:title>Reference Corpus Middle Low German/Low Rhenish (1200–1650); Referenzkorpus Mittelniederdeutsch/Niederrheinisch (1200–1650)</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:12250</identifier>
        <datestamp>2023-10-11T07:32:59Z</datestamp>
        <setSpec>user-uhh</setSpec>
        <setSpec>user-hzsk</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:creator>Schröder, Ingrid</dc:creator>
          <dc:creator>Elmentaler, Michael</dc:creator>
          <dc:creator>Gessinger, Joachim</dc:creator>
          <dc:creator>Macha, Jürgen</dc:creator>
          <dc:creator>Rosenberg, Peter</dc:creator>
          <dc:creator>Wirrer, Jan</dc:creator>
          <dc:date>2011-01-01</dc:date>
          <dc:description>Die Daten wurden im Rahmen des Projekts „Sprachvariation in Norddeutschland“ (SiN) erhoben. Sie umfassen die unterschiedlichen Sprachlagen zwischen hochdeutscher Standardsprache und nieder­deutschen Dialekten und repräsentieren den alltäglichen Sprachgebrauch Norddeutschlands mit seinen regionalen und lokalen Besonderheiten, auch in Hinblick auf die Verwendungsweisen in unterschiedlichen Situationen. Überdies dokumentieren sie Spracheinstellungen und Spracherfahrungen.

Die Aufnahmen stammen aus 36 Orten in 18 Regionen Norddeutschlands. Insgesamt wurden 144 Frauen im Alter von 40 bis 55 Jahren aufgenommen, die am Ort aufgewachsen sind und überwiegend dort gelebt haben. Gewählt wurden fünf Aufnahmesituationen mit unterschiedlichem Formalitätsgrad: Vorlesen (hochdeutsch), Interview (hochdeutsch), Tischgespräch (hochdeutsch und niederdeutsch), Erzählung (niederdeutsch), Übersetzung in den Dialekt (niederdeutsch). Tests zum Sprachwissen und zur Spracheinstellung (Salienz-, Normativitäts-, Arealitätstest) ergänzen das empirische Design.</dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/12250</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.12250</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:12250</dc:identifier>
          <dc:language>nds</dc:language>
          <dc:relation>doi:10.1515/9783110363449-018</dc:relation>
          <dc:relation>doi:10.25592/uhhfdm.10797</dc:relation>
          <dc:rights>info:eu-repo/semantics/restrictedAccess</dc:rights>
          <dc:subject>Variationslinguistik</dc:subject>
          <dc:subject>Wahrnehmungsdialektologie</dc:subject>
          <dc:subject>Niederdeutsch</dc:subject>
          <dc:subject>Dialekt</dc:subject>
          <dc:subject>Regionalsprache</dc:subject>
          <dc:subject>Varietäten</dc:subject>
          <dc:subject>gesprochene Sprache</dc:subject>
          <dc:subject>Interaktion</dc:subject>
          <dc:subject>Sprachvariation</dc:subject>
          <dc:subject>Sprachwandel</dc:subject>
          <dc:subject>Sprachbiographie</dc:subject>
          <dc:subject>Spracheinstellungen</dc:subject>
          <dc:subject>Salienz</dc:subject>
          <dc:subject>Wenker-Sätze</dc:subject>
          <dc:subject>EXMARaLDA</dc:subject>
          <dc:subject>regional variety</dc:subject>
          <dc:subject>adult bilingualism</dc:subject>
          <dc:title>Korpus Sprachvariation in Norddeutschland (SiN)</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:1483</identifier>
        <datestamp>2023-02-13T13:04:17Z</datestamp>
        <setSpec>user-hzsk</setSpec>
        <setSpec>user-uhh</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:creator>HZSK Hamburger Zentrum für Sprachkorpora</dc:creator>
          <dc:date>2020-09-02</dc:date>
          <dc:description>Audio recordings of a film retelling task with adult L2 users of German. The speakers' L1 and their L2 proficiencies vary. 24 communications + 1 German reference communication, duration between 2 and 16 minutes. For each speaker, a language learner biography (audio and freely transcribes) is available.

The Hamburg Modern Times Corpus (HaMoTiC) consists of transcribed audio recordings of learners of German at different proficiency levels who renarrate a few scenes from the silent film "Modern Times" (USA 1936, Charles Chaplin). The corpus was compiled at the Hamburg Centre for Language Corpora (HZSK) in 2012 and 2013 with the intent to create a corpus that is analogous to data created within the ESF project 'Second language acquisition by adult immigrants' (cf. Perdue &amp; Klein, 1992). For this purpose the original video clip was made available to us by the IMDI Corpus Archive of the Max Planck-Institute for Psycholinguistics in Nijmegen, Netherlands. The main objective was thus both to create a comparable linguistic resource based on existing data and also to demonstrate the functionality of the EXMARaLDA tools for transcription, annotation and analysis of spoken language corpora. The structural design of HaMoTiC is derived from the Hamburg Map Task Corpus (HAMATAC), that was developed in 2010. In terms of their content, HAMATAC and HaMoTiC complement each other with regard to their authenticity and the degree of control of learner language.

 

CLARIN Metadata summary for Hamburg Modern Times Corpus (HaMoTiC) (CMDI-based)

Title: Hamburg Modern Times Corpus (HaMoTiC)
Description: Audio recordings of a film retelling task with adult L2 users of German. The speakers' L1 and their L2 proficiencies vary. 24 communications + 1 German reference communication, duration between 2 and 16 minutes. For each speaker, a language learner biography (audio and freely transcribes) is available.
Publication date: 2013-11-01
Data owner:  HZSK Hamburger Zentrum für Sprachkorpora, Max-Brauer-Allee 60 / D-22765 Hamburg, corpora@uni-hamburg.de
Contributors:  HZSK Hamburger Zentrum für Sprachkorpora, Max-Brauer-Allee 60 / D-22765 Hamburg, corpora@uni-hamburg.de (compiler)
Keywords:  adult L2 acquisition, learner corpus, task-oriented communication, successive bilingualism, L2 data, adult bilingualism, film retelling task, multilingualism, EXMARaLDA
Language:  German (deu)
Size:  29 speakers (20 female, 9 male), 25 communications, 25 recordings, 186 minutes, 25 transcriptions, 24464 words
Annotation types:  transcription (manual): HIAT (modified), pho: manual annotation of phonetic phenomena, k: free comment, akz: accentuation/stress
Temporal Coverage:  2012-07-05/2012-12-13
Spatial Coverage:  Hamburg, DE
Genre:  discourse
Modality:  spoken

 </dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/1483</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.1483</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:1483</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.1482</dc:relation>
          <dc:rights>info:eu-repo/semantics/restrictedAccess</dc:rights>
          <dc:subject>adult L2 acquisition</dc:subject>
          <dc:subject>learner corpus</dc:subject>
          <dc:subject>task-oriented communication</dc:subject>
          <dc:subject>successive bilingualism</dc:subject>
          <dc:subject>L2 data</dc:subject>
          <dc:subject>adult bilingualism</dc:subject>
          <dc:subject>film retelling task</dc:subject>
          <dc:subject>multilingualism</dc:subject>
          <dc:subject>EXMARaLDA</dc:subject>
          <dc:subject>linguistics</dc:subject>
          <dc:subject>German</dc:subject>
          <dc:title>Hamburg Modern Times Corpus (HaMoTiC)</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:1425</identifier>
        <datestamp>2025-11-28T10:51:33Z</datestamp>
        <setSpec>user-hzsk</setSpec>
        <setSpec>user-uhh</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:creator>Bernhard Brehmer</dc:creator>
          <dc:date>2009-04-17</dc:date>
          <dc:description>This corpus version is deprecated for version 0.2. </dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/1425</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.1425</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:1425</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.1426</dc:relation>
          <dc:relation>doi:10.25592/uhhfdm.1424</dc:relation>
          <dc:rights>info:eu-repo/semantics/restrictedAccess</dc:rights>
          <dc:subject>language attrition</dc:subject>
          <dc:subject>adult bilingualism</dc:subject>
          <dc:subject>cross-sectional data</dc:subject>
          <dc:subject>simultaneous bilingualism</dc:subject>
          <dc:subject>successive bilingualism</dc:subject>
          <dc:subject>child L2 acquisition</dc:subject>
          <dc:subject>adult L2 acquisition</dc:subject>
          <dc:subject>L1 data</dc:subject>
          <dc:subject>EXMARaLDA</dc:subject>
          <dc:subject>linguistics</dc:subject>
          <dc:subject>Polish</dc:subject>
          <dc:title>Hamburg Corpus of Polish in Germany (HamCoPoliG)</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:18672</identifier>
        <datestamp>2026-05-26T06:55:32Z</datestamp>
        <setSpec>user-uhh</setSpec>
        <setSpec>user-hzsk</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:creator>Berlage, Eva</dc:creator>
          <dc:creator>Frömberg, Marcel</dc:creator>
          <dc:date>2026-05-04</dc:date>
          <dc:description>A Corpus of German and English Weather Forecasts from 2025 to 2026

In this research project, we compiled a corpus of German and English weather forecasts, containing forecasts from 2025 to 2026. The weather forecasts have been collected from nine websites over a period of nine months (starting on 11 July 2025 and running up until 31 March 2026). Four of these websites are hosted in the UK, providing weather forecasts on the UK and Europe, with another five websites hosted in Germany and therefore providing weather forecasts either (and most frequently) on different regions across Germany or on Europe. Overall, the size of the British English component of the corpus amounts to 794,396 tokens and the German one to 756,393 tokens.

Since the English and German parts of the corpus are directly comparable in terms of a) size (see above) b) genre (weather forecasts) and c) time  (2025-2026), the corpus allows for a series of structural comparisons across the two Germanic languages. Given that English meteorological predicates boast a series of non-agentive subjects –  i.e. subjects which denote a time or a location, as in Sunday will be dry or London will be rainy, – this corpus provides a sound quantitative basis for comparing the subject-forming possibilities of English and German (going back to  Rohdenburg 1974).

The corpus can be searched with AntConc. The files are available as plain txt files and as tagged files, using either upos or xpos tagging. Additionally, parsed versions of the texts are available, utilizing the CoNLL-U format. We are happy to grant researchers and interested students access to the corpus upon request.

 

This project was funded by the Hamburgische Wissenschaftliche Stiftung, Antrag 07/25, Kontrastive Linguistik interdisziplinär gedacht: die Erstellung eines digitalen Wetterkorpus zum Deutschen und Englischen.

References:

Berlage, Eva and Marcel Frömberg. 2026. A Corpus of German and English Weather Forecasts from 2025/2026. Universität Hamburg.

Rohdenburg, Günter (1974). Sekundäre Subjektivierungen im Englischen und Deutschen. Bielefeld: Corneslen-Velhagen &amp;Klasing.</dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/18672</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.18672</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:18672</dc:identifier>
          <dc:language>eng</dc:language>
          <dc:relation>doi:10.25592/uhhfdm.18671</dc:relation>
          <dc:rights>info:eu-repo/semantics/restrictedAccess</dc:rights>
          <dc:subject>Linguistics</dc:subject>
          <dc:subject>Corpus Linguistics</dc:subject>
          <dc:subject>Contrastive Linguistics</dc:subject>
          <dc:subject>subject-forming possibilities of English and German</dc:subject>
          <dc:title>A Corpus of German and English Weather Forecasts from 2025/2026</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:1460</identifier>
        <datestamp>2020-08-25T21:02:28Z</datestamp>
        <setSpec>user-hzsk</setSpec>
        <setSpec>user-uhh</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:creator>Lleó, Conxita</dc:creator>
          <dc:date>2020-08-25</dc:date>
          <dc:description>Additionals data from the PhonBLA Querschnittsstudie Madrid project.</dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/1460</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.1460</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:1460</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.1459</dc:relation>
          <dc:rights>info:eu-repo/semantics/closedAccess</dc:rights>
          <dc:subject>German</dc:subject>
          <dc:subject>Spanish</dc:subject>
          <dc:subject>Linguistics</dc:subject>
          <dc:subject>Language Acquisition</dc:subject>
          <dc:subject>Bilingualism</dc:subject>
          <dc:title>PhonBLA Querschnittsstudie Madrid Recordings</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:1464</identifier>
        <datestamp>2020-12-03T10:43:02Z</datestamp>
        <setSpec>user-hzsk</setSpec>
        <setSpec>user-uhh</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:creator>Meisel, Jürgen M.</dc:creator>
          <dc:date>2020-08-25</dc:date>
          <dc:description>Audio recordings of five adult learners of German as an L2 with L1s Spanish, Italian and Portuguese. Recording sessions (interview/conversation) in German once or twice a month over approx. two years starting 3-14 weeks after their arrival in Germany. Only transcripts and metadata are available. Orthographic transcription according to project internal conventions.</dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/1464</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.1464</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:1464</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.1074</dc:relation>
          <dc:relation>doi:10.25592/uhhfdm.1463</dc:relation>
          <dc:rights>info:eu-repo/semantics/restrictedAccess</dc:rights>
          <dc:subject>German</dc:subject>
          <dc:subject>Portuguese</dc:subject>
          <dc:subject>Spanish</dc:subject>
          <dc:subject>Italian</dc:subject>
          <dc:subject>linguistics</dc:subject>
          <dc:subject>language acquisition</dc:subject>
          <dc:subject>bilingualism</dc:subject>
          <dc:subject>learner corpus</dc:subject>
          <dc:subject>adult L2 acquisition</dc:subject>
          <dc:subject>adult bilingualism</dc:subject>
          <dc:subject>longitudinal data</dc:subject>
          <dc:subject>successive bilingualism</dc:subject>
          <dc:subject>l2 data</dc:subject>
          <dc:subject>EXMARaLDA</dc:subject>
          <dc:title>ZISA</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:1478</identifier>
        <datestamp>2022-08-23T07:27:42Z</datestamp>
        <setSpec>user-hzsk</setSpec>
        <setSpec>user-uhh</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:creator>Angermeyer, Philipp</dc:creator>
          <dc:creator>Bührig, Kristin</dc:creator>
          <dc:creator>Meyer, Bernd</dc:creator>
          <dc:date>2010-03-09</dc:date>
          <dc:description>Audio and video recordings of various types of community interpreted discourse (doctor-patient communication, simulated doctor-patient communication, courtroom communication) in German (simulated and authentic doctor-patient communication) and US (courtroom communication) institutions with varying community languages. Video recordings only exist for the simulated communication. For the authentic interpreted doctor-patient communication, no audio files will be made available.

The ComInDat pilot corpus contains sample data from three different projects: the DiK corpus of Portuguese/German and Turkish/German interpreted doctor-patient communication in hospitals (Bührig &amp; Meyer 2004), he IiSCC-corpus, a corpus of interpreted court proceedings in different language constellations (Spanish/English, Russian/English, Haitian Creole/English and Polish/English) (Angermeyer 2006), a corpus of simulated interpreted doctor-patient interactions in different language constellations (Russian/German, Polish/German and Romanian/German) from a training seminar for bilingual nursing staff ("SimDiK", Bührig, Kliche, Meyer &amp; Pawlack 2012). More information about the background of the corpus and the details of its design can be found in (Angermeyer, Meyer &amp; Schmidt 2012). For more information about the project, please contact Philipp Angermeyer.

Angermeyer, P., Meyer, B. and Schmidt, T. (2012). Sharing Community Interpreting Corpora: A pilot study. In: Schmidt, T. and Wörner, K. (eds.) Multilingual Corpora and Multilingual Corpus Analysis. Amsterdam: Benjamins, 275-294.

 

CLARIN Metadata summary for Community Interpreting Database Pilot Corpus (ComInDat) (CMDI-based)

Title: Community Interpreting Database Pilot Corpus (ComInDat)
Description: Audio and video recordings of various types of community interpreted discourse (doctor-patient communication, simulated doctor-patient communication, courtroom communication) in German (simulated and authentic doctor-patient communication) and US (courtroom communication) institutions with varying community languages. Video recordings only exist for the simulated communication. For the authentic interpreted doctor-patient communication, no audio files will be made available.
Publication date: 2013-06-10
Data owner:  Philipp Angermeyer, Department of Languages, Literatures and Linguistics / York University / 4700 Keele Street / Canada M3J 1P3, pangerme@yorku.ca, Kristin Bührig, Institut für Germanistik I / Von-Melle-Park 6 / D-20146 Hamburg, kristin.buehrig@uni-hamburg.de, Bernd Meyer, Arbeitsbereich Interkulturelle Kommunikation / Fachbereich 06: Translations-, Sprach- und Kulturwissenschaft / Johannes Gutenberg-Universität Mainz / An der Hochschule 2 / D-76726 Germersheim, meyerb@uni-mainz.de
Contributors:  Philipp Angermeyer, Department of Languages, Literatures and Linguistics / York University / 4700 Keele Street / Canada M3J 1P3, pangerme@yorku.ca (compiler), Kristin Bührig, Institut für Germanistik I / Von-Melle-Park 6 / D-20146 Hamburg, kristin.buehrig@uni-hamburg.de (compiler), Bernd Meyer, Arbeitsbereich Interkulturelle Kommunikation / Fachbereich 06: Translations-, Sprach- und Kulturwissenschaft / Johannes Gutenberg-Universität Mainz / An der Hochschule 2 / D-76726 Germersheim, meyerb@uni-mainz.de (compiler)
Project:  The Integration of Text, Sound, and Image into the Corpus-Based Analysis of Interpreter-Mediated Interaction
Keywords:  community interpreting, doctor-patient communication, courtroom communication, EXMARaLDA
Languages:  German (deu), English (eng), Spanish (spa), Turkish (tur), Polish (pol), Portuguese (por), Romanian (ron), Russian (rus), Haitian (hat)
Size:  54 speakers (35 female, 16 male, 3 unknown), 14 communications, 12 recordings, 83 minutes, 17 transcriptions, 35051 words
Annotation types:  transcription (manual): HIAT/CHAT, deu: German translation, eng: English translation, k: free comment, lang: utterance language, sup: suprasegmental information, trans: utterance translation status, akz: accentuation/stress, pol: Polish translation
Temporal Coverage:  1999-07-01/2010-03-09
Spatial Coverage:  Hamburg, DE; New York, US; Neumünster, DE
Genre:  discourse
Modality:  spoken
References:  Angermeyer, P., Meyer, B. and Schmidt, T. (2012). Sharing Community Interpreting Corpora: A pilot study. In: Schmidt, T. and Wörner, K. (eds.) Multilingual Corpora and Multilingual Corpus Analysis. Amsterdam: Benjamins, 275-294.

 </dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/1478</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.1478</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:1478</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.1477</dc:relation>
          <dc:rights>info:eu-repo/semantics/restrictedAccess</dc:rights>
          <dc:subject>community interpreting</dc:subject>
          <dc:subject>doctor-patient communication</dc:subject>
          <dc:subject>courtroom communication</dc:subject>
          <dc:subject>EXMARaLDA</dc:subject>
          <dc:subject>linguistics</dc:subject>
          <dc:subject>German</dc:subject>
          <dc:subject>English</dc:subject>
          <dc:subject>Spanish</dc:subject>
          <dc:subject>Turkish</dc:subject>
          <dc:subject>Polish</dc:subject>
          <dc:subject>Portuguese</dc:subject>
          <dc:subject>Romanian</dc:subject>
          <dc:subject>Russian</dc:subject>
          <dc:subject>Haitian</dc:subject>
          <dc:title>Community Interpreting Database Pilot Corpus (ComInDat)</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:1480</identifier>
        <datestamp>2020-11-06T12:36:43Z</datestamp>
        <setSpec>user-hzsk</setSpec>
        <setSpec>user-uhh</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:creator>Hamburger Zentrum für Sprachkorpora</dc:creator>
          <dc:date>2010-01-01</dc:date>
          <dc:description>Audio and two video recordings of map tasks with adult L2 users of German and one L1 speaker. The speakers' L1 and their L2 proficiencies vary. The maps used for the tasks are available.

The Hamburg MapTask Corpus (HAMATAC) is a spoken language corpus documenting the performance of 24 L2 learners of German in a map task. HAMATAC was recorded and transcribed in project Z2 at the Research Centre on Multilingualism. The current version 0.3 contains a new communication with video recording as well as the resources known from the previous version, e.g. orthographic transcriptions of the recordings, manual annotation of disfluencies and automatic annotation of part-of-speech and lemmas.

 

CLARIN Metadata summary for The Hamburg MapTask Corpus (HAMATAC) (CMDI-based)

Title: The Hamburg MapTask Corpus (HAMATAC)
Description: Audio and two video recordings of map tasks with adult L2 users of German and one L1 speaker. The speakers' L1 and their L2 proficiencies vary. The maps used for the tasks are available.
Publication date: 2010-09-16
Data owner:  Hamburger Zentrum für Sprachkorpora, Max-Brauer-Allee 60 / D-22765 Hamburg, corpora@uni-hamburg.de
Contributors:  Hamburger Zentrum für Sprachkorpora, Max-Brauer-Allee 60 / D-22765 Hamburg, corpora@uni-hamburg.de (compiler)
Project:  Z2 "Computer Assisted Methods for the creation and analysis of multilingual data", German Research Foundation (DFG)
Keywords:  adult L2 acquisition, learner corpus, task-oriented communication, successive bilingualism, L2 data, adult bilingualism, simultaneous bilingualism, map task, EXMARaLDA
Language:  German (deu)
Size:  28 speakers (16 female, 12 male), 26 communications, 26 recordings, 208 minutes, 26 transcriptions, 22898 words
Annotation types:  transcription (manual): orthographic transcription/simplified HIAT, pos: Fine-grained part of speech tagging using TreeTagger and the STTS tagset., pos-sup: superordinate part of Speech (manual, STTS tagset), c: indicates that the automatic pos-annotation is incorrect, lemma: lemma (TreeTagger), disfluency: manual annotation of disfluency phenomena, pho: manual annotation of phonetic phenomena
Temporal Coverage:  2009-10-28/2013-06-19
Spatial Coverage:  Hamburg, DE
Genre:  discourse
Modality:  spoken

 </dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/1480</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.1480</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:1480</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.1479</dc:relation>
          <dc:rights>info:eu-repo/semantics/restrictedAccess</dc:rights>
          <dc:subject>adult L2 acquisition</dc:subject>
          <dc:subject>learner corpus</dc:subject>
          <dc:subject>task-oriented communication</dc:subject>
          <dc:subject>successive bilingualism</dc:subject>
          <dc:subject>L2 data</dc:subject>
          <dc:subject>adult bilingualism</dc:subject>
          <dc:subject>simultaneous bilingualism</dc:subject>
          <dc:subject>map task</dc:subject>
          <dc:subject>EXMARaLDA</dc:subject>
          <dc:subject>linguistics</dc:subject>
          <dc:subject>German</dc:subject>
          <dc:title>The Hamburg MapTask Corpus (HAMATAC)</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:1485</identifier>
        <datestamp>2021-05-27T11:10:23Z</datestamp>
        <setSpec>user-hzsk</setSpec>
        <setSpec>user-uhh</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:creator>Lleó, Conxita</dc:creator>
          <dc:date>2020-09-02</dc:date>
          <dc:description>Audio recordings of five German and five Spanish speaking monolingual children. For the German children there are 175 recordings (interviewer/child interaction), on an average starting at 9 months and ending at 3 years; for the Spanish children there are 77 recordings, on average starting at 9 months and ending at 2;6 years.

PAIDUS is a phonetically and orthographically transcribed corpus of German and Spanish child language, finalized in the project Prosodic constraints on phonological and morphological development in bilingual first language acquisition at the Research Center on Multilingualism, University of Hamburg.

Lleó, Conxita. 2012. "Monolingual and Bilingual Phonoprosodic Corpora of Child German and Child Spanish." In Multilingual Corpora and Multilingual Corpus Analysis, edited by Thomas Schmidt and Kai Wörner, 14:pp. 107–22. Hamburg Studies in Multilingualism. John Benjamins.

 

CLARIN Metadata summary for Parameterfixierung im Deutschen und Spanischen (PAIDUS) (CMDI-based)

Title: Parameterfixierung im Deutschen und Spanischen (PAIDUS)
Description: Audio recordings of five German and five Spanish speaking monolingual children. For the German children there are 175 recordings (interviewer/child interaction), on an average starting at 9 months and ending at 3 years; for the Spanish children there are 77 recordings, on average starting at 9 months and ending at 2;6 years.
Publication date: 2010-09-30
Data owner:  Conxita Lleó, Institut für Romanistik / Von-Melle-Park 6 / D-20146 Hamburg, lleo@uni-hamburg.de
Contributors:  Conxita Lleó, Institut für Romanistik / Von-Melle-Park 6 / D-20146 Hamburg, lleo@uni-hamburg.de (compiler)
Project:  E3 "Prosodic Constraints on Phonological and Morphological development in Bilingual First Language Acquisition", German Research Foundation (DFG)
Keywords:  child language acquisition, longitudinal data, monolingual data, EXMARaLDA
Languages:  German (deu), Spanish (spa)
Size:  66 speakers (43 female, 23 male), 253 communications, 252 recordings, 8217 minutes, 253 transcriptions, 166976 words
Genre:  discourse
Modality:  spoken
References:  Lleó, Conxita. 2012. "Monolingual and Bilingual Phonoprosodic Corpora of Child German and Child Spanish." In Multilingual Corpora and Multilingual Corpus Analysis, edited by Thomas Schmidt and Kai Wörner, 14:pp. 107–22. Hamburg Studies in Multilingualism. John Benjamins.

 </dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/1485</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.1485</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:1485</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.1484</dc:relation>
          <dc:rights>info:eu-repo/semantics/restrictedAccess</dc:rights>
          <dc:subject>child language acquisition</dc:subject>
          <dc:subject>longitudinal data</dc:subject>
          <dc:subject>monolingual data</dc:subject>
          <dc:subject>EXMARaLDA</dc:subject>
          <dc:subject>linguistics</dc:subject>
          <dc:subject>German</dc:subject>
          <dc:subject>Spanish</dc:subject>
          <dc:title>Parameterfixierung im Deutschen und Spanischen (PAIDUS)</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:1502</identifier>
        <datestamp>2021-05-27T11:10:24Z</datestamp>
        <setSpec>user-hzsk</setSpec>
        <setSpec>user-uhh</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:creator>Lleó, Conxita</dc:creator>
          <dc:date>2011-01-01</dc:date>
          <dc:description>Audio recordings of eight German participants in Spain, with Spanish as L2, children first exposed to their L2 at or before the age of three, and children exposed to their L2 at or after the age of five. Recording sessions in Spanish based on picture naming and story telling etc. Rich metadata on language use and attitudes in the family submitted by the parents.

Phon-CL2 is a phonetically and orthographically transcribed corpus of German and Spanish, finalized in the project Prosodic constraints on phonological and morphological development in bilingual first language acquisition at the Research Center on Multilingualism, University of Hamburg.

Ulloa, Marta Saceda, Conxita Lleó, and Izarbe Garcia Sanchez. 2012. "Corpora of Spoken Spanish by Simultaneous and Successive German-Spanish Bilingual and Spanish Monolingual Children." In Multilingual Corpora and Multilingual Corpus Analysis, edited by Thomas Schmidt and Kai Wörner, pp. 97–106. Hamburg Studies in Multilingualism, 14. John Benjamins.

 

CLARIN Metadata summary for Phon-CL2 (CMDI-based)

Title: Phon-CL2
Description: Audio recordings of eight German participants in Spain, with Spanish as L2, children first exposed to their L2 at or before the age of three, and children exposed to their L2 at or after the age of five. Recording sessions in Spanish based on picture naming and story telling etc. Rich metadata on language use and attitudes in the family submitted by the parents.
Data owner:  Conxita Lleó, Institut für Romanistik / Von-Melle-Park 6 / D-20146 Hamburg, lleo@uni-hamburg.de
Contributors:  Conxita Lleó, Institut für Romanistik / Von-Melle-Park 6 / D-20146 Hamburg, lleo@uni-hamburg.de (compiler)
Project:  E3 "Prosodic Constraints on Phonological and Morphological development in Bilingual First Language Acquisition", German Research Foundation (DFG)
Keywords:  child language acquisition, child bilingualism, successive bilingualism, child L2 acquisition, L2 data, adult bilingualism, EXMARaLDA
Languages:  German (deu), Spanish (spa)
Size:  22 speakers (8 female, 14 male), 26 communications, 0 recordings, 0 minutes, 26 transcriptions, 17412 words
Genre:  discourse
Modality:  spoken
References:  Ulloa, Marta Saceda, Conxita Lleó, and Izarbe Garcia Sanchez. 2012. "Corpora of Spoken Spanish by Simultaneous and Successive German-Spanish Bilingual and Spanish Monolingual Children." In Multilingual Corpora and Multilingual Corpus Analysis, edited by Thomas Schmidt and Kai Wörner, pp. 97–106. Hamburg Studies in Multilingualism, 14. John Benjamins.

 </dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/1502</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.1502</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:1502</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.1501</dc:relation>
          <dc:rights>info:eu-repo/semantics/restrictedAccess</dc:rights>
          <dc:subject>child language acquisition</dc:subject>
          <dc:subject>child bilingualism</dc:subject>
          <dc:subject>successive bilingualism</dc:subject>
          <dc:subject>child L2 acquisition</dc:subject>
          <dc:subject>L2 data</dc:subject>
          <dc:subject>adult bilingualism</dc:subject>
          <dc:subject>EXMARaLDA</dc:subject>
          <dc:subject>linguistics</dc:subject>
          <dc:subject>German</dc:subject>
          <dc:subject>Spanish</dc:subject>
          <dc:title>Phon-CL2</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:1557</identifier>
        <datestamp>2020-09-07T19:41:35Z</datestamp>
        <setSpec>user-hzsk</setSpec>
        <setSpec>user-uhh</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:creator>Braunmüller, Kurt</dc:creator>
          <dc:date>2020-09-07</dc:date>
          <dc:description>Data from the SkandSemiko project.</dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/1557</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.1557</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:1557</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.1556</dc:relation>
          <dc:rights>info:eu-repo/semantics/closedAccess</dc:rights>
          <dc:subject>Danish</dc:subject>
          <dc:subject>English</dc:subject>
          <dc:subject>Norwegian Bokmål</dc:subject>
          <dc:subject>Swedish</dc:subject>
          <dc:subject>Linguistics</dc:subject>
          <dc:subject>Semi-communication</dc:subject>
          <dc:subject>Recepteive multilingualism</dc:subject>
          <dc:subject>Adult language acquisition</dc:subject>
          <dc:title>SkandSemikoAdmin</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:1595</identifier>
        <datestamp>2020-09-14T20:09:43Z</datestamp>
        <setSpec>user-hzsk</setSpec>
        <setSpec>user-uhh</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:creator>Fiedler, Ines</dc:creator>
          <dc:date>2020-09-14</dc:date>
          <dc:description>The data sets for each language consist of a small number of mini-dialogues, chosen out of the 189 entries within the Focus Translation Task (cf. Skopeteas et al. 2006: 209ff.) in order to get a basic set of utterances for comparison between the languages dealt with in the project.

Fiedler, Ines (2011) QUIS Data from Yom, Aja, Anii and Foodo. With Notes on Genetic and Areal Relations. In: Petrova &amp; M. Grubic (eds.): Interdisciplinary Studies on Information Structure 16, Linguistic Fieldnotes III: Information structure in Gur and Kwa language. 49-96.

Schwarz, Anne (2011) QUIS Data from Buli, Kɔnni and Baatɔnum. With Notes on the Comparative Approach. In: Petrova &amp; M. Grubic (eds.): Interdisciplinary Studies on Information Structure 16, Linguistic Fieldnotes III: Information structure in Gur and Kwa language. 1-48.

 

CLARIN Metadata summary for B1 Foodo (CMDI-based)

Title: B1 Foodo
Description: The data sets for each language consist of a small number of mini-dialogues, chosen out of the 189 entries within the Focus Translation Task (cf. Skopeteas et al. 2006: 209ff.) in order to get a basic set of utterances for comparison between the languages dealt with in the project.
Publication date: 2015
Data owner:  Dr. phil. Ines Fiedler
Contributors:  Manfred Krifka (editor), Katharina Hartmann (editor), Brigitte Reineke (editor)
Project:  Special Research Centre 632 Information structure, German Research Foundation
Keywords:  focus, Fokus
Language:  Foodo (fod)
Size:  257 Token
Segmentation units:  other
Temporal Coverage:  2005-03-14/2005-03-18
Spatial Coverage:  Séméré, BJ
Genre:  discourse
Modality:  spoken
References:  Fiedler, Ines (2011) QUIS Data from Yom, Aja, Anii and Foodo. With Notes on Genetic and Areal Relations. In: Petrova &amp; M. Grubic (eds.): Interdisciplinary Studies on Information Structure 16, Linguistic Fieldnotes III: Information structure in Gur and Kwa language. 49-96. References:  Schwarz, Anne (2011) QUIS Data from Buli, Kɔnni and Baatɔnum. With Notes on the Comparative Approach. In: Petrova &amp; M. Grubic (eds.): Interdisciplinary Studies on Information Structure 16, Linguistic Fieldnotes III: Information structure in Gur and Kwa language. 1-48.

 </dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/1595</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.1595</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:1595</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.1594</dc:relation>
          <dc:rights>info:eu-repo/semantics/openAccess</dc:rights>
          <dc:rights>https://creativecommons.org/licenses/by/3.0/legalcode</dc:rights>
          <dc:subject>focus</dc:subject>
          <dc:subject>Fokus</dc:subject>
          <dc:subject>linguistics</dc:subject>
          <dc:subject>information structure</dc:subject>
          <dc:subject>Foodo</dc:subject>
          <dc:title>B1 Foodo</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:1599</identifier>
        <datestamp>2020-09-14T20:14:00Z</datestamp>
        <setSpec>user-hzsk</setSpec>
        <setSpec>user-uhh</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:creator>Linde, Sonja</dc:creator>
          <dc:date>2015-09-14</dc:date>
          <dc:description>Heliand 1, 4 and 5: complete text, status: final, digitalization, translation to Modern German, manually annotated with parts of speech, syntactic categories, grammatical functions, clause status, numbers of syllables (per constituent), alliteration, information status, topic/comment, position of phrase in sentence, definiteness, focus/background, focus-marker, comments on context, source (bibliography).

 

CLARIN Metadata summary for B4 Heliand (CMDI-based)

Title: B4 Heliand
Description: Heliand 1, 4 and 5: complete text, status: final, digitalization, translation to Modern German, manually annotated with parts of speech, syntactic categories, grammatical functions, clause status, numbers of syllables (per constituent), alliteration, information status, topic/comment, position of phrase in sentence, definiteness, focus/background, focus-marker, comments on context, source (bibliography).
Publication date: 2015
Data owner:  Sonja Linde
Contributors:  Svetlana Petrova (editor), Karin Donhauser (editor), Eva Schlachter (annotator), Marco Coniglio (annotator), Oxana Rasskazova (annotator), Anke Gehrlein (annotator)
Project:  Special Research Centre 632 Information structure, German Research Foundation
Keywords:  diachronic, historical texts, religious texts, information structure
Language:  Old High German (goh)
Size:  3495 Token
Segmentation units:  other
Genre:  historic manuscript
Modality:  written

 </dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/1599</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.1599</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:1599</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.1598</dc:relation>
          <dc:rights>info:eu-repo/semantics/openAccess</dc:rights>
          <dc:rights>https://creativecommons.org/licenses/by/3.0/legalcode</dc:rights>
          <dc:subject>diachronic</dc:subject>
          <dc:subject>historical texts</dc:subject>
          <dc:subject>religious texts</dc:subject>
          <dc:subject>information structure</dc:subject>
          <dc:subject>linguistics</dc:subject>
          <dc:subject>Old High German</dc:subject>
          <dc:title>B4 Heliand</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:972</identifier>
        <datestamp>2020-06-30T13:54:59Z</datestamp>
        <setSpec>user-hzsk</setSpec>
        <setSpec>user-uhh</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:creator>Meisel, Jürgen M.</dc:creator>
          <dc:date>2020-05-25</dc:date>
          <dc:description>Additional data from the CHILD-L2 project.</dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/972</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.972</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:972</dc:identifier>
          <dc:language>fra</dc:language>
          <dc:relation>doi:10.25592/uhhfdm.1064</dc:relation>
          <dc:relation>doi:10.25592/uhhfdm.971</dc:relation>
          <dc:rights>info:eu-repo/semantics/closedAccess</dc:rights>
          <dc:title>CHILD-L2 Media</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>video</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:1074</identifier>
        <datestamp>2020-09-10T08:47:49Z</datestamp>
        <setSpec>user-hzsk</setSpec>
        <setSpec>user-uhh</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:creator>Meisel, Jürgen M.</dc:creator>
          <dc:date>2020-06-24</dc:date>
          <dc:description>Additional data from the ZISA project.</dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/1074</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.1074</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:1074</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.1464</dc:relation>
          <dc:relation>doi:10.25592/uhhfdm.1073</dc:relation>
          <dc:rights>info:eu-repo/semantics/closedAccess</dc:rights>
          <dc:subject>Portuguese</dc:subject>
          <dc:subject>Spanish</dc:subject>
          <dc:subject>Italian</dc:subject>
          <dc:subject>Linguistics</dc:subject>
          <dc:subject>Language Acquisition</dc:subject>
          <dc:subject>Bilingualism</dc:subject>
          <dc:title>ZISA Media</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:1076</identifier>
        <datestamp>2020-11-10T20:29:38Z</datestamp>
        <setSpec>user-hzsk</setSpec>
        <setSpec>user-uhh</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:creator>Meisel, Jürgen M.</dc:creator>
          <dc:date>2020-06-24</dc:date>
          <dc:description>Additional data from the BIPODE project.</dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/1076</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.1076</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:1076</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.1535</dc:relation>
          <dc:relation>doi:10.25592/uhhfdm.1075</dc:relation>
          <dc:rights>info:eu-repo/semantics/closedAccess</dc:rights>
          <dc:subject>German</dc:subject>
          <dc:subject>Portuguese</dc:subject>
          <dc:subject>Linguistics</dc:subject>
          <dc:subject>Language Acquisition</dc:subject>
          <dc:subject>Bilingualism</dc:subject>
          <dc:title>BIPODE Media</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>video</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:9002</identifier>
        <datestamp>2021-04-30T06:23:09Z</datestamp>
        <setSpec>user-hzsk</setSpec>
        <setSpec>user-uhh</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:contributor>Deutsch Forschungsgemeinschaft (DFG)</dc:contributor>
          <dc:contributor>Hundt, Markus</dc:contributor>
          <dc:contributor>Hundt, Markus</dc:contributor>
          <dc:contributor>Hundt, Markus</dc:contributor>
          <dc:contributor>Naths, Saskia (geb. Schröder)</dc:contributor>
          <dc:contributor>Naths, Saskia (geb. Schröder)</dc:contributor>
          <dc:contributor>Palliwoda, Nicole</dc:contributor>
          <dc:contributor>Palliwoda, Nicole</dc:contributor>
          <dc:contributor>Beuge, Patrick</dc:contributor>
          <dc:contributor>Hannemann, Timo</dc:contributor>
          <dc:contributor>Hoffmeister, Toke</dc:contributor>
          <dc:creator>Hundt, Markus</dc:creator>
          <dc:date>2017-07-10</dc:date>
          <dc:description>Im Forschungsprojekt "Der deutsche Sprachraum aus der Sicht linguistischer Laien" wurden die laienlinguistischen Konzeptualisierungen zur deutschen Sprache untersucht. Wesentliche Ziele dieses Forschungsprojektes sind die Grundlagenforschung zur laienbezogenen Sprachkonzeption und die empirische Erhebung von Wissensbeständen linguistischer Laien. Dazu wurden in Deutschland, Österreich, der deutschsprachigen Schweiz, Liechtenstein, Luxemburg, Ostbelgien und Südtirol jeweils sechs Personen dreier Altersgruppen befragt. Momentan steht ein Korpus von sprachlichen und metasprachlichen Daten von 139 Gewährspersonen zur Verfügung.

Die produzierten mental maps sind durch eine WebGIS-Anwendung abrufbar.

Die im August 2017 erschienene Abschlusspublikation mit ausgewählten Ergebnissen beschließt das DFG-Projekt.  

Hundt, Markus; Palliwoda, Nicole; Schröder, Saskia (Hrsg.) (2017) Der deutsche Sprachraum aus der Sicht linguistischer Laien. Ergebnisse des Kieler DFG-Projektes. De Gruyter. DOI: https://doi.org/10.1515/9783110554212

Rezension: Stoltmann, Kai (2018): Rezension. In: Zeitschrift für Dialektologie und Linguistik (ZDL), 85/3. S. 365-368.

Homepage: Wahrnehmungsdialektologie

 </dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/9002</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.9002</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:9002</dc:identifier>
          <dc:language>deu</dc:language>
          <dc:publisher>Hundt, Markus; Palliwoda, Nicole; Schröder, Saskia</dc:publisher>
          <dc:relation>doi:10.25592/uhhfdm.8320</dc:relation>
          <dc:rights>info:eu-repo/semantics/restrictedAccess</dc:rights>
          <dc:subject>EXMARaLDA</dc:subject>
          <dc:subject>perceptual dialectology</dc:subject>
          <dc:subject>Wahrnehmungsdialektologie</dc:subject>
          <dc:subject>German-speaking area</dc:subject>
          <dc:subject>Deutscher Sprachraum</dc:subject>
          <dc:subject>mental maps</dc:subject>
          <dc:subject>Mikrokartierung</dc:subject>
          <dc:subject>Makrokartierung</dc:subject>
          <dc:subject>Pilesorte-Methode</dc:subject>
          <dc:subject>Dialektrometrie</dc:subject>
          <dc:subject>Laienlinguistik</dc:subject>
          <dc:subject>Normkonzept</dc:subject>
          <dc:subject>Regionalsprache</dc:subject>
          <dc:subject>linguistics</dc:subject>
          <dc:subject>German</dc:subject>
          <dc:title>Der deutsche Sprachraum aus der Sicht linguistischer Laien - wahrnehmungsdialektologische Grundlagenforschung und die Rekonstruktion von Laienkonzeptualisierungen zur deutschen Sprache</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:1684</identifier>
        <datestamp>2022-07-15T11:46:06Z</datestamp>
        <setSpec>user-hzsk</setSpec>
        <setSpec>user-uhh</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:contributor>Wallmeier, Nadine</dc:contributor>
          <dc:contributor>Hütter, Julia</dc:contributor>
          <dc:contributor>Schilling, Elmar</dc:contributor>
          <dc:contributor>Ihden, Sarah</dc:contributor>
          <dc:contributor>Nagel, Norbert</dc:contributor>
          <dc:contributor>Schroeder, Meile-Andrea</dc:contributor>
          <dc:contributor>Schröder, Katharina</dc:contributor>
          <dc:contributor>Kleymann, Verena</dc:contributor>
          <dc:contributor>Wollenschläger, Anna</dc:contributor>
          <dc:contributor>Schmitt, Eleonore</dc:contributor>
          <dc:contributor>Eichhorn-Hartmeyer, Christina</dc:contributor>
          <dc:contributor>Hübener, Carlotta</dc:contributor>
          <dc:contributor>Dreessen, Katharina</dc:contributor>
          <dc:contributor>Tews, Sebastian</dc:contributor>
          <dc:contributor>Lehmberg, Timm</dc:contributor>
          <dc:contributor>Bischoff, Annemarie</dc:contributor>
          <dc:contributor>Birr, Katharina</dc:contributor>
          <dc:contributor>Kahre, Paul</dc:contributor>
          <dc:contributor>Liebl, Mattias</dc:contributor>
          <dc:contributor>Schnee, Lena</dc:contributor>
          <dc:contributor>Barteld, Fabian</dc:contributor>
          <dc:contributor>Mirahmadi, Hamasa</dc:contributor>
          <dc:contributor>Glawe, Meike</dc:contributor>
          <dc:contributor>Dücker, Lisa</dc:contributor>
          <dc:contributor>Nasielski, Vanessa</dc:contributor>
          <dc:contributor>Sluyter-Gäthje, Henny</dc:contributor>
          <dc:contributor>Recker, Anabel</dc:contributor>
          <dc:contributor>Matthies, Sarah</dc:contributor>
          <dc:creator>ReN-Team</dc:creator>
          <dc:date>2017-03-08</dc:date>
          <dc:description>- This version is deprecated because it contains the wrong handbook files - 

The project "Reference Corpus Middle Low German/Low Rhenish (1200–1650)", abbreviated as "ReN", is part of the "Corpus of Historical German Texts", which includes the projects "Old German Reference Corpus (750–1050)", the "Reference Corpus Middle High German (1050–1350)" and the "Reference Corpus EarlyModern New High German (1350–1650)". The project "Reference Corpus Middle Low German/Low Rhenish (1200–1650)" deals with a structured selection of Middle Low German and Low Rhenish monuments of speech from 1200 to 1650. This selection is based on the parameters "space", "time" and "field of writing".

The reference corpus of Middle Low German and Low Rhenish texts is based on manuscripts, prints and inscriptions. It is intended to provide an insight into the culture of speech and writing in Middle Low German and Low Rhenish regions. This spectrum of texttypes can be used to trace the linguistic development on the base of diatopic and diacronic subcategorisation. The aim of the project is the publication of diplomatic transcribed, lemmatised and grammatically annotated texts. The processed data – especially on the grammatical level – enables a linguistic analysis of the Middle Low German and Low Rhenish language, which goes far beyond what was possible until now.

Das Referenzkorpus mittelniederdeutscher und niederrheinischer Texte basiert auf Handschriften, Drucken und Inschriften. Es soll einen Einblick in die Sprach- und Textkultur des niederdeutschen und niederrheinischen Raums geben. Die historische Sprachentwicklung soll in ihrer diatopischen und diachronischen Untergliederung anhand des Textsortenspektrums nachgezeichnet werden. Das Ziel des Projektes besteht in der Veröffentlichung diplomatisch transkribierter, lemmatisierter und grammatisch annotierter Texte. Die so bearbeiteten Daten ermöglichen - insbesondere auf grammatischer Ebene - sprachwissenschaftliche Analysen des Mittelniederdeutschen und Niederrheinischen, die weit über das bisher Mögliche hinausgehen.

Ingrid Schröder. 2014. "Das Referenzkorpus: Neue Perspektiven für die mittelniederdeutsche Grammatikographie." Jahrbuch für Germanistische Sprachgeschichte, 5(1), pp. 150-164. http://www.degruyter.com/view/j/jbgsg.2014.5.issue-1/jbgsg-2014-0011/jbgsg-2014-0011.xml .

Robert Peters, Norbert Nagel. 2014. "Das digitale 'Referenzkorpus Mittelniederdeutsch / Niederrheinisch (ReN)'." Jahrbuch für Germanistische Sprachgeschichte, 5(1), pp. 165-175. https://www.degruyter.com/view/j/jbgsg.2014.5.issue-1/jbgsg-2014-0012/jbgsg-2014-0012.xml .

 

CLARIN Metadata summary for Referenzkorpus Mittelniederdeutsch/Niederrheinisch (1200-1650)Reference Corpus Middle Low German/Low Rhenish (1200-1650) (CMDI-based)

Title: Referenzkorpus Mittelniederdeutsch/Niederrheinisch (1200-1650)
Title: Reference Corpus Middle Low German/Low Rhenish (1200-1650)
Description: The reference corpus of Middle Low German and Low Rhenish texts is based on manuscripts, prints and inscriptions. It is intended to provide an insight into the culture of speech and writing in Middle Low German and Low Rhenish regions. This spectrum of texttypes can be used to trace the linguistic development on the base of diatopic and diacronic subcategorisation. The aim of the project is the publication of diplomatic transcribed, lemmatised and grammatically annotated texts. The processed data – especially on the grammatical level – enables a linguistic analysis of the Middle Low German and Low Rhenish language, which goes far beyond what was possible until now.
Description: Das Referenzkorpus mittelniederdeutscher und niederrheinischer Texte basiert auf Handschriften, Drucken und Inschriften. Es soll einen Einblick in die Sprach- und Textkultur des niederdeutschen und niederrheinischen Raums geben. Die historische Sprachentwicklung soll in ihrer diatopischen und diachronischen Untergliederung anhand des Textsortenspektrums nachgezeichnet werden. Das Ziel des Projektes besteht in der Veröffentlichung diplomatisch transkribierter, lemmatisierter und grammatisch annotierter Texte. Die so bearbeiteten Daten ermöglichen - insbesondere auf grammatischer Ebene - sprachwissenschaftliche Analysen des Mittelniederdeutschen und Niederrheinischen, die weit über das bisher Mögliche hinausgehen.
Data owner:  ReN-Team
Contributors:  Sarah Ihden (transcriber), Sarah Ihden (annotator), Sarah Ihden (researcher), Norbert Nagel (transcriber), Norbert Nagel (researcher), Katharina Dreessen (transcriber), Katharina Dreessen (annotator), Katharina Dreessen (researcher), Annemarie Bischoff (annotator), Annemarie Bischoff (researcher), Meike Glawe (transcriber), Meike Glawe (researcher), Verena Kleymann (transcriber), Verena Kleymann (researcher), Elmar Schilling (annotator), Elmar Schilling (researcher), Nadine Wallmeier (annotator), Nadine Wallmeier (researcher), Fabian Barteld (developer), Fabian Barteld (researcher), Timm Lehmberg (developer), Timm Lehmberg (researcher), Katharina Birr (annotator), Carlotta Hübener (transcriber), Paul Kahre (annotator), Vanessa Nasielski (annotator), Anabel Recker (annotator), Lena Schnee (annotator), Henny Sluyter-Gäthje (developer), Sebastian Tews (annotator), Anna Wollenschläger (transcriber), Lisa Dücker (annotator), Sarah Matthies (transcriber), Hamasa Mirahmadi (transcriber), Eleonore Schmitt (annotator), Christina Eichhorn-Hartmeyer (annotator), Julia Hütter (annotator), Meile-Andrea Schroeder (annotator), Mattias Liebl (transcriber), Katharina Schröder (transcriber)
Project:  Referenzkorpus Mittelniederdeutsch/ Niederrheinisch (1200-1650), Deutsche Forschungsgemeinschaft (DFG)
Languages:  Middle Low German (gml), Low Rhenish ()
Size:  18 Text, 101,403 tok_anno, 102,867 tok_dipl
Annotation types:  Transcription (manual), Part-of-speech (semi-automatic), Lemma (semi-automatic), Morphology (semi-automatic)
Temporal Coverage:  1200-1650
References:  Ingrid Schröder (2014) Das Referenzkorpus: Neue Perspektiven für die mittelniederdeutsche Grammatikographie References:  Robert Peters; Norbert Nagel (2014) Das digitale 'Referenzkorpus Mittelniederdeutsch / Niederrheinisch (ReN)'

 </dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/1684</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.1684</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:1684</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.1668</dc:relation>
          <dc:rights>info:eu-repo/semantics/openAccess</dc:rights>
          <dc:rights>https://creativecommons.org/licenses/by/4.0/legalcode</dc:rights>
          <dc:subject>Middle Low German</dc:subject>
          <dc:subject>Low Rhenish</dc:subject>
          <dc:subject>ANNIS</dc:subject>
          <dc:subject>linguistics</dc:subject>
          <dc:subject>diacronic</dc:subject>
          <dc:subject>historical</dc:subject>
          <dc:title>Reference Corpus Middle Low German/Low Rhenish (1200-1650); Referenzkorpus Mittelniederdeutsch/Niederrheinisch (1200-1650)</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:1669</identifier>
        <datestamp>2022-07-15T11:46:05Z</datestamp>
        <setSpec>user-hzsk</setSpec>
        <setSpec>user-uhh</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:contributor>Tews, Sebastian</dc:contributor>
          <dc:contributor>Glawe, Meike</dc:contributor>
          <dc:contributor>Hübener, Carlotta</dc:contributor>
          <dc:contributor>Nasielski, Vanessa</dc:contributor>
          <dc:contributor>Kleymann, Verena</dc:contributor>
          <dc:contributor>Schmitt, Eleonore</dc:contributor>
          <dc:contributor>Nagel, Norbert</dc:contributor>
          <dc:contributor>Hütter, Julia</dc:contributor>
          <dc:contributor>Barteld, Fabian</dc:contributor>
          <dc:contributor>Mirahmadi, Hamasa</dc:contributor>
          <dc:contributor>Dücker, Lisa</dc:contributor>
          <dc:contributor>Ihden, Sarah</dc:contributor>
          <dc:contributor>Schilling, Elmar</dc:contributor>
          <dc:contributor>Matthies, Sarah</dc:contributor>
          <dc:contributor>Birr, Katharina</dc:contributor>
          <dc:contributor>Lehmberg, Timm</dc:contributor>
          <dc:contributor>Liebl, Mattias</dc:contributor>
          <dc:contributor>Schroeder, Meile-Andrea</dc:contributor>
          <dc:contributor>Bischoff, Annemarie</dc:contributor>
          <dc:contributor>Wallmeier, Nadine</dc:contributor>
          <dc:contributor>Sluyter-Gäthje, Henny</dc:contributor>
          <dc:contributor>Schröder, Katharina</dc:contributor>
          <dc:contributor>Kahre, Paul</dc:contributor>
          <dc:contributor>Dreessen, Katharina</dc:contributor>
          <dc:contributor>Recker, Anabel</dc:contributor>
          <dc:contributor>Eichhorn-Hartmeyer, Christina</dc:contributor>
          <dc:contributor>Schnee, Lena</dc:contributor>
          <dc:contributor>Wollenschläger, Anna</dc:contributor>
          <dc:creator>Schröder, Ingrid</dc:creator>
          <dc:creator>Peters, Robert</dc:creator>
          <dc:date>2016-08-23</dc:date>
          <dc:description>The project "Reference Corpus Middle Low German/Low Rhenish (1200–1650)", abbreviated as "ReN", is part of the "Corpus of Historical German Texts", which includes the projects "Old German Reference Corpus (750–1050)", the "Reference Corpus Middle High German (1050–1350)" and the "Reference Corpus EarlyModern New High German (1350–1650)". The project "Reference Corpus Middle Low German/Low Rhenish (1200–1650)" deals with a structured selection of Middle Low German and Low Rhenish monuments of speech from 1200 to 1650. This selection is based on the parameters "space", "time" and "field of writing".

The reference corpus of Middle Low German and Low Rhenish texts is based on manuscripts, prints and inscriptions. It is intended to provide an insight into the culture of speech and writing in Middle Low German and Low Rhenish regions. This spectrum of texttypes can be used to trace the linguistic development on the base of diatopic and diacronic subcategorisation. The aim of the project is the publication of diplomatic transcribed, lemmatised and grammatically annotated texts. The processed data – especially on the grammatical level – enables a linguistic analysis of the Middle Low German and Low Rhenish language, which goes far beyond what was possible until now.

Das Referenzkorpus mittelniederdeutscher und niederrheinischer Texte basiert auf Handschriften, Drucken und Inschriften. Es soll einen Einblick in die Sprach- und Textkultur des niederdeutschen und niederrheinischen Raums geben. Die historische Sprachentwicklung soll in ihrer diatopischen und diachronischen Untergliederung anhand des Textsortenspektrums nachgezeichnet werden. Das Ziel des Projektes besteht in der Veröffentlichung diplomatisch transkribierter, lemmatisierter und grammatisch annotierter Texte. Die so bearbeiteten Daten ermöglichen - insbesondere auf grammatischer Ebene - sprachwissenschaftliche Analysen des Mittelniederdeutschen und Niederrheinischen, die weit über das bisher Mögliche hinausgehen.

Ingrid Schröder. 2014. "Das Referenzkorpus: Neue Perspektiven für die mittelniederdeutsche Grammatikographie." Jahrbuch für Germanistische Sprachgeschichte, 5(1), pp. 150-164. http://www.degruyter.com/view/j/jbgsg.2014.5.issue-1/jbgsg-2014-0011/jbgsg-2014-0011.xml .

Robert Peters, Norbert Nagel. 2014. "Das digitale 'Referenzkorpus Mittelniederdeutsch / Niederrheinisch (ReN)'." Jahrbuch für Germanistische Sprachgeschichte, 5(1), pp. 165-175. https://www.degruyter.com/view/j/jbgsg.2014.5.issue-1/jbgsg-2014-0012/jbgsg-2014-0012.xml .

 

CLARIN Metadata summary for Referenzkorpus Mittelniederdeutsch/Niederrheinisch (1200-1650)Reference Corpus Middle Low German/Low Rhenish (1200-1650) (CMDI-based)

Title: Referenzkorpus Mittelniederdeutsch/Niederrheinisch (1200-1650)
Title: Reference Corpus Middle Low German/Low Rhenish (1200-1650)
Description: The reference corpus of Middle Low German and Low Rhenish texts is based on manuscripts, prints and inscriptions. It is intended to provide an insight into the culture of speech and writing in Middle Low German and Low Rhenish regions. This spectrum of texttypes can be used to trace the linguistic development on the base of diatopic and diacronic subcategorisation. The aim of the project is the publication of diplomatic transcribed, lemmatised and grammatically annotated texts. The processed data – especially on the grammatical level – enables a linguistic analysis of the Middle Low German and Low Rhenish language, which goes far beyond what was possible until now.
Description: Das Referenzkorpus mittelniederdeutscher und niederrheinischer Texte basiert auf Handschriften, Drucken und Inschriften. Es soll einen Einblick in die Sprach- und Textkultur des niederdeutschen und niederrheinischen Raums geben. Die historische Sprachentwicklung soll in ihrer diatopischen und diachronischen Untergliederung anhand des Textsortenspektrums nachgezeichnet werden. Das Ziel des Projektes besteht in der Veröffentlichung diplomatisch transkribierter, lemmatisierter und grammatisch annotierter Texte. Die so bearbeiteten Daten ermöglichen - insbesondere auf grammatischer Ebene - sprachwissenschaftliche Analysen des Mittelniederdeutschen und Niederrheinischen, die weit über das bisher Mögliche hinausgehen.
Publication date: 2016-08-23
Data owner:  Ingrid Schröder, Robert Peters
Contributors:  Sarah Ihden (transcriber), Sarah Ihden (annotator), Sarah Ihden (researcher), Norbert Nagel (transcriber), Norbert Nagel (researcher), Katharina Dreessen (transcriber), Katharina Dreessen (annotator), Katharina Dreessen (researcher), Annemarie Bischoff (annotator), Annemarie Bischoff (researcher), Meike Glawe (transcriber), Meike Glawe (researcher), Verena Kleymann (transcriber), Verena Kleymann (researcher), Elmar Schilling (annotator), Elmar Schilling (researcher), Nadine Wallmeier (annotator), Nadine Wallmeier (researcher), Fabian Barteld (developer), Fabian Barteld (researcher), Timm Lehmberg (developer), Timm Lehmberg (researcher), Katharina Birr (annotator), Carlotta Hübener (transcriber), Paul Kahre (annotator), Vanessa Nasielski (annotator), Anabel Recker (annotator), Lena Schnee (annotator), Henny Sluyter-Gäthje (developer), Sebastian Tews (annotator), Anna Wollenschläger (transcriber), Lisa Dücker (annotator), Sarah Matthies (transcriber), Hamasa Mirahmadi (transcriber), Eleonore Schmitt (annotator), Christina Eichhorn-Hartmeyer (annotator), Julia Hütter (annotator), Meile-Andrea Schroeder (annotator), Mattias Liebl (transcriber), Katharina Schröder (transcriber)
Project:  Referenzkorpus Mittelniederdeutsch/ Niederrheinisch (1200-1650), Deutsche Forschungsgemeinschaft (DFG)
Languages:  Middle Low German (gml), Low Rhenish ()
Size:  5 Text, 36.269 Modernized token (tok_mod), 37.215 Diplomatic token (tok_dipl)
Annotation types:  Transcription (manual), Part-of-speech (semi-automatic), Lemma (semi-automatic), Morphology (semi-automatic)
Temporal Coverage:  1200-1650
References:  Ingrid Schröder (2014) Das Referenzkorpus: Neue Perspektiven für die mittelniederdeutsche Grammatikographie References:  Robert Peters; Norbert Nagel (2014) Das digitale 'Referenzkorpus Mittelniederdeutsch / Niederrheinisch (ReN)'

 </dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/1669</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.1669</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:1669</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.1668</dc:relation>
          <dc:rights>info:eu-repo/semantics/openAccess</dc:rights>
          <dc:rights>https://creativecommons.org/licenses/by/4.0/legalcode</dc:rights>
          <dc:subject>Middle Low German</dc:subject>
          <dc:subject>Low Rhenish</dc:subject>
          <dc:subject>ANNIS</dc:subject>
          <dc:subject>linguistics</dc:subject>
          <dc:subject>diacronic</dc:subject>
          <dc:subject>historical</dc:subject>
          <dc:title>Reference Corpus Middle Low German/Low Rhenish (1200-1650); Referenzkorpus Mittelniederdeutsch/Niederrheinisch (1200-1650)</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:8249</identifier>
        <datestamp>2021-09-21T14:05:40Z</datestamp>
        <setSpec>user-uhh</setSpec>
        <setSpec>user-hzsk</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:creator>Wiese, Heike</dc:creator>
          <dc:date>2016-11-21</dc:date>
          <dc:description>A corpus including E-Mails and reader comments of the public debate about "Kiezdeutsch" ("Einstellungen" (Settings)- Additional to the KiezDeutsch- corpus, KiDKo). KiDKo/E captures spontaneous data from the public discussion on "Kiezdeutsch" (neighborhood German): The body gathers emails and letters to the editor,written in response to Kiezdeutsch media reports. KiDKo / E thus provides data on attitudes and perceptions and language ideologies that became clear in the context of the Kiezdeutsch discussion, however, often refer to more extensive domains such as multilingualism, standard language, language prestige and social class.

Ein Korpus aus Emails und Leserkommentaren in der öffentlichen Debatte zu Kiezdeutsch ("Einstellungen"-Ergänzung zum KiezDeutsch-Korpus, KiDKo). KiDKo/E erfasst Spontandaten aus der öffentlichen Diskussion zu Kiezdeutsch: Das Korpus versammelt Email-Zuschriften und Leserbriefe, die in Reaktion auf Medienberichte zu Kiezdeutsch verfasst wurden. KiDKo/E liefert damit Daten zu Einstellungen, Wahrnehmungen und Sprachideologien, die im Rahmen der Diskussion zu Kiezdeutsch deutlich wurden, sich dabei jedoch oft auf weiter gehende Domänen wie Mehrsprachigkeit, Standardsprache, Sprachprestige und soziale Schicht beziehen.</dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/8249</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.8249</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:8249</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.8247</dc:relation>
          <dc:relation>doi:10.25592/uhhfdm.8248</dc:relation>
          <dc:rights>info:eu-repo/semantics/restrictedAccess</dc:rights>
          <dc:subject>written language</dc:subject>
          <dc:subject>urban youth language</dc:subject>
          <dc:subject>Kiezdeutsch</dc:subject>
          <dc:subject>urban vernaculars</dc:subject>
          <dc:subject>German</dc:subject>
          <dc:subject>Sprachliche Entwicklung im Gegenwartsdeutschen</dc:subject>
          <dc:subject>informeller Sprachgebrauch</dc:subject>
          <dc:subject>Jugendsprache im urbanen Raum</dc:subject>
          <dc:subject>ANNIS</dc:subject>
          <dc:subject>EXMARaLDA</dc:subject>
          <dc:title>Das Kiezdeutschkorpus "KiDKo": Einstellungen (KiDKo/E)</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:8323</identifier>
        <datestamp>2021-05-13T10:26:33Z</datestamp>
        <setSpec>user-uhh</setSpec>
        <setSpec>user-hzsk</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:contributor>Andresen, Melanie</dc:contributor>
          <dc:contributor>Knorr, Dagmar</dc:contributor>
          <dc:creator>Knorr, Dagmar</dc:creator>
          <dc:date>2015-11-18</dc:date>
          <dc:description>Authentic texts written by students of the University of Hamburg as part of their studies, the students have various L1 languages and study various subjects, all of the texts were subject of a writing counseling at the Writing Center Multilingualism (Schreibwerkstatt Mehrsprachigkeit), for some of the texts comments by peer tutors and several versions are available.

Authentische Texte Studierender der Universität Hamburg unterschiedlicher Herkunftssprachen und Studiengänge, die im Rahmen ihres Studiums als Prüfungsleistung entstanden sind und in einer individuellen Schreibberatung an der Schreibwerkstatt Mehrsprachigkeit besprochen wurden, zum Teil mit Kommentaren von Peer Tutoren und mehreren Versionen aus dem Schreibprozess.

Im Institut für Interkulturelle Bildung der Fakultät für Erziehungswissenschaft an der Universität Hamburg wird seit mehreren Jahrzehnten der Zusammenhang zwischen Bildungserfolg und Bildungssprache erforscht. Es wurde beobachtet, dass Studierende mit Migrationshintergrund Probleme beim Verfassen akademischer Texte haben. Die – naheliegende – Hypothese war, dass dies an den sprachlichen Fähigkeiten der Studierenden liegt. Allerdings ist eine Überprüfung der Hypothese ohne entsprechendes empirisches Datenmaterial schwierig. Die Fachliteratur zum deutschsprachigen akademischen Schreiben weist hier eine Lücke auf: Zugängliche Korpora von authentischen, deutschsprachigen akademischen Texten von Studierenden, die möglichst auch noch longitudinal untersucht werden können, existierten nicht. Mit dem Korpus KoLaS möchte die Schreibwerkstatt Mehrsprachigkeit einen Beitrag zur Forschung leisten, indem authentische Texte von Studierenden für Forschungszwecke zur Verfügung gestellt werden.

Andresen, Melanie; Knorr, Dagmar (2015) KoLaS: Kommentiertes Lernendenkorpus akademisches Schreiben

 

CLARIN Metadata summary for Commented Learner Corpus Academic WritingKommentiertes Lerndenkorpus akademisches Schreiben (CMDI-based)

Title: Commented Learner Corpus Academic Writing
Title: Kommentiertes Lerndenkorpus akademisches Schreiben
Description: Authentic texts written by students of the University of Hamburg as part of their studies, the students have various L1 languages and study various subjects, all of the texts were subject of a writing counseling at the Writing Center Multilingualism (Schreibwerkstatt Mehrsprachigkeit), for some of the texts comments by peer tutors and several versions are available.
Description: Authentische Texte Studierender der Universität Hamburg unterschiedlicher Herkunftssprachen und Studiengänge, die im Rahmen ihres Studiums als Prüfungsleistung entstanden sind und in einer individuellen Schreibberatung an der Schreibwerkstatt Mehrsprachigkeit besprochen wurden, zum Teil mit Kommentaren von Peer Tutoren und mehreren Versionen aus dem Schreibprozess.
Publication date: 2015
Data owner:  Dagmar Knorr, Universität Hamburg / Universitätskolleg / Schreibwerkstatt Mehrsprachigkeit / Von-Melle-Park 8 / 20146 Hamburg, dagmar.knorr@uni-hamburg.de
Contributors:  Dagmar Knorr (compiler), Melanie Andresen (compiler)
Project:  Schreibwerkstatt Mehrsprachigkeit, ZEIT-Stiftung Ebelin und Gerd Bucerius (2011–2014), Schreibwerkstatt Mehrsprachigkeit, Federal Ministry of Education and Research (2012–2016)
Keywords:  academic writing, student texts, written assignments, L2 data, L1 data, aquisition of academic writing, text comments, writing counselling, EXMARaLDA, akademisches Schreiben, wissenschaftlichen Schreiben, Studierendentexte, studentische Hausarbeiten, L2-Daten, L1-Daten, Erwerb der Wissenschaftssprache, Textkommentare, Schreibberatung, EXMARaLDA
Language:  German (deu)
Size:  454 texts
Segmentation units:  text
Temporal Coverage:  2011-01-01/2013-12-31
Spatial Coverage:  Hamburg, DE
Genre:  academic writing, akademisches Schreiben
Modality:  written
References:  Andresen, Melanie; Knorr, Dagmar (2015) KoLaS: Kommentiertes Lernendenkorpus akademisches Schreiben

 </dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/8323</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.8323</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:8323</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.8322</dc:relation>
          <dc:rights>info:eu-repo/semantics/closedAccess</dc:rights>
          <dc:subject>academic writing</dc:subject>
          <dc:subject>student texts</dc:subject>
          <dc:subject>written assignments</dc:subject>
          <dc:subject>L2 data</dc:subject>
          <dc:subject>L1 data</dc:subject>
          <dc:subject>aquisition of academic writing</dc:subject>
          <dc:subject>text comments</dc:subject>
          <dc:subject>writing counselling</dc:subject>
          <dc:subject>EXMARaLDA</dc:subject>
          <dc:subject>akademisches Schreiben</dc:subject>
          <dc:subject>wissenschaftlichen Schreiben</dc:subject>
          <dc:subject>Studierendentexte</dc:subject>
          <dc:subject>studentische Hausarbeiten</dc:subject>
          <dc:subject>L2-Daten</dc:subject>
          <dc:subject>L1-Daten</dc:subject>
          <dc:subject>Erwerb der Wissenschaftssprache</dc:subject>
          <dc:subject>Textkommentare</dc:subject>
          <dc:subject>Schreibberatung</dc:subject>
          <dc:subject>EXMARaLDA</dc:subject>
          <dc:subject>linguistics</dc:subject>
          <dc:subject>German</dc:subject>
          <dc:title>Commented Learner Corpus Academic Writing; Kommentiertes Lernendenkorpus akademisches Schreiben</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:8269</identifier>
        <datestamp>2021-05-26T07:49:35Z</datestamp>
        <setSpec>user-uhh</setSpec>
        <setSpec>user-hzsk</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:creator>Gabriel, Christoph</dc:creator>
          <dc:date>2020-01-01</dc:date>
          <dc:description>Subcorpus from the HaCASpa project with material used for the Atlas interactivo de la entonación del español.</dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/8269</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.8269</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:8269</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.1438</dc:relation>
          <dc:relation>doi:10.25592/uhhfdm.8268</dc:relation>
          <dc:rights>info:eu-repo/semantics/closedAccess</dc:rights>
          <dc:subject>Spanish</dc:subject>
          <dc:subject>Linguistics</dc:subject>
          <dc:subject>Regional Variety</dc:subject>
          <dc:subject>Language Contact</dc:subject>
          <dc:title>Hamburg Corpus of Argentinean Spanish (HaCASpa) Subcorpus</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:8247</identifier>
        <datestamp>2024-03-08T09:25:35Z</datestamp>
        <setSpec>user-uhh</setSpec>
        <setSpec>user-hzsk</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:contributor>Visser, Emiel</dc:contributor>
          <dc:contributor>Schumann, Kathleen</dc:contributor>
          <dc:contributor>Bunk, Oliver</dc:contributor>
          <dc:contributor>Lestmann, Nadine</dc:contributor>
          <dc:contributor>Hamm, Sophie</dc:contributor>
          <dc:contributor>Junghans, Anne</dc:contributor>
          <dc:contributor>Tjona, Kristina</dc:contributor>
          <dc:contributor>Kostka, Julia</dc:contributor>
          <dc:contributor>Rohland, Franziska</dc:contributor>
          <dc:contributor>Kiolbassa, Jana</dc:contributor>
          <dc:contributor>Mayr, Katharina</dc:contributor>
          <dc:contributor>Freywald, Ulrike</dc:contributor>
          <dc:contributor>Özçelik, Tiner</dc:contributor>
          <dc:contributor>Popova, Gergana</dc:contributor>
          <dc:contributor>Schalowski, Sören</dc:contributor>
          <dc:contributor>Pauli, Charlotte</dc:contributor>
          <dc:contributor>Leisner, Marlen</dc:contributor>
          <dc:contributor>Reinold, Nadja</dc:contributor>
          <dc:contributor>Rehbein, Ines</dc:contributor>
          <dc:contributor>Wiese,, Heike</dc:contributor>
          <dc:contributor>Hueck, Banu</dc:contributor>
          <dc:creator>Wiese, Heike</dc:creator>
          <dc:date>2016-11-21</dc:date>
          <dc:description>A multi-modal digital corpus of spontaneous discourse data from informal, oral peer group in multi- and monoethnic speech communities.

Multimodales, digitales Korpus spontansprachlicher Gesprächsdaten aus informellen, mündlichen Peer-Group-Situationen in multi- und monoethnischen Sprechergemeinschaften.

 

CLARIN Metadata summary for Das Kiezdeutschkorpus (KiDKo) (CMDI-based)

Title: Das Kiezdeutschkorpus (KiDKo)
Description:  A multi-modal digital corpus of spontaneous discourse data from informal, oral peer group situations in multi- and monoethnic speech communities.
Description:  Multimodales, digitales Korpus spontansprachlicher Gesprächsdaten aus informellen, mündlichen Peer-Group-Situationen in multi- und monoethnischen Sprechergemeinschaften.
Publication date: 2016-11-21
Data owner:  Heike Wiese
Contributors:  Heike Wiese, heike.wiese@uni-potsdam.de) (compiler), Oliver Bunk (compiler), Ulrike Freywald (compiler), Sophie Hamm (compiler), Banu Hueck (compiler), Anne Junghans (compiler), Jana Kiolbassa (compiler), Julia Kostka (compiler), Marlen Leisner (compiler), Nadine Lestmann (compiler), Katharina Mayr (compiler), Tiner Özçelik (compiler), Charlotte Pauli (compiler), Gergana Popova (compiler), Ines Rehbein (compiler), Nadja Reinold (compiler), Franziska Rohland (compiler), Sören Schalowski (compiler), Kathleen Schumann (compiler), Kristina Tjona Sommer (compiler), Emiel Visser (compiler)
Project:  B6: Analysis on the periphery, German Research Foundation (DFG)
Keywords:  spoken language, urban youth language, Kiezdeutsch, Sprachliche Entwicklung im Gegenwartsdeutschen, informeller Sprachgebrauch, Jugendsprache im urbanen Raum, Kiezdeutsch
Language:  German (deu)
Size:  23 speakers (8 female, 15 male), 270 communications, 270 recordings, 66 hours, 270 transcriptions, 333000 words
Segmentation units:  lexeme
Annotation types:  non-verbal layer, transcription (manual): literary transcription for spoken language/GAT2, n: normalisation (automatic, dictionary lookup)orthographic norminalisation of non-canonical pronunciations, punctuations and capitalisations to Standard German, pos: automated part of speech tagging using adapted SSTS-tagset for spoken language developed for KiDKo, macro: marking of repairs, tr: transcription of Turkish language material, trnorm: norminalisation to Standard Turkish, trdtwwue: literal translation of Turkish to Standard German, trdtue: free translation of Turkish to Standard German
Temporal Coverage:  2008/2011
Spatial Coverage:  Berlin-Kreuzberg, DE
Genre:  discourse
Modality:  spoken

 </dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/8247</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.8247</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:8247</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.8246</dc:relation>
          <dc:rights>info:eu-repo/semantics/restrictedAccess</dc:rights>
          <dc:subject>spoken language</dc:subject>
          <dc:subject>urban youth language</dc:subject>
          <dc:subject>Kiezdeutsch</dc:subject>
          <dc:subject>Sprachliche Entwicklung im Gegenwartsdeutschen</dc:subject>
          <dc:subject>informeller Sprachgebrauch</dc:subject>
          <dc:subject>Jugendsprache im urbanen Raum</dc:subject>
          <dc:subject>Kiezdeutsch</dc:subject>
          <dc:subject>linguistics</dc:subject>
          <dc:subject>German</dc:subject>
          <dc:title>Das Kiezdeutschkorpus (KiDKo)</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:1446</identifier>
        <datestamp>2020-12-03T10:43:01Z</datestamp>
        <setSpec>user-hzsk</setSpec>
        <setSpec>user-uhh</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:creator>Meisel, Jürgen M.</dc:creator>
          <dc:date>1992-01-01</dc:date>
          <dc:description>Video recordings of seven German/French simultaneous bilingual children, starting at approx. 1 year and 6 months. One or two recordings eachmonth until approx. 5 years and 6 months (and some later recordings for some subjects). In each recording session (interviewer/child interaction) the child is addressed in both languages in one French and one German part. For privacy protection reasons, only the transcripts can be made available. Video recordings cannot be published.

Meisel, Jürgen (ed.) (1994): Bilingual First Language Acquistion: French and German Grammatical Development (Language Acquisition &amp; Language Disorders). Amsterdam: John Benjamins.</dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/1446</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.1446</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:1446</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.1445</dc:relation>
          <dc:rights>info:eu-repo/semantics/restrictedAccess</dc:rights>
          <dc:subject>German</dc:subject>
          <dc:subject>French</dc:subject>
          <dc:subject>linguistics</dc:subject>
          <dc:subject>language acquisition</dc:subject>
          <dc:subject>bilingualism</dc:subject>
          <dc:subject>child language acquisition</dc:subject>
          <dc:subject>longitudinal data</dc:subject>
          <dc:subject>child bilingualism</dc:subject>
          <dc:subject>simultaneous bilingualism</dc:subject>
          <dc:subject>L1 data</dc:subject>
          <dc:subject>L2 data</dc:subject>
          <dc:subject>EXMARaLDA</dc:subject>
          <dc:title>DUFDE</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:1450</identifier>
        <datestamp>2020-12-03T10:43:01Z</datestamp>
        <setSpec>user-hzsk</setSpec>
        <setSpec>user-uhh</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:creator>Meisel, Jürgen M.</dc:creator>
          <dc:date>1992-01-01</dc:date>
          <dc:description>Subcorpus (two children) from the DUFDE project.

Video recordings of two German/French simultaneous bilingual children, starting at approx. 1 year and 6 months. One or two recordings eachmonth until approx. 5 years and 6 months (and some later recordings for some subjects). In each recording session (interviewer/child interaction) the child is addressed in both languages in one French and one German part. For privacy protection reasons, only the transcripts can be made available. Video recordings cannot be published.

Meisel, Jürgen (ed.) (1994): Bilingual First Language Acquistion: French and German Grammatical Development (Language Acquisition &amp; Language Disorders). Amsterdam: John Benjamins.</dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/1450</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.1450</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:1450</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.1446</dc:relation>
          <dc:relation>doi:10.25592/uhhfdm.1449</dc:relation>
          <dc:rights>info:eu-repo/semantics/restrictedAccess</dc:rights>
          <dc:subject>German</dc:subject>
          <dc:subject>French</dc:subject>
          <dc:subject>linguistics</dc:subject>
          <dc:subject>language acquisition</dc:subject>
          <dc:subject>bilingualism</dc:subject>
          <dc:subject>child language acquisition</dc:subject>
          <dc:subject>longitudinal data</dc:subject>
          <dc:subject>child bilingualism</dc:subject>
          <dc:subject>simultaneous bilingualism</dc:subject>
          <dc:subject>L1 data</dc:subject>
          <dc:subject>L2 data</dc:subject>
          <dc:subject>EXMARaLDA</dc:subject>
          <dc:title>DUFDE_AN_PA</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:1476</identifier>
        <datestamp>2020-08-31T11:59:56Z</datestamp>
        <setSpec>user-hzsk</setSpec>
        <setSpec>user-uhh</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:creator>Rehbein, Jochen</dc:creator>
          <dc:date>2020-08-31</dc:date>
          <dc:description>Audio recordings of evocative field experiments (picture story, retelling, spontaneous discourse etc.) with Turkish/German bilingual children and monolingual Turkish / monolingual German children as control data.

ENDFAS and SKOBI are two corpora of Turkish-German bilingual children, created for the research projects ENDFAS ("Die Entwicklung narrativer Diskursfähigkeiten im Deutschen und Türkischen in Familie und Schule") and SKOBI ("Sprachliche Konnektivität bei bilingual türkisch-deutsch aufwachsenden Kindern und Jugendlichen") at the University of Hamburg. The corpus creators are Jochen Rehbein, Annette Herkenrath and Birsel Karakoç. Post-editing and publication of the data was done by project Z2 at the Research Center on Multilingualism, University of Hamburg.

Herkenrath, Annette, and Jochen Rehbein. 2012. "Pragmatic Corpus Analysis, Exemplified by Turkish-German Bilingual and Monolingual Data." In Multilingual Corpora and Multilingual Corpus Analysis, edited by Thomas Schmidt and Kai Wörner, pp. 123–52. Hamburg Studies in Multilingualism, 14. John Benjamins.

 

CLARIN Metadata summary for Rehbein-ENDFAS/Rehbein-SKOBI-Korpus (CMDI-based)

Title: Rehbein-ENDFAS/Rehbein-SKOBI-Korpus
Description: Audio recordings of evocative field experiments (picture story, retelling, spontaneous discourse etc.) with Turkish/German bilingual children and monolingual Turkish / monolingual German children as control data.
Publication date: 2009-03-04
Data owner:  Jochen Rehbein, Institut für Germanistik I / Von Melle Park 6 / D-20146 Hamburg, rehbein@uni-hamburg.de
Contributors:  Jochen Rehbein, Institut für Germanistik I / Von Melle Park 6 / D-20146 Hamburg, rehbein@uni-hamburg.de (compiler)
Project:  E5 "Linguistic connectivity in bilingual Turkish-German children", German Research Foundation (DFG)
Keywords:  child language acquisition, child bilingualism, monolingual data, L1 data, L2 data, contact variety, EXMARaLDA
Languages:  German (deu), Turkish (tur)
Size:  523 speakers (284 female, 200 male, 36 unknown), 1017 communications, 371 recordings, 14033 minutes, 836 transcriptions, 748074 words
Temporal Coverage:  1987-10-10/2009-09-19
Spatial Coverage:  Bozyazı, TR; Anamur, TR; Istanbul, TR; Hamburg, DE; Ankara, TR; TR; Mersin, TR; istanbul, TR; Hamburg, TR; Anakara, TR; Icel, Bozyazi, TR; Icel, Anamur, TR; Icel, Anamar, TR; Icel (Anamur), TR; Icel, (Anamur), TR; Icel, (Bozyazi), TR; Icel (Bozyazi), TR; Eskisehir, TR; Tokat, TR; Izmir, TR; Istan, TR; Istanbul, TR
Genre:  discourse
Modality:  spoken

 </dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/1476</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.1476</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:1476</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.1475</dc:relation>
          <dc:rights>info:eu-repo/semantics/restrictedAccess</dc:rights>
          <dc:subject>child language acquisition</dc:subject>
          <dc:subject>child bilingualism</dc:subject>
          <dc:subject>monolingual data</dc:subject>
          <dc:subject>L1 data</dc:subject>
          <dc:subject>L2 data</dc:subject>
          <dc:subject>contact variety</dc:subject>
          <dc:subject>EXMARaLDA</dc:subject>
          <dc:subject>linguistics</dc:subject>
          <dc:subject>German</dc:subject>
          <dc:subject>Turkish</dc:subject>
          <dc:subject>SFB 538</dc:subject>
          <dc:title>Rehbein-ENDFAS/Rehbein-SKOBI-Korpus</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:1531</identifier>
        <datestamp>2021-05-27T11:10:24Z</datestamp>
        <setSpec>user-hzsk</setSpec>
        <setSpec>user-uhh</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:creator>Lleó, Conxita</dc:creator>
          <dc:date>2020-09-07</dc:date>
          <dc:description>Audio recordings in Spanish with 23 German/Spanish simultaneous bilingual children living in Germany and attending the Spanish complementary school at the first level. 1-6 recordings with each child, with 11 children also before the children attended the Spanish complementary school. All recordings feature elicited speech: A picture naming task, a story telling task, a morphosyntactic test, a lexical test, and the HAVAS 5. Rich metadata on language use and attitudes in the family submitted by the parents.

ALCEBLA is a phonetically and orthographically transcribed corpus of German and Spanish, finalized in the project Research based support of the complementary Spanish school in Germany (FUSED) at the Research Center on Multilingualism, University of Hamburg.

Ulloa Saceda, Marta; Lleó, Conxita and García Sánchez, Izarbe (2012): Corpora of spoken Spanish by simultaneous and successive German-Spanish bilingual and Spanish monolingual children. In: Schmidt, Thomas and Wörner, Kai (eds.): Multilingual Corpora and Multilingual Corpus Analysis. Hamburg Studies in Multilingualism (14), 97-106. Amsterdam: John Benjamins.

 

CLARIN Metadata summary for ALCEBLA (CMDI-based)

Title: ALCEBLA
Description: Audio recordings in Spanish with 23 German/Spanish simultaneous bilingual children living in Germany and attending the Spanish complementary school at the first level. 1-6 recordings with each child, with 11 children also before the children attended the Spanish complementary school. All recordings feature elicited speech: A picture naming task, a story telling task, a morphosyntactic test, a lexical test, and the HAVAS 5. Rich metadata on language use and attitudes in the family submitted by the parents.
Publication date: 2020
Data owner:  Conxita Lleó, lleo@uni-hamburg.de
Contributors:  Conxita Lleó, Institut für Romanistik / Von-Melle-Park 6 / D-20146 Hamburg, lleo@uni-hamburg.de (compiler)
Project:  T4 "Research based support of the complementary Spanish school in Germany (FUSED)", German Research Foundation (DFG)
Keywords:  child language acquisition, child bilingualism, simultaneous bilingualism, EXMARaLDA
Languages:  German (deu), Spanish (spa)
Size:  23 speakers (14 female, 9 male), 86 communications, 84 recordings, 2434 minutes, 66 transcriptions, 36717 words
Spatial Coverage:  DE
Genre:  discourse
Modality:  spoken
References:  Ulloa Saceda, Marta; Lleó, Conxita and García Sánchez, Izarbe (2012): Corpora of spoken Spanish by simultaneous and successive German-Spanish bilingual and Spanish monolingual children. In: Schmidt, Thomas and Wörner, Kai (eds.): Multilingual Corpora and Multilingual Corpus Analysis. Hamburg Studies in Multilingualism (14), 97-106. Amsterdam: John Benjamins.</dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/1531</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.1531</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:1531</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.1529</dc:relation>
          <dc:rights>info:eu-repo/semantics/restrictedAccess</dc:rights>
          <dc:subject>child language acquisition</dc:subject>
          <dc:subject>child bilingualism</dc:subject>
          <dc:subject>simultaneous bilingualism</dc:subject>
          <dc:subject>EXMARaLDA</dc:subject>
          <dc:subject>linguistics</dc:subject>
          <dc:subject>German</dc:subject>
          <dc:subject>Spanish</dc:subject>
          <dc:title>ALCEBLA</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:1537</identifier>
        <datestamp>2020-12-03T10:43:00Z</datestamp>
        <setSpec>user-hzsk</setSpec>
        <setSpec>user-uhh</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:creator>Meisel, Jürgen M.</dc:creator>
          <dc:date>2020-09-07</dc:date>
          <dc:description>Subcorpus 'Brazilian Portuguese' (two children) from the BIPODE project.

Video recordings and transcriptions of two German/Portuguese simultaneous bilingual children, starting at approx. 1 year and 6 months. One or two recordings each month until approx. 5 years and 6 months. In each recording session (interviewer/child interaction) the child is addressed in both languages in one Portuguese and one German part. For privacy protection reasons, only the transcripts can be made available. Video recordings cannot be published.</dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/1537</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.1537</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:1537</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.1535</dc:relation>
          <dc:relation>doi:10.25592/uhhfdm.1536</dc:relation>
          <dc:rights>info:eu-repo/semantics/restrictedAccess</dc:rights>
          <dc:subject>German</dc:subject>
          <dc:subject>Portuguese</dc:subject>
          <dc:subject>Brazilian Portuguese</dc:subject>
          <dc:subject>Linguistics</dc:subject>
          <dc:subject>Language Acquisition</dc:subject>
          <dc:subject>Bilingualism</dc:subject>
          <dc:subject>simultaneous bilingualism</dc:subject>
          <dc:subject>child language acquisition</dc:subject>
          <dc:subject>child bilingualism</dc:subject>
          <dc:subject>longitudinal data</dc:subject>
          <dc:subject>L1 data</dc:subject>
          <dc:subject>L2 data</dc:subject>
          <dc:subject>EXMARaLDA</dc:subject>
          <dc:title>BIPODE_BR</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:1548</identifier>
        <datestamp>2020-09-07T08:57:17Z</datestamp>
        <setSpec>user-hzsk</setSpec>
        <setSpec>user-uhh</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:creator>Wörner, Kai</dc:creator>
          <dc:creator>Schmidt, Jara</dc:creator>
          <dc:creator>Bührig, Kristin</dc:creator>
          <dc:date>2020-09-07</dc:date>
          <dc:description>Sammlung von Nutzungsrichtlinien zur Freigabe von Korpora aus dem Bestand des HZSK.</dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/1548</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.1548</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:1548</dc:identifier>
          <dc:language>deu</dc:language>
          <dc:relation>doi:10.25592/uhhfdm.1547</dc:relation>
          <dc:rights>info:eu-repo/semantics/openAccess</dc:rights>
          <dc:rights>https://creativecommons.org/publicdomain/zero/1.0/legalcode</dc:rights>
          <dc:title>Nutzungsrichtlinien zur Korpusfreigabe</dc:title>
          <dc:type>info:eu-repo/semantics/technicalDocumentation</dc:type>
          <dc:type>publication-technicalnote</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:1546</identifier>
        <datestamp>2021-05-27T11:10:22Z</datestamp>
        <setSpec>user-hzsk</setSpec>
        <setSpec>user-uhh</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:creator>Lleó, Conxita</dc:creator>
          <dc:date>2020-09-07</dc:date>
          <dc:description>Audio and Video recordings of four German/Spanish bilingual children starting at approx. 1 year and 6 months and ending at age 6-7 years with about 392 recordings (interviewer/child interaction), 185 in a German and 207 in a Spanish speaking environment.

PhonBLA Longitudinalstudie Hamburg is a phonetically and orthographically transcribed corpus of German and Spanish, finalized in the project Prosodic constraints on phonological and morphological development in bilingual first language acquisition at the Research Center on Multilingualism, University of Hamburg.

Lleó, Conxita. 2012. "Monolingual and Bilingual Phonoprosodic Corpora of Child German and Child Spanish." In Multilingual Corpora and Multilingual Corpus Analysis, edited by Thomas Schmidt and Kai Wörner, 14:pp. 107–22. Hamburg Studies in Multilingualism. John Benjamins.

 

CLARIN Metadata summary for PhonBLA Longitudinalstudie Hamburg (CMDI-based)

Title: PhonBLA Longitudinalstudie Hamburg
Description: Audio and Video recordings of four German/Spanish bilingual children starting at approx. 1 year and 6 months and ending at age 6-7 years with about 392 recordings (interviewer/child interaction), 185 in a German and 207 in a Spanish speaking environment.
Data owner:  Conxita Lleó, Institut für Romanistik / Von-Melle-Park 6 / D-20146 Hamburg, lleo@uni-hamburg.de
Contributors:  Conxita Lleó, Institut für Romanistik / Von-Melle-Park 6 / D-20146 Hamburg, lleo@uni-hamburg.de (compiler)
Project:  E3 "Prosodic Constraints on Phonological and Morphological development in Bilingual First Language Acquisition", German Research Foundation (DFG)
Keywords:  child language acquisition, child bilingualism, longitudinal data, simultaneous bilingualism, L2 data, L1 data, EXMARaLDA
Languages:  German (deu), Spanish (spa)
Size:  61 speakers (42 female, 18 male, 1 unknown), 413 communications, 392 recordings, 12725 minutes, 413 transcriptions, 303792 words
Genre:  discourse
Modality:  spoken
References:  Lleó, Conxita. 2012. "Monolingual and Bilingual Phonoprosodic Corpora of Child German and Child Spanish." In Multilingual Corpora and Multilingual Corpus Analysis, edited by Thomas Schmidt and Kai Wörner, 14:pp. 107–22. Hamburg Studies in Multilingualism. John Benjamins.

 </dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/1546</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.1546</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:1546</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.1545</dc:relation>
          <dc:rights>info:eu-repo/semantics/restrictedAccess</dc:rights>
          <dc:subject>child language acquisition</dc:subject>
          <dc:subject>child bilingualism</dc:subject>
          <dc:subject>longitudinal data</dc:subject>
          <dc:subject>simultaneous bilingualism</dc:subject>
          <dc:subject>L2 data</dc:subject>
          <dc:subject>L1 data</dc:subject>
          <dc:subject>EXMARaLDA</dc:subject>
          <dc:subject>linguistics</dc:subject>
          <dc:subject>German</dc:subject>
          <dc:subject>Spanish</dc:subject>
          <dc:title>PhonBLA Longitudinalstudie Hamburg</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:1551</identifier>
        <datestamp>2021-05-27T11:10:22Z</datestamp>
        <setSpec>user-hzsk</setSpec>
        <setSpec>user-uhh</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:creator>Lleó, Conxita</dc:creator>
          <dc:date>2020-09-07</dc:date>
          <dc:description>Audio and Video recordings of five German/Spanish bilingual children starting at approx. 1 year and 6 months and ending at age 6-7 years with about 419 recordings (interviewer/child interaction), 203 in a German and 216 in a Spanish speaking environment.

PhonBLA Longitudinalstudie Hamburg is a phonetically and orthographically transcribed corpus of German and Spanish, finalized in the project Prosodic constraints on phonological and morphological development in bilingual first language acquisition at the Research Center on Multilingualism, University of Hamburg.

Lleó, Conxita. 2012. "Monolingual and Bilingual Phonoprosodic Corpora of Child German and Child Spanish." In Multilingual Corpora and Multilingual Corpus Analysis, edited by Thomas Schmidt and Kai Wörner, 14:pp. 107–22. Hamburg Studies in Multilingualism. John Benjamins.

 

CLARIN Metadata summary for PhonBLA Longitudinalstudie Hamburg (CMDI-based)

Title: PhonBLA Longitudinalstudie Hamburg
Description: Audio and Video recordings of five German/Spanish bilingual children starting at approx. 1 year and 6 months and ending at age 6-7 years with about 419 recordings (interviewer/child interaction), 203 in a German and 216 in a Spanish speaking environment.
Data owner:  Conxita Lleó, Institut für Romanistik / Von-Melle-Park 6 / D-20146 Hamburg, lleo@uni-hamburg.de
Contributors:  Conxita Lleó, Institut für Romanistik / Von-Melle-Park 6 / D-20146 Hamburg, lleo@uni-hamburg.de (depositor), Conxita Lleó, Institut für Romanistik / Von-Melle-Park 6 / D-20146 Hamburg, lleo@uni-hamburg.de (compiler), Conxita Lleó, Institut für Romanistik / Von-Melle-Park 6 / D-20146 Hamburg, lleo@uni-hamburg.de (researcher)
Project:  E3 "Prosodic Constraints on Phonological and Morphological development in Bilingual First Language Acquisition", German Research Foundation (DFG)
Keywords:  child language acquisition, child bilingualism, longitudinal data, simultaneous bilingualism, L2 data, L1 data, EXMARaLDA
Languages:  German (deu), Spanish (spa)
Size:  62 speakers (43 female, 18 male, 1 unknown), 422 communications, 225.96 hours, 13558 minutes, 419 recordings, 413 transcriptions, 303792 words
Genre:  discourse
Modality:  spoken
References:  Lleó, Conxita. 2012. "Monolingual and Bilingual Phonoprosodic Corpora of Child German and Child Spanish." In Multilingual Corpora and Multilingual Corpus Analysis, edited by Thomas Schmidt and Kai Wörner, 14:pp. 107–22. Hamburg Studies in Multilingualism. John Benjamins.</dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/1551</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.1551</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:1551</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.1545</dc:relation>
          <dc:rights>info:eu-repo/semantics/restrictedAccess</dc:rights>
          <dc:subject>child language acquisition</dc:subject>
          <dc:subject>child bilingualism</dc:subject>
          <dc:subject>longitudinal data</dc:subject>
          <dc:subject>simultaneous bilingualism</dc:subject>
          <dc:subject>L2 data</dc:subject>
          <dc:subject>L1 data</dc:subject>
          <dc:subject>EXMARaLDA</dc:subject>
          <dc:subject>linguistics</dc:subject>
          <dc:subject>German</dc:subject>
          <dc:subject>Spanish</dc:subject>
          <dc:title>PhonBLA Longitudinalstudie Hamburg</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:1559</identifier>
        <datestamp>2020-09-07T19:55:38Z</datestamp>
        <setSpec>user-hzsk</setSpec>
        <setSpec>user-uhh</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:creator>Braunmüller, Kurt</dc:creator>
          <dc:date>2020-09-07</dc:date>
          <dc:description>Data from the SkandSemiko project.</dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/1559</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.1559</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:1559</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.1558</dc:relation>
          <dc:rights>info:eu-repo/semantics/closedAccess</dc:rights>
          <dc:subject>Danish</dc:subject>
          <dc:subject>Swedish</dc:subject>
          <dc:subject>Linguistics</dc:subject>
          <dc:subject>Semi-communication</dc:subject>
          <dc:subject>Recepteive multilingualism</dc:subject>
          <dc:subject>Adult language acquisition</dc:subject>
          <dc:subject>Child bilingualism</dc:subject>
          <dc:subject>Youth language</dc:subject>
          <dc:title>SkandSemikoTeach</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:1597</identifier>
        <datestamp>2020-09-14T20:11:58Z</datestamp>
        <setSpec>user-hzsk</setSpec>
        <setSpec>user-uhh</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:creator>Fiedler, Ines</dc:creator>
          <dc:date>2020-09-14</dc:date>
          <dc:description>The data sets for each language consist of a small number of mini-dialogues, chosen out of the 189 entries within the Focus Translation Task (cf. Skopeteas et al. 2006: 209ff.) in order to get a basic set of utterances for comparison between the languages dealt with in the project.

Fiedler, Ines (2011) QUIS Data from Yom, Aja, Anii and Foodo. With Notes on Genetic and Areal Relations. In: Petrova &amp; M. Grubic (eds.): Interdisciplinary Studies on Information Structure 16, Linguistic Fieldnotes III: Information structure in Gur and Kwa language. 49-96.

Schwarz, Anne (2011) QUIS Data from Buli, Kɔnni and Baatɔnum. With Notes on the Comparative Approach. In: Petrova &amp; M. Grubic (eds.): Interdisciplinary Studies on Information Structure 16, Linguistic Fieldnotes III: Information structure in Gur and Kwa language. 1-48.

 

CLARIN Metadata summary for B1 Yom (CMDI-based)

Title: B1 Yom
Description: The data sets for each language consist of a small number of mini-dialogues, chosen out of the 189 entries within the Focus Translation Task (cf. Skopeteas et al. 2006: 209ff.) in order to get a basic set of utterances for comparison between the languages dealt with in the project.
Publication date: 2015
Data owner:  Dr. phil. Ines Fiedler
Contributors:  Manfred Krifka (editor), Katharina Hartmann (editor), Brigitte Reineke (editor)
Project:  Special Research Centre 632 Information structure, German Research Foundation
Keywords:  focus, Fokus
Language:  Yom (pil)
Size:  257 Token
Segmentation units:  other
Temporal Coverage:  2005-02-28/2005-03-04
Spatial Coverage:  Djougou, BJ
Genre:  discourse
Modality:  spoken
References:  Fiedler, Ines (2011) QUIS Data from Yom, Aja, Anii and Foodo. With Notes on Genetic and Areal Relations. In: Petrova &amp; M. Grubic (eds.): Interdisciplinary Studies on Information Structure 16, Linguistic Fieldnotes III: Information structure in Gur and Kwa language. 49-96. References:  Schwarz, Anne (2011) QUIS Data from Buli, Kɔnni and Baatɔnum. With Notes on the Comparative Approach. In: Petrova &amp; M. Grubic (eds.): Interdisciplinary Studies on Information Structure 16, Linguistic Fieldnotes III: Information structure in Gur and Kwa language. 1-48.

 </dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/1597</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.1597</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:1597</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.1596</dc:relation>
          <dc:rights>info:eu-repo/semantics/openAccess</dc:rights>
          <dc:rights>https://creativecommons.org/licenses/by/3.0/legalcode</dc:rights>
          <dc:subject>focus</dc:subject>
          <dc:subject>Fokus</dc:subject>
          <dc:subject>linguistics</dc:subject>
          <dc:subject>information structure</dc:subject>
          <dc:subject>Yom</dc:subject>
          <dc:title>B1 Yom</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:1603</identifier>
        <datestamp>2020-09-14T20:24:27Z</datestamp>
        <setSpec>user-hzsk</setSpec>
        <setSpec>user-uhh</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:creator>Fiedler, Ines</dc:creator>
          <dc:date>2015-09-14</dc:date>
          <dc:description>The corpus comprises out of a collection of texts from the Wolof Wikipedia, randomly chosen for their near-standard like orthography and language, and treating different topics. The texts are translated manually by a mother tongue speaker and automatically tagged by a part-of-speech tagger. No further annotation is provided.

 

CLARIN Metadata summary for B7 Wolof (Wikipedia) (CMDI-based)

Title: B7 Wolof (Wikipedia)
Description: The corpus comprises out of a collection of texts from the Wolof Wikipedia, randomly chosen for their near-standard like orthography and language, and treating different topics. The texts are translated manually by a mother tongue speaker and automatically tagged by a part-of-speech tagger. No further annotation is provided.
Publication date: 2015
Data owner:  Dr. phil. Ines Fiedler
Contributors:  Tom Güldemann (editor), Ines Fiedler (researcher), Peggy Jacob (researcher), Yokiko Morimoto (researcher), Anne Schwarz (researcher), Andreas Wetter (researcher)
Project:  Special Research Centre 632 Information structure, German Research Foundation
Keywords:  predicate-centered focus types, focus
Language:  Wolof (wol)
Size:  12725 Token
Segmentation units:  other
Genre:  wiki-article
Modality:  written

 </dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/1603</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.1603</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:1603</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.1602</dc:relation>
          <dc:rights>info:eu-repo/semantics/openAccess</dc:rights>
          <dc:rights>https://creativecommons.org/licenses/by/3.0/legalcode</dc:rights>
          <dc:subject>predicate-centered focus types</dc:subject>
          <dc:subject>focus</dc:subject>
          <dc:subject>linguistics</dc:subject>
          <dc:subject>information structure</dc:subject>
          <dc:subject>Wolof</dc:subject>
          <dc:title>B7 Wolof (Wikipedia)</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:1695</identifier>
        <datestamp>2022-07-15T11:46:05Z</datestamp>
        <setSpec>user-hzsk</setSpec>
        <setSpec>user-uhh</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:contributor>Wollenschläger, Anna</dc:contributor>
          <dc:contributor>Barteld, Fabian</dc:contributor>
          <dc:contributor>Schröder, Katharina</dc:contributor>
          <dc:contributor>Schroeder, Meile-Andrea</dc:contributor>
          <dc:contributor>Schnee, Lena</dc:contributor>
          <dc:contributor>Bischoff, Annemarie</dc:contributor>
          <dc:contributor>Matthies, Sarah</dc:contributor>
          <dc:contributor>Peterberns, Hellen</dc:contributor>
          <dc:contributor>Sluyter-Gäthje, Henny</dc:contributor>
          <dc:contributor>Mirahmadi, Hamasa</dc:contributor>
          <dc:contributor>Lehmberg, Timm</dc:contributor>
          <dc:contributor>Sturm, John</dc:contributor>
          <dc:contributor>Kahre, Paul</dc:contributor>
          <dc:contributor>Sandmann, Lena</dc:contributor>
          <dc:contributor>Dreessen, Katharina</dc:contributor>
          <dc:contributor>Eichhorn-Hartmeyer, Christina</dc:contributor>
          <dc:contributor>Kleymann, Verena</dc:contributor>
          <dc:contributor>Tiedemann, Meike</dc:contributor>
          <dc:contributor>Nagel, Norbert</dc:contributor>
          <dc:contributor>Schilling, Elmar</dc:contributor>
          <dc:contributor>Schmitt, Eleonore</dc:contributor>
          <dc:contributor>Hübener, Carlotta</dc:contributor>
          <dc:contributor>Dücker, Lisa</dc:contributor>
          <dc:contributor>Hütter, Julia</dc:contributor>
          <dc:contributor>Wallmeier, Nadine</dc:contributor>
          <dc:contributor>Tews, Sebastian</dc:contributor>
          <dc:contributor>Nasielski, Vanessa</dc:contributor>
          <dc:contributor>Birr, Katharina</dc:contributor>
          <dc:contributor>Tran, Ilka</dc:contributor>
          <dc:contributor>Liebl, Mattias</dc:contributor>
          <dc:contributor>Recker, Anabel</dc:contributor>
          <dc:contributor>Ihden, Sarah</dc:contributor>
          <dc:creator>ReN-Team</dc:creator>
          <dc:date>2018-12-18</dc:date>
          <dc:description>The project "Reference Corpus Middle Low German/Low Rhenish (1200–1650)", abbreviated as "ReN", is part of the "Corpus of Historical German Texts", which includes the projects "Old German Reference Corpus (750–1050)", the "Reference Corpus Middle High German (1050–1350)" and the "Reference Corpus EarlyModern New High German (1350–1650)". The project "Reference Corpus Middle Low German/Low Rhenish (1200–1650)" deals with a structured selection of Middle Low German and Low Rhenish monuments of speech from 1200 to 1650. This selection is based on the parameters "space", "time" and "field of writing".

The reference corpus of Middle Low German and Low Rhenish texts is based on manuscripts, prints and inscriptions. It is intended to provide an insight into the culture of speech and writing in Middle Low German and Low Rhenish regions. This spectrum of texttypes can be used to trace the linguistic development on the base of diatopic and diacronic subcategorisation. The aim of the project is the publication of diplomatic transcribed, lemmatised and grammatically annotated texts. The processed data – especially on the grammatical level – enables a linguistic analysis of the Middle Low German and Low Rhenish language, which goes far beyond what was possible until now.

Das Referenzkorpus mittelniederdeutscher und niederrheinischer Texte basiert auf Handschriften, Drucken und Inschriften. Es soll einen Einblick in die Sprach- und Textkultur des niederdeutschen und niederrheinischen Raums geben. Die historische Sprachentwicklung soll in ihrer diatopischen und diachronischen Untergliederung anhand des Textsortenspektrums nachgezeichnet werden. Das Ziel des Projektes besteht in der Veröffentlichung diplomatisch transkribierter, lemmatisierter und grammatisch annotierter Texte. Die so bearbeiteten Daten ermöglichen - insbesondere auf grammatischer Ebene - sprachwissenschaftliche Analysen des Mittelniederdeutschen und Niederrheinischen, die weit über das bisher Mögliche hinausgehen.

Ingrid Schröder. 2014. "Das Referenzkorpus: Neue Perspektiven für die mittelniederdeutsche Grammatikographie." Jahrbuch für Germanistische Sprachgeschichte, 5(1), pp. 150-164. http://www.degruyter.com/view/j/jbgsg.2014.5.issue-1/jbgsg-2014-0011/jbgsg-2014-0011.xml .

Robert Peters, Norbert Nagel. 2014. "Das digitale 'Referenzkorpus Mittelniederdeutsch / Niederrheinisch (ReN)'." Jahrbuch für Germanistische Sprachgeschichte, 5(1), pp. 165-175. https://www.degruyter.com/view/j/jbgsg.2014.5.issue-1/jbgsg-2014-0012/jbgsg-2014-0012.xml .

 

CLARIN Metadata summary for Referenzkorpus Mittelniederdeutsch/Niederrheinisch (1200-1650)Reference Corpus Middle Low German/Low Rhenish (1200-1650) (CMDI-based)

Title: Referenzkorpus Mittelniederdeutsch/Niederrheinisch (1200-1650)
Title: Reference Corpus Middle Low German/Low Rhenish (1200-1650)
Description: The reference corpus of Middle Low German and Low Rhenish texts is based on manuscripts, prints and inscriptions. It is intended to provide an insight into the culture of speech and writing in Middle Low German and Low Rhenish regions. This spectrum of texttypes can be used to trace the linguistic development on the base of diatopic and diacronic subcategorisation. The aim of the project is the publication of diplomatic transcribed, lemmatised and grammatically annotated texts. The processed data – especially on the grammatical level – enables a linguistic analysis of the Middle Low German and Low Rhenish language, which goes far beyond what was possible until now.
Description: Das Referenzkorpus mittelniederdeutscher und niederrheinischer Texte basiert auf Handschriften, Drucken und Inschriften. Es soll einen Einblick in die Sprach- und Textkultur des niederdeutschen und niederrheinischen Raums geben. Die historische Sprachentwicklung soll in ihrer diatopischen und diachronischen Untergliederung anhand des Textsortenspektrums nachgezeichnet werden. Das Ziel des Projektes besteht in der Veröffentlichung diplomatisch transkribierter, lemmatisierter und grammatisch annotierter Texte. Die so bearbeiteten Daten ermöglichen - insbesondere auf grammatischer Ebene - sprachwissenschaftliche Analysen des Mittelniederdeutschen und Niederrheinischen, die weit über das bisher Mögliche hinausgehen.
Data owner:  ReN-Team
Contributors:  Sarah Ihden (transcriber), Sarah Ihden (annotator), Sarah Ihden (researcher), Norbert Nagel (transcriber), Norbert Nagel (researcher), Katharina Dreessen (transcriber), Katharina Dreessen (annotator), Katharina Dreessen (researcher), Annemarie Bischoff (annotator), Annemarie Bischoff (researcher), Meike Tiedemann (transcriber), Meike Tiedemann (annotator), Meike Tiedemann (researcher), Verena Kleymann (transcriber), Verena Kleymann (annotator), Verena Kleymann (researcher), Elmar Schilling (annotator), Elmar Schilling (researcher), Nadine Wallmeier (annotator), Nadine Wallmeier (researcher), Fabian Barteld (developer), Fabian Barteld (researcher), Timm Lehmberg (developer), Timm Lehmberg (researcher), Katharina Birr (annotator), Carlotta Hübener (transcriber), Carlotta Hübener (annotator), Paul Kahre (annotator), Vanessa Nasielski (annotator), Anabel Recker (annotator), Lena Schnee (annotator), Henny Sluyter-Gäthje (developer), Lena Sandmann (developer), Hellen Peterberns (annotator), Sebastian Tews (annotator), Ilka Tran (transcriber), Anna Wollenschläger (transcriber), Lisa Dücker (annotator), Sarah Matthies (transcriber), Hamasa Mirahmadi (transcriber), Eleonore Schmitt (annotator), Christina Eichhorn-Hartmeyer (annotator), Julia Hütter (annotator), Meile-Andrea Schroeder (annotator), Mattias Liebl (transcriber), Katharina Schröder (transcriber), John Sturm (annotator)
Project:  Referenzkorpus Mittelniederdeutsch/ Niederrheinisch (1200-1650), Deutsche Forschungsgemeinschaft (DFG)
Languages:  Middle Low German (gml), Low Rhenish (mis)
Size:  62 Text, 571325 tok_anno, 582340 tok_dipl
Annotation types:  Transcription (manual), Part-of-speech (semi-automatic), Lemma (semi-automatic), Morphology (semi-automatic)
Temporal Coverage:  1200/1650
Genre:  various
Modality:  written
References:  Ingrid Schröder (2014) Das Referenzkorpus: Neue Perspektiven für die mittelniederdeutsche Grammatikographie References:  Robert Peters; Norbert Nagel (2014) Das digitale 'Referenzkorpus Mittelniederdeutsch / Niederrheinisch (ReN)'

 </dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/1695</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.1695</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:1695</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.1668</dc:relation>
          <dc:rights>info:eu-repo/semantics/openAccess</dc:rights>
          <dc:rights>https://creativecommons.org/licenses/by/4.0/legalcode</dc:rights>
          <dc:subject>Middle Low German</dc:subject>
          <dc:subject>Low Rhenish</dc:subject>
          <dc:subject>ANNIS</dc:subject>
          <dc:subject>linguistics</dc:subject>
          <dc:subject>diacronic</dc:subject>
          <dc:subject>historical</dc:subject>
          <dc:title>Reference Corpus Middle Low German/Low Rhenish (1200-1650); Referenzkorpus Mittelniederdeutsch/Niederrheinisch (1200-1650)</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:1687</identifier>
        <datestamp>2022-07-15T11:46:05Z</datestamp>
        <setSpec>user-hzsk</setSpec>
        <setSpec>user-uhh</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:contributor>Lehmberg, Timm</dc:contributor>
          <dc:contributor>Schröder, Katharina</dc:contributor>
          <dc:contributor>Schnee, Lena</dc:contributor>
          <dc:contributor>Liebl, Mattias</dc:contributor>
          <dc:contributor>Dreessen, Katharina</dc:contributor>
          <dc:contributor>Hütter, Julia</dc:contributor>
          <dc:contributor>Wallmeier, Nadine</dc:contributor>
          <dc:contributor>Wollenschläger, Anna</dc:contributor>
          <dc:contributor>Nasielski, Vanessa</dc:contributor>
          <dc:contributor>Recker, Anabel</dc:contributor>
          <dc:contributor>Bischoff, Annemarie</dc:contributor>
          <dc:contributor>Hübener, Carlotta</dc:contributor>
          <dc:contributor>Tran, Ilka</dc:contributor>
          <dc:contributor>Ihden, Sarah</dc:contributor>
          <dc:contributor>Kahre, Paul</dc:contributor>
          <dc:contributor>Schilling, Elmar</dc:contributor>
          <dc:contributor>Barteld, Fabian</dc:contributor>
          <dc:contributor>Eichhorn-Hartmeyer, Christina</dc:contributor>
          <dc:contributor>Tews, Sebastian</dc:contributor>
          <dc:contributor>Matthies, Sarah</dc:contributor>
          <dc:contributor>Mirahmadi, Hamasa</dc:contributor>
          <dc:contributor>Nagel, Norbert</dc:contributor>
          <dc:contributor>Kleymann, Verena</dc:contributor>
          <dc:contributor>Dücker, Lisa</dc:contributor>
          <dc:contributor>Schroeder, Meile-Andrea</dc:contributor>
          <dc:contributor>Birr, Katharina</dc:contributor>
          <dc:contributor>Schmitt, Eleonore</dc:contributor>
          <dc:contributor>Sluyter-Gäthje, Henny</dc:contributor>
          <dc:contributor>Glawe, Meike</dc:contributor>
          <dc:creator>ReN-Team</dc:creator>
          <dc:date>2017-06-15</dc:date>
          <dc:description>The project "Reference Corpus Middle Low German/Low Rhenish (1200–1650)", abbreviated as "ReN", is part of the "Corpus of Historical German Texts", which includes the projects "Old German Reference Corpus (750–1050)", the "Reference Corpus Middle High German (1050–1350)" and the "Reference Corpus EarlyModern New High German (1350–1650)". The project "Reference Corpus Middle Low German/Low Rhenish (1200–1650)" deals with a structured selection of Middle Low German and Low Rhenish monuments of speech from 1200 to 1650. This selection is based on the parameters "space", "time" and "field of writing".

The reference corpus of Middle Low German and Low Rhenish texts is based on manuscripts, prints and inscriptions. It is intended to provide an insight into the culture of speech and writing in Middle Low German and Low Rhenish regions. This spectrum of texttypes can be used to trace the linguistic development on the base of diatopic and diacronic subcategorisation. The aim of the project is the publication of diplomatic transcribed, lemmatised and grammatically annotated texts. The processed data – especially on the grammatical level – enables a linguistic analysis of the Middle Low German and Low Rhenish language, which goes far beyond what was possible until now.

Das Referenzkorpus mittelniederdeutscher und niederrheinischer Texte basiert auf Handschriften, Drucken und Inschriften. Es soll einen Einblick in die Sprach- und Textkultur des niederdeutschen und niederrheinischen Raums geben. Die historische Sprachentwicklung soll in ihrer diatopischen und diachronischen Untergliederung anhand des Textsortenspektrums nachgezeichnet werden. Das Ziel des Projektes besteht in der Veröffentlichung diplomatisch transkribierter, lemmatisierter und grammatisch annotierter Texte. Die so bearbeiteten Daten ermöglichen - insbesondere auf grammatischer Ebene - sprachwissenschaftliche Analysen des Mittelniederdeutschen und Niederrheinischen, die weit über das bisher Mögliche hinausgehen.

Ingrid Schröder. 2014. "Das Referenzkorpus: Neue Perspektiven für die mittelniederdeutsche Grammatikographie." Jahrbuch für Germanistische Sprachgeschichte, 5(1), pp. 150-164. http://www.degruyter.com/view/j/jbgsg.2014.5.issue-1/jbgsg-2014-0011/jbgsg-2014-0011.xml .

Robert Peters, Norbert Nagel. 2014. "Das digitale 'Referenzkorpus Mittelniederdeutsch / Niederrheinisch (ReN)'." Jahrbuch für Germanistische Sprachgeschichte, 5(1), pp. 165-175. https://www.degruyter.com/view/j/jbgsg.2014.5.issue-1/jbgsg-2014-0012/jbgsg-2014-0012.xml .

 

CLARIN Metadata summary for Referenzkorpus Mittelniederdeutsch/Niederrheinisch (1200-1650)Reference Corpus Middle Low German/Low Rhenish (1200-1650) (CMDI-based)

Title: Referenzkorpus Mittelniederdeutsch/Niederrheinisch (1200-1650)
Title: Reference Corpus Middle Low German/Low Rhenish (1200-1650)
Description: The reference corpus of Middle Low German and Low Rhenish texts is based on manuscripts, prints and inscriptions. It is intended to provide an insight into the culture of speech and writing in Middle Low German and Low Rhenish regions. This spectrum of texttypes can be used to trace the linguistic development on the base of diatopic and diacronic subcategorisation. The aim of the project is the publication of diplomatic transcribed, lemmatised and grammatically annotated texts. The processed data – especially on the grammatical level – enables a linguistic analysis of the Middle Low German and Low Rhenish language, which goes far beyond what was possible until now.
Description: Das Referenzkorpus mittelniederdeutscher und niederrheinischer Texte basiert auf Handschriften, Drucken und Inschriften. Es soll einen Einblick in die Sprach- und Textkultur des niederdeutschen und niederrheinischen Raums geben. Die historische Sprachentwicklung soll in ihrer diatopischen und diachronischen Untergliederung anhand des Textsortenspektrums nachgezeichnet werden. Das Ziel des Projektes besteht in der Veröffentlichung diplomatisch transkribierter, lemmatisierter und grammatisch annotierter Texte. Die so bearbeiteten Daten ermöglichen - insbesondere auf grammatischer Ebene - sprachwissenschaftliche Analysen des Mittelniederdeutschen und Niederrheinischen, die weit über das bisher Mögliche hinausgehen.
Publication date: 2017-06-15
Data owner:  ReN-Team
Contributors:  Sarah Ihden (transcriber), Sarah Ihden (annotator), Sarah Ihden (researcher), Norbert Nagel (transcriber), Norbert Nagel (researcher), Katharina Dreessen (transcriber), Katharina Dreessen (annotator), Katharina Dreessen (researcher), Annemarie Bischoff (annotator), Annemarie Bischoff (researcher), Meike Glawe (transcriber), Meike Glawe (researcher), Verena Kleymann (transcriber), Verena Kleymann (researcher), Elmar Schilling (annotator), Elmar Schilling (researcher), Nadine Wallmeier (annotator), Nadine Wallmeier (researcher), Fabian Barteld (developer), Fabian Barteld (researcher), Timm Lehmberg (developer), Timm Lehmberg (researcher), Katharina Birr (annotator), Carlotta Hübener (transcriber), Paul Kahre (annotator), Vanessa Nasielski (annotator), Anabel Recker (annotator), Lena Schnee (annotator), Henny Sluyter-Gäthje (developer), Sebastian Tews (annotator), Ilka Tran (transcriber), Anna Wollenschläger (transcriber), Lisa Dücker (annotator), Sarah Matthies (transcriber), Hamasa Mirahmadi (transcriber), Eleonore Schmitt (annotator), Christina Eichhorn-Hartmeyer (annotator), Julia Hütter (annotator), Meile-Andrea Schroeder (annotator), Mattias Liebl (transcriber), Katharina Schröder (transcriber)
Project:  Referenzkorpus Mittelniederdeutsch/ Niederrheinisch (1200-1650), Deutsche Forschungsgemeinschaft (DFG)
Languages:  Middle Low German (gml), Low Rhenish ()
Size:  32 Text, 200.664 tok_anno, 204.249 tok_dipl
Annotation types:  Transcription (manual), Part-of-speech (semi-automatic), Lemma (semi-automatic), Morphology (semi-automatic)
Temporal Coverage:  1200-1650
Genre:  various
Modality:  written
References:  Ingrid Schröder (2014) Das Referenzkorpus: Neue Perspektiven für die mittelniederdeutsche Grammatikographie References:  Robert Peters; Norbert Nagel (2014) Das digitale 'Referenzkorpus Mittelniederdeutsch / Niederrheinisch (ReN)'

 </dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/1687</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.1687</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:1687</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.1668</dc:relation>
          <dc:rights>info:eu-repo/semantics/openAccess</dc:rights>
          <dc:rights>https://creativecommons.org/licenses/by/4.0/legalcode</dc:rights>
          <dc:subject>Middle Low German</dc:subject>
          <dc:subject>Low Rhenish</dc:subject>
          <dc:subject>ANNIS</dc:subject>
          <dc:subject>linguistics</dc:subject>
          <dc:subject>diacronic</dc:subject>
          <dc:subject>historical</dc:subject>
          <dc:title>Reference Corpus Middle Low German/Low Rhenish (1200-1650); Referenzkorpus Mittelniederdeutsch/Niederrheinisch (1200-1650)</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:1705</identifier>
        <datestamp>2020-09-29T19:04:18Z</datestamp>
        <setSpec>user-hzsk</setSpec>
        <setSpec>user-uhh</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:contributor>Hartmann, Katharina</dc:contributor>
          <dc:contributor>Jacob, Peggy</dc:contributor>
          <dc:creator>Hartmann, Katharina</dc:creator>
          <dc:date>2015-09-29</dc:date>
          <dc:description>Hausa:  complete set, status: final, manually transcribed, glossed and translated to English, annotated wrt. morphology, parts of speech, syntax, gramm. function, sem. roles, focus and focus position (e.g. ex situ) in EXMARaLDA.CLARIN Metadata summary for B2 Hausa (CMDI-based)    	    	    	    		Title: B2 Hausa    	        	Description: Hausa:  complete set, status: final, manually transcribed, glossed and translated to English, annotated wrt. morphology, parts of speech, syntax, gramm. function, sem. roles, focus and focus position (e.g. ex situ) in EXMARaLDA.					Publication date: 2015					Data owner: 			Univ.-Prof. Dr. Katharina Hartmann											                		Contributors:                 	Katharina Hartmann (editor), Peggy Jacob (researcher)                			                		            	Project:             	Special Research Centre 632 Information structure, German Research Foundation            				            			            	Keywords:             	focus, Fokus            				            			    	Language:     	Hausa (hau)    						    					            	Size:             	6991 Token            					            				                		Segmentation units:                 	other                			                		            	Genre:             	discourse            				            			            	Modality:             	written            				            					</dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/1705</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.1705</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:1705</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.1704</dc:relation>
          <dc:rights>info:eu-repo/semantics/openAccess</dc:rights>
          <dc:rights>https://creativecommons.org/licenses/by-nc/3.0/legalcode</dc:rights>
          <dc:subject>focus</dc:subject>
          <dc:subject>Fokus</dc:subject>
          <dc:subject>linguistics</dc:subject>
          <dc:subject>information structure</dc:subject>
          <dc:subject>Hausa</dc:subject>
          <dc:title>B2 Hausa</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:1725</identifier>
        <datestamp>2020-09-29T21:10:59Z</datestamp>
        <setSpec>user-hzsk</setSpec>
        <setSpec>user-uhh</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:creator>House, Juliane</dc:creator>
          <dc:date>2011-01-01</dc:date>
          <dc:description>Translation corpora of original texts with translations and comparable texts from the genre popular scientific prose.

Übersetzungs- und Vergleichskorpus mit authentischen populärwissenschaftlichen Texten. Texte aus deutschsprachigen und englischsprachigen populärwissenschaftlichen Zeitschriften jeweils im Original und der jeweiligen Übersetzung. Zeitfenster 1978-1982 (E-&gt;D), 1978-1982 (D), 1999-2002 (E-&gt;D), 1999-2002 (D)</dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/1725</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.1725</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:1725</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.1724</dc:relation>
          <dc:rights>info:eu-repo/semantics/restrictedAccess</dc:rights>
          <dc:subject>translated texts</dc:subject>
          <dc:subject>popular science texts</dc:subject>
          <dc:subject>parallel corpus</dc:subject>
          <dc:subject>comparable corpus</dc:subject>
          <dc:subject>übersetzte Texte</dc:subject>
          <dc:subject>Populärwissenschaftstexte</dc:subject>
          <dc:subject>Parallelkorpus</dc:subject>
          <dc:subject>Vergleichskorpus</dc:subject>
          <dc:subject>linguistics</dc:subject>
          <dc:subject>German</dc:subject>
          <dc:subject>English</dc:subject>
          <dc:title>Covert translation: popular science</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:1721</identifier>
        <datestamp>2020-09-29T20:40:48Z</datestamp>
        <setSpec>user-hzsk</setSpec>
        <setSpec>user-uhh</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:contributor>Schlachter, Eva</dc:contributor>
          <dc:contributor>Coniglio, Marco</dc:contributor>
          <dc:contributor>Rasskazova, Oxana</dc:contributor>
          <dc:contributor>Donhauser, Karin</dc:contributor>
          <dc:contributor>Gehrlein, Anke</dc:contributor>
          <dc:contributor>Petrova, Svetlana</dc:contributor>
          <dc:creator>Petrova, Svetlana</dc:creator>
          <dc:date>2015-09-29</dc:date>
          <dc:description>The corpus contains a chronic from the 13th century in Middle Low German.Es handelt sich um eine Chronik, in Mittelniederdeutsch, 13 Jh. Beschreibung der Textzeugen usw. in: Herkommer, Hubert (1985): Sächsische Weltchronik. In: Kurt Ruh, Gundolf Keil, Werner Schröder, Burghart Wachinger and Franz Josef Worstbrock (Hg.): Die deutsche Literatur des Mittelalters. Verfasserlexikon. Band 8. Zweite, völlig neu bearbeitete Auflage. Berlin/New York, Sp. 473-500. In Exmaralda eingefügt sind 50 Seiten von dieser Textausgabe: Weiland, Ludwig (Hg.) (1980): Sächsische Weltchronik. Unveränderter Nachdruck der Ausgabe 1877. München (= Monumenta Germaniae Historica).CLARIN Metadata summary for B4 Sächsische Weltchronik (CMDI-based)    	    	    	    		Title: B4 Sächsische Weltchronik    	        	Description: The corpus contains a chronic from the 13th century in Middle Low German.		        	Description: Es handelt sich um eine Chronik, in Mittelniederdeutsch, 13 Jh. Beschreibung der Textzeugen usw. in: Herkommer, Hubert (1985): Sächsische Weltchronik. In: Kurt Ruh, Gundolf Keil, Werner Schröder, Burghart Wachinger and Franz Josef Worstbrock (Hg.): Die deutsche Literatur des Mittelalters. Verfasserlexikon. Band 8. Zweite, völlig neu bearbeitete Auflage. Berlin/New York, Sp. 473-500. In Exmaralda eingefügt sind 50 Seiten von dieser Textausgabe: Weiland, Ludwig (Hg.) (1980): Sächsische Weltchronik. Unveränderter Nachdruck der Ausgabe 1877. München (= Monumenta Germaniae Historica).					Publication date: 2015					Data owner: 			Prof. Dr. Svetlana Petrova											                		Contributors:                 	Svetlana Petrova (editor), Karin Donhauser (editor), Eva Schlachter (annotator), Marco Coniglio (annotator), Oxana Rasskazova (annotator), Anke Gehrlein (annotator)                			                		            	Project:             	Special Research Centre 632 Information structure, German Research Foundation            				            			            	Keywords:             	historical texts, religious texts, information structure            				            			    	Language:     	Old High German (goh)    						    					            	Size:             	4300 Token            					            				                		Segmentation units:                 	other                			                		            	Genre:             	historic manuscript            				            			            	Modality:             	written            				            					</dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/1721</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.1721</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:1721</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.1720</dc:relation>
          <dc:rights>info:eu-repo/semantics/openAccess</dc:rights>
          <dc:rights>https://creativecommons.org/licenses/by-nc/3.0/legalcode</dc:rights>
          <dc:subject>historical texts</dc:subject>
          <dc:subject>religious texts</dc:subject>
          <dc:subject>information structure</dc:subject>
          <dc:subject>linguistics</dc:subject>
          <dc:subject>Old High German</dc:subject>
          <dc:title>B4 Sächsische Weltchronik</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:8364</identifier>
        <datestamp>2020-11-26T12:30:18Z</datestamp>
        <setSpec>user-uhh</setSpec>
        <setSpec>user-hzsk</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:contributor>Schütte, Wilfried</dc:contributor>
          <dc:contributor>Stachowicz, Roman</dc:contributor>
          <dc:contributor>Fuchs, Florian</dc:contributor>
          <dc:contributor>Schwalm, Martina</dc:contributor>
          <dc:contributor>Görlich, Maria</dc:contributor>
          <dc:contributor>Sambale, Heidemarie</dc:contributor>
          <dc:contributor>Kaminska, Karolina</dc:contributor>
          <dc:contributor>Hamze, Kim-Chi</dc:contributor>
          <dc:contributor>Loehr,, Dan</dc:contributor>
          <dc:contributor>Schmidt, Thomas</dc:contributor>
          <dc:contributor>Watzke, Franziska</dc:contributor>
          <dc:contributor>M., Peter</dc:contributor>
          <dc:contributor>Rolle, Andrea</dc:contributor>
          <dc:contributor>Hedeland, Hanna</dc:contributor>
          <dc:contributor>Al-Jaraf, Tara</dc:contributor>
          <dc:contributor>Schnieder, Annette</dc:contributor>
          <dc:contributor>Merkel, Silke</dc:contributor>
          <dc:contributor>Yusun, Secil</dc:contributor>
          <dc:contributor>Wörner, Kai</dc:contributor>
          <dc:contributor>Stäwen, Nicole</dc:contributor>
          <dc:creator>Hamburger Zentrum für Sprachkorpora</dc:creator>
          <dc:date>2020-11-26</dc:date>
          <dc:description>A selection of short audio and video recordings in various languages to be used for instruction or demonstration of the EXMARaLDA system.

The EXMARaLDA Demo Corpus is a small corpus which you can use to try out the functionality of the EXMARaLDA system. Please note that this corpus is for demonstration purposes only and will be changed occasionally. Further information can be found in a PDF document that describes the online and offline use of the EXMARaLDA Demo Corpus.

 

CLARIN Metadata summary for EXMARaLDA Demo corpus (CMDI-based)

Title: EXMARaLDA Demo corpus
Description: A selection of short audio and video recordings in various languages to be used for instruction or demonstration of the EXMARaLDA system.
Publication date: 2020
Data owner:  Hamburger Zentrum für Sprachkorpora, Max-Brauer-Allee 60 / D-22765 Hamburg, corpora@uni-hamburg.de
Contributors:  Hamburger Zentrum für Sprachkorpora, Max-Brauer-Allee 60 / D-22765 Hamburg, corpora@uni-hamburg.de (depositor), Hamburger Zentrum für Sprachkorpora, Max-Brauer-Allee 60 / D-22765 Hamburg, corpora@uni-hamburg.de (compiler), Wilfried Schütte, schuette@ids-mannheim.de (compiler), Dan Loehr, loehrd@georgetown.edu (compiler), Hamburger Zentrum für Sprachkorpora, Max-Brauer-Allee 60 / D-22765 Hamburg, corpora@uni-hamburg.de (compiler), Secil Yusun (data_inputter), Annette Schnieder (data_inputter), Andrea Rolle (data_inputter), Silke Merkel (data_inputter), Thomas Schmidt (data_inputter), Martina Schwalm (data_inputter), Peter M. Fischer (data_inputter), Kim-Chi Hamze (data_inputter), Franziska Watzke (data_inputter), Roman Stachowicz (data_inputter), Karolina Kaminska (data_inputter), Nicole Stäwen (data_inputter), Maria Görlich (data_inputter), Tara Al-Jaraf (data_inputter), Florian Fuchs (data_inputter), Heidemarie Sambale (data_inputter), Hamburger Zentrum für Sprachkorpora, Max-Brauer-Allee 60 / D-22765 Hamburg, corpora@uni-hamburg.de (developer), Thomas Schmidt (researcher), Kai Wörner (researcher), Hanna Hedeland (researcher), Deutsche Forschungsgemeinschaft (DFG) (sponsor)
Project:  Z2 "Computer Assisted Methods for the creation and analysis of multilingual data", German Research Foundation (DFG)
Keywords:  L1 data, EXMARaLDA
Languages:  German (deu), English (eng), French (fra), Spanish (spa), Turkish (tur), Polish (pol), Vietnamese (vie), Swedish (swe), Norwegian (nor), Italian (ita), Russian (rus), Afrikaans (afr), Portuguese (por)
Size:  69 speakers (23 female, 46 male), 26 communications, 1.89 hours, 113 minutes, 26 recordings, 26 transcriptions, 19918 words
Annotation types: 


	transcription (manual): HIAT (simplified)
	HIAT Mimik und Gestik der Sprecher werden nur ansatzweise angedeutet. Abkürzungen: LA= linker Arm, RA= rechter Arm, LH= linke Hand, RH= rechte Hand, KO= Kopf, OK= Oberkörper.
	cs: code-switch
	de: German translation
	en: English translation
	k: free comment
	akz: accentuation/stress
	nv: non-verbal
	sup: suprasegmental information
	hd: Standard German translation


Temporal Coverage:  1970-01-07/2013-04-01
Spatial Coverage:  Hamburg, DE; DE; Lisbon, PT; London, GB; St. Aegidien, DE; ES; IT; Mülheimer Straße 36, 46045 Oberhausen, DE; Finnentroper Str. 39, 57439 Attendorn, DE; Reeperbahn, 20359 Hamburg, DE; 21 Jump Street, 41610 Virginia, US; Theaterplatz 2, 01067 Dresden, DE; GB; Moscow, RU; VN; US; Reykjavik, IS; PL; Paris, FR; SE
Genre:  discourse
Modality:  spoken

 </dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/8364</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.8364</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:8364</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.8363</dc:relation>
          <dc:rights>info:eu-repo/semantics/openAccess</dc:rights>
          <dc:rights>https://creativecommons.org/licenses/by-nc/4.0/legalcode</dc:rights>
          <dc:subject>L1 data</dc:subject>
          <dc:subject>EXMARaLDA</dc:subject>
          <dc:subject>linguistics</dc:subject>
          <dc:subject>German</dc:subject>
          <dc:subject>English</dc:subject>
          <dc:subject>French</dc:subject>
          <dc:subject>Spanish</dc:subject>
          <dc:subject>Turkish</dc:subject>
          <dc:subject>Polish</dc:subject>
          <dc:subject>Vietnamese</dc:subject>
          <dc:subject>Swedish</dc:subject>
          <dc:subject>Norwegian</dc:subject>
          <dc:subject>Italian</dc:subject>
          <dc:subject>Russian</dc:subject>
          <dc:subject>Afrikaans</dc:subject>
          <dc:subject>Portuguese</dc:subject>
          <dc:title>EXMARaLDA Demo corpus 1.1</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:12574</identifier>
        <datestamp>2023-06-22T07:36:44Z</datestamp>
        <setSpec>user-uhh</setSpec>
        <setSpec>user-hzsk</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:creator>Wörner, Kai</dc:creator>
          <dc:date>2023-06-21</dc:date>
          <dc:description>This record serves as a documentation and code archive for webservices developed for the CLARIN infrastructure that are no longer maintained by the HZSK. Most of them were deployed on https://corpora.uni-hamburg.de/apps, which is no longer in service.</dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/12574</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.12574</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:12574</dc:identifier>
          <dc:language>eng</dc:language>
          <dc:relation>doi:10.25592/uhhfdm.12573</dc:relation>
          <dc:rights>info:eu-repo/semantics/openAccess</dc:rights>
          <dc:rights>https://creativecommons.org/licenses/by/4.0/legalcode</dc:rights>
          <dc:subject>webservices</dc:subject>
          <dc:subject>clarin</dc:subject>
          <dc:subject>tombstone</dc:subject>
          <dc:subject>deprecated</dc:subject>
          <dc:title>Resources for deprecated CLARIN webservices</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:11121</identifier>
        <datestamp>2023-12-29T15:14:36Z</datestamp>
        <setSpec>user-hzsk</setSpec>
        <setSpec>user-uhh</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:creator>Wagner-Nagy, Beáta</dc:creator>
          <dc:creator>Brykina, Maria</dc:creator>
          <dc:creator>Gusev, Valentin</dc:creator>
          <dc:creator>Szeverényi, Sándor</dc:creator>
          <dc:date>2018-06-12</dc:date>
          <dc:description>Corpus Citation

Brykina, Maria - Valentin Gusev - Sándor Szeverényi - Beáta Wagner-Nagy 2018: “Nganasan Spoken Language Corpus (NSLC).” Archived in Hamburger Zentrum für Sprachkorpora. Version 0.2. Publication date 2018-06-12. http://hdl.handle.net/11022/0000-0007-C6F2-8

Corpus Description

The Nganasan Spoken Language Corpus, Version 0.2 has been created as part of Corpus based grammatical studies on Nganasan project (supported by the German Research Grant; WA3153/2-1) whose primary goal is to generate a digital, machine-searchable corpus of spoken Nganasan and based on this corpus, to prepare a corpus-based reference grammar of the language.

This project fills basic gaps in the existing research into Nganasan descriptive grammar creating new and more widely accessible materials and information on this lesser known and severely endangered Uralic language.

This second version 0.2 of the corpus is a subcorpus that comprises 177 communications, 136 of which contain an aligned audio recording, with glossed (Toolbox/FLEx) and annotated (EXMARaLDA) transcripts from 57 speakers. All texts have been translated into Russian and English, some also into German. The corpus also contains rich metadata on the communications and speakers.</dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/11121</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.11121</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:11121</dc:identifier>
          <dc:language>nio</dc:language>
          <dc:relation>info:eu-repo/semantics/altIdentifier/handle/11022/0000-0007-C6F2-8</dc:relation>
          <dc:relation>doi:10.25592/uhhfdm.11120</dc:relation>
          <dc:rights>info:eu-repo/semantics/openAccess</dc:rights>
          <dc:rights>https://creativecommons.org/licenses/by/4.0/legalcode</dc:rights>
          <dc:title>Nganasan Spoken Language Corpus (NSLC)</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:14222</identifier>
        <datestamp>2024-06-05T09:39:33Z</datestamp>
        <setSpec>user-hzsk</setSpec>
        <setSpec>user-uhh</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:creator>Wagner-Nagy, Beáta</dc:creator>
          <dc:creator>Budzisch, Josefina</dc:creator>
          <dc:creator>Behnke, Anja</dc:creator>
          <dc:date>2024-04-23</dc:date>
          <dc:description>Corpus Citation

Budzisch, Josefina – Anja Behnke – Beáta Wagner-Nagy 2024. Selkup Language Corpus (SLC). Archived at Universität Hamburg. Version 2.0. Publication date 2024-04-23. https://hdl.handle.net/11022/0000-0007-FE08-3.

Corpus Description

The corpus has been created within the projects supported by the German Research Grant (WA 3153/3-1: Syntactic description of the Central and Southern Selkup dialects: a corpus based analyses and WA 3153/7-1: Konverbale &amp; Äquivalente Strukturen: Eine komparative Untersuchung des Selkupischen und der samojedischen Sprachen). The corpus contains 154 texts already published in written form with glosses and annotations. All texts have been translated into English, and mostly into Russian and German. The corpus also contains rich metadata on the communications and speakers.</dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/14222</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.14222</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:14222</dc:identifier>
          <dc:language>sel</dc:language>
          <dc:relation>info:eu-repo/semantics/altIdentifier/handle/11022/0000-0007-FE08-3</dc:relation>
          <dc:relation>doi:10.25592/uhhfdm.11058</dc:relation>
          <dc:rights>info:eu-repo/semantics/openAccess</dc:rights>
          <dc:rights>https://creativecommons.org/licenses/by-sa/4.0/legalcode</dc:rights>
          <dc:title>Selkup Language Corpus (SLC)</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:1875</identifier>
        <datestamp>2025-06-05T12:13:23Z</datestamp>
        <setSpec>user-hzsk</setSpec>
        <setSpec>user-uhh</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:contributor>Stegen, Florian</dc:contributor>
          <dc:contributor>Baumann, Timo</dc:contributor>
          <dc:contributor>Köhn, Arne</dc:contributor>
          <dc:creator>Baumann, Timo</dc:creator>
          <dc:date>2017-10-27</dc:date>
          <dc:description>The Spoken Wikipedia project unites volunteer readers of Wikipedia articles. Hundreds of spoken articles in multiple languages are available to users who are – for one reason or another – unable or unwilling to consume the written version of the article. Our resource, the Spoken Wikipedia Corpus, consolidates the Spoken Wikipediae, adding text segmentation, normalization, time-alignment and further annotations, making it accessible for research and fostering new ways of interacting with the material.

Timo Baumann and Arne Köhn and Felix Hennig. 2018. The Spoken Wikipedia Corpus Collection: Harvesting, Alignment and an Application to Hyperlistening, in Language Resources and Evaluation, Special Issue representing significant contributions of LREC 2016.

Arne Köhn, Florian Stegen, Timo Baumann. 2016. Mining the Spoken Wikipedia for Speech Data and Beyond, in Proceedings of the Tenth International Conference on Language Resources and Evaluation (LREC 2016).

 

CLARIN Metadata summary for The Spoken Wikipedia Corpora (CMDI-based)

Title: The Spoken Wikipedia Corpora
Description:  The Spoken Wikipedia project unites volunteer readers of Wikipedia articles. Hundreds of spoken articles in multiple languages are available to users who are – for one reason or another – unable or unwilling to consume the written version of the article. Our resource, the Spoken Wikipedia Corpus, consolidates the Spoken Wikipediae, adding text segmentation, normalization, time-alignment and further annotations, making it accessible for research and fostering new ways of interacting with the material.
Publication date: 2017
Data owner:  Timo Baumann - Universität Hamburg
Contributors:  Timo Baumann (author), Arne Köhn (author), Florian Stegen (author)
Languages:  English (eng), German (deu), Dutch (nld)
Size:  5397 article, 1005 hour
Segmentation units:  other
Genre:  encyclopedia
Modality:  spoken
References:  Timo Baumann; Arne Köhn; Felix Hennig (2018) The Spoken Wikipedia Corpus Collection: Harvesting, Alignment and an Application to Hyperlistening References:  Arne Köhn; Florian Stegen; Timo Baumann (2016) Mining the Spoken Wikipedia for Speech Data and Beyond

 </dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/1875</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.1875</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:1875</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.1874</dc:relation>
          <dc:rights>info:eu-repo/semantics/openAccess</dc:rights>
          <dc:rights>https://creativecommons.org/licenses/by-sa/4.0/legalcode</dc:rights>
          <dc:source>Language Resources and Evaluation 53(2) 303–329</dc:source>
          <dc:subject>linguistics</dc:subject>
          <dc:subject>English</dc:subject>
          <dc:subject>German</dc:subject>
          <dc:subject>Dutch</dc:subject>
          <dc:title>The Spoken Wikipedia Corpora</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:1440</identifier>
        <datestamp>2022-12-07T10:12:05Z</datestamp>
        <setSpec>user-hzsk</setSpec>
        <setSpec>user-uhh</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:creator>Bührig, Kristin</dc:creator>
          <dc:creator>Meyer, Bernd</dc:creator>
          <dc:date>1997-07-11</dc:date>
          <dc:description>Transcriptions of audio recordings of various kinds of doctor-patient communication in hospitals. There are both monolingual conversations in German, Portuguese and Turkish, recorded in the respective country, and interpreted conversations recorded in Germany (i.e. in German-Turkish, German-Portuguese, and German-Portuguese/Spanish), about 15-20 recordings of each kind. The persons interpreting are bilingual hospital employees or relatives of the patients, who are all adults living in Germany but with varying knowledge of German.

The corpus Dolmetschen im Krankenhaus (DiK - Interpreting in Hospitals) was compiled between July 1999 and June 2005 in the research project "Interpreting in Hospitals" (K2, principal investigator: Kristin Bührig), part of the Collaborative Research Centre "Multilingualism", funded by the German Research Foundation (Deutsche Forschungsgemeinschaft, DFG) and hosted by the University of Hamburg. The corpus is based on doctor-patient-communications between German doctors or nursing staff and patients with Turkish or Portuguese as their mother tongue. These conversations were translated by laypersons (nursing staff, relatives). Furthermore, the corpus comprises monolingual doctor-patient-communications from Germany, Portugal and Turkey as comparative data.

Bührig, Kristin; Kliche, Ortrun; Meyer, Bernd and Pawlack, Birte. 2012. "The Corpus 'Interpreting in Hospitals'. Possible Applications for Research and Communication Training." In Multilingual Corpora and Multilingual Corpus Analysis, edited by Thomas Schmidt and Kai Wörner, pp. 305–15. Hamburg Studies in Multilingualism. John Benjamins.

 

CLARIN Metadata summary for Dolmetschen im Krankenhaus (DiK) (CMDI-based)

Title: Dolmetschen im Krankenhaus (DiK)
Description: Audio recordings of various kinds of doctor-patient communication in hospitals. There are both monolingual conversations in German, Portuguese and Turkish, recorded in the respective country, and interpreted conversations recorded in Germany (i.e. in German-Turkish, German-Portuguese, and German-Portuguese/Spanish), about 15-20 recordings of each kind. The persons interpreting are bilingual hospital employees or relatives of the patients, who are all adults living in Germany but with varying knowledge of German.
Publication date: 2009-01-05
Data owner:  Kristin Bührig, kristin.buehrig@uni-hamburg.de, Bernd Meyer, meyerb@uni-mainz.de
Contributors:  Kristin Bührig, Institut für Germanistik I / Von-Melle-Park 6 / D-20146 Hamburg, kristin.buehrig@uni-hamburg.de (compiler), Bernd Meyer, Arbeitsbereich Interkulturelle Kommunikation / Fachbereich 06: Translations-, Sprach- und Kulturwissenschaft / Johannes Gutenberg-Universität Mainz / An der Hochschule 2 / D-76726 Germersheim, meyerb@uni-mainz.de (compiler)
Project:  K2 "Interpreting in Hospitals", German Research Foundation (DFG)
Keywords:  community interpreting, consecutive interpreting, interpreted communication, doctor-patient communication, communication in institutions, EXMARaLDA
Languages:  German (deu), Portuguese (por), Spanish (spa), Turkish (tur)
Size:  187 speakers (98 female, 89 male), 91 communications, 92 transcriptions, 170925 words
Annotation types:  transcription (manual): HIAT, k: free comment, sup: suprasegmental information, akz: accentuation/stress, de: German translation, en: English translation, mt: morphological transliteration
Temporal Coverage:  1997-07-11/2005-01-12
Spatial Coverage:  Hamburg, DE; Viana do Castelo, PT; Vila do Conde, PT; Viana do castelo, PT; Ankara, TR
Genre:  discourse
Modality:  spoken
References:  Bührig, Kristin; Kliche, Ortrun; Meyer, Bernd; Pawlack, Birte (2012) The corpus "Interpreting in Hospitals": Possible applications for research and communication training. In: Schmidt, Thomas and Wörner, Kai (eds.): Multilingual Corpora and Multilingual Corpus Analysis. Hamburg Studies in Multilingualism (14), 305-315. Amsterdam: John Benjamins.

 </dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/1440</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.1440</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:1440</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.1439</dc:relation>
          <dc:rights>info:eu-repo/semantics/restrictedAccess</dc:rights>
          <dc:subject>community interpreting</dc:subject>
          <dc:subject>consecutive interpreting</dc:subject>
          <dc:subject>interpreted communication</dc:subject>
          <dc:subject>doctor-patient communication</dc:subject>
          <dc:subject>communication in institutions</dc:subject>
          <dc:subject>EXMARaLDA</dc:subject>
          <dc:subject>linguistics</dc:subject>
          <dc:subject>German</dc:subject>
          <dc:subject>Portuguese</dc:subject>
          <dc:subject>Spanish</dc:subject>
          <dc:subject>Turkish</dc:subject>
          <dc:title>Dolmetschen im Krankenhaus (DiK)</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:1458</identifier>
        <datestamp>2020-12-03T10:43:02Z</datestamp>
        <setSpec>user-hzsk</setSpec>
        <setSpec>user-uhh</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:creator>Meisel, Jürgen M.</dc:creator>
          <dc:date>2020-08-25</dc:date>
          <dc:description>Video recordings of French and German children that start acquiring German or French as an L2 with varying AOAs, mainly of approx. 3 years. Some families use additional languages, mainly English. The child is addressed in the L2 in the recording sessions, which are of the type interviewer/child interaction. On an average, for each child there is data from two occasions each year during two years. Only transcripts and metadata are available. Orthographic transcription according to project internal conventions.</dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/1458</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.1458</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:1458</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.1457</dc:relation>
          <dc:rights>info:eu-repo/semantics/restrictedAccess</dc:rights>
          <dc:subject>German</dc:subject>
          <dc:subject>French</dc:subject>
          <dc:subject>linguistics</dc:subject>
          <dc:subject>language acquisition</dc:subject>
          <dc:subject>bilingualism</dc:subject>
          <dc:subject>child language acquisition</dc:subject>
          <dc:subject>child L2 acquisition</dc:subject>
          <dc:subject>child bilingualism</dc:subject>
          <dc:subject>longitudinal data</dc:subject>
          <dc:subject>successive bilingualism</dc:subject>
          <dc:subject>l2 data</dc:subject>
          <dc:subject>EXMARaLDA</dc:subject>
          <dc:title>CHILD-L2</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:1535</identifier>
        <datestamp>2020-12-03T22:00:26Z</datestamp>
        <setSpec>user-hzsk</setSpec>
        <setSpec>user-uhh</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:creator>Meisel, Jürgen M.</dc:creator>
          <dc:date>2020-09-07</dc:date>
          <dc:description>Video recordings and transcriptions of three German/Portuguese simultaneous bilingual children, starting at approx. 1 year and 6 months. One or two recordings each month until approx. 5 years and 6 months. In each recording session (interviewer/child interaction) the child is addressed in both languages in one Portuguese and one German part. For privacy protection reasons, only the transcripts can be made available. Video recordings cannot be published.</dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/1535</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.1535</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:1535</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.1534</dc:relation>
          <dc:rights>info:eu-repo/semantics/restrictedAccess</dc:rights>
          <dc:subject>German</dc:subject>
          <dc:subject>Portuguese</dc:subject>
          <dc:subject>European Portuguese</dc:subject>
          <dc:subject>Brazilian Portuguese</dc:subject>
          <dc:subject>Linguistics</dc:subject>
          <dc:subject>Language Acquisition</dc:subject>
          <dc:subject>Bilingualism</dc:subject>
          <dc:subject>simultaneous bilingualism</dc:subject>
          <dc:subject>child language acquisition</dc:subject>
          <dc:subject>child bilingualism</dc:subject>
          <dc:subject>longitudinal data</dc:subject>
          <dc:subject>L1 data</dc:subject>
          <dc:subject>L2 data</dc:subject>
          <dc:subject>EXMARaLDA</dc:subject>
          <dc:title>BIPODE</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:1539</identifier>
        <datestamp>2020-12-03T10:43:01Z</datestamp>
        <setSpec>user-hzsk</setSpec>
        <setSpec>user-uhh</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:creator>Meisel, Jürgen M.</dc:creator>
          <dc:date>2020-09-07</dc:date>
          <dc:description>Subcorpus 'Iberian Portuguese' (one child) from the BIPODE project.

Video recordings of one German/Portuguese simultaneous bilingual child, starting at approx. 1 year and 6 months. One or two recordings each month until approx. 5 years and 6 months. In each recording session (interviewer/child interaction) the child is addressed in both languages in one Portuguese and one German part. For privacy protection reasons, only the transcripts can be made available. Video recordings cannot be published.</dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/1539</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.1539</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:1539</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.1075</dc:relation>
          <dc:relation>doi:10.25592/uhhfdm.1535</dc:relation>
          <dc:relation>doi:10.25592/uhhfdm.1538</dc:relation>
          <dc:rights>info:eu-repo/semantics/restrictedAccess</dc:rights>
          <dc:subject>German</dc:subject>
          <dc:subject>Portuguese</dc:subject>
          <dc:subject>European Portuguese</dc:subject>
          <dc:subject>Linguistics</dc:subject>
          <dc:subject>Language Acquisition</dc:subject>
          <dc:subject>Bilingualism</dc:subject>
          <dc:subject>simultaneous bilingualism</dc:subject>
          <dc:subject>child language acquisition</dc:subject>
          <dc:subject>child bilingualism</dc:subject>
          <dc:subject>longitudinal data</dc:subject>
          <dc:subject>L1 data</dc:subject>
          <dc:subject>L2 data</dc:subject>
          <dc:subject>EXMARaLDA</dc:subject>
          <dc:title>BIPODE_PT</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:1541</identifier>
        <datestamp>2021-05-27T11:10:23Z</datestamp>
        <setSpec>user-hzsk</setSpec>
        <setSpec>user-uhh</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:creator>Lleó, Conxita</dc:creator>
          <dc:date>2020-09-07</dc:date>
          <dc:description>Audio recordings of three German/Spanish simultaneous bilingual children starting at approx. 1 year and ending between the ages 2;4 and 3 years. There are 125 recording sessions (interviewer/child interaction), half of them conducted in a German and half in a Spanish speaking environment.

Phonologie-Erwerb Deutsch-Spanisch als Erste Sprachen (PEDSES) is a phonetically and orthographically transcribed corpus of German and Spanish, finalized in the project Prosodic constraints on phonological and morphological development in bilingual first language acquisition at the Research Center on Multilingualism, University of Hamburg.

Lleó, Conxita. 2012. "Monolingual and Bilingual Phonoprosodic Corpora of Child German and Child Spanish." In Multilingual Corpora and Multilingual Corpus Analysis, edited by Thomas Schmidt and Kai Wörner, 14:pp. 107–22. Hamburg Studies in Multilingualism. John Benjamins.

 

CLARIN Metadata summary for Phonologie-Erwerb Deutsch-Spanisch als Erste Sprachen (PEDSES) (CMDI-based)

Title: Phonologie-Erwerb Deutsch-Spanisch als Erste Sprachen (PEDSES)
Description: Audio recordings of three German/Spanish simultaneous bilingual children starting at approx. 1 year and ending between the ages 2;4 and 3 years. There are 125 recording sessions (interviewer/child interaction), half of them conducted in a German and half in a Spanish speaking environment.
Data owner:  Conxita Lleó, Institut für Romanistik / Von-Melle-Park 6 / D-20146 Hamburg, lleo@uni-hamburg.de
Contributors:  Conxita Lleó, Institut für Romanistik / Von-Melle-Park 6 / D-20146 Hamburg, lleo@uni-hamburg.de (compiler)
Project:  E3 "Prosodic Constraints on Phonological and Morphological development in Bilingual First Language Acquisition", German Research Foundation (DFG)
Keywords:  child language acquisition, child bilingualism, longitudinal data, simultaneous bilingualism, L2 data, L1 data, L1-Daten, L2-Daten, EXMARaLDA
Languages:  German (deu), Spanish (spa)
Size:  21 speakers (15 female, 6 male), 127 communications, 125 recordings, 3897 minutes, 127 transcriptions, 101292 words
Genre:  discourse
Modality:  spoken
References:  Lleó, Conxita. 2012. "Monolingual and Bilingual Phonoprosodic Corpora of Child German and Child Spanish." In Multilingual Corpora and Multilingual Corpus Analysis, edited by Thomas Schmidt and Kai Wörner, 14:pp. 107–22. Hamburg Studies in Multilingualism. John Benjamins.

 </dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/1541</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.1541</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:1541</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.1540</dc:relation>
          <dc:rights>info:eu-repo/semantics/restrictedAccess</dc:rights>
          <dc:subject>child language acquisition</dc:subject>
          <dc:subject>child bilingualism</dc:subject>
          <dc:subject>longitudinal data</dc:subject>
          <dc:subject>simultaneous bilingualism</dc:subject>
          <dc:subject>L2 data</dc:subject>
          <dc:subject>L1 data</dc:subject>
          <dc:subject>L1-Daten</dc:subject>
          <dc:subject>L2-Daten</dc:subject>
          <dc:subject>EXMARaLDA</dc:subject>
          <dc:subject>linguistics</dc:subject>
          <dc:subject>German</dc:subject>
          <dc:subject>Spanish</dc:subject>
          <dc:title>Phonologie-Erwerb Deutsch-Spanisch als Erste Sprachen (PEDSES)</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:1555</identifier>
        <datestamp>2020-09-07T19:00:06Z</datestamp>
        <setSpec>user-hzsk</setSpec>
        <setSpec>user-uhh</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:creator>Lleó, Conxita</dc:creator>
          <dc:date>2020-09-07</dc:date>
          <dc:description>Between 2003 and 2006 the project E3 collected the corpus Madrid-PhonBLA (Bilingual Language Acquisition), which contains cross-sectional data from 71 Spanish-German bilingual children, aged 2 to 6, growing up in Madrid (Spain), and the PhonMAS (Monolingual Acquisition of Spanish) corpus, which contains cross-sectional data from 45 monolingual Spanish children, aged 2 to 9. The Madrid-PhonBLA corpus contains 46 hours of bilingual (Spanish and German) spontaneous speech, and the PhonMAS contains 45 hours of monolingual Spanish speech. The recordings were made in Madrid, in various preschools. The purpose of such data was to use them as a larger basis for comparison with the bilingual corpus on sound phenomena a total of 45 monolingual children (between 2 and 6 years of age). Four adults were recorded by the E3 project in Madrid, too. The audio files of PhonMAS contain 17 hours of spontaneous speech. Two of the adults' corpora in these files have been orthographically transcribed and transferred to EXMARaLDA. From the 168 available audio files (from 116 bilinguals and 52 monolinguals), 37 are provided with orthographic and phonetic transcriptions, and 34 files have been transferred to EXMARaLDA. For the phonetic transcription of those 34 files the International Phonetic Alphabet (IPA) was used. The criteria used for transcribing these files were the same as those used for the transcription of the Hamburg-PhonBLA corpus.

Lleó, Conxita. 2012. "Monolingual and Bilingual Phonoprosodic Corpora of Child German and Child Spanish." In Multilingual Corpora and Multilingual Corpus Analysis, edited by Thomas Schmidt and Kai Wörner, 14:pp. 107–22. Hamburg Studies in Multilingualism. John Benjamins.</dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/1555</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.1555</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:1555</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.1554</dc:relation>
          <dc:rights>info:eu-repo/semantics/restrictedAccess</dc:rights>
          <dc:subject>German</dc:subject>
          <dc:subject>Spanish</dc:subject>
          <dc:subject>Linguistics</dc:subject>
          <dc:subject>Language Acquisition</dc:subject>
          <dc:subject>Bilingualism</dc:subject>
          <dc:title>PhonBLA Querschnittsstudie Madrid</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:1591</identifier>
        <datestamp>2020-09-14T19:57:09Z</datestamp>
        <setSpec>user-hzsk</setSpec>
        <setSpec>user-uhh</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:creator>Fiedler, Ines</dc:creator>
          <dc:date>2020-09-14</dc:date>
          <dc:description>The data sets for each language consist of a small number of mini-dialogues, chosen out of the 189 entries within the Focus Translation Task (cf. Skopeteas et al. 2006: 209ff.) in order to get a basic set of utterances for comparison between the languages dealt with in the project.

Fiedler, Ines (2011) QUIS Data from Yom, Aja, Anii and Foodo. With Notes on Genetic and Areal Relations. In: Petrova &amp; M. Grubic (eds.): Interdisciplinary Studies on Information Structure 16, Linguistic Fieldnotes III: Information structure in Gur and Kwa language. 49-96.

Schwarz, Anne (2011) QUIS Data from Buli, Kɔnni and Baatɔnum. With Notes on the Comparative Approach. In: Petrova &amp; M. Grubic (eds.): Interdisciplinary Studies on Information Structure 16, Linguistic Fieldnotes III: Information structure in Gur and Kwa language. 1-48.

 

CLARIN Metadata summary for B1 Aja (CMDI-based)

Title: B1 Aja
Description: The data sets for each language consist of a small number of mini-dialogues, chosen out of the 189 entries within the Focus Translation Task (cf. Skopeteas et al. 2006: 209ff.) in order to get a basic set of utterances for comparison between the languages dealt with in the project.
Publication date: 2015
Data owner:  Dr. phil. Ines Fiedler
Contributors:  Manfred Krifka (editor), Katharina Hartmann (editor), Brigitte Reineke (editor)
Project:  Special Research Centre 632 Information structure, German Research Foundation
Keywords:  focus, Fokus
Language:  Aja (aja)
Size:  257 Token
Segmentation units:  other
Temporal Coverage:  2005-02-09/2005-02-12
Spatial Coverage:  Hlassame, BJ
Genre:  discourse
Modality:  spoken
References:  Fiedler, Ines (2011) QUIS Data from Yom, Aja, Anii and Foodo. With Notes on Genetic and Areal Relations. In: Petrova &amp; M. Grubic (eds.): Interdisciplinary Studies on Information Structure 16, Linguistic Fieldnotes III: Information structure in Gur and Kwa language. 49-96. References:  Schwarz, Anne (2011) QUIS Data from Buli, Kɔnni and Baatɔnum. With Notes on the Comparative Approach. In: Petrova &amp; M. Grubic (eds.): Interdisciplinary Studies on Information Structure 16, Linguistic Fieldnotes III: Information structure in Gur and Kwa language. 1-48.

 </dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/1591</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.1591</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:1591</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.1590</dc:relation>
          <dc:rights>info:eu-repo/semantics/openAccess</dc:rights>
          <dc:rights>https://creativecommons.org/licenses/by/3.0/legalcode</dc:rights>
          <dc:subject>focus</dc:subject>
          <dc:subject>Fokus</dc:subject>
          <dc:subject>linguistics</dc:subject>
          <dc:subject>information structure</dc:subject>
          <dc:subject>Aja</dc:subject>
          <dc:title>B1 Aja</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:1593</identifier>
        <datestamp>2020-09-14T20:01:40Z</datestamp>
        <setSpec>user-hzsk</setSpec>
        <setSpec>user-uhh</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:creator>Fiedler, Ines</dc:creator>
          <dc:date>2020-09-14</dc:date>
          <dc:description>The data sets for each language consist of a small number of mini-dialogues, chosen out of the 189 entries within the Focus Translation Task (cf. Skopeteas et al. 2006: 209ff.) in order to get a basic set of utterances for comparison between the languages dealt with in the project.

Fiedler, Ines (2011) QUIS Data from Yom, Aja, Anii and Foodo. With Notes on Genetic and Areal Relations. In: Petrova &amp; M. Grubic (eds.): Interdisciplinary Studies on Information Structure 16, Linguistic Fieldnotes III: Information structure in Gur and Kwa language. 49-96.

Schwarz, Anne (2011) QUIS Data from Buli, Kɔnni and Baatɔnum. With Notes on the Comparative Approach. In: Petrova &amp; M. Grubic (eds.): Interdisciplinary Studies on Information Structure 16, Linguistic Fieldnotes III: Information structure in Gur and Kwa language. 1-48.

 

CLARIN Metadata summary for B1 Fon (CMDI-based)

Title: B1 Fon
Description: The data sets for each language consist of a small number of mini-dialogues, chosen out of the 189 entries within the Focus Translation Task (cf. Skopeteas et al. 2006: 209ff.) in order to get a basic set of utterances for comparison between the languages dealt with in the project.
Publication date: 2015
Data owner:  Dr. phil. Ines Fiedler
Contributors:  Ines Fiedler (editor)
Project:  Special Research Centre 632 Information structure, German Research Foundation
Keywords:  focus, Fokus
Language:  Fon (fon)
Size:  264 Token
Segmentation units:  other
Temporal Coverage:  2005-02-05/2005-02-06
Spatial Coverage:  Cotonou, BJ
Genre:  discourse
Modality:  spoken
References:  Fiedler, Ines (2011) QUIS Data from Yom, Aja, Anii and Foodo. With Notes on Genetic and Areal Relations. In: Petrova &amp; M. Grubic (eds.): Interdisciplinary Studies on Information Structure 16, Linguistic Fieldnotes III: Information structure in Gur and Kwa language. 49-96. References:  Schwarz, Anne (2011) QUIS Data from Buli, Kɔnni and Baatɔnum. With Notes on the Comparative Approach. In: Petrova &amp; M. Grubic (eds.): Interdisciplinary Studies on Information Structure 16, Linguistic Fieldnotes III: Information structure in Gur and Kwa language. 1-48.

 </dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/1593</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.1593</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:1593</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.1592</dc:relation>
          <dc:rights>info:eu-repo/semantics/openAccess</dc:rights>
          <dc:rights>https://creativecommons.org/licenses/by/3.0/legalcode</dc:rights>
          <dc:subject>focus</dc:subject>
          <dc:subject>Fokus</dc:subject>
          <dc:subject>linguistics</dc:subject>
          <dc:subject>information structure</dc:subject>
          <dc:subject>Fon</dc:subject>
          <dc:title>B1 Fon</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:1701</identifier>
        <datestamp>2020-09-29T18:46:49Z</datestamp>
        <setSpec>user-hzsk</setSpec>
        <setSpec>user-uhh</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:contributor>Hartmann, Katharina</dc:contributor>
          <dc:contributor>Jacob, Peggy</dc:contributor>
          <dc:creator>Hartmann, Katharina</dc:creator>
          <dc:date>2015-09-29</dc:date>
          <dc:description>Full set: all focus related experiments, status: work in progress, large parts elicited, most of the data transcribed, partly annotatedCLARIN Metadata summary for B2 Bura (CMDI-based)    	    	    	    		Title: B2 Bura    	        	Description: Full set: all focus related experiments, status: work in progress, large parts elicited, most of the data transcribed, partly annotated					Publication date: 2015					Data owner: 			Univ.-Prof. Dr. Katharina Hartmann											                		Contributors:                 	Katharina Hartmann (editor), Peggy Jacob (researcher)                			                		            	Project:             	Special Research Centre 632 Information structure, German Research Foundation            				            			            	Keywords:             	focus, Fokus            				            			    	Language:     	Bura (bwr)    						    					            	Size:             	818 Token            					            				                		Segmentation units:                 	other                			                		            	Genre:             	discourse            				            			            	Modality:             	spoken            				            					</dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/1701</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.1701</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:1701</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.1700</dc:relation>
          <dc:rights>info:eu-repo/semantics/openAccess</dc:rights>
          <dc:rights>https://creativecommons.org/licenses/by-nc/3.0/legalcode</dc:rights>
          <dc:subject>focus</dc:subject>
          <dc:subject>Fokus</dc:subject>
          <dc:subject>linguistics</dc:subject>
          <dc:subject>information structure</dc:subject>
          <dc:subject>Bura</dc:subject>
          <dc:title>B2 Bura</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:1707</identifier>
        <datestamp>2020-09-29T19:05:33Z</datestamp>
        <setSpec>user-hzsk</setSpec>
        <setSpec>user-uhh</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:contributor>Hartmann, Katharina</dc:contributor>
          <dc:contributor>Jacob, Peggy</dc:contributor>
          <dc:creator>Hartmann, Katharina</dc:creator>
          <dc:date>2015-09-29</dc:date>
          <dc:description>Full set: all focus related experiments, status: work in progress, large parts elicited, most of the data transcribed, partly annotated.CLARIN Metadata summary for B2 Marghi (CMDI-based)    	    	    	    		Title: B2 Marghi    	        	Description: Full set: all focus related experiments, status: work in progress, large parts elicited, most of the data transcribed, partly annotated.					Publication date: 2015					Data owner: 			Univ.-Prof. Dr. Katharina Hartmann											                		Contributors:                 	Katharina Hartmann (editor), Peggy Jacob (researcher)                			                		            	Project:             	Special Research Centre 632 Information structure, German Research Foundation            				            			            	Keywords:             	focus, Fokus            				            			    	Language:     	Marghi (mrt)    						    					            	Size:             	831 Token            					            				                		Segmentation units:                 	other                			                		            	Genre:             	discourse            				            			            	Modality:             	spoken            				            					</dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/1707</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.1707</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:1707</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.1706</dc:relation>
          <dc:rights>info:eu-repo/semantics/openAccess</dc:rights>
          <dc:rights>https://creativecommons.org/licenses/by-nc/3.0/legalcode</dc:rights>
          <dc:subject>focus</dc:subject>
          <dc:subject>Fokus</dc:subject>
          <dc:subject>linguistics</dc:subject>
          <dc:subject>information structure</dc:subject>
          <dc:subject>Marghi</dc:subject>
          <dc:title>B2 Marghi</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:1730</identifier>
        <datestamp>2020-09-29T21:10:58Z</datestamp>
        <setSpec>user-hzsk</setSpec>
        <setSpec>user-uhh</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:contributor>House, Juliane</dc:contributor>
          <dc:creator>House, Juliane</dc:creator>
          <dc:date>2011-06-30</dc:date>
          <dc:description>Translation corpora of original texts with translations and comparable texts from the genre popular scientific prose.Übersetzungs- und Vergleichskorpus mit authentischen populärwissenschaftlichen Texten. Texte aus deutschsprachigen und englischsprachigen populärwissenschaftlichen Zeitschriften jeweils im Original und der jeweiligen Übersetzung. Zeitfenster 1978-1982 (E-&gt;D), 1978-1982 (D), 1999-2002 (E-&gt;D), 1999-2002 (D)CLARIN Metadata summary for Covert translation: popular science (CMDI-based)    	    	    	    		Title: Covert translation: popular science    	        	Description:           Translation corpora of original texts with translations and comparable          texts from the genre popular scientific prose.        		        	Description:           Übersetzungs- und Vergleichskorpus mit authentischen          populärwissenschaftlichen Texten.          Texte aus deutschsprachigen und englischsprachigen          populärwissenschaftlichen Zeitschriften jeweils im Original und der          jeweiligen Übersetzung. Zeitfenster 1978-1982 (E-&gt;D), 1978-1982          (D), 1999-2002 (E-&gt;D), 1999-2002 (D)        					Publication date: 2011-06-30					Data owner: 			Juliane House											                		Contributors:                 	Juliane House (compiler)                			                		            	Project:             	A4 "Covert Translation" A4 "Verdecktes Übersetzen - Covert Translation", German Research Foundation (DFG), K4 "Covert Translation" K4 "Verdecktes Übersetzen - Covert Translation", German Research Foundation (DFG)            				            			            	Keywords:             	translated texts, popular science texts, parallel corpus, comparable corpus, übersetzte Texte, Populärwissenschaftstexte, Parallelkorpus, Vergleichskorpus            				            			    	Languages:     	German (deu), English (eng)    						    					            	Size:             	114 texts, 500466 words            					            				                		Segmentation units:                 	text unit, orthographic sentence, lexeme                			                		                		Annotation types:                 	annotation (automatic), pos.susanne: This annotation scheme reflects the markup employed                in the original corpus files and corresponds to the Susanne                Tagset., pos.stts: This annotation scheme reflects the markup employed                in the original corpus files and corresponds to the                Stuttgart-Tübingen Tagset (STTS).                			                		            	Temporal Coverage:             	1978/2002            	                        			Spatial Coverage:             		US; DE            				            			            	Genre:             	popular science texts, populärwissenschaftliche Texte            				            			            	Modality:             	written            				            					</dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/1730</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.1730</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:1730</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.1724</dc:relation>
          <dc:rights>info:eu-repo/semantics/restrictedAccess</dc:rights>
          <dc:subject>translated texts</dc:subject>
          <dc:subject>popular science texts</dc:subject>
          <dc:subject>parallel corpus</dc:subject>
          <dc:subject>comparable corpus</dc:subject>
          <dc:subject>übersetzte Texte</dc:subject>
          <dc:subject>Populärwissenschaftstexte</dc:subject>
          <dc:subject>Parallelkorpus</dc:subject>
          <dc:subject>Vergleichskorpus</dc:subject>
          <dc:subject>linguistics</dc:subject>
          <dc:subject>German</dc:subject>
          <dc:subject>English</dc:subject>
          <dc:title>Covert translation: popular science</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:1713</identifier>
        <datestamp>2020-09-29T19:04:37Z</datestamp>
        <setSpec>user-hzsk</setSpec>
        <setSpec>user-uhh</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:contributor>Petrova, Svetlana</dc:contributor>
          <dc:contributor>Battefeld, Malte</dc:contributor>
          <dc:creator>Petrova, Svetlana</dc:creator>
          <dc:date>2015-09-29</dc:date>
          <dc:description>The texts of this corpus, Ludolf von Sudheims Reise ins Heilige Land (Ludolf of Sudheim's Journey to the Holy Land), is a journey diary describing the adventures of a group of pilgrims, written in Middle Low German and dated back to 1350. For information on the properties of the text, including the manuscripts, see Blust-Thiele (1985). This corpus uses the text edition by Stapelmohr (1937). The first 20 pages of it are tagged for clause type and grammatical function. The corpus includes 6,690 tokens.CLARIN Metadata summary for B4 Ludolf (CMDI-based)    	    	    	    		Title: B4 Ludolf    	        	Description: The texts of this corpus, Ludolf von Sudheims Reise ins Heilige Land (Ludolf of Sudheim's Journey to the Holy Land), is a journey diary describing the adventures of a group of pilgrims, written in Middle Low German and dated back to 1350. For information on the properties of the text, including the manuscripts, see Blust-Thiele (1985). This corpus uses the text edition by Stapelmohr (1937). The first 20 pages of it are tagged for clause type and grammatical function. The corpus includes 6,690 tokens.					Publication date: 2015					Data owner: 			Prof. Dr. Svetlana Petrova											                		Contributors:                 	Svetlana Petrova (editor), Malte Battefeld (annotator)                			                		            	Project:             	Special Research Centre 632 Information structure, German Research Foundation            				            			            	Keywords:             	historical texts, journey diary, information structure            				            			    	Language:     	German Middle Low (gml)    						    					            	Size:             	6690 Token            					            				                		Segmentation units:                 	other                			                		            	Genre:             	historic manuscript            				            			            	Modality:             	written            				            					</dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/1713</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.1713</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:1713</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.1712</dc:relation>
          <dc:rights>info:eu-repo/semantics/openAccess</dc:rights>
          <dc:rights>https://creativecommons.org/licenses/by-nc/3.0/legalcode</dc:rights>
          <dc:subject>historical texts</dc:subject>
          <dc:subject>journey diary</dc:subject>
          <dc:subject>information structure</dc:subject>
          <dc:subject>linguistics</dc:subject>
          <dc:subject>German Middle Low</dc:subject>
          <dc:title>B4 Ludolf</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:1711</identifier>
        <datestamp>2020-09-29T19:05:07Z</datestamp>
        <setSpec>user-hzsk</setSpec>
        <setSpec>user-uhh</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:contributor>Rasskazova, Oxana</dc:contributor>
          <dc:contributor>Petrova, Svetlana</dc:contributor>
          <dc:contributor>Gehrlein, Anke</dc:contributor>
          <dc:contributor>Coniglio, Marco</dc:contributor>
          <dc:contributor>Donhauser, Karin</dc:contributor>
          <dc:contributor>Schlachter, Eva</dc:contributor>
          <dc:creator>Schlachter, Eva</dc:creator>
          <dc:date>2014-09-29</dc:date>
          <dc:description>HIPKON is the first corpus based on only one text type (sermons) and on one dialect area, Upper German (Bavarian-Alemannic). The sermons cover the time from Middle High German to the beginning of the New High German period. They were accurately selected so that each of them is representative of one century. Among others, syntax, information structure and discourse structure were annotated in the corpus.Das Korpus beinhaltet Predigttexte aus dem Oberdeutschen (bairisch-alemannisch). Jedes Teilkorpus hat einen Umfang von etwa 12.000 bis max. 17.000 Wortformen, der zeitliche Abstand zwischen den Texten beträgt etwa 100 Jahre, so dass jedes Jahrhundert durch ein Teilkorpus repräsentiert wird. Lediglich für das 15. Jahrhundert liegt kein Teilkorpus vor, da aus dieser Epoche generell sehr wenige Predigten überliefert sind. Das Korpus ist ein multi-layer Korpus, das nur in den für die Forschungfrage relevanten Belegen (Hauptsätze mit komplexem Verbgefüge mit Nachfeldbesetzung) annotiert wurde.CLARIN Metadata summary for B4 Historisches Predigtenkorpus zum Nachfeld (CMDI-based)    	    	    	    		Title: B4 Historisches Predigtenkorpus zum Nachfeld    	        	Description: HIPKON is the first corpus based on only one text type (sermons) and on one dialect area, Upper German (Bavarian-Alemannic). The sermons cover the time from Middle High German to the beginning of the New High German period. They were accurately selected so that each of them is representative of one century. Among others, syntax, information structure and discourse structure were annotated in the corpus.		        	Description: Das Korpus beinhaltet Predigttexte aus dem Oberdeutschen (bairisch-alemannisch). Jedes Teilkorpus hat einen Umfang von etwa 12.000 bis max. 17.000 Wortformen, der zeitliche Abstand zwischen den Texten beträgt etwa 100 Jahre, so dass jedes Jahrhundert durch ein Teilkorpus repräsentiert wird. Lediglich für das 15. Jahrhundert liegt kein Teilkorpus vor, da aus dieser Epoche generell sehr wenige Predigten überliefert sind. Das Korpus ist ein multi-layer Korpus, das nur in den für die Forschungfrage relevanten Belegen (Hauptsätze mit komplexem Verbgefüge mit Nachfeldbesetzung) annotiert wurde.					Publication date: 2014					Data owner: 			Dr. Eva Schlachter											                		Contributors:                 	Svetlana Petrova (editor), Karin Donhauser (editor), Eva Schlachter (annotator), Marco Coniglio (annotator), Oxana Rasskazova (annotator), Anke Gehrlein (annotator)                			                		            	Project:             	Special Research Centre 632 Information structure, German Research Foundation            				            			            	Keywords:             	historical texts, religious texts, information structure            				            			    	Language:     	New High German (deu)    						    					            	Size:             	92500 Token            					            				                		Segmentation units:                 	other                			                		            	Genre:             	historic manuscript            				            			            	Modality:             	written            				            					</dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/1711</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.1711</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:1711</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.1710</dc:relation>
          <dc:rights>info:eu-repo/semantics/openAccess</dc:rights>
          <dc:rights>https://creativecommons.org/licenses/by-nc/3.0/legalcode</dc:rights>
          <dc:subject>historical texts</dc:subject>
          <dc:subject>religious texts</dc:subject>
          <dc:subject>information structure</dc:subject>
          <dc:subject>linguistics</dc:subject>
          <dc:subject>New High German</dc:subject>
          <dc:title>B4 Historisches Predigtenkorpus zum Nachfeld</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:1719</identifier>
        <datestamp>2020-09-29T20:40:12Z</datestamp>
        <setSpec>user-hzsk</setSpec>
        <setSpec>user-uhh</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:contributor>Petrova, Svetlana</dc:contributor>
          <dc:creator>Petrova, Svetlana</dc:creator>
          <dc:date>2015-09-29</dc:date>
          <dc:description>Das Referenzkorpus Altdeutsch erfasst und annotiert die ältesten Sprachdenkmäler des Deutschen vom Beginn der kontinuierlichen schriftlichen Überlieferung um 750 bis etwa 1050 mit einem Umfang von ca. 650 000 Textwörtern. Aufgenommen werden alle in dieser Zeit überlieferten Texte des Althochdeutschen und des Altsächsischen in einer möglichst genauen Wiedergabestufe. Dabei werden die handschriftengetreuesten gedruckten Texteditionen zugrundegelegt. Die Annotation erfasst Header-Informationen, strukturelle (Wort, Satz, Zeile, Absatz etc.) und linguistische Annotationen (Part of Speech-Tagging, Flexionsmorphologie) sowie syntaktische Satzinformationen und erfolgt mit Unterstützung einer semi-automatischen Vorannotation, die mit Hilfe der digitalisierten Sprachstufen- und Textwörterbücher und Glossare zum Althochdeutschen und zum Altsächsischen erzeugt wurde. Die verschiedenen Stufen der Annotation werden in Form einer Mehrebenenarchitektur aufeinander bezogen.The reference corpus Old German contains (annotated) data from the oldest language monuments of German before the continuous written transduction around 750 until 1050 with approx. 650,000 text words.CLARIN Metadata summary for B4 Otfrid (CMDI-based)    	    	    	    		Title: B4 Otfrid    	        	Description: Das Referenzkorpus Altdeutsch erfasst und annotiert die ältesten Sprachdenkmäler des Deutschen vom Beginn der kontinuierlichen schriftlichen Überlieferung um 750 bis etwa 1050 mit einem Umfang von ca. 650 000 Textwörtern. Aufgenommen werden alle in dieser Zeit überlieferten Texte des Althochdeutschen und des Altsächsischen in einer möglichst genauen Wiedergabestufe. Dabei werden die handschriftengetreuesten gedruckten Texteditionen zugrundegelegt. Die Annotation erfasst Header-Informationen, strukturelle (Wort, Satz, Zeile, Absatz etc.) und linguistische Annotationen (Part of Speech-Tagging, Flexionsmorphologie) sowie syntaktische Satzinformationen und erfolgt mit Unterstützung einer semi-automatischen Vorannotation, die mit Hilfe der digitalisierten Sprachstufen- und Textwörterbücher und Glossare zum Althochdeutschen und zum Altsächsischen erzeugt wurde. Die verschiedenen Stufen der Annotation werden in Form einer Mehrebenenarchitektur aufeinander bezogen.		        	Description: The reference corpus Old German contains (annotated) data from the oldest language monuments of German before the continuous written transduction around 750 until 1050 with approx. 650,000 text words.					Publication date: 2015					Data owner: 			Prof. Dr. Svetlana Petrova											                		Contributors:                 	Svetlana Petrova (editor)                			                		            	Project:             	Special Research Centre 632 Information structure, German Research Foundation            				            			            	Keywords:             	historical texts, religious texts, information structure            				            			    	Language:     	Old High German (goh)    						    					            	Size:             	300000 Token            					            				                		Segmentation units:                 	other                			                		            	Genre:             	historic manuscript            				            			            	Modality:             	written            				            					</dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/1719</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.1719</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:1719</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.1718</dc:relation>
          <dc:rights>info:eu-repo/semantics/openAccess</dc:rights>
          <dc:rights>https://creativecommons.org/licenses/by-nc/3.0/legalcode</dc:rights>
          <dc:subject>historical texts</dc:subject>
          <dc:subject>religious texts</dc:subject>
          <dc:subject>information structure</dc:subject>
          <dc:subject>linguistics</dc:subject>
          <dc:subject>Old High German</dc:subject>
          <dc:title>B4 Otfrid</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:1723</identifier>
        <datestamp>2020-09-29T20:40:38Z</datestamp>
        <setSpec>user-hzsk</setSpec>
        <setSpec>user-uhh</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:contributor>Petrova, Svetlana</dc:contributor>
          <dc:contributor>Chun, Yen</dc:contributor>
          <dc:contributor>Odebrecht, Carolin</dc:contributor>
          <dc:contributor>Battefeld, Malte</dc:contributor>
          <dc:contributor>Linde, Sonja</dc:contributor>
          <dc:contributor>Donhauser, Karin</dc:contributor>
          <dc:contributor>Solf, Michael</dc:contributor>
          <dc:contributor>Kullick, Axel</dc:contributor>
          <dc:contributor>Gehrlein, Anke</dc:contributor>
          <dc:creator>Petrova, Svetlana</dc:creator>
          <dc:date>2014-12-01</dc:date>
          <dc:description>The present corpus, the Tatian Corpus of Deviating Examples T-CODEX 2.1, provides morpho-syntactic and information structural annotation of parts of the Old High German translation attested in the MS St. Gallen Cod. 56, traditionally called the OHG Tatian, one of the largest prose texts from the classical OHG period. This corpus was designed and annotated by Project B4 of Collaborative Research Center on Information Structure at Humboldt University Berlin. The present corpus compiles ca. 2.000 deviating examples found in the text portions of the scribes α, β, γ and ε. Each clause structure represents an extra file annotated with the annotation tool EXMARaLDA and searchable via ANNIS, a general-purpose tool for the publication, visualisation and querying of linguistic data collections, developed by Project D1 of the Collaborative Research Center on Information Structure at Potsdam University.CLARIN Metadata summary for B4 Tatian Corpus of Deviating Examples 2.1 (CMDI-based)    	    	    	    		Title: B4 Tatian Corpus of Deviating Examples 2.1    	        	Description: The present corpus, the Tatian Corpus of Deviating Examples T-CODEX 2.1, provides morpho-syntactic and information structural annotation of parts of the Old High German translation attested in the MS St. Gallen Cod. 56, traditionally called the OHG Tatian, one of the largest prose texts from the classical OHG period. This corpus was designed and annotated by Project B4 of Collaborative Research Center on Information Structure at Humboldt University Berlin. The present corpus compiles ca. 2.000 deviating examples found in the text portions of the scribes α, β, γ and ε. Each clause structure represents an extra file annotated with the annotation tool EXMARaLDA and searchable via ANNIS, a general-purpose tool for the publication, visualisation and querying of linguistic data collections, developed by Project D1 of the Collaborative Research Center on Information Structure at Potsdam University.					Publication date: 2014-12-01					Data owner: 			Prof. Dr. Svetlana Petrova											                		Contributors:                 	Svetlana Petrova (editor), Karin Donhauser (editor), Carolin Odebrecht (editor), Svetlana Petrova (annotator), Carolin Odebrecht (annotator), Michael Solf (annotator), Yen Chun Chen (annotator), Axel Kullick (annotator), Malte Battefeld (annotator), Sonja Linde (annotator), Anke Gehrlein (annotator)                			                		            	Project:             	Special Research Centre 632 Information structure, German Research Foundation            				            			            	Keywords:             	historical texts, religious texts, information structure            				            			    	Languages:     	Latin (lat), Old High German (goh)    						    					            	Size:             	11295 Token            					            				                		Segmentation units:                 	other                			                		                		Annotation types:                 	aboutness (manual), tok (manual), LAT (manual), align (manual), pos (manual), cat (manual), clause-status (manual), gf (manual), syl_no (manual), givenness (manual), top-comm (manual), position (manual), topic-marker (manual), definiteness (manual), foc-bg (manual), foc-marker (manual), context (manual), comment (manual), bibl (manual), meta::writer (manual), meta::corpus-code (manual), meta::page (manual), X::abbreviation (manual), X::sex (manual)                			                		            	Temporal Coverage:             	830-01-01/830-12-31            	                        			Spatial Coverage:             		Fulda, DE            				            			            	Genre:             	religious text            				            			            	Modality:             	written            				            					</dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/1723</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.1723</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:1723</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.1722</dc:relation>
          <dc:rights>info:eu-repo/semantics/openAccess</dc:rights>
          <dc:rights>https://creativecommons.org/licenses/by-nc/3.0/legalcode</dc:rights>
          <dc:subject>historical texts</dc:subject>
          <dc:subject>religious texts</dc:subject>
          <dc:subject>information structure</dc:subject>
          <dc:subject>linguistics</dc:subject>
          <dc:subject>Latin</dc:subject>
          <dc:subject>Old High German</dc:subject>
          <dc:title>B4 Tatian Corpus of Deviating Examples 2.1</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:8236</identifier>
        <datestamp>2020-12-09T13:18:26Z</datestamp>
        <setSpec>user-uhh</setSpec>
        <setSpec>user-hzsk</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:contributor>Neumann, Lara</dc:contributor>
          <dc:contributor>Heinicke, Vivien</dc:contributor>
          <dc:contributor>Wagner, Jonas</dc:contributor>
          <dc:contributor>VolkswagenStiftung</dc:contributor>
          <dc:contributor>Dilger, Yael</dc:contributor>
          <dc:contributor>Schulze, Barbara</dc:contributor>
          <dc:contributor>Bochmann, Carina</dc:contributor>
          <dc:contributor>Thielmann, Winfried</dc:contributor>
          <dc:contributor>Uhlenberg, Marlene</dc:contributor>
          <dc:contributor>DiMaio, Claudia</dc:contributor>
          <dc:contributor>Hawranke, Maike</dc:contributor>
          <dc:contributor>Krause, Arne</dc:contributor>
          <dc:contributor>Loch, Moritz</dc:contributor>
          <dc:contributor>Carobbio, Gabriella</dc:contributor>
          <dc:contributor>Wittwer, Laura</dc:contributor>
          <dc:contributor>Makarenko, Vera</dc:contributor>
          <dc:contributor>Breitsprecher, Christoph</dc:contributor>
          <dc:contributor>Redder, Angelika</dc:contributor>
          <dc:contributor>Gerstner, Julia</dc:contributor>
          <dc:contributor>Brücke, Linda</dc:contributor>
          <dc:contributor>Storz, Coretta</dc:contributor>
          <dc:contributor>Dörschner, Marieke</dc:contributor>
          <dc:contributor>Süß, Franziska</dc:contributor>
          <dc:contributor>Weidenhöffer, Jessica</dc:contributor>
          <dc:contributor>Alghisi, Alessandra</dc:contributor>
          <dc:contributor>Bolpagni, Elisa</dc:contributor>
          <dc:contributor>Heller, Dorothee</dc:contributor>
          <dc:creator>Redder, Angelika</dc:creator>
          <dc:creator>Thielmann, Winfried</dc:creator>
          <dc:creator>Heller, Dorothee</dc:creator>
          <dc:date>2016-07-09</dc:date>
          <dc:description>Subcorpus 1 presents part of the euroWiss-Corpus covering communication in teaching/learning discourses in instruction at German and Italian universities, in the humanities as well as the technical and natural sciences; it offers access to transcriptions of lectures and seminars aligned with audio recordings and the text types used for instruction. The corpus comprises 18 Communications, 24 audio recordings, 24 transcriptions, 140,000 transcribed words, 19 identified speakers, 18 students' notes, 2 lecture scripts, 24 chalkboard presentions, 2 powerpoint presentations, 3 overhead slides, 3 handouts, 14 schedules/descriptions of recorded lecture/seminar

Das Projekt euroWiss basiert auf einem eigenen Korpus authentischer mündlicher Hochschulkommunikation an deutschen und italienischen Universitäten. Es konnten in insgesamt 58 Lehrveranstaltungen vom Typus Vorlesung und Seminar, Übung sowie Kolloquium je mindestens drei aufeinander folgende Sitzungen aufgezeichnet und so ein Videokorpus mit einem Umfang von ca. 357 Stunden erstellt werden. Die Aufnahmen dokumentieren verschiedene Disziplinen der drei Fakultäten Geisteswissenschaften (GW), Wirtschafts- und Sozialwissenschaften (WiSo) sowie Naturwissenschaften und Technik (MINT). Zusätzlich wurden begleitende Textarten, insbesondere studentische Mitschriften, PowerPoint-Präsentationen und Overhead-Folien, Handouts sowie einzelne Seminararbeiten bzw. Qualifikationsschriften wie tesine di laurea in das Korpus aufgenommen. Eine umfangreich angelegte Fragebogenstudie sowie gezielt erhobene narrative Interviews bieten methodisch ergänzende Daten. So entstand ein umfassendes, transnationales Korpus universitärer Wissensvermittlung, von dem ausgewählte Diskurssequenzen transkribiert wurden. Einen Einblick in die Transkripte bieten die euroWiss-Publikationen. Nach Abschluss des Projekts wurde das Korpus an das Hamburger Zentrum für Sprachkorpora (HZSK) übergeben, Teile davon wurden hier für Forschungszwecke digital zugänglich gemacht.

 

CLARIN Metadata summary for euroWiss - Linguistic Profiling of European Academic Education (Subcorpus 1)euroWiss - Linguistische Profilierung einer europäischen Wissenschaftsbildung (Subkorpus 1) (CMDI-based)

Title: euroWiss - Linguistic Profiling of European Academic Education (Subcorpus 1)
Title: euroWiss - Linguistische Profilierung einer europäischen Wissenschaftsbildung (Subkorpus 1)
Description: Subcorpus 1 presents part of the euroWiss-Corpus covering communication in teaching/learning discourses in instruction at German and Italian universities, in the humanities as well as the technical and natural sciences; it offers access to transcriptions of lectures and seminars aligned with audio recordings and the text types used for instruction. The corpus comprises 18 Communications, 24 audio recordings, 24 transcriptions, 140,000 transcribed words, 19 identified speakers, 18 students' notes, 2 lecture scripts, 24 chalkboard presentions, 2 powerpoint presentations, 3 overhead slides, 3 handouts, 14 schedules/descriptions of recorded lecture/seminar
Publication date: 2016-07
Data owner:  Angelika Redder, Winfried Thielmann, Dorothee Heller
Contributors:  Angelika Redder (depositor), Winfried Thielmann (depositor), Dorothee Heller (depositor), Christoph Breitsprecher (compiler), Yael Dilger (data_inputter), Marlene Uhlenberg (data_inputter), Maike Hawranke (data_inputter), Jessica Weidenhöffer (data_inputter), Julia Gerstner (data_inputter), Vera Makarenko (data_inputter), Lara Neumann (data_inputter), Barbara Schulze (data_inputter), Carina Bochmann (data_inputter), Linda Brücke (data_inputter), Marieke Dörschner (data_inputter), Laura Wittwer (data_inputter), Moritz Loch (data_inputter), Vivien Heinicke (data_inputter), Coretta Storz (data_inputter), Franziska Süß (data_inputter), Alessandra Alghisi (data_inputter), Elisa Bolpagni (data_inputter), Angelika Redder (researcher), Winfried Thielmann (researcher), Dorothee Heller (researcher), Christoph Breitsprecher (researcher), Jonas Wagner (researcher), Claudia DiMaio (researcher), Arne Krause (researcher), Gabriella Carobbio (researcher), VolkswagenStiftung (sponsor)
Project:  Linguistic Profiling of European Academic Education, VolkswagenStiftung
Keywords:  academic writing, communication in institutions, cross-sectional data, aquisition of academic language, L1 data, learner corpus, monolingual data, student texts, task-oriented communication, Economics, Philosophy, academic discourse, knowledge mediation, business administration, mathematics, sociology, literary studies, political science, physics, engineering, akademisches Schreiben, Kommunikation in Institutionen, Querschnittsdaten, Erwerb der Wissenschaftssprache, L1-Daten, Studierendentexte, Volkswirtschaftslehre, Philosophie, Hochschulkommunikation, Wissensvermittlung, Betriebswirtschaftslehre, Mathematik, Soziologie, Literaturwissenschaft, Politikwissenschaft, Physik, Ingenieurwissenschaften, EXMARaLDA
Languages:  German (deu), Italian (ita)
Size:  19 speakers (7 female, 12 male), 18 communications, 22 recordings, 1268 minutes, 23 transcriptions, 140541 words
Annotation types:  transcription (manual), sup: suprasegmental information, akz: Accentuation/stress, nv: Non-verbal, anno: Anonymisation, nn: Action by unspecified source, k: Free Comment, no: Numbering, de: German translation
Temporal Coverage:  2011-01-01/2014-01-01
Spatial Coverage:  DE; IT
Genre:  discourse
Modality:  spoken, written

 </dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/8236</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.8236</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:8236</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.8235</dc:relation>
          <dc:rights>info:eu-repo/semantics/restrictedAccess</dc:rights>
          <dc:subject>academic writing</dc:subject>
          <dc:subject>communication in institutions</dc:subject>
          <dc:subject>cross-sectional data</dc:subject>
          <dc:subject>aquisition of academic language</dc:subject>
          <dc:subject>L1 data</dc:subject>
          <dc:subject>learner corpus</dc:subject>
          <dc:subject>monolingual data</dc:subject>
          <dc:subject>student texts</dc:subject>
          <dc:subject>task-oriented communication</dc:subject>
          <dc:subject>Economics</dc:subject>
          <dc:subject>Philosophy</dc:subject>
          <dc:subject>academic discourse</dc:subject>
          <dc:subject>knowledge mediation</dc:subject>
          <dc:subject>business administration</dc:subject>
          <dc:subject>mathematics</dc:subject>
          <dc:subject>sociology</dc:subject>
          <dc:subject>literary studies</dc:subject>
          <dc:subject>political science</dc:subject>
          <dc:subject>physics</dc:subject>
          <dc:subject>engineering</dc:subject>
          <dc:subject>akademisches Schreiben</dc:subject>
          <dc:subject>Kommunikation in Institutionen</dc:subject>
          <dc:subject>Querschnittsdaten</dc:subject>
          <dc:subject>Erwerb der Wissenschaftssprache</dc:subject>
          <dc:subject>L1-Daten</dc:subject>
          <dc:subject>Studierendentexte</dc:subject>
          <dc:subject>Volkswirtschaftslehre</dc:subject>
          <dc:subject>Philosophie</dc:subject>
          <dc:subject>Hochschulkommunikation</dc:subject>
          <dc:subject>Wissensvermittlung</dc:subject>
          <dc:subject>Betriebswirtschaftslehre</dc:subject>
          <dc:subject>Mathematik</dc:subject>
          <dc:subject>Soziologie</dc:subject>
          <dc:subject>Literaturwissenschaft</dc:subject>
          <dc:subject>Politikwissenschaft</dc:subject>
          <dc:subject>Physik</dc:subject>
          <dc:subject>Ingenieurwissenschaften</dc:subject>
          <dc:subject>EXMARaLDA</dc:subject>
          <dc:subject>linguistics</dc:subject>
          <dc:subject>German</dc:subject>
          <dc:subject>Italian</dc:subject>
          <dc:title>euroWiss - Linguistic Profiling of European Academic Education (Subcorpus 1) euroWiss - Linguistische Profilierung einer europäischen Wissenschaftsbildung (Subkorpus 1)</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:8313</identifier>
        <datestamp>2022-06-28T11:28:54Z</datestamp>
        <setSpec>user-uhh</setSpec>
        <setSpec>user-hzsk</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:creator>Bührig, Kristin</dc:creator>
          <dc:creator>Meyer, Bernd</dc:creator>
          <dc:date>2020-01-01</dc:date>
          <dc:description>Data from the SimDiK project.</dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/8313</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.8313</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:8313</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.8312</dc:relation>
          <dc:rights>info:eu-repo/semantics/restrictedAccess</dc:rights>
          <dc:subject>German</dc:subject>
          <dc:subject>Polish</dc:subject>
          <dc:subject>Romanian</dc:subject>
          <dc:subject>Russian</dc:subject>
          <dc:subject>Linguistics</dc:subject>
          <dc:subject>community Interpreting</dc:subject>
          <dc:subject>Doctor-Patient Communication</dc:subject>
          <dc:title>SimDiK</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:8319</identifier>
        <datestamp>2020-11-18T09:45:23Z</datestamp>
        <setSpec>user-uhh</setSpec>
        <setSpec>user-hzsk</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:creator>Rothweiler, Monika</dc:creator>
          <dc:date>2011-01-01</dc:date>
          <dc:description>The TÜ_DE-L1-Korpus is a corpus of spoken child language that has been collected in the project Specific Language Impairment and Early Successive Language Acquisition (Project E4) at the Research Center on Multilingualism, University of Hamburg.

Audio recordings (spontaneous and elicited language) in Turkish with twelve bilingual children with L1 Turkish and L2 German with AOA of 3-4 years. Comparable data for the TÜ_DE-cL2-Korpus.</dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/8319</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.8319</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:8319</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.8318</dc:relation>
          <dc:rights>info:eu-repo/semantics/restrictedAccess</dc:rights>
          <dc:subject>German</dc:subject>
          <dc:subject>Turkish</dc:subject>
          <dc:subject>linguistics</dc:subject>
          <dc:subject>language acquisition</dc:subject>
          <dc:subject>bilingualism</dc:subject>
          <dc:subject>child language acquisition</dc:subject>
          <dc:subject>child bilingualism</dc:subject>
          <dc:subject>successive bilingualism</dc:subject>
          <dc:subject>child L2 acquisition</dc:subject>
          <dc:subject>L1 data</dc:subject>
          <dc:subject>EXMARaLDA</dc:subject>
          <dc:title>TU_DE_L1-Korpus</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:8325</identifier>
        <datestamp>2021-05-13T10:26:33Z</datestamp>
        <setSpec>user-uhh</setSpec>
        <setSpec>user-hzsk</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:contributor>Andresen, Melanie</dc:contributor>
          <dc:contributor>Knorr, Dagmar</dc:contributor>
          <dc:creator>Knorr, Dagmar</dc:creator>
          <dc:date>2017-11-18</dc:date>
          <dc:description>Authentic texts written by students of the University of Hamburg as part of their studies, the students have various L1 languages and study various subjects, all of the texts were subject of a writing counseling at the Writing Center Multilingualism (Schreibwerkstatt Mehrsprachigkeit), for some of the texts comments by peer tutors and several versions are available.

Authentische Texte Studierender der Universität Hamburg unterschiedlicher Herkunftssprachen und Studiengänge, die im Rahmen ihres Studiums als Prüfungsleistung entstanden sind und in einer individuellen Schreibberatung an der Schreibwerkstatt Mehrsprachigkeit besprochen wurden, zum Teil mit Kommentaren von Peer Tutoren und mehreren Versionen aus dem Schreibprozess.

Im Institut für Interkulturelle Bildung der Fakultät für Erziehungswissenschaft an der Universität Hamburg wird seit mehreren Jahrzehnten der Zusammenhang zwischen Bildungserfolg und Bildungssprache erforscht. Es wurde beobachtet, dass Studierende mit Migrationshintergrund Probleme beim Verfassen akademischer Texte haben. Die – naheliegende – Hypothese war, dass dies an den sprachlichen Fähigkeiten der Studierenden liegt. Allerdings ist eine Überprüfung der Hypothese ohne entsprechendes empirisches Datenmaterial schwierig. Die Fachliteratur zum deutschsprachigen akademischen Schreiben weist hier eine Lücke auf: Zugängliche Korpora von authentischen, deutschsprachigen akademischen Texten von Studierenden, die möglichst auch noch longitudinal untersucht werden können, existierten nicht. Mit dem Korpus KoLaS möchte die Schreibwerkstatt Mehrsprachigkeit einen Beitrag zur Forschung leisten, indem authentische Texte von Studierenden für Forschungszwecke zur Verfügung gestellt werden.

Andresen, Melanie; Knorr, Dagmar (2015) KoLaS: Kommentiertes Lernendenkorpus akademisches Schreiben

 

CLARIN Metadata summary for Commented Learner Corpus Academic WritingKommentiertes Lerndenkorpus akademisches Schreiben (CMDI-based)

Title: Commented Learner Corpus Academic Writing
Title: Kommentiertes Lerndenkorpus akademisches Schreiben
Description: Authentic texts written by students of the University of Hamburg as part of their studies, the students have various L1 languages and study various subjects, all of the texts were subject of a writing counseling at the Writing Center Multilingualism (Schreibwerkstatt Mehrsprachigkeit), for some of the texts comments by peer tutors and several versions are available.
Description: Authentische Texte Studierender der Universität Hamburg unterschiedlicher Herkunftssprachen und Studiengänge, die im Rahmen ihres Studiums als Prüfungsleistung entstanden sind und in einer individuellen Schreibberatung an der Schreibwerkstatt Mehrsprachigkeit besprochen wurden, zum Teil mit Kommentaren von Peer Tutoren und mehreren Versionen aus dem Schreibprozess.
Publication date: 2017
Data owner:  Dagmar Knorr, Universität Hamburg / Universitätskolleg / Schreibwerkstatt Mehrsprachigkeit / Von-Melle-Park 8 / 20146 Hamburg, dagmar.knorr@uni-hamburg.de
Contributors:  Dagmar Knorr (compiler), Melanie Andresen (compiler)
Project:  Schreibwerkstatt Mehrsprachigkeit, ZEIT-Stiftung Ebelin und Gerd Bucerius (2011–2014), Schreibwerkstatt Mehrsprachigkeit, Federal Ministry of Education and Research (2012–2016)
Keywords:  academic writing, student texts, written assignments, L2 data, L1 data, aquisition of academic writing, text comments, writing counselling, EXMARaLDA, akademisches Schreiben, wissenschaftlichen Schreiben, Studierendentexte, studentische Hausarbeiten, L2-Daten, L1-Daten, Erwerb der Wissenschaftssprache, Textkommentare, Schreibberatung, EXMARaLDA
Language:  German (deu)
Size:  854 texts
Segmentation units:  text
Temporal Coverage:  2011-01-01/2016-12-31
Spatial Coverage:  Hamburg, DE
Genre:  academic writing, akademisches Schreiben
Modality:  written

 </dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/8325</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.8325</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:8325</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.8322</dc:relation>
          <dc:rights>info:eu-repo/semantics/closedAccess</dc:rights>
          <dc:subject>academic writing</dc:subject>
          <dc:subject>student texts</dc:subject>
          <dc:subject>written assignments</dc:subject>
          <dc:subject>L2 data</dc:subject>
          <dc:subject>L1 data</dc:subject>
          <dc:subject>aquisition of academic writing</dc:subject>
          <dc:subject>text comments</dc:subject>
          <dc:subject>writing counselling</dc:subject>
          <dc:subject>EXMARaLDA</dc:subject>
          <dc:subject>akademisches Schreiben</dc:subject>
          <dc:subject>wissenschaftlichen Schreiben</dc:subject>
          <dc:subject>Studierendentexte</dc:subject>
          <dc:subject>studentische Hausarbeiten</dc:subject>
          <dc:subject>L2-Daten</dc:subject>
          <dc:subject>L1-Daten</dc:subject>
          <dc:subject>Erwerb der Wissenschaftssprache</dc:subject>
          <dc:subject>Textkommentare</dc:subject>
          <dc:subject>Schreibberatung</dc:subject>
          <dc:subject>EXMARaLDA</dc:subject>
          <dc:subject>linguistics</dc:subject>
          <dc:subject>German</dc:subject>
          <dc:title>Commented Learner Corpus Academic Writing; Kommentiertes Lernendenkorpus akademisches Schreiben</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:8985</identifier>
        <datestamp>2023-08-22T09:31:20Z</datestamp>
        <setSpec>user-uhh</setSpec>
        <setSpec>user-hzsk</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:creator>Andresen, Melanie</dc:creator>
          <dc:date>2021-03-30</dc:date>
          <dc:description>For this upload, all Word files (.doc and .docx) in the original KoLaS corpus were converted to plain text. For more information see https://github.com/melandresen/Ich-Daten.</dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/8985</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.8985</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:8985</dc:identifier>
          <dc:language>deu</dc:language>
          <dc:relation>doi:10.25592/uhhfdm.8326</dc:relation>
          <dc:relation>doi:10.25592/uhhfdm.8984</dc:relation>
          <dc:rights>info:eu-repo/semantics/restrictedAccess</dc:rights>
          <dc:subject>academic writing</dc:subject>
          <dc:subject>student texts</dc:subject>
          <dc:subject>written assignments</dc:subject>
          <dc:subject>L2 data</dc:subject>
          <dc:subject>L1 data</dc:subject>
          <dc:subject>aquisition of academic writing</dc:subject>
          <dc:subject>text comments</dc:subject>
          <dc:subject>writing counselling</dc:subject>
          <dc:subject>akademisches Schreiben</dc:subject>
          <dc:subject>wissenschaftliches Schreiben</dc:subject>
          <dc:subject>Studierendentexte</dc:subject>
          <dc:subject>studentische Hausarbeiten</dc:subject>
          <dc:subject>L2-Daten</dc:subject>
          <dc:subject>L1-Daten</dc:subject>
          <dc:subject>Erwerb der Wissenschaftssprache</dc:subject>
          <dc:subject>Textkommentare</dc:subject>
          <dc:subject>Schreibberatung</dc:subject>
          <dc:subject>linguistics</dc:subject>
          <dc:subject>German</dc:subject>
          <dc:title>Subset of KoLaS (Commented Learner Corpus Academic Writing), Plain Text Version</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:17731</identifier>
        <datestamp>2025-07-18T18:39:42Z</datestamp>
        <setSpec>user-hzsk</setSpec>
        <setSpec>user-uhh</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:contributor>Berger, Paul</dc:contributor>
          <dc:contributor>Malleyeck, Herman William</dc:contributor>
          <dc:contributor>Merjaŋda</dc:contributor>
          <dc:contributor>Gidayiɲi</dc:contributor>
          <dc:creator>Kießling, Roland</dc:creator>
          <dc:date>2025-07-18</dc:date>
          <dc:description>This is a test bundle installed to serve as a model in the reviewing process of the project “Tanzanian Rift Valley languages and cultures: documentary infrastructures, digital linguistics, local repatriation (TRiVLaC)”.</dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/17731</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.17731</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:17731</dc:identifier>
          <dc:language>eng</dc:language>
          <dc:relation>doi:10.25592/uhhfdm.17730</dc:relation>
          <dc:rights>info:eu-repo/semantics/openAccess</dc:rights>
          <dc:rights>https://creativecommons.org/licenses/by/4.0/legalcode</dc:rights>
          <dc:subject>Datooga</dc:subject>
          <dc:subject>oral literature</dc:subject>
          <dc:subject>narrative</dc:subject>
          <dc:subject>Southern Nilotic</dc:subject>
          <dc:subject>Gisamjanga</dc:subject>
          <dc:subject>Swahili</dc:subject>
          <dc:subject>Tanzania</dc:subject>
          <dc:title>Heydesh - The giant white bull (Datooga)</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:16609</identifier>
        <datestamp>2025-09-06T14:19:35Z</datestamp>
        <setSpec>user-uhh</setSpec>
        <setSpec>user-hzsk</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:contributor>Dr. Maxim Makartsev</dc:contributor>
          <dc:contributor>Dr. Elena Uzeneva</dc:contributor>
          <dc:contributor>Dr. Timofey Arkhangelskiy</dc:contributor>
          <dc:creator>Makartsev, Maxim</dc:creator>
          <dc:creator>Arkhangelskiy, Timofey</dc:creator>
          <dc:date>2024-12-31</dc:date>
          <dc:description>A corpus of Slavic dialects in Albania

The user-friendly version of the Corpus with search options is available here.

These are the main parameters of the corpus:


	Korça Macedonian (KM): Rural locations: Boboshtica | Urban locations: Korça | Size: 34.0 thousand words | Morphological analysis: Classla 2.1.1 for Macedonian, partial manual postprocessing.
	Prespa Macedonian (PM): Rural locations: Pustec, Gorna Gorica, Dolna Gorica, Shulin | Urban locations: Elbasan, Korça | Size: 171.3 thousand words | Morphological analysis: Classla 2.1.1 for Macedonian, partial manual postprocessing.
	Golloborda Macedonian (GM): Rural locations: Trebisht, Vërnica, Malestreni | Urban locations: Durrës, Elbasan, Tirana | Size: 239.7 thousand words | Morphological analysis: Classla 2.1.1 for Macedonian, partial manual postprocessing.
	Myzeqe Štokavian (MŠ): Rural locations: Rreth Libofsha, Petova | Urban locations: Fier | Size: 58.8 thousand words | Morphological analysis: Classla 2.1.1 for Serbo-Croatian, partial manual postprocessing.
	Shijak Štokavian (SŠ): Rural locations: Borake, Koxhas | Urban locations: Shijak, Sukth | Size: 68.8 thousand words | Morphological analysis: Classla 2.1.1 for Serbo-Croatian, partial manual postprocessing.
	Albanian: Rural locations: All of the above | Urban locations: All of the above | Size: 34.7 thousand words | Morphological analysis: uniparser-albanian.
	Other languages: Bulgarian, English, French, German, Greek, BCMS, Italian, Russian, Turkish | Size: 4.5 thousand words | Morphological analysis: not analyzed.
	Total size: 611.8 thousand words.


Previous research on Slavic dialects in Albania

This project is not the first study of Slavic dialects in Albania (SDAs), but we did consider varieties that have never been studied (e.g., Slavic speech in urban Albanian settings; Myzeqe Štokavian).

Starting from the groundbreaking monograph Slavic populations of Albania by A. Seliščev (1931), there have been several publications on SDAs that covered both the history of the Slavic population in this country (e.g., Ylli’s (1997, 2000) monograph on Slavic borrowings in Albanian toponymics) and its current state (Bojović 1991; Tončeva 2014; Vidoeski 1998). Of the utmost importance is the four-volume work “Die slavischen Minderheiten in Albanien” by Steinke and Ylli (2007, 2008, 2010, 2013), supported by a Deutsche Forschungsgemeinschaft (DFG) grant between 2002 and 2011.

In the selected publications listed, as well as in those dedicated to separate dialects (see the outline of dialects below), you can find descriptions of the language systems of the SDAs, information about the current status of the communities that use the SDAs, and dialectal transcripts.

Selected language varieties

The labels for the language varieties included in the corpus do not make any claims about the national, ethnic, or other identities of the speakers; they are purely provided for an orientation in terms of the respective dialectologies. The labels are not necessarily those that the speakers used, either. In fact, some speakers do not use any labels at all, while others use a variety of labels, often non-terminologically. The language issue within some of the ethnolinguistic minorities in Albania is seriously politicized; however, this corpus does not carry any political claims of any political organization, party, group of individuals, or state, etc. The language labels used in the external sources quoted here are as in the original for identification purposes only.

Five dialects were chosen for this project, as shown on the Google Map.

Golloborda Macedonian

Golloborda Macedonian is a peripheral Balkan Slavic dialect that continues West Macedonian Debar dialects in Albanian territory. It is spoken in 15 villages in the Albanian regions of Dibra and Elbasan, as well as in migrant communities in the cities of Durrës, Tirana, and Elbasan. It has been estimated that it has more than 7,000 speakers in total. In its rural centers, this community has been studied thoroughly by a team of researchers from the Institute of Linguistic Studies (Russian Academy of Sciences), Saint Petersburg State University, and Peter the Great Museum of Anthropology and Ethnography (Kunstkammer); their research resulted in a valuable monograph that was translated into Albanian and Macedonian. Notably, our corpus-based research focused on the sociolinguistic variation and changes within this dialect, specifically between its rural and urban centers, so our methods and data differ from those in the research of our peers from Saint Petersburg.

Selected literature: Steinke &amp; Ylli (2008); Sobolev &amp; Novik (2013, 2017, 2018).

Korça Macedonian

Korça Macedonian is an apparently extinct Balkan Slavic island dialect (structurally close to the dialectical area of Southeastern Macedonian). The corpus includes speech samples from the last six speakers (three of whom lived in the village of Boboshtica and three of whom lived in the town of Korça in Southeastern Albania but were originally from Drenova). Despite the community being so small, this dialect was crucial for the project, as it has been subjected to the most prolonged and intensive Albanian influence. However, family discussions could not be organized in this community.

Selected literature: Mazon (1936); Mazon and Filipova-Bajrova (1965); Steinke and Ylli (2007)

Prespa Macedonian

Prespa Macedonian is a peripheral Balkan Slavic dialect that continues the West Macedonian Ohrid-Prespa dialects on the Albanian side of Great Prespa Lake, transitional to Southeastern Macedonian. According to estimates, it has around 4,500 speakers in nine villages of the region and two large towns, namely Korça and Bilisht.

Selected literature: Steinke and Ylli (2007); Cvetanovski (2010)

Myzeqe Štokavian

Spoken in several quarters in Fier and several villages around this town, Myzeqe Štokavianan is a Štokavian island dialect spoken among recent (1920s) migrants from the Sandžak region (a Novi Pazar-Sjenica dialect of the Zeta-Sjenica dialectal zone) of what is now Southwestern Serbia and the bordering region of Montenegro.

Selected literature: Makartsev and Kikilo (2022); Makartsev (2023)

Shijak Štokavian

Spoken in the village of Borake and its satellite village of Koxhas, Shijak Štokavian is a Štokavian island dialect spoken among relatively recent (from the 1880s) migrants from the Mostar region in what is now Bosnia and Herzegovina (a central Herzegovinian subdialect of the East Bosnian dialectal zone, spoken in the Mostar—Čapljina—Stolac triangle). It is spoken by 150 to 220 families in both villages. Speakers of this dialect also live in the town of Shijak and Sukth.

Selected literature: Steinke and Ylli (2013); Makartsev and Kikilo (2022); Makartsev (2023)

Sociolinguistic diversity among the SDAs

These dialects have varying degrees of structural affinity with Albanian due to their differing connections to the Balkan sprachbund. Structurally, the closest to Albanian are the Balkan Slavic dialects of Korça, Golloborda, and Prespa. The Štokavian dialects of Myzeqe and Shijak are not included in the Balkan sprachbund and show less structural affinity with Albanian.

The selected dialects do not represent all the SDAs. One of the most complete lists of SDAs can be found in Steinke and Ylli’s monograph. However, they contain the variety of elements and parameters that have caused the diversity in the SDAs.

Two of the dialects are Štokavian (Myzeqe and Shijak), but their speakers have differing ethnopolitical and linguistic orientations: Our interviewees in Shijak usually articulated their Bosniak identity; Myzeqe speakers usually clarified their Bosniak or Serbian identity. Three dialects are Balkan Slavic (Golloborda, Korça, Prespa). The orientation of the speakers of these dialects toward a standard language (Macedonian or Bulgarian) is usually individual. The ethnopolitical and linguistic orientations mentioned by the speakers were thus not interpreted as making any political claims, but they allowed us to thematize the orientation of the speakers toward one of the standard Southern Slavic languages and better explain certain features in their speech.

Four of the communities have a rural (more conservative) and an urban (less conservative) center (except Korça Macedonian, whose number of speakers did not allow us to construct this opposition). For Shijak, the labor activity of the speakers was important, especially in terms of whether their jobs were connected to work at the national road services. (Such workers have daily contact with the language of TIR truck drivers who speak BCMS.) The opposition of rural versus urban was not relevant for the Shijak data because of the short distance between the settlements and the small size of the towns of Sukth and Shijak. The city of Durrës, which is where many of the dialectal speakers work daily, is also located too close to Shijak to allow for the shaping of a significant urban colony with distinct features.

Religion was another factor that we considered might influence linguistic identity choices (cf. the linguo-confessional situation in Bosnia and Herzegovina and regions of Montenegro and Serbia populated by a Štokavian-speaking but traditionally Muslim population) since it is relevant to numerous distinctive features of the traditional culture. All Myzeqe and Shijak Štokavian speakers that we interviewed culturally and traditionally belong to Sunni Islam. Among the Balkan Slavic communities, all our Korça and Prespa speakers culturally belong to Orthodox Christianity, while Golloborda is heterogenous with the domination of Sunni Islam.

Topics of our interviews

1. Narratives and memorates: Type: Unstructured 

2. Ethnographical and ethnolinguistic interviews:

2.1. Calendar, rites of passage (birth, marriage, death), demonology: Type: Semi-structured | Bibliographical reference: (Plotnikova 2009; see the online publication).

2.2. Rites and beliefs connected to the moon: Type: Semi-structured | Bibliographical reference: (Čëxa 2009).

2.3. Rites and beliefs connected to the cuckoo: Type: Semi-structured | Bibliographical reference: (Makartsev 2017).

3. Frog, Where Are You?: Bibliographical reference: (Berman et al. 1994–2004; Mayer 1969; see preview).

3.1. Conducted by the researchers

3.2. Conducted by trained local assistants

4. Family talks: Type: Unstructured | Bibliographical reference: (Hentschel and Zeller 2013)

The narratives and memorates (T. 1) were unstructured discussions about the oral history and current problems of a given community that also provided insights into the identity and politics of memory of the community. The researchers led these discussions.

The ethnographical and ethnolinguistic interviews (T. 2) were conducted to collect ethnographical and ethnolinguistic information. They comprised the informants’ answers to our questions and covered various aspects of the traditional culture. We mainly followed the structure of the questionnaires (or interview designs) listed in the table, with slight adaptations.

Frog, Where Are You? (T. 3) is a book with 24 pictures that combine to form a visual narrative. This section was structured as a questionnaire, ensuring that the researcher was minimally involved. We also asked our trained local assistants to record themselves or their relatives and friends answering this questionnaire; therefore, the data that we collected here resembled real-life language use.

Our trained local assistants organized the family talks (T. 4) in our absence. The aim was to record spontaneous speech, so the topics were irrelevant. Since the same assistants prepared the transcripts, they could omit any sections that contained potentially harmful information or could have been used to identify the speakers.

Collecting this type of data was most successful for Golloborda Macedonian speakers since we had a network of trained local assistants upon whom we could rely.

We managed to arrange a few family talks among Prespa Macedonian speakers and just one family talk with Shijak Štokavian speakers.

For Myzeqe Štokavian, arranging family talks has not yet been successful.

For Korça Macedonian, such discussions were impossible since none of our speakers still used the dialect daily, although they could still speak it with us. The epigraph above was spoken by one of the speakers from Drenova (Dre01), wherein he described the frustration he felt while witnessing the attrition and loss of his native dialect.

Speakers

The speakers in the transcripts were divided into the three main categories:

1) Native speakers of the respective SDAs. They were anonymized. All information that could be used for their identification was manually removed from the corpus (tagged as ((ERASED))). All speakers of this category were referenced with indices comprising three letters (for the settlement) and two digits. We also referred to them by these indices in our publications based on the corpus.

2) Researchers. They were only referenced with letter indices. Their names are provided in the Acknowledgments section.

3) SPK. This abbreviation was used to refer to all other speakers whose speech was transcribed for context but was not annotated for various reasons (an unknown neighbor passing by the window and saying hello, an Albanian-speaking waiter in a village café, some unidentified background voices, etc.).

See list of speakers

Transcripts

Our team manually prepared all transcripts. When possible, our trained local assistants, speakers of the dialects who also organized the family discussions, prepared the draft versions. Our editors (specialists in their respective philology) proofread these draft versions, following which Dr. Maxim Makartsev double-proofread them. If trained local assistants were unavailable, our editors prepared the transcripts. We used EXMARaLDA Partitur Editor to match the transcripts with the recordings. The original scripts for Albanian and other languages were used. Only transcripts in Slavic dialects were proofread, while transcripts for Albanian and other languages were not for contextual purposes. (Hence, the spelling ranges from standard to non-orthographical semi-phonetic spelling.)

Transcription

See transcription conventions

Annotation

We followed several steps for the annotation.

First, we formulated the rules to define the language of the respective word form based on the tags provided in EXMARaLDA Partitur Manager manually by the transcribers and editors and on language-specific scripts (e.g., Greek, Russian), symbols (e.g., ë, special for Albanian), and combinations of symbols (e.g., ll, rr, initial ng and mb, special or unique for Albanian).

Second, the respective parsers and taggers were applied depending on the language of the word form (see the table above). Following this, only those parts of the transcripts that the speakers uttered (not the researchers) in Slavic dialects were manually and semi-automatically checked, proofread, and edited. The parts of the transcripts that the researchers uttered or that speakers uttered in other language varieties were not proofread. In such cases, we kept the automatic annotation.

Third, lemmatization was manually checked (for Slavic); the lemmas automatically marked as Albanian were selectively checked and corrected, if needed. For Golloborda, Korça, and Prespa Macedonian, the lemmatization was performed in standard Macedonian—as the closest standard language structurally. For Myzeqe and Shijak, lemmatization was based on Ijekavian standards. For dialectal lexemes that did not exist in the respective standards, standard phonology was applied, resulting in the creation of dummy lemmata that followed standard phonology but cannot be found in standard dictionaries. The only function of these lemmata was to allow for the trans-dialectal search of word forms. If a standard cognate could not be established, we adopted any suitable word form attested in our transcripts.

Fourth, the morphological tags for Macedonian and Štokavian were harmonized, since Classla 2.1.1 uses slightly different MULTEXT-East conventions for varieties. The resulting tag set is provided below.

Fifth, the results of the morphological tagging were selectively checked. We focused on the word forms with the greatest homonymy and the lexemes with the most frequent tokens.

This is the beta-version of our corpus, so the manual editing of the morphological tags is ongoing. If you notice an error, please feel free to contact us. When working with the corpus, manual checking of the search results is highly recommended.

Considering the principally bilingual nature of our data and frequent code-switches, we would like to highlight two instruments here:

1) The corpus allows for searching sets of word forms that are specifically ordered (one after another or with one or more irrelevant word forms in between). You may compose the search entry by marking one of the word forms as Slavic (you can choose the dialect) and another as Albanian, which will show all cases of code-switches that follow your chosen parameters.

2) The special field Foreign includes all lexical matter borrowings and congruent lexicalizations from Albanian. (They cannot be formally distinguished since both types have Albanian stems and Slavic inflectional morphology.)

We also distinguished direct speech (tag OWN) and quotations (XENO), while several other tags provide additional information about the intonation and context of the interview (((LAUGH)), ((COUGH)), ((NOISE))).

Metadata


	Transcript ID
	Year and location of the recording
	Sub-corpus (dialect)
	Code of the speaker
	Code of the researcher who participated in the recording and transcribing of the interview
	The type of interview (whether the researchers were present or absent)
	The speaker’s birth place, birth year, gender, and occupation; other sociolinguistic data when relevant; familial relations when relevant
	Current place of residence of the speaker
	Genre


List of transcripts

See full list

Tag set for SDAs

Originally, the tag set was based on MULTEXT-East morphosyntactic specifications. Korça Macedonian, Prespa Macedonian, and Golloborda Macedonian were based on Macedonian specifications; Myzeqe Štokavian and Shijak Štokavian were based on Serbo-Croatian specifications.

We introduced several changes to 1) harmonize the Macedonian and Štokavian parsers and taggers that otherwise followed somewhat different principles and conventions; 2) adapt the tag set to the terminology most widely used in Slavic linguistics (e.g., the use of the term “imperfective aspect” instead of “progressive aspect”); 3) unify all other possible idiosyncrasies (e.g., MULTEXT-East morphosyntactic specifications for Serbo-Croatian do not have verbal aspects, so these had to be introduced for our corpus).

The tag set for the Albanian language section was developed by Maria Morozova, Alexander Rusakov, and Timofey Arkhangelskiy for the Albanian National Corpus and can be found here. Albanian tags are preceded by the prefix sq: to avoid confusion with homonymous Slavic tags.

The grammatical features of the words in the corpus are marked with short tags. In tags, abbreviations are capitalized, while full words are not.

See the tag set

Frequently asked questions

— What is the Corpus of Slavic dialects in Albania?

This is a language corpus or collection of non-adapted transcripts of interviews done in Slavic dialects that are spoken in Albania. Each word form in these dialects included in the corpus are enriched with additional linguistic information or annotations. We also have a user-friendly interface that allows for writing search queries.

— Who needs corpora?

Corpora are used by linguists. The search engines and annotations of corpora are designed to allow for easily making linguistic queries such as “find all pronouns in the accusative case” or “find all forms of the word mačka followed by a verb” or “find all instances of a noun followed by an adjective” so that you can retrieve relevant information from the provided linguistic varieties in seconds. Further analyses of this type of data allow linguists to determine how linguistic varieties have changed, how Albanian has influenced these varieties, what the limits of variation are, or whether there are any new and interesting linguistic phenomena that are not found in Macedonian and Štokavian dialects that have had no contact with Albanian.

Aside from linguists, corpora can be useful tools for language teachers, language learners, and even native speakers.

A corpus documents linguistic varieties in a given period. For example, one of the varieties included (Korça Macedonian) appeared to have gone extinct during our project (or its last speakers became unavailable to the researchers due to their old age). To the best of our knowledge, our corpus includes the last speech examples of this dialect available. It preserves this dialect and other included varieties for future generations and can be used by language activists for language revitalization.

— Can I use the corpus for other things beyond purely linguistic research?

The Corpus of Slavic dialects in Albania makes full transcripts of our recordings available. Aside from merely linguistic interests, the content of the transcripts can also be analyzed since the transcripts are so diverse and include many narratives containing oral history, identity, and anonymized personal biographies. Our transcripts also include much ethnographic and ethnolinguistic information on the traditional culture of the communities, which can be relevant for ethnolinguists, ethnographers, and members of the communities. There are also many examples of oral folk traditions (songs, tales, proverbs, etc.) available for researchers and the general public.

— Can I use the corpus as a dictionary?

You might not be able to use this corpus like you would a traditional dictionary because it does not provide translations or explanations of the included words. You may, however, discover in which context the word is used, which you can then use to clarify the word’s meaning.

— What is a morphological annotation, and how is it obtained?

Our corpus was lemmatized and morphologically annotated. Lemmatization means that each word in the texts was annotated with its lemma, i.e., its dictionary or citation form. Morphological annotation means that the grammatical features of each word were annotated, including its part of speech, number, case, tense, etc. Since the corpus was too large for manual annotations, it was annotated automatically with programs called morphological analyzers.

We used analyzers compiled for standard Macedonian since it is closest to Golloborda,  Korça, and Prespa Macedonian structurally, as well as analyzers compiled for Štokavian-based standard languages for Myzeqe and Shijak Štokavian (mostly the Croatian analyzer since it could account for dialectal variations in phonology and morphology within the Štokavian dialects). The language denominators for the respective analyzers do not make any claims about the identities of our speakers and were only used as references in external sources.

The results of the automatic annotation were partially proofread and edited manually. Our corpus still has homonymy, i.e., when one word form may have several possible morphological analyses. For example, ja in Macedonian dialects can mean ‘I’ (first-person singular personal pronoun in the nominative case), ‘her’ (third-person singular feminine personal pronoun in the accusative case), ‘here’ (a deictic particle), etc. Hence, when looking for anything within the corpus, you will receive false positive results. Manually checking the data you find in the corpus is thus strongly recommended.

Acknowledgments

This corpus is the main research instrument developed for the project “Contact-induced language change in situations of non-stable bilingualism—Its limits and modelling: Slavic (social) dialects in Albania,” funded by the DFG (German Research Foundation), project number 8750/1-1 (October 16th, 2019–April 30th, 2024). The principal investigator was Dr. Maxim Makartsev.

The concept, development, and realization of this project would not have been possible without the constant support of Prof. Dr. Gerd Hentschel, my deepest gratitude to whom words cannot express. I am deeply indebted to Prof. Dr. Jan Patrick Zeller and my other colleagues from the Institute for Slavistics (Carl von Ossietzky Universität Oldenburg) for their support and thorough feedback on my project during its various stages.

Authors

The corpus was developed and is maintained by:


	Dr. Maxim Makartsev (Institut für Slavistik, Carl von Ossietzky Universität Oldenburg), maxim.makartsev@gmail.com
	Dr. Timofey Arkhangelskiy (Institut für Finno-Ugristik/Uralistik, Universität Hamburg), timarkh@gmail.com


The current version of the corpus uses the platform tsakorpus developed by Dr. Timofey Arkhangelsky. It is stored on the server of Carl von Ossietzky University Oldenburg.

I am deeply grateful to Dr. Elena Uzeneva, my cooperation partner (between January 1st, 2020, and March 31st, 2022), without whom this project would not have been possible.

Several field trips were undertaken to collect the speech samples that were included in the corpus. In 2010–2019, Dr. Maxim Makartsev organized these trips using his own resources (see his legacy page on the site of his previous host institution for publications based on those data). The participants in these fieldtrips were:


	Dr. Mikhail Chivarzin, Moscow—Shenzhen (MC)
	Alexandra Chivarzina, Moscow—Shenzhen (AC)
	Renata Hamidullina, Perm—Vienna (RH)
	Marina Mihajlova, Sofia—Calgary (MMI)


In 2020–2022, Dr. Maxim Makartsev and Dr. Elena Uzeneva organized the field trips within the framework of the aforementioned project funded by the DFG. The participants in these fieldtrips were:


	Dr. Natalia Kikilo, Moscow (NK)
	Dr. Anna Leontyeva, Moscow (AL)
	Nora Muheim, Zürich—Helsinki (NM)
	Aino Väänänen, Oulu—Berlin (AV)


The photo on the start page was taken in Trebisht/Требишта, Albania, by Aino Väänänen, an independent documentary photographer, during our joint research field trip for the aforementioned project in 2020 and was used with her kind permission.

The transcripts for the corpus were prepared by:


	Hristina Angeleska—Prilep
	Bojana Damnjanović—Helsinki
	Pavel Falaleev—Helsinki
	Đorđe Genović—Belgrade
	Violeta Jordanova—Skopje
	Dr. Natalia Kikilo—Moscow
	Dr. Maxim Makartsev—Oldenburg
	Milan Milenović—Belgrade
	Ekaterina Panova—Saint Petersburg
	Uliana Putilina—Moscow
	Anđela Redžić—Belgrade
	Maria Stryszewska—Wrocław
	Ekaterina Titova—Moscow


We are deeply grateful to our speakers and local assistants whose hard work and passion made it possible to make this corpus available. We cannot disclose their names and personal details for their protection.

References

Berman, Ruth A., Dan I. Slobin, Sven Stromqvist, and Ludo T. Verhoeven. 1994–2004. Relating Events in Narrative. Hillsdale, N.J. L. Erlbaum Associates.

Bojović, Jovan R., ed. 1991. Stanovništvo slovenskog porijekla u Albaniji : zbornik radova sa međunarodnog naučnog skupa održanog u Cetinju 21, 22. i 23. juna 1990. Titograd: Stručna knjiga.

Čëxa, Oksana V. 2009. “Novogrečeskaja leksika narodnoj astronomii v sopostavlenii s balkanoslavjanskoj: Luna i lunnoe vremja (ėtnolingvističeskij aspekt).” Ph.D., Institute of Slavic Studies, Russian academy of sciences. https://inslav.ru/event/chyoha-oksana-vladimirovna-novogrecheskaya-leksika-narodnoy-astronomii-v-sopostavlenii-s.

Cvetanovski, Goce. 2010. Govorot na makedoncite vo Mala Prespa: zapadnoprespanski govor. Skopje: Institut za makedonski jazik “Krste Misirkov”.

Hentschel, Gerd, and Jan P. Zeller. 2013. “Gemischte Rede, gemischter Diskurs, Sprechertypen: Weißrussisch, Russisch und gemischte Rede in der Kommunikation weißrussischer Familien.” In Wiener Slawistischer Almanach, edited by Aage A. Hansen-Löve and Tilmann Reuther, 127–55 70. München, Berlin, Wien: Peter Lang.

Makartsev, Maxim. 2017. “Ėtjudy k balkanskomu bestiariju: Kukuška.” Živaja starina 95 (3): 46.

———. 2023. “Razvoj balkanoslavenskoga tipa futura u štokavskim iseljeničkim dijalektima u Albaniji i jezički kontakti.” Književni jezik (34): 41–69.

Makartsev, Maxim, and Natalia Kikilo. 2022. “Some Tendencies in the Morphosyntax of the Migrational Shtokavian Dialects in Albania (Shijak and Myzeqe) And Slavic-Albanian Language Contact.” Slavic World in the Third Millennium 17 (1-2): 120–41. doi:10.31168/2412-6446.2022.17.1-2.07.

Mayer, Mercer. 1969. Frog, Where Are You? Sequel to a Boy, a Dog and a Frog. New York: Dial Books for Young Readers (a division of Penguin Putnam Inc.).

Mazon, André. 1936. Documents, contes et chansons slaves de l’Albanie du Sud. Bibliothèque d’études balkaniques 5. Paris: Librarie Droz.

Mazon, André, and Maria Filipova-Bajrova. 1965. Documents slaves de l’Albanie du Sud: II. Pièces complémentaires. Bibliothèque d’études balkaniques 8. Paris: Institut d’études slaves.

Plotnikova, Anna A. 2009. Materialy dlja ėtnolingvističeskogo izučenija balkanoslavjanskogo areala. 2, revised. Moskva: Institut slavjanovedenija RAN.

Seliščev, Afanasij M. 1931. Slavjanskoe naselenie v Albanii (s illjustracijami v tekste i s kartoju Albanii). Sofia.

Sobolev, Andrey N., and Aleksandr Novik. 2013. Golo Bordo (Gollobordë), Albanija: Iz materialov balkanskoj ėkspedicij RAN i SPbGU 2008-2010 gg. Materialien zum Südosteuropasprachatlas Bd. 6. Sankt-Peterburg: Nauka.

———. 2017. Gollobordë (Golo Bordo), Shqipëri: Nga materialet e ekspeditës ballkanike të AShR-së dhe UShSt-P-së në vitet 2008-2010. Translated by Ligor Cullufe. Tiranë: Botimet Toena.

———. 2018. Golo Brdo: Od materijalite na balkanskata ekspedicija na RAN i SPbDU vo 2008-2010 godina. Materialien zum Südosteuropasprachatlas Band 6. Skopje, Sankt Peterburg: Univerzitet "Sv. Kiril i Metodij"; Institut za makedonski jazik "Krste Misirkov"; "Nauka".

Steinke, Klaus, and Xhelal Ylli. 2007. Die slavischen Minderheiten in Albanien (SMA): 1. Teil. Prespa-Vërnik-Boboshtica. Slavistische Beiträge 458. München: Otto Sagner.

———. 2008. Die slavischen Minderheiten in Albanien (SMA): 2. Teil. Golloborda-Herbel-Kërçishti i Epërm. Slavistische Beiträge 462. München: Otto Sagner.

———. 2010. Die slavischen Minderheiten in Albanien (SMA): 3. Teil. Gora. Slavistische Beiträge 474. München: Otto Sagner.

———. 2013. Die slavischen Minderheiten in Albanien (SMA): 4. Teil. Vraka-Borakaj. Slavistische Beiträge 491. München, Berlin: Sagner.

Tončeva, Veselka. 2014. Našencite v Albanija: Istorija, ezik, tradicii. Sofia: Ongăl.

Vidoeski, Božidar. 1998. Dijalektite na makedonskiot jazik. Vol. 1. Skopje: Makedonska Akademija na naukite i umetnostite.

Ylli, Xhelal. 1997. Das slavische Lehngut im Albanischen. Teil 1 : Lehnwörter. Slavistische Beiträge. Digitale Ausgabe 350. München: Verlag Otto Sagner.

———. 2000. Das slavische Lehngut im Albanischen. Teil 2 : Ortsnamen. Slavistische Beiträge. Digitale Ausgabe 395. München: Verlag Otto Sagner.

Contact

If you have questions, would like to propose collaboration, or noticed an error in the corpus, please contact Dr. Maxim Makartsev.

 </dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/16609</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.16609</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:16609</dc:identifier>
          <dc:language>eng</dc:language>
          <dc:relation>doi:10.5281/zenodo.14191622</dc:relation>
          <dc:relation>doi:10.25592/uhhfdm.16608</dc:relation>
          <dc:rights>info:eu-repo/semantics/openAccess</dc:rights>
          <dc:rights>https://creativecommons.org/licenses/by/4.0/legalcode</dc:rights>
          <dc:subject>Balkan linguistics</dc:subject>
          <dc:subject>language contact</dc:subject>
          <dc:subject>sociolinguistics</dc:subject>
          <dc:subject>Slavic dialects in Albania</dc:subject>
          <dc:subject>Slavic ethnography</dc:subject>
          <dc:subject>Macedonian dialectology</dc:subject>
          <dc:subject>Štokavian dialectology</dc:subject>
          <dc:subject>Albanian language</dc:subject>
          <dc:title>A corpus of Slavic dialects in Albania</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:8311</identifier>
        <datestamp>2025-11-28T10:50:56Z</datestamp>
        <setSpec>user-uhh</setSpec>
        <setSpec>user-hzsk</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:contributor>Lehmberg, Timm</dc:contributor>
          <dc:contributor>Czachór, Agnieszka</dc:contributor>
          <dc:contributor>Hajduk, Maja</dc:contributor>
          <dc:contributor>Kulesz, Magdalena</dc:contributor>
          <dc:contributor>Brehmer, Bernhard</dc:contributor>
          <dc:creator>Bernhard Brehmer</dc:creator>
          <dc:date>2011-09-02</dc:date>
          <dc:description>Original Data: Audio recordings of German/Polish bilingual and Polish monolingual adults (16-46 years). Recordings of semi-spontaneous data (3 topics) and renarration of a picture story.

Corpus: The Hamburg Corpus of Polish in Germany (HamCoPoliG) was compiled between July 2008 and June 2011 in the research project "Current Polish-German Bilingualism in Germany" (H8, principal investigator: Bernhard Brehmer), part of the Collaborative Research Centre "Multilingalism", funded by the German Research Foundation (Deutsche Forschungsgemeinschaft, DFG) and hosted by the University of Hamburg.

Czachor, Agnieszka (2012): Polish-German bilingual corpus: collecting and analysing written and spoken data for investigating contact-induced change. Submitted to: Schmidt, Thomas &amp; Wörner, Kai (eds.): Multilingual Corpora and Multilingual Corpus Analysis. Hamburg Studies in Multilingualism (14). Amsterdam: Benjamins.

 

CLARIN Metadata summary for Hamburg Corpus of Polish in Germany (HamCoPoliG) (CMDI-based)

Title: Hamburg Corpus of Polish in Germany (HamCoPoliG)
Description: Audio recordings of German/Polish bilingual and Polish monolingual adults (16-46 years). Recordings of semi-spontaneous data (3 topics) and renarration of a picture story.
Publication date: 2020-09-30
Data owner:  Bernhard Brehmer, Institut für Slavistik / Von-Melle-Park 6 / D-20146 Hamburg, bernhard.brehmer@uni-hamburg.de
Contributors:  Bernhard Brehmer, Institut für Slavistik / Von-Melle-Park 6 / D-20146 Hamburg, bernhard.brehmer@uni-hamburg.de (depositor), Bernhard Brehmer, Institut für Slavistik / Von-Melle-Park 6 / D-20146 Hamburg, bernhard.brehmer@uni-hamburg.de (compiler), Magdalena Kulesz (data_inputter), Maja Hajduk (data_inputter), Timm Lehmberg, timm.lehmberg@uni-hamburg.de (developer), Bernhard Brehmer, Institut für Slavistik / Von-Melle-Park 6 / D-20146 Hamburg, bernhard.brehmer@uni-hamburg.de (researcher), Agnieszka Czachór (researcher), Deutsche Forschungsgemeinschaft (DFG) (sponsor)
Project:  H8 "Current Polish-German Bilingualism in Germany", German Research Foundation (DFG)
Keywords:  language attrition, adult bilingualism, cross-sectional data, simultaneous bilingualism, successive bilingualism, child L2 acquisition, adult L2 acquisition, L1 data, EXMARaLDA
Language:  Polish (pol)
Size:  93 speakers (64 female, 29 male), 360 communications, 37.73 hours, 2264 minutes, 358 recordings, 358 transcriptions, 294663 words
Temporal Coverage:  2009-04-17/2010-12-17
Spatial Coverage:  Max-Brauer-Allee 60, 22765 Hamburg, DE; Warszawa, PL
Genre:  discourse
Modality:  spoken

 </dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/8311</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.8311</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:8311</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.1424</dc:relation>
          <dc:rights>info:eu-repo/semantics/restrictedAccess</dc:rights>
          <dc:subject>language attrition</dc:subject>
          <dc:subject>adult bilingualism</dc:subject>
          <dc:subject>cross-sectional data</dc:subject>
          <dc:subject>simultaneous bilingualism</dc:subject>
          <dc:subject>successive bilingualism</dc:subject>
          <dc:subject>child L2 acquisition</dc:subject>
          <dc:subject>adult L2 acquisition</dc:subject>
          <dc:subject>L1 data</dc:subject>
          <dc:subject>EXMARaLDA</dc:subject>
          <dc:subject>linguistics</dc:subject>
          <dc:subject>Polish</dc:subject>
          <dc:title>Hamburg Corpus of Polish in Germany (HamCoPoliG)</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:17950</identifier>
        <datestamp>2026-03-03T16:16:51Z</datestamp>
        <setSpec>user-uhh</setSpec>
        <setSpec>user-hzsk</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:creator>Bührig, Kristin</dc:creator>
          <dc:date>2025-09-01</dc:date>
          <dc:description> 

Diese Curation Policy legt die Kriterien und Verfahren für die Aufnahme, Pflege und langfristige Bereitstellung von Forschungsdaten in der HZSK-Community des Repositoriums der Universität Hamburg fest. Ziel ist eine nachhaltige, interoperable und nachvollziehbare Archivierung sowie die Unterstützung der Forschungscommunity in der Sprachwissenschaft.</dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/17950</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.17950</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:17950</dc:identifier>
          <dc:language>deu</dc:language>
          <dc:relation>doi:10.25592/uhhfdm.17949</dc:relation>
          <dc:rights>info:eu-repo/semantics/openAccess</dc:rights>
          <dc:rights>https://creativecommons.org/licenses/by/4.0/legalcode</dc:rights>
          <dc:subject>Forschungsdaten</dc:subject>
          <dc:subject>Datenkuratierung</dc:subject>
          <dc:subject>HZSK</dc:subject>
          <dc:subject>Nachhaltigkeit</dc:subject>
          <dc:subject>Metadaten</dc:subject>
          <dc:subject>Digital Preservation</dc:subject>
          <dc:subject>Data Curation</dc:subject>
          <dc:title>Curation Policy für das Forschungsdatenrepositorium der Universität Hamburg – Community des Hamburger Zentrums für Sprachkorpora (HZSK)</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>publication-other</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:1350</identifier>
        <datestamp>2021-09-20T11:45:09Z</datestamp>
        <setSpec>user-hzsk</setSpec>
        <setSpec>user-uhh</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:creator>Kupisch, Tanja</dc:creator>
          <dc:date>2009-07-01</dc:date>
          <dc:description>This version of the corpus is deprecated and missing files. Please refer to the 0.2 version available here: https://doi.org/10.25592/uhhfdm.1351</dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/1350</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.1350</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:1350</dc:identifier>
          <dc:language>fra</dc:language>
          <dc:relation>doi:10.25592/uhhfdm.1349</dc:relation>
          <dc:rights>info:eu-repo/semantics/restrictedAccess</dc:rights>
          <dc:subject>German</dc:subject>
          <dc:subject>French</dc:subject>
          <dc:subject>Italian</dc:subject>
          <dc:subject>Linguistics</dc:subject>
          <dc:subject>cross-sectional data</dc:subject>
          <dc:subject>simultaneous bilingualism</dc:subject>
          <dc:subject>successive bilingualism</dc:subject>
          <dc:subject>L2 data</dc:subject>
          <dc:subject>language attrition</dc:subject>
          <dc:subject>adult bilingualism</dc:subject>
          <dc:subject>EXMARaLDA</dc:subject>
          <dc:title>Hamburg Adult Bilingual LAnguage (HABLA)</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:1438</identifier>
        <datestamp>2021-05-26T07:49:36Z</datestamp>
        <setSpec>user-hzsk</setSpec>
        <setSpec>user-uhh</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:creator>Gabriel, Christoph</dc:creator>
          <dc:date>2011-06-30</dc:date>
          <dc:description>Audio and video recordings of experimental/read and spontaneous speech from adult speakers of Porteño Spanish in Argentina. Speakers are 18-69 years old and from two geographic areas. For the intonational experiments, there are audio recordings only, whereas some of the free interviews and map tasks feature video recordings. The material used as stimuli in the experiments is available with references encoded in the transcriptions.

The Hamburg Corpus of Argentinean Spanish (HaCASpa) was compiled in December 2008 and November/December 2009 within the context of the research project The intonation of Spanish in Argentina (H9, director: Christoph Gabriel), part of the Collaborative Research Centre "Multilingalism", funded by the German Research Foundation (Deutsche Forschungsgemeinschaft, DFG) and hosted by the University of Hamburg. It comprises data from two varieties of Argentinean Spanish, i.e. a) the dialect spoken in the capital of Buenos Aires (also called Porteño, derived from puerto 'harbor') and b) the variety of the Neuquén/Comahue area (Northern Patagonia). The seven parts of HaCASpa correspond to the seven tasks described below in more detail:

Five experiments were carried out in order to elicit specific data for research in prosody, with a main focus on (Task 1–5); in addition, several speakers took part in a free interview (Task 6) and a map task experiment (Task 7). The Task is encoded as a metadata attribute for each communication. HaCASpa comprises three different types of spoken data, depending on the Task, i.e. spontaneous, semi-spontaneous, and scripted speech. This information corresponds to the metadata attribute Speech type.

The regional dimension of the corpus is represented through the attribute Area (i.e. Buenos Aires or Neuquén/Comahue), its diachronic dimension through the attribute Age group (i.e. Under 25/Over 25). The subjects are 60 native speakers of the relevant variety of Argentinean Spanish, i.e. Buenos Aires (Porteño) or Nequén/Comahue Spanish. For each speaker, the following information is available: Age, Education, Occupation, Year of school enrollment, Year of school graduation and Parents' mother tongue.

The current version 0.2 contains mainly orthographic transcriptions of verbal behaviour (141,000 transcribed words) and codes that relate utterances to the materials used for the experimental tasks. Experimental design:

Task (1) consists of two subparts: reading a story (1a) and retelling it (1b). For (1a), the subjects were asked to read the short story "The North Wind and the Sun", which was presented on a computer screen, two times. The fable is well known for its use of phonetic descriptions of different languages (see Handbook of the International Phonetic Association, International Phonetic Association. Cambridge: Cambridge University Press, 2005); the Latin American version we used in our data stems from the Dialectoteca del español, (coordination: C.-E. Piñeros). For (1b), the speakers were instructed to retell the story in their own words without being able to consult the text. With the help of these two parts, data of scripted (part 1a) as well as of semi-spontaneous speech (part 1b) could be collected.

Task (2) was designed to collect data of semi-spontaneous speech by asking the subjects to answer questions pertaining to a given picture story. In a first step, the speakers were familiarized with the story, which was presented as two pictures displayed on a computer screen. In a second step, they were asked to answer specific questions about the story. The questions were also presented on the computer screen and varied in their design in order to elicit answers with different information-structural readings (such as broad vs. narrow focus or different focus types). In general, the speakers were free to answer as they wished. However, in order to avoid single word answers, they were asked to utter complete sentences.

Task (3) consisted of reading question-answer pairs, the content of which was based on the picture stories already familiar from task (2). The answers were given together with the questions on the computer screen (i.e. one question / one answer) and the speakers simply had to read both the question and the answer.

Task (4) was a reading task in which the subjects were asked to utter 10 simple subject-verb-object (SVO) sentences, presented on a computer screen. The speakers were instructed to read them at both normal and fast speech rate. Along the lines proposed in D´Imperio et al. 2005 ("Intonational Phrasing in Romance: The Role of Syntactic and Prosodic Structure", in: Prosodies: With Special Reference to Iberian Languages, ed. by Frota, S. et al., Berlin: Mouton de Gruyter, 59-97), the subject and object constituents differed in their syntactic and prosodic complexity (e.g. determiner plus noun or determiner plus noun plus adjective and one or three prosodic words, respectively). The participants were instructed to read the sentences as if they contained new information. The complete experiment design is described in Gabriel, C. et al. 2011 ("Prosodic phrasing in Porteño Spanish", in: Intonational Phrasing in Romance and Germanic: Cross-Linguistic and Bilingual Studies, ed. by Gabriel, C. &amp; Lleó, C., Amsterdam: Benjamins, 153-182).

Task (5), the so-called intonation survey, consisted of 48 situations designed to elicit various intonational contours with specific pragmatic meanings. In this inductive method, the researcher confronts the speaker with a series of hypothetical situations to which he or she is supposed to react verbally. In the Argentinean version of the questionnaire, the hypothetical situations were illustrated by appropriate pictures. The experimental design is described in more detail in Prieto, P. &amp; Roseano, P. 2010 (eds). Transcription of Intonation of the Spanish Language. Munich: Lincom; see also the Interactive atlas of Spanish intonation (coordination: P. Prieto &amp; P. Roseano).

Task (6) was conducted to collect spontaneous speech data by conducting free interviews. In this task, the subjects were asked to tell the interviewer something about a past experience, be it a vacation or memories of Argentina as it was decades ago. Even though the interviewer was still part of the conversation, it was mainly the subjects who spoke during the recordings.

Task (7) consists of Map Task dialogs. Map Task is a technique employed to collect data of spontaneous speech in which two subjects cooperate to complete a specified task. It is designed to lead the subjects to produce particular interrogative patterns. Each of the two subjects receives a map of an imaginary town marked with buildings and other specific elements. A route is marked on the map of one of the two participants, who assumes the role of the instruction-giver. The version of the same map given to the other participant, who assumes the role of the instruction-follower, differs from that of the instruction-giver in that it does not show the route to be followed. The instruction-follower therefore must ask the instruction-giver questions in order to be able to reproduce the same route on his or her own map (see also the Interactive atlas of Spanish intonation).

 

CLARIN Metadata summary for Hamburg Corpus of Argentinean Spanish (HaCASpa) (CMDI-based)

Title: Hamburg Corpus of Argentinean Spanish (HaCASpa)
Description: Audio and video recordings of experimental/read and spontaneous speech from adult speakers of Porteño Spanish in Argentina. Speakers are 18-69 years old and from two geographic areas. For the intonational experiments, there are audio recordings only, whereas some of the free interviews and map tasks feature video recordings. The material used as stimuli in the experiments is available with references encoded in the transcriptions.
Publication date: 2011-06-30
Data owner:  Christoph Gabriel, Institut für Romanistik / Von-Melle-Park 6 / D-20146 Hamburg, christoph.gabriel@uni-hamburg.de
Contributors:  Christoph Gabriel, Institut für Romanistik / Von-Melle-Park 6 / D-20146 Hamburg, christoph.gabriel@uni-hamburg.de (compiler)
Project:  H9 "The intonation of Spanish in Argentina", German Research Foundation (DFG)
Keywords:  contact variety, cross-sectional data, regional variety, language contact, EXMARaLDA
Language:  Spanish (spa)
Size:  63 speakers (39 female, 24 male), 259 communications, 261 recordings, 1119 minutes, 261 transcriptions, 141321 words
Annotation types:  transcription (manual): mainly orthographic, project-specific conventions, code: reference to underlying prompts
Temporal Coverage:  2008-11-01/2009-12-01
Spatial Coverage:  Buenos Aires, AR; Neuquén/Comahue, AR
Genre:  discourse
Modality:  spoken

 </dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/1438</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.1438</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:1438</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.1437</dc:relation>
          <dc:rights>info:eu-repo/semantics/restrictedAccess</dc:rights>
          <dc:subject>contact variety</dc:subject>
          <dc:subject>cross-sectional data</dc:subject>
          <dc:subject>regional variety</dc:subject>
          <dc:subject>language contact</dc:subject>
          <dc:subject>EXMARaLDA</dc:subject>
          <dc:subject>linguistics</dc:subject>
          <dc:subject>Spanish</dc:subject>
          <dc:title>Hamburg Corpus of Argentinean Spanish (HaCASpa)</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:1452</identifier>
        <datestamp>2020-08-25T07:03:29Z</datestamp>
        <setSpec>user-hzsk</setSpec>
        <setSpec>user-uhh</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:creator>Lleó, Conxita</dc:creator>
          <dc:date>2020-08-24</dc:date>
          <dc:description>Additional data from the PAIDUS project.</dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/1452</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.1452</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:1452</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.1451</dc:relation>
          <dc:rights>info:eu-repo/semantics/closedAccess</dc:rights>
          <dc:subject>German</dc:subject>
          <dc:subject>Spanish</dc:subject>
          <dc:subject>Linguistics</dc:subject>
          <dc:subject>Language Acquisition</dc:subject>
          <dc:subject>Bilingualism</dc:subject>
          <dc:title>PAIDUS Media</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:1456</identifier>
        <datestamp>2020-08-25T20:08:37Z</datestamp>
        <setSpec>user-hzsk</setSpec>
        <setSpec>user-uhh</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:creator>Lleó, Conxita</dc:creator>
          <dc:date>2020-08-25</dc:date>
          <dc:description>Additional data from the PhonBLA Longitudinalstudie Hamburg project.</dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/1456</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.1456</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:1456</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.1455</dc:relation>
          <dc:rights>info:eu-repo/semantics/closedAccess</dc:rights>
          <dc:subject>German</dc:subject>
          <dc:subject>Spanish</dc:subject>
          <dc:subject>Linguistics</dc:subject>
          <dc:subject>Language Acquisition</dc:subject>
          <dc:subject>Bilingualism</dc:subject>
          <dc:title>PhonBLA Longitudinalstudie Hamburg Media</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:1462</identifier>
        <datestamp>2020-12-03T10:43:02Z</datestamp>
        <setSpec>user-hzsk</setSpec>
        <setSpec>user-uhh</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:creator>Meisel, Jürgen M.</dc:creator>
          <dc:date>2020-08-25</dc:date>
          <dc:description>Sub-corpus of the ZISA project with one Italian and one Portuguese learner. The ZISA project contains audio recordings of five adult learners of German as an L2 with L1s Spanish, Italian and Portuguese. Recording sessions (interview/conversation) in German once or twice a month over approx. two years starting 3-14 weeks after their arrival in Germany. Only transcripts and metadata are available. Orthographic transcription according to project internal conventions.</dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/1462</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.1462</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:1462</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.1461</dc:relation>
          <dc:rights>info:eu-repo/semantics/restrictedAccess</dc:rights>
          <dc:subject>German</dc:subject>
          <dc:subject>Portuguese</dc:subject>
          <dc:subject>Italian</dc:subject>
          <dc:subject>linguistics</dc:subject>
          <dc:subject>language acquisition</dc:subject>
          <dc:subject>bilingualism</dc:subject>
          <dc:subject>learner corpus</dc:subject>
          <dc:subject>adult L2 acquisition</dc:subject>
          <dc:subject>adult bilingualism</dc:subject>
          <dc:subject>longitudinal data</dc:subject>
          <dc:subject>successive bilingualism</dc:subject>
          <dc:subject>l2 data</dc:subject>
          <dc:subject>EXMARaLDA</dc:subject>
          <dc:title>ZISA_BR_ZI</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:1468</identifier>
        <datestamp>2020-11-10T20:10:56Z</datestamp>
        <setSpec>user-hzsk</setSpec>
        <setSpec>user-uhh</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:creator>Rehbein, Jochen</dc:creator>
          <dc:creator>Herkenrath, Annette</dc:creator>
          <dc:creator>Karakoç, Birsel</dc:creator>
          <dc:date>2020-08-27</dc:date>
          <dc:description>Subcorpus of the ENDFAS/SKOBI Korpus.</dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/1468</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.1468</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:1468</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.1467</dc:relation>
          <dc:rights>info:eu-repo/semantics/restrictedAccess</dc:rights>
          <dc:subject>German</dc:subject>
          <dc:subject>Turkish</dc:subject>
          <dc:subject>Linguistics</dc:subject>
          <dc:subject>Language Acquisition</dc:subject>
          <dc:subject>Bilingualism</dc:subject>
          <dc:title>ENDFAS_SKOBI_Gold</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:1500</identifier>
        <datestamp>2021-05-27T11:10:23Z</datestamp>
        <setSpec>user-hzsk</setSpec>
        <setSpec>user-uhh</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:creator>Lleó, Conxita</dc:creator>
          <dc:date>1993-01-01</dc:date>
          <dc:description>Audio recordings of prompted, read and spontaneous speech data from L1 Catalan speakers from Barcelona. The data is stratified according to three different city districts and three age groups. Speakers' age vary from approx. 5 to 45 years.

The Phonoprosodic corpus of spoken Catalan (PhonCAT) was compiled between July 2006 and June 2011 in the research project Phonoprosodic development of Catalan in its current bilingual context (H6, director: Conxita Lleò), part of the Collaborative Research Centre "Multilingualism", funded by the German Research Foundation (Deutsche Forschungsgemeinschaft, DFG) and hosted by the University of Hamburg. It comprises read and elicited as well as spontaneous speech data from speakers of Catalan in Barcelona. The speakers were divided into nine groups according to their age and to the district in Barcelona in which they live. For most recordings, selected words were transcribed phonetically by one to three transcribers and annotated with their orthographic target form. A smaller set of recordings was transcribed orthographically.

Benet, Ariadna, Susana Cortés, and Conxita Lleó. 2012. "Phonoprosodic Corpus of Spoken Catalan (PhonCAT)." In Multilingual Corpora and Multilingual Corpus Analysis, edited by Thomas Schmidt and Kai Wörner, pp. 215–29. Hamburg Studies in Multilingualism, 14. John Benjamins.

 

CLARIN Metadata summary for Catalan in a bilingual context (PhonCAT) (CMDI-based)

Title: Catalan in a bilingual context (PhonCAT)
Description: Audio recordings of prompted, read and spontaneous speech data from L1 Catalan speakers from Barcelona. The data is stratified according to three different city districts and three age groups. Speakers' age vary from approx. 5 to 45 years.
Data owner:  Conxita Lleó, Institut für Romanistik / Von-Melle-Park 6 / D-20146 Hamburg, lleo@uni-hamburg.de
Contributors:  Conxita Lleó, Institut für Romanistik / Von-Melle-Park 6 / D-20146 Hamburg, lleo@uni-hamburg.de (compiler)
Project:  H6 "Phonoprosodic development of Catalan in its current bilingual context", German Research Foundation (DFG)
Keywords:  adult bilingualism, bilingual society, cross-sectional data, simultaneous bilingualism, successive bilingualism, child L2 acquisition, L1 data, child bilingualism, language contact, EXMARaLDA
Language:  Catalan (cat)
Size:  234 speakers (0 female, 234 male), 225 communications, 205 recordings, 8719 minutes, 875 transcriptions, 187967 words
Genre:  discourse
Modality:  spoken
References:  Benet, Ariadna, Susana Cortés, and Conxita Lleó. 2012. "Phonoprosodic Corpus of Spoken Catalan (PhonCAT)." In Multilingual Corpora and Multilingual Corpus Analysis, edited by Thomas Schmidt and Kai Wörner, pp. 215–29. Hamburg Studies in Multilingualism, 14. John Benjamins.

 </dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/1500</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.1500</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:1500</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.1499</dc:relation>
          <dc:rights>info:eu-repo/semantics/restrictedAccess</dc:rights>
          <dc:subject>adult bilingualism</dc:subject>
          <dc:subject>bilingual society</dc:subject>
          <dc:subject>cross-sectional data</dc:subject>
          <dc:subject>simultaneous bilingualism</dc:subject>
          <dc:subject>successive bilingualism</dc:subject>
          <dc:subject>child L2 acquisition</dc:subject>
          <dc:subject>L1 data</dc:subject>
          <dc:subject>child bilingualism</dc:subject>
          <dc:subject>language contact</dc:subject>
          <dc:subject>EXMARaLDA</dc:subject>
          <dc:subject>linguistics</dc:subject>
          <dc:subject>Catalan</dc:subject>
          <dc:title>Catalan in a bilingual context (PhonCAT)</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:1530</identifier>
        <datestamp>2021-05-27T11:10:24Z</datestamp>
        <setSpec>user-hzsk</setSpec>
        <setSpec>user-uhh</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:creator>Lleó, Conxita</dc:creator>
          <dc:date>2020-09-07</dc:date>
          <dc:description>Audio recordings in Spanish with 23 German/Spanish simultaneous bilingual children living in Germany and attending the Spanish complementary school at the first level. 1-6 recordings with each child, with 11 children also before the children attended the Spanish complementary school. All recordings feature elicited speech: A picture naming task, a story telling task, a morphosyntactic test, a lexical test, and the HAVAS 5. Rich metadata on language use and attitudes in the family submitted by the parents.

ALCEBLA is a phonetically and orthographically transcribed corpus of German and Spanish, finalized in the project Research based support of the complementary Spanish school in Germany (FUSED) at the Research Center on Multilingualism, University of Hamburg.

Ulloa Saceda, Marta; Lleó, Conxita and García Sánchez, Izarbe (2012): Corpora of spoken Spanish by simultaneous and successive German-Spanish bilingual and Spanish monolingual children. In: Schmidt, Thomas and Wörner, Kai (eds.): Multilingual Corpora and Multilingual Corpus Analysis. Hamburg Studies in Multilingualism (14), 97-106. Amsterdam: John Benjamins.

 

CLARIN Metadata summary for ALCEBLA (CMDI-based)

Title: ALCEBLA
Description: Audio recordings in Spanish with 23 German/Spanish simultaneous bilingual children living in Germany and attending the Spanish complementary school at the first level. 1-6 recordings with each child, with 11 children also before the children attended the Spanish complementary school. All recordings feature elicited speech: A picture naming task, a story telling task, a morphosyntactic test, a lexical test, and the HAVAS 5. Rich metadata on language use and attitudes in the family submitted by the parents.
Publication date: 2011-06-30
Data owner:  Conxita Lleó, lleo@uni-hamburg.de
Contributors:  Conxita Lleó, Institut für Romanistik / Von-Melle-Park 6 / D-20146 Hamburg, lleo@uni-hamburg.de (compiler)
Project:  T4 "Research based support of the complementary Spanish school in Germany (FUSED)", German Research Foundation (DFG)
Keywords:  child language acquisition, child bilingualism, simultaneous bilingualism, EXMARaLDA
Languages:  German (deu), Spanish (spa)
Size:  23 speakers (14 female, 9 male), 66 communications, 64 recordings, 2122 minutes, 66 transcriptions, 36717 words
Spatial Coverage:  DE
Genre:  discourse
Modality:  spoken
References:  Ulloa Saceda, Marta; Lleó, Conxita and García Sánchez, Izarbe (2012): Corpora of spoken Spanish by simultaneous and successive German-Spanish bilingual and Spanish monolingual children. In: Schmidt, Thomas and Wörner, Kai (eds.): Multilingual Corpora and Multilingual Corpus Analysis. Hamburg Studies in Multilingualism (14), 97-106. Amsterdam: John Benjamins.

 </dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/1530</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.1530</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:1530</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.1529</dc:relation>
          <dc:rights>info:eu-repo/semantics/restrictedAccess</dc:rights>
          <dc:subject>child language acquisition</dc:subject>
          <dc:subject>child bilingualism</dc:subject>
          <dc:subject>simultaneous bilingualism</dc:subject>
          <dc:subject>EXMARaLDA</dc:subject>
          <dc:subject>linguistics</dc:subject>
          <dc:subject>German</dc:subject>
          <dc:subject>Spanish</dc:subject>
          <dc:title>ALCEBLA</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:1581</identifier>
        <datestamp>2020-09-14T12:28:51Z</datestamp>
        <setSpec>user-hzsk</setSpec>
        <setSpec>user-uhh</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:creator>Zimmermann, Malte</dc:creator>
          <dc:date>2015-09-14</dc:date>
          <dc:description>This corpus of news articles from the online news service of Deutsche Welle contains 4 texts with a total of 2017 tokens.

 

CLARIN Metadata summary for A5 Hausa News (CMDI-based)

Title: A5 Hausa News
Description: This corpus of news articles from the online news service of Deutsche Welle contains 4 texts with a total of 2017 tokens.
Publication date: 2015
Data owner:  Sonderforschungsbereich 632 / Institut für Linguistik, Universität Potsdam
Contributors:  Malte Zimmermann (editor), Thuan Tran (researcher), Agata Renans (researcher)
Project:  Special Research Centre 632 Information structure, German Research Foundation
Language:  Hausa (hau)
Size:  2017 Token
Segmentation units:  other
Genre:  news articles
Modality:  unknown

 </dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/1581</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.1581</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:1581</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.1580</dc:relation>
          <dc:rights>info:eu-repo/semantics/openAccess</dc:rights>
          <dc:rights>https://creativecommons.org/licenses/by-nc/3.0/legalcode</dc:rights>
          <dc:subject>linguistics</dc:subject>
          <dc:subject>information structure</dc:subject>
          <dc:subject>Hausa</dc:subject>
          <dc:title>A5 Hausa News</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:1583</identifier>
        <datestamp>2020-09-14T12:26:19Z</datestamp>
        <setSpec>user-hzsk</setSpec>
        <setSpec>user-uhh</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:creator>Zimmermann, Malte</dc:creator>
          <dc:date>2015-09-14</dc:date>
          <dc:description>This corpus of Umarnin Uwa film transcripts contains 47 transcripts with a total of 10194 tokens. It provides information including automatic POS tagging, speaker and extralinguistic information, foreign words and code-switching.

 

CLARIN Metadata summary for A5 Hausa Umarnin Uwa (CMDI-based)

Title: A5 Hausa Umarnin Uwa
Description: This corpus of Umarnin Uwa film transcripts contains 47 transcripts with a total of 10194 tokens. It provides information including automatic POS tagging, speaker and extralinguistic information, foreign words and code-switching.
Publication date: 2015
Data owner:  Sonderforschungsbereich 632 / Institut für Linguistik, Universität Potsdam
Contributors:  Malte Zimmermann (editor), Thuan Tran (researcher), Agata Renans (researcher)
Project:  Special Research Centre 632 Information structure, German Research Foundation
Language:  Hausa (hau)
Size:  10194 Token
Segmentation units:  other
Genre:  discourse
Modality:  spoken

 </dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/1583</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.1583</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:1583</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.1582</dc:relation>
          <dc:rights>info:eu-repo/semantics/openAccess</dc:rights>
          <dc:rights>https://creativecommons.org/licenses/by-nc/3.0/legalcode</dc:rights>
          <dc:subject>linguistics</dc:subject>
          <dc:subject>information structure</dc:subject>
          <dc:subject>Hausa</dc:subject>
          <dc:title>A5 Hausa Umarnin Uwa</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:1064</identifier>
        <datestamp>2020-09-10T08:48:01Z</datestamp>
        <setSpec>user-hzsk</setSpec>
        <setSpec>user-uhh</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:creator>Meisel, Jürgen M.</dc:creator>
          <dc:date>2020-06-22</dc:date>
          <dc:description>Data from the CHILD-L2 project.</dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/1064</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.1064</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:1064</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.972</dc:relation>
          <dc:relation>doi:10.25592/uhhfdm.1063</dc:relation>
          <dc:rights>info:eu-repo/semantics/closedAccess</dc:rights>
          <dc:subject>German</dc:subject>
          <dc:subject>French</dc:subject>
          <dc:subject>Linguistics</dc:subject>
          <dc:subject>Language Acquisition</dc:subject>
          <dc:subject>Bilingualism</dc:subject>
          <dc:title>Child-L2</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:1697</identifier>
        <datestamp>2022-07-15T11:46:06Z</datestamp>
        <setSpec>user-hzsk</setSpec>
        <setSpec>user-uhh</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:contributor>Birr, Katharina</dc:contributor>
          <dc:contributor>Nagel, Norbert</dc:contributor>
          <dc:contributor>Mirahmadi, Hamasa</dc:contributor>
          <dc:contributor>Ihden, Sarah</dc:contributor>
          <dc:contributor>Nasielski, Vanessa</dc:contributor>
          <dc:contributor>Dücker, Lisa</dc:contributor>
          <dc:contributor>Tews, Sebastian</dc:contributor>
          <dc:contributor>Wallmeier, Nadine</dc:contributor>
          <dc:contributor>Hübener, Carlotta</dc:contributor>
          <dc:contributor>Recker, Anabel</dc:contributor>
          <dc:contributor>Peterberns, Hellen</dc:contributor>
          <dc:contributor>Dreessen, Katharina</dc:contributor>
          <dc:contributor>Bischoff, Annemarie</dc:contributor>
          <dc:contributor>Wollenschläger, Anna</dc:contributor>
          <dc:contributor>Schmitt, Eleonore</dc:contributor>
          <dc:contributor>Liebl, Mattias</dc:contributor>
          <dc:contributor>Schilling, Elmar</dc:contributor>
          <dc:contributor>Schroeder, Meile-Andrea</dc:contributor>
          <dc:contributor>Lehmberg, Timm</dc:contributor>
          <dc:contributor>Kahre, Paul</dc:contributor>
          <dc:contributor>Kleymann, Verena</dc:contributor>
          <dc:contributor>Sandmann, Lena</dc:contributor>
          <dc:contributor>Sturm, John</dc:contributor>
          <dc:contributor>Matthies, Sarah</dc:contributor>
          <dc:contributor>Sluyter-Gäthje, Henny</dc:contributor>
          <dc:contributor>Tran, Ilka</dc:contributor>
          <dc:contributor>Hütter, Julia</dc:contributor>
          <dc:contributor>Barteld, Fabian</dc:contributor>
          <dc:contributor>Schröder, Katharina</dc:contributor>
          <dc:contributor>Eichhorn-Hartmeyer, Christina</dc:contributor>
          <dc:contributor>Tiedemann, Meike</dc:contributor>
          <dc:contributor>Schnee, Lena</dc:contributor>
          <dc:contributor>Schröder, Ingrid</dc:contributor>
          <dc:contributor>Peters,  Robert</dc:contributor>
          <dc:creator>ReN-Team</dc:creator>
          <dc:date>2019-08-14</dc:date>
          <dc:description>The project "Reference Corpus Middle Low German/Low Rhenish (1200–1650)", abbreviated as "ReN", is part of the "Corpus of Historical German Texts", which includes the projects "Old German Reference Corpus (750–1050)", the "Reference Corpus Middle High German (1050–1350)" and the "Reference Corpus EarlyModern New High German (1350–1650)". The project "Reference Corpus Middle Low German/Low Rhenish (1200–1650)" deals with a structured selection of Middle Low German and Low Rhenish monuments of speech from 1200 to 1650. This selection is based on the parameters "space", "time" and "field of writing".

The reference corpus of Middle Low German and Low Rhenish texts is based on manuscripts, prints and inscriptions. It is intended to provide an insight into the culture of speech and writing in Middle Low German and Low Rhenish regions. This spectrum of texttypes can be used to trace the linguistic development on the base of diatopic and diacronic subcategorisation. The aim of the project is the publication of diplomatic transcribed, lemmatised and grammatically annotated texts. The processed data – especially on the grammatical level – enables a linguistic analysis of the Middle Low German and Low Rhenish language, which goes far beyond what was possible until now.

Das „Referenzkorpus Mittelniederdeutsch/Niederrheinisch (1200–1650)“, kurz „ReN“, ist Teil des „Korpus historischer Texte des Deutschen“, zu welchem außerdem die Referenzkorpora Altdeutsch, Mittelhochdeutsch und Frühneuhochdeutsch zählen. Das ReN umfasst mittelniederdeutsche und niederrheinische Sprachdenkmäler von 1200 bis 1650 in einer strukturierten Auswahl. Diese ergibt sich aus den Parametern „Raum“, „Zeit“ und „Feld der Schriftlichkeit“.

Das Referenzkorpus mittelniederdeutscher und niederrheinischer Texte basiert auf Handschriften, Drucken und Inschriften. Es soll einen Einblick in die Sprach- und Textkultur des niederdeutschen und niederrheinischen Raums geben. Die historische Sprachentwicklung soll in ihrer diatopischen und diachronischen Untergliederung anhand des Textsortenspektrums nachgezeichnet werden. Das Ziel des Projektes besteht in der Veröffentlichung diplomatisch transkribierter, lemmatisierter und grammatisch annotierter Texte. Die so bearbeiteten Daten ermöglichen - insbesondere auf grammatischer Ebene - sprachwissenschaftliche Analysen des Mittelniederdeutschen und Niederrheinischen, die weit über das bisher Mögliche hinausgehen.

Ingrid Schröder. 2014. "Das Referenzkorpus: Neue Perspektiven für die mittelniederdeutsche Grammatikographie." Jahrbuch für Germanistische Sprachgeschichte, 5(1), pp. 150-164. http://www.degruyter.com/view/j/jbgsg.2014.5.issue-1/jbgsg-2014-0011/jbgsg-2014-0011.xml .

Robert Peters, Norbert Nagel. 2014. "Das digitale 'Referenzkorpus Mittelniederdeutsch / Niederrheinisch (ReN)'." Jahrbuch für Germanistische Sprachgeschichte, 5(1), pp. 165-175. https://www.degruyter.com/view/j/jbgsg.2014.5.issue-1/jbgsg-2014-0012/jbgsg-2014-0012.xml .

 

CLARIN Metadata summary for Referenzkorpus Mittelniederdeutsch/Niederrheinisch (1200-1650)Reference Corpus Middle Low German/Low Rhenish (1200-1650) (CMDI-based)

Title: Referenzkorpus Mittelniederdeutsch/Niederrheinisch (1200-1650)
Title: Reference Corpus Middle Low German/Low Rhenish (1200-1650)
Description: The reference corpus of Middle Low German and Low Rhenish texts is based on manuscripts, prints and inscriptions. It is intended to provide an insight into the culture of speech and writing in Middle Low German and Low Rhenish regions. This spectrum of texttypes can be used to trace the linguistic development on the base of diatopic and diacronic subcategorisation. The aim of the project is the publication of diplomatic transcribed, lemmatised and grammatically annotated texts. The processed data – especially on the grammatical level – enables a linguistic analysis of the Middle Low German and Low Rhenish language, which goes far beyond what was possible until now.
Description: Das Referenzkorpus mittelniederdeutscher und niederrheinischer Texte basiert auf Handschriften, Drucken und Inschriften. Es soll einen Einblick in die Sprach- und Textkultur des niederdeutschen und niederrheinischen Raums geben. Die historische Sprachentwicklung soll in ihrer diatopischen und diachronischen Untergliederung anhand des Textsortenspektrums nachgezeichnet werden. Das Ziel des Projektes besteht in der Veröffentlichung diplomatisch transkribierter, lemmatisierter und grammatisch annotierter Texte. Die so bearbeiteten Daten ermöglichen - insbesondere auf grammatischer Ebene - sprachwissenschaftliche Analysen des Mittelniederdeutschen und Niederrheinischen, die weit über das bisher Mögliche hinausgehen.
Data owner:  ReN-Team
Contributors:  Sarah Ihden (transcriber), Sarah Ihden (annotator), Sarah Ihden (researcher), Norbert Nagel (transcriber), Norbert Nagel (researcher), Katharina Dreessen (transcriber), Katharina Dreessen (annotator), Katharina Dreessen (researcher), Annemarie Bischoff (annotator), Annemarie Bischoff (researcher), Meike Tiedemann (transcriber), Meike Tiedemann (annotator), Meike Tiedemann (researcher), Verena Kleymann (transcriber), Verena Kleymann (annotator), Verena Kleymann (researcher), Elmar Schilling (annotator), Elmar Schilling (researcher), Nadine Wallmeier (annotator), Nadine Wallmeier (researcher), Fabian Barteld (developer), Fabian Barteld (researcher), Timm Lehmberg (developer), Timm Lehmberg (researcher), Katharina Birr (annotator), Carlotta Hübener (transcriber), Carlotta Hübener (annotator), Paul Kahre (annotator), Vanessa Nasielski (annotator), Anabel Recker (annotator), Lena Schnee (annotator), Henny Sluyter-Gäthje (developer), Lena Sandmann (developer), Hellen Peterberns (annotator), Sebastian Tews (annotator), Ilka Tran (transcriber), Anna Wollenschläger (transcriber), Lisa Dücker (annotator), Sarah Matthies (transcriber), Hamasa Mirahmadi (transcriber), Eleonore Schmitt (annotator), Christina Eichhorn-Hartmeyer (annotator), Julia Hütter (annotator), Meile-Andrea Schroeder (annotator), Mattias Liebl (transcriber), Katharina Schröder (transcriber), John Sturm (annotator)
Project:  Referenzkorpus Mittelniederdeutsch/ Niederrheinisch (1200-1650), Deutsche Forschungsgemeinschaft (DFG)
Languages:  Middle Low German (gml), Low Rhenish (mis)
Size:  146 Text (annotated), 1415362 tok_anno (annotated), 1450562 tok_dipl (annotated), 89 Text (transcribed), 908682 tok_anno (transcribed), 924860 tok_dipl (transcribed)
Annotation types:  Transcription (manual), Part-of-speech (semi-automatic), Lemma (semi-automatic), Morphology (semi-automatic)
Temporal Coverage:  1200/1650
References:  Ingrid Schröder (2014) Das Referenzkorpus: Neue Perspektiven für die mittelniederdeutsche Grammatikographie References:  Robert Peters; Norbert Nagel (2014) Das digitale 'Referenzkorpus Mittelniederdeutsch / Niederrheinisch (ReN)'

 </dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/1697</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.1697</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:1697</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.1668</dc:relation>
          <dc:rights>info:eu-repo/semantics/openAccess</dc:rights>
          <dc:rights>https://creativecommons.org/licenses/by/4.0/legalcode</dc:rights>
          <dc:subject>Middle Low German</dc:subject>
          <dc:subject>Low Rhenish</dc:subject>
          <dc:subject>ANNIS</dc:subject>
          <dc:subject>linguistics</dc:subject>
          <dc:subject>diacronic</dc:subject>
          <dc:subject>historical</dc:subject>
          <dc:title>Reference Corpus Middle Low German/Low Rhenish (1200–1650); Referenzkorpus Mittelniederdeutsch/Niederrheinisch (1200–1650)</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:1691</identifier>
        <datestamp>2022-07-15T11:46:05Z</datestamp>
        <setSpec>user-hzsk</setSpec>
        <setSpec>user-uhh</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:contributor>Tews, Sebastian</dc:contributor>
          <dc:contributor>Wollenschläger, Anna</dc:contributor>
          <dc:contributor>Ihden, Sarah</dc:contributor>
          <dc:contributor>Hübener, Carlotta</dc:contributor>
          <dc:contributor>Kleymann, Verena</dc:contributor>
          <dc:contributor>Sluyter-Gäthje, Henny</dc:contributor>
          <dc:contributor>Mirahmadi, Hamasa</dc:contributor>
          <dc:contributor>Liebl, Mattias</dc:contributor>
          <dc:contributor>Tran, Ilka</dc:contributor>
          <dc:contributor>Kahre, Paul</dc:contributor>
          <dc:contributor>Schilling, Elmar</dc:contributor>
          <dc:contributor>Nagel, Norbert</dc:contributor>
          <dc:contributor>Hütter, Julia</dc:contributor>
          <dc:contributor>Schröder, Katharina</dc:contributor>
          <dc:contributor>Schroeder, Meile-Andrea</dc:contributor>
          <dc:contributor>Barteld, Fabian</dc:contributor>
          <dc:contributor>Wallmeier, Nadine</dc:contributor>
          <dc:contributor>Schnee, Lena</dc:contributor>
          <dc:contributor>Recker, Anabel</dc:contributor>
          <dc:contributor>Dücker, Lisa</dc:contributor>
          <dc:contributor>Matthies, Sarah</dc:contributor>
          <dc:contributor>Lehmberg, Timm</dc:contributor>
          <dc:contributor>Schmitt, Eleonore</dc:contributor>
          <dc:contributor>Eichhorn-Hartmeyer, Christina</dc:contributor>
          <dc:contributor>Bischoff, Annemarie</dc:contributor>
          <dc:contributor>Birr, Katharina</dc:contributor>
          <dc:contributor>Dreessen, Katharina</dc:contributor>
          <dc:contributor>Nasielski, Vanessa</dc:contributor>
          <dc:contributor>Sturm, John</dc:contributor>
          <dc:contributor>Tiedemann, Meike</dc:contributor>
          <dc:creator>ReN-Team</dc:creator>
          <dc:date>2018-03-07</dc:date>
          <dc:description>The project "Reference Corpus Middle Low German/Low Rhenish (1200–1650)", abbreviated as "ReN", is part of the "Corpus of Historical German Texts", which includes the projects "Old German Reference Corpus (750–1050)", the "Reference Corpus Middle High German (1050–1350)" and the "Reference Corpus EarlyModern New High German (1350–1650)". The project "Reference Corpus Middle Low German/Low Rhenish (1200–1650)" deals with a structured selection of Middle Low German and Low Rhenish monuments of speech from 1200 to 1650. This selection is based on the parameters "space", "time" and "field of writing".

The reference corpus of Middle Low German and Low Rhenish texts is based on manuscripts, prints and inscriptions. It is intended to provide an insight into the culture of speech and writing in Middle Low German and Low Rhenish regions. This spectrum of texttypes can be used to trace the linguistic development on the base of diatopic and diacronic subcategorisation. The aim of the project is the publication of diplomatic transcribed, lemmatised and grammatically annotated texts. The processed data – especially on the grammatical level – enables a linguistic analysis of the Middle Low German and Low Rhenish language, which goes far beyond what was possible until now.

Das Referenzkorpus mittelniederdeutscher und niederrheinischer Texte basiert auf Handschriften, Drucken und Inschriften. Es soll einen Einblick in die Sprach- und Textkultur des niederdeutschen und niederrheinischen Raums geben. Die historische Sprachentwicklung soll in ihrer diatopischen und diachronischen Untergliederung anhand des Textsortenspektrums nachgezeichnet werden. Das Ziel des Projektes besteht in der Veröffentlichung diplomatisch transkribierter, lemmatisierter und grammatisch annotierter Texte. Die so bearbeiteten Daten ermöglichen - insbesondere auf grammatischer Ebene - sprachwissenschaftliche Analysen des Mittelniederdeutschen und Niederrheinischen, die weit über das bisher Mögliche hinausgehen.

Ingrid Schröder. 2014. "Das Referenzkorpus: Neue Perspektiven für die mittelniederdeutsche Grammatikographie." Jahrbuch für Germanistische Sprachgeschichte, 5(1), pp. 150-164. http://www.degruyter.com/view/j/jbgsg.2014.5.issue-1/jbgsg-2014-0011/jbgsg-2014-0011.xml .

Robert Peters, Norbert Nagel. 2014. "Das digitale 'Referenzkorpus Mittelniederdeutsch / Niederrheinisch (ReN)'." Jahrbuch für Germanistische Sprachgeschichte, 5(1), pp. 165-175. https://www.degruyter.com/view/j/jbgsg.2014.5.issue-1/jbgsg-2014-0012/jbgsg-2014-0012.xml .

 

CLARIN Metadata summary for Referenzkorpus Mittelniederdeutsch/Niederrheinisch (1200-1650)Reference Corpus Middle Low German/Low Rhenish (1200-1650) (CMDI-based)

Title: Referenzkorpus Mittelniederdeutsch/Niederrheinisch (1200-1650)
Title: Reference Corpus Middle Low German/Low Rhenish (1200-1650)
Description: The reference corpus of Middle Low German and Low Rhenish texts is based on manuscripts, prints and inscriptions. It is intended to provide an insight into the culture of speech and writing in Middle Low German and Low Rhenish regions. This spectrum of texttypes can be used to trace the linguistic development on the base of diatopic and diacronic subcategorisation. The aim of the project is the publication of diplomatic transcribed, lemmatised and grammatically annotated texts. The processed data – especially on the grammatical level – enables a linguistic analysis of the Middle Low German and Low Rhenish language, which goes far beyond what was possible until now.
Description: Das Referenzkorpus mittelniederdeutscher und niederrheinischer Texte basiert auf Handschriften, Drucken und Inschriften. Es soll einen Einblick in die Sprach- und Textkultur des niederdeutschen und niederrheinischen Raums geben. Die historische Sprachentwicklung soll in ihrer diatopischen und diachronischen Untergliederung anhand des Textsortenspektrums nachgezeichnet werden. Das Ziel des Projektes besteht in der Veröffentlichung diplomatisch transkribierter, lemmatisierter und grammatisch annotierter Texte. Die so bearbeiteten Daten ermöglichen - insbesondere auf grammatischer Ebene - sprachwissenschaftliche Analysen des Mittelniederdeutschen und Niederrheinischen, die weit über das bisher Mögliche hinausgehen.
Data owner:  ReN-Team
Contributors:  Sarah Ihden (transcriber), Sarah Ihden (annotator), Sarah Ihden (researcher), Norbert Nagel (transcriber), Norbert Nagel (researcher), Katharina Dreessen (transcriber), Katharina Dreessen (annotator), Katharina Dreessen (researcher), Annemarie Bischoff (annotator), Annemarie Bischoff (researcher), Meike Tiedemann (transcriber), Meike Tiedemann (annotator), Meike Tiedemann (researcher), Verena Kleymann (transcriber), Verena Kleymann (annotator), Verena Kleymann (researcher), Elmar Schilling (annotator), Elmar Schilling (researcher), Nadine Wallmeier (annotator), Nadine Wallmeier (researcher), Fabian Barteld (developer), Fabian Barteld (researcher), Timm Lehmberg (developer), Timm Lehmberg (researcher), Katharina Birr (annotator), Carlotta Hübener (transcriber), Carlotta Hübener (annotator), Paul Kahre (annotator), Vanessa Nasielski (annotator), Anabel Recker (annotator), Lena Schnee (annotator), Henny Sluyter-Gäthje (developer), Sebastian Tews (annotator), Ilka Tran (transcriber), Anna Wollenschläger (transcriber), Lisa Dücker (annotator), Sarah Matthies (transcriber), Hamasa Mirahmadi (transcriber), Eleonore Schmitt (annotator), Christina Eichhorn-Hartmeyer (annotator), Julia Hütter (annotator), Meile-Andrea Schroeder (annotator), Mattias Liebl (transcriber), Katharina Schröder (transcriber), John Sturm (annotator)
Project:  Referenzkorpus Mittelniederdeutsch/ Niederrheinisch (1200-1650), Deutsche Forschungsgemeinschaft (DFG)
Languages:  Middle Low German (gml), Low Rhenish (mis)
Size:  50 Text, 339664 tok_anno, 346793 tok_dipl
Annotation types:  Transcription (manual), Part-of-speech (semi-automatic), Lemma (semi-automatic), Morphology (semi-automatic)
Temporal Coverage:  1200/1650
Genre:  various
Modality:  written
References:  Ingrid Schröder (2014) Das Referenzkorpus: Neue Perspektiven für die mittelniederdeutsche Grammatikographie References:  Robert Peters; Norbert Nagel (2014) Das digitale 'Referenzkorpus Mittelniederdeutsch / Niederrheinisch (ReN)'

 </dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/1691</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.1691</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:1691</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.1668</dc:relation>
          <dc:rights>info:eu-repo/semantics/openAccess</dc:rights>
          <dc:rights>https://creativecommons.org/licenses/by/4.0/legalcode</dc:rights>
          <dc:subject>Middle Low German</dc:subject>
          <dc:subject>Low Rhenish</dc:subject>
          <dc:subject>ANNIS</dc:subject>
          <dc:subject>linguistics</dc:subject>
          <dc:subject>diacronic</dc:subject>
          <dc:subject>historical</dc:subject>
          <dc:title>Reference Corpus Middle Low German/Low Rhenish (1200-1650); Referenzkorpus Mittelniederdeutsch/Niederrheinisch (1200-1650)</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:1690</identifier>
        <datestamp>2022-07-15T11:46:05Z</datestamp>
        <setSpec>user-hzsk</setSpec>
        <setSpec>user-uhh</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:contributor>Schroeder, Meile-Andrea</dc:contributor>
          <dc:contributor>Wallmeier, Nadine</dc:contributor>
          <dc:contributor>Lehmberg, Timm</dc:contributor>
          <dc:contributor>Schröder, Katharina</dc:contributor>
          <dc:contributor>Hübener, Carlotta</dc:contributor>
          <dc:contributor>Tews, Sebastian</dc:contributor>
          <dc:contributor>Kleymann, Verena</dc:contributor>
          <dc:contributor>Nagel, Norbert</dc:contributor>
          <dc:contributor>Tiedemann, Meike</dc:contributor>
          <dc:contributor>Schilling, Elmar</dc:contributor>
          <dc:contributor>Bischoff, Annemarie</dc:contributor>
          <dc:contributor>Eichhorn-Hartmeyer, Christina</dc:contributor>
          <dc:contributor>Liebl, Mattias</dc:contributor>
          <dc:contributor>Birr, Katharina</dc:contributor>
          <dc:contributor>Ihden, Sarah</dc:contributor>
          <dc:contributor>Barteld, Fabian</dc:contributor>
          <dc:contributor>Dreessen, Katharina</dc:contributor>
          <dc:contributor>Sturm, John</dc:contributor>
          <dc:contributor>Hütter, Julia</dc:contributor>
          <dc:contributor>Mirahmadi, Hamasa</dc:contributor>
          <dc:contributor>Dücker, Lisa</dc:contributor>
          <dc:contributor>Schmitt, Eleonore</dc:contributor>
          <dc:contributor>Sluyter-Gäthje, Henny</dc:contributor>
          <dc:contributor>Tran, Ilka</dc:contributor>
          <dc:contributor>Matthies, Sarah</dc:contributor>
          <dc:contributor>Nasielski, Vanessa</dc:contributor>
          <dc:contributor>Wollenschläger, Anna</dc:contributor>
          <dc:contributor>Kahre, Paul</dc:contributor>
          <dc:contributor>Recker, Anabel</dc:contributor>
          <dc:contributor>Schnee, Lena</dc:contributor>
          <dc:creator>ReN-Team</dc:creator>
          <dc:date>2017-12-06</dc:date>
          <dc:description>The project "Reference Corpus Middle Low German/Low Rhenish (1200–1650)", abbreviated as "ReN", is part of the "Corpus of Historical German Texts", which includes the projects "Old German Reference Corpus (750–1050)", the "Reference Corpus Middle High German (1050–1350)" and the "Reference Corpus EarlyModern New High German (1350–1650)". The project "Reference Corpus Middle Low German/Low Rhenish (1200–1650)" deals with a structured selection of Middle Low German and Low Rhenish monuments of speech from 1200 to 1650. This selection is based on the parameters "space", "time" and "field of writing".

The reference corpus of Middle Low German and Low Rhenish texts is based on manuscripts, prints and inscriptions. It is intended to provide an insight into the culture of speech and writing in Middle Low German and Low Rhenish regions. This spectrum of texttypes can be used to trace the linguistic development on the base of diatopic and diacronic subcategorisation. The aim of the project is the publication of diplomatic transcribed, lemmatised and grammatically annotated texts. The processed data – especially on the grammatical level – enables a linguistic analysis of the Middle Low German and Low Rhenish language, which goes far beyond what was possible until now.

Das Referenzkorpus mittelniederdeutscher und niederrheinischer Texte basiert auf Handschriften, Drucken und Inschriften. Es soll einen Einblick in die Sprach- und Textkultur des niederdeutschen und niederrheinischen Raums geben. Die historische Sprachentwicklung soll in ihrer diatopischen und diachronischen Untergliederung anhand des Textsortenspektrums nachgezeichnet werden. Das Ziel des Projektes besteht in der Veröffentlichung diplomatisch transkribierter, lemmatisierter und grammatisch annotierter Texte. Die so bearbeiteten Daten ermöglichen - insbesondere auf grammatischer Ebene - sprachwissenschaftliche Analysen des Mittelniederdeutschen und Niederrheinischen, die weit über das bisher Mögliche hinausgehen.

Ingrid Schröder. 2014. "Das Referenzkorpus: Neue Perspektiven für die mittelniederdeutsche Grammatikographie." Jahrbuch für Germanistische Sprachgeschichte, 5(1), pp. 150-164. http://www.degruyter.com/view/j/jbgsg.2014.5.issue-1/jbgsg-2014-0011/jbgsg-2014-0011.xml .

Robert Peters, Norbert Nagel. 2014. "Das digitale 'Referenzkorpus Mittelniederdeutsch / Niederrheinisch (ReN)'." Jahrbuch für Germanistische Sprachgeschichte, 5(1), pp. 165-175. https://www.degruyter.com/view/j/jbgsg.2014.5.issue-1/jbgsg-2014-0012/jbgsg-2014-0012.xml .

 

CLARIN Metadata summary for Referenzkorpus Mittelniederdeutsch/Niederrheinisch (1200-1650)Reference Corpus Middle Low German/Low Rhenish (1200-1650) (CMDI-based)

Title: Referenzkorpus Mittelniederdeutsch/Niederrheinisch (1200-1650)
Title: Reference Corpus Middle Low German/Low Rhenish (1200-1650)
Description: The reference corpus of Middle Low German and Low Rhenish texts is based on manuscripts, prints and inscriptions. It is intended to provide an insight into the culture of speech and writing in Middle Low German and Low Rhenish regions. This spectrum of texttypes can be used to trace the linguistic development on the base of diatopic and diacronic subcategorisation. The aim of the project is the publication of diplomatic transcribed, lemmatised and grammatically annotated texts. The processed data – especially on the grammatical level – enables a linguistic analysis of the Middle Low German and Low Rhenish language, which goes far beyond what was possible until now.
Description: Das Referenzkorpus mittelniederdeutscher und niederrheinischer Texte basiert auf Handschriften, Drucken und Inschriften. Es soll einen Einblick in die Sprach- und Textkultur des niederdeutschen und niederrheinischen Raums geben. Die historische Sprachentwicklung soll in ihrer diatopischen und diachronischen Untergliederung anhand des Textsortenspektrums nachgezeichnet werden. Das Ziel des Projektes besteht in der Veröffentlichung diplomatisch transkribierter, lemmatisierter und grammatisch annotierter Texte. Die so bearbeiteten Daten ermöglichen - insbesondere auf grammatischer Ebene - sprachwissenschaftliche Analysen des Mittelniederdeutschen und Niederrheinischen, die weit über das bisher Mögliche hinausgehen.
Data owner:  ReN-Team
Contributors:  Sarah Ihden (transcriber), Sarah Ihden (annotator), Sarah Ihden (researcher), Norbert Nagel (transcriber), Norbert Nagel (researcher), Katharina Dreessen (transcriber), Katharina Dreessen (annotator), Katharina Dreessen (researcher), Annemarie Bischoff (annotator), Annemarie Bischoff (researcher), Meike Tiedemann (transcriber), Meike Tiedemann (annotator), Meike Tiedemann (researcher), Verena Kleymann (transcriber), Verena Kleymann (annotator), Verena Kleymann (researcher), Elmar Schilling (annotator), Elmar Schilling (researcher), Nadine Wallmeier (annotator), Nadine Wallmeier (researcher), Fabian Barteld (developer), Fabian Barteld (researcher), Timm Lehmberg (developer), Timm Lehmberg (researcher), Katharina Birr (annotator), Carlotta Hübener (transcriber), Carlotta Hübener (annotator), Paul Kahre (annotator), Vanessa Nasielski (annotator), Anabel Recker (annotator), Lena Schnee (annotator), Henny Sluyter-Gäthje (developer), Sebastian Tews (annotator), Ilka Tran (transcriber), Anna Wollenschläger (transcriber), Lisa Dücker (annotator), Sarah Matthies (transcriber), Hamasa Mirahmadi (transcriber), Eleonore Schmitt (annotator), Christina Eichhorn-Hartmeyer (annotator), Julia Hütter (annotator), Meile-Andrea Schroeder (annotator), Mattias Liebl (transcriber), Katharina Schröder (transcriber), John Sturm (annotator)
Project:  Referenzkorpus Mittelniederdeutsch/ Niederrheinisch (1200-1650), Deutsche Forschungsgemeinschaft (DFG)
Languages:  Middle Low German (gml), Low Rhenish (mis)
Size:  47 Text, 324785 tok_anno, 332673 tok_dipl
Annotation types:  Transcription (manual), Part-of-speech (semi-automatic), Lemma (semi-automatic), Morphology (semi-automatic)
Temporal Coverage:  1200/1650
Genre:  various
Modality:  written
References:  Ingrid Schröder (2014) Das Referenzkorpus: Neue Perspektiven für die mittelniederdeutsche Grammatikographie References:  Robert Peters; Norbert Nagel (2014) Das digitale 'Referenzkorpus Mittelniederdeutsch / Niederrheinisch (ReN)'

 </dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/1690</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.1690</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:1690</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.1668</dc:relation>
          <dc:rights>info:eu-repo/semantics/openAccess</dc:rights>
          <dc:rights>https://creativecommons.org/licenses/by/4.0/legalcode</dc:rights>
          <dc:subject>Middle Low German</dc:subject>
          <dc:subject>Low Rhenish</dc:subject>
          <dc:subject>ANNIS</dc:subject>
          <dc:subject>linguistics</dc:subject>
          <dc:subject>diacronic</dc:subject>
          <dc:subject>hstorical</dc:subject>
          <dc:title>Reference Corpus Middle Low German/Low Rhenish (1200-1650); Referenzkorpus Mittelniederdeutsch/Niederrheinisch (1200-1650)</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:1694</identifier>
        <datestamp>2022-07-15T11:46:05Z</datestamp>
        <setSpec>user-hzsk</setSpec>
        <setSpec>user-uhh</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:contributor>Hübener, Carlotta</dc:contributor>
          <dc:contributor>Birr, Katharina</dc:contributor>
          <dc:contributor>Tran, Ilka</dc:contributor>
          <dc:contributor>Barteld, Fabian</dc:contributor>
          <dc:contributor>Tews, Sebastian</dc:contributor>
          <dc:contributor>Schilling, Elmar</dc:contributor>
          <dc:contributor>Hütter, Julia</dc:contributor>
          <dc:contributor>Schnee, Lena</dc:contributor>
          <dc:contributor>Wallmeier, Nadine</dc:contributor>
          <dc:contributor>Kleymann, Verena</dc:contributor>
          <dc:contributor>Wollenschläger, Anna</dc:contributor>
          <dc:contributor>Tiedemann, Meike</dc:contributor>
          <dc:contributor>Dücker, Lisa</dc:contributor>
          <dc:contributor>Nagel, Norbert</dc:contributor>
          <dc:contributor>Schröder, Katharina</dc:contributor>
          <dc:contributor>Matthies, Sarah</dc:contributor>
          <dc:contributor>Mirahmadi, Hamasa</dc:contributor>
          <dc:contributor>Ihden, Sarah</dc:contributor>
          <dc:contributor>Eichhorn-Hartmeyer, Christina</dc:contributor>
          <dc:contributor>Sluyter-Gäthje, Henny</dc:contributor>
          <dc:contributor>Liebl, Mattias</dc:contributor>
          <dc:contributor>Bischoff, Annemarie</dc:contributor>
          <dc:contributor>Sturm, John</dc:contributor>
          <dc:contributor>Sandmann, Lena</dc:contributor>
          <dc:contributor>Kahre, Paul</dc:contributor>
          <dc:contributor>Dreessen, Katharina</dc:contributor>
          <dc:contributor>Lehmberg, Timm</dc:contributor>
          <dc:contributor>Peterberns, Hellen</dc:contributor>
          <dc:contributor>Nasielski, Vanessa</dc:contributor>
          <dc:contributor>Schmitt, Eleonore</dc:contributor>
          <dc:contributor>Recker, Anabel</dc:contributor>
          <dc:contributor>Schroeder, Meile-Andrea</dc:contributor>
          <dc:creator>ReN-Team</dc:creator>
          <dc:date>2018-07-23</dc:date>
          <dc:description>The project "Reference Corpus Middle Low German/Low Rhenish (1200–1650)", abbreviated as "ReN", is part of the "Corpus of Historical German Texts", which includes the projects "Old German Reference Corpus (750–1050)", the "Reference Corpus Middle High German (1050–1350)" and the "Reference Corpus EarlyModern New High German (1350–1650)". The project "Reference Corpus Middle Low German/Low Rhenish (1200–1650)" deals with a structured selection of Middle Low German and Low Rhenish monuments of speech from 1200 to 1650. This selection is based on the parameters "space", "time" and "field of writing".

The reference corpus of Middle Low German and Low Rhenish texts is based on manuscripts, prints and inscriptions. It is intended to provide an insight into the culture of speech and writing in Middle Low German and Low Rhenish regions. This spectrum of texttypes can be used to trace the linguistic development on the base of diatopic and diacronic subcategorisation. The aim of the project is the publication of diplomatic transcribed, lemmatised and grammatically annotated texts. The processed data – especially on the grammatical level – enables a linguistic analysis of the Middle Low German and Low Rhenish language, which goes far beyond what was possible until now.

Das Referenzkorpus mittelniederdeutscher und niederrheinischer Texte basiert auf Handschriften, Drucken und Inschriften. Es soll einen Einblick in die Sprach- und Textkultur des niederdeutschen und niederrheinischen Raums geben. Die historische Sprachentwicklung soll in ihrer diatopischen und diachronischen Untergliederung anhand des Textsortenspektrums nachgezeichnet werden. Das Ziel des Projektes besteht in der Veröffentlichung diplomatisch transkribierter, lemmatisierter und grammatisch annotierter Texte. Die so bearbeiteten Daten ermöglichen - insbesondere auf grammatischer Ebene - sprachwissenschaftliche Analysen des Mittelniederdeutschen und Niederrheinischen, die weit über das bisher Mögliche hinausgehen.

Ingrid Schröder. 2014. "Das Referenzkorpus: Neue Perspektiven für die mittelniederdeutsche Grammatikographie." Jahrbuch für Germanistische Sprachgeschichte, 5(1), pp. 150-164. http://www.degruyter.com/view/j/jbgsg.2014.5.issue-1/jbgsg-2014-0011/jbgsg-2014-0011.xml .

Robert Peters, Norbert Nagel. 2014. "Das digitale 'Referenzkorpus Mittelniederdeutsch / Niederrheinisch (ReN)'." Jahrbuch für Germanistische Sprachgeschichte, 5(1), pp. 165-175. https://www.degruyter.com/view/j/jbgsg.2014.5.issue-1/jbgsg-2014-0012/jbgsg-2014-0012.xml .

 

CLARIN Metadata summary for Referenzkorpus Mittelniederdeutsch/Niederrheinisch (1200-1650)Reference Corpus Middle Low German/Low Rhenish (1200-1650) (CMDI-based)

Title: Referenzkorpus Mittelniederdeutsch/Niederrheinisch (1200-1650)
Title: Reference Corpus Middle Low German/Low Rhenish (1200-1650)
Description: The reference corpus of Middle Low German and Low Rhenish texts is based on manuscripts, prints and inscriptions. It is intended to provide an insight into the culture of speech and writing in Middle Low German and Low Rhenish regions. This spectrum of texttypes can be used to trace the linguistic development on the base of diatopic and diacronic subcategorisation. The aim of the project is the publication of diplomatic transcribed, lemmatised and grammatically annotated texts. The processed data – especially on the grammatical level – enables a linguistic analysis of the Middle Low German and Low Rhenish language, which goes far beyond what was possible until now.
Description: Das Referenzkorpus mittelniederdeutscher und niederrheinischer Texte basiert auf Handschriften, Drucken und Inschriften. Es soll einen Einblick in die Sprach- und Textkultur des niederdeutschen und niederrheinischen Raums geben. Die historische Sprachentwicklung soll in ihrer diatopischen und diachronischen Untergliederung anhand des Textsortenspektrums nachgezeichnet werden. Das Ziel des Projektes besteht in der Veröffentlichung diplomatisch transkribierter, lemmatisierter und grammatisch annotierter Texte. Die so bearbeiteten Daten ermöglichen - insbesondere auf grammatischer Ebene - sprachwissenschaftliche Analysen des Mittelniederdeutschen und Niederrheinischen, die weit über das bisher Mögliche hinausgehen.
Data owner:  ReN-Team
Contributors:  Sarah Ihden (transcriber), Sarah Ihden (annotator), Sarah Ihden (researcher), Norbert Nagel (transcriber), Norbert Nagel (researcher), Katharina Dreessen (transcriber), Katharina Dreessen (annotator), Katharina Dreessen (researcher), Annemarie Bischoff (annotator), Annemarie Bischoff (researcher), Meike Tiedemann (transcriber), Meike Tiedemann (annotator), Meike Tiedemann (researcher), Verena Kleymann (transcriber), Verena Kleymann (annotator), Verena Kleymann (researcher), Elmar Schilling (annotator), Elmar Schilling (researcher), Nadine Wallmeier (annotator), Nadine Wallmeier (researcher), Fabian Barteld (developer), Fabian Barteld (researcher), Timm Lehmberg (developer), Timm Lehmberg (researcher), Katharina Birr (annotator), Carlotta Hübener (transcriber), Carlotta Hübener (annotator), Paul Kahre (annotator), Vanessa Nasielski (annotator), Anabel Recker (annotator), Lena Schnee (annotator), Henny Sluyter-Gäthje (developer), Lena Sandmann (developer), Hellen Peterberns (annotator), Sebastian Tews (annotator), Ilka Tran (transcriber), Anna Wollenschläger (transcriber), Lisa Dücker (annotator), Sarah Matthies (transcriber), Hamasa Mirahmadi (transcriber), Eleonore Schmitt (annotator), Christina Eichhorn-Hartmeyer (annotator), Julia Hütter (annotator), Meile-Andrea Schroeder (annotator), Mattias Liebl (transcriber), Katharina Schröder (transcriber), John Sturm (annotator)
Project:  Referenzkorpus Mittelniederdeutsch/ Niederrheinisch (1200-1650), Deutsche Forschungsgemeinschaft (DFG)
Languages:  Middle Low German (gml), Low Rhenish (mis)
Size:  56 Text, 460130 tok_anno, 468954 tok_dipl
Annotation types:  Transcription (manual), Part-of-speech (semi-automatic), Lemma (semi-automatic), Morphology (semi-automatic)
Temporal Coverage:  1200/1650
Genre:  various
Modality:  written
References:  Ingrid Schröder (2014) Das Referenzkorpus: Neue Perspektiven für die mittelniederdeutsche Grammatikographie References:  Robert Peters; Norbert Nagel (2014) Das digitale 'Referenzkorpus Mittelniederdeutsch / Niederrheinisch (ReN)'

 </dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/1694</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.1694</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:1694</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.1668</dc:relation>
          <dc:rights>info:eu-repo/semantics/openAccess</dc:rights>
          <dc:rights>https://creativecommons.org/licenses/by/4.0/legalcode</dc:rights>
          <dc:subject>Middle Low German</dc:subject>
          <dc:subject>Low Rhenish</dc:subject>
          <dc:subject>ANNIS</dc:subject>
          <dc:subject>linguistics</dc:subject>
          <dc:subject>diacronic</dc:subject>
          <dc:subject>historical</dc:subject>
          <dc:title>Reference Corpus Middle Low German/Low Rhenish (1200-1650); Referenzkorpus Mittelniederdeutsch/Niederrheinisch (1200-1650)</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:1689</identifier>
        <datestamp>2022-07-15T11:46:05Z</datestamp>
        <setSpec>user-hzsk</setSpec>
        <setSpec>user-uhh</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:contributor>Nagel, Norbert</dc:contributor>
          <dc:contributor>Schilling, Elmar</dc:contributor>
          <dc:contributor>Barteld, Fabian</dc:contributor>
          <dc:contributor>Ihden, Sarah</dc:contributor>
          <dc:contributor>Liebl, Mattias</dc:contributor>
          <dc:contributor>Schnee, Lena</dc:contributor>
          <dc:contributor>Dücker, Lisa</dc:contributor>
          <dc:contributor>Matthies, Sarah</dc:contributor>
          <dc:contributor>Hütter, Julia</dc:contributor>
          <dc:contributor>Eichhorn-Hartmeyer, Christina</dc:contributor>
          <dc:contributor>Schmitt, Eleonore</dc:contributor>
          <dc:contributor>Glawe, Meike</dc:contributor>
          <dc:contributor>Sluyter-Gäthje, Henny</dc:contributor>
          <dc:contributor>Birr, Katharina</dc:contributor>
          <dc:contributor>Kleymann, Verena</dc:contributor>
          <dc:contributor>Wallmeier, Nadine</dc:contributor>
          <dc:contributor>Nasielski, Vanessa</dc:contributor>
          <dc:contributor>Tran, Ilka</dc:contributor>
          <dc:contributor>Hübener, Carlotta</dc:contributor>
          <dc:contributor>Schroeder, Meile-Andrea</dc:contributor>
          <dc:contributor>Schröder, Katharina</dc:contributor>
          <dc:contributor>Recker, Anabel</dc:contributor>
          <dc:contributor>Dreessen, Katharina</dc:contributor>
          <dc:contributor>Kahre, Paul</dc:contributor>
          <dc:contributor>Bischoff, Annemarie</dc:contributor>
          <dc:contributor>Lehmberg, Timm</dc:contributor>
          <dc:contributor>Wollenschläger, Anna</dc:contributor>
          <dc:contributor>Mirahmadi, Hamasa</dc:contributor>
          <dc:contributor>Tews, Sebastian</dc:contributor>
          <dc:creator>ReN-Team</dc:creator>
          <dc:date>2017-09-05</dc:date>
          <dc:description>The project "Reference Corpus Middle Low German/Low Rhenish (1200–1650)", abbreviated as "ReN", is part of the "Corpus of Historical German Texts", which includes the projects "Old German Reference Corpus (750–1050)", the "Reference Corpus Middle High German (1050–1350)" and the "Reference Corpus EarlyModern New High German (1350–1650)". The project "Reference Corpus Middle Low German/Low Rhenish (1200–1650)" deals with a structured selection of Middle Low German and Low Rhenish monuments of speech from 1200 to 1650. This selection is based on the parameters "space", "time" and "field of writing".

The reference corpus of Middle Low German and Low Rhenish texts is based on manuscripts, prints and inscriptions. It is intended to provide an insight into the culture of speech and writing in Middle Low German and Low Rhenish regions. This spectrum of texttypes can be used to trace the linguistic development on the base of diatopic and diacronic subcategorisation. The aim of the project is the publication of diplomatic transcribed, lemmatised and grammatically annotated texts. The processed data – especially on the grammatical level – enables a linguistic analysis of the Middle Low German and Low Rhenish language, which goes far beyond what was possible until now.

Das Referenzkorpus mittelniederdeutscher und niederrheinischer Texte basiert auf Handschriften, Drucken und Inschriften. Es soll einen Einblick in die Sprach- und Textkultur des niederdeutschen und niederrheinischen Raums geben. Die historische Sprachentwicklung soll in ihrer diatopischen und diachronischen Untergliederung anhand des Textsortenspektrums nachgezeichnet werden. Das Ziel des Projektes besteht in der Veröffentlichung diplomatisch transkribierter, lemmatisierter und grammatisch annotierter Texte. Die so bearbeiteten Daten ermöglichen - insbesondere auf grammatischer Ebene - sprachwissenschaftliche Analysen des Mittelniederdeutschen und Niederrheinischen, die weit über das bisher Mögliche hinausgehen.

Ingrid Schröder. 2014. "Das Referenzkorpus: Neue Perspektiven für die mittelniederdeutsche Grammatikographie." Jahrbuch für Germanistische Sprachgeschichte, 5(1), pp. 150-164. http://www.degruyter.com/view/j/jbgsg.2014.5.issue-1/jbgsg-2014-0011/jbgsg-2014-0011.xml .

Robert Peters, Norbert Nagel. 2014. "Das digitale 'Referenzkorpus Mittelniederdeutsch / Niederrheinisch (ReN)'." Jahrbuch für Germanistische Sprachgeschichte, 5(1), pp. 165-175. https://www.degruyter.com/view/j/jbgsg.2014.5.issue-1/jbgsg-2014-0012/jbgsg-2014-0012.xml .

 

CLARIN Metadata summary for Referenzkorpus Mittelniederdeutsch/Niederrheinisch (1200-1650)Reference Corpus Middle Low German/Low Rhenish (1200-1650) (CMDI-based)

Title: Referenzkorpus Mittelniederdeutsch/Niederrheinisch (1200-1650)
Title: Reference Corpus Middle Low German/Low Rhenish (1200-1650)
Description: The reference corpus of Middle Low German and Low Rhenish texts is based on manuscripts, prints and inscriptions. It is intended to provide an insight into the culture of speech and writing in Middle Low German and Low Rhenish regions. This spectrum of texttypes can be used to trace the linguistic development on the base of diatopic and diacronic subcategorisation. The aim of the project is the publication of diplomatic transcribed, lemmatised and grammatically annotated texts. The processed data – especially on the grammatical level – enables a linguistic analysis of the Middle Low German and Low Rhenish language, which goes far beyond what was possible until now.
Description: Das Referenzkorpus mittelniederdeutscher und niederrheinischer Texte basiert auf Handschriften, Drucken und Inschriften. Es soll einen Einblick in die Sprach- und Textkultur des niederdeutschen und niederrheinischen Raums geben. Die historische Sprachentwicklung soll in ihrer diatopischen und diachronischen Untergliederung anhand des Textsortenspektrums nachgezeichnet werden. Das Ziel des Projektes besteht in der Veröffentlichung diplomatisch transkribierter, lemmatisierter und grammatisch annotierter Texte. Die so bearbeiteten Daten ermöglichen - insbesondere auf grammatischer Ebene - sprachwissenschaftliche Analysen des Mittelniederdeutschen und Niederrheinischen, die weit über das bisher Mögliche hinausgehen.
Publication date: 2017-09-05
Data owner:  ReN-Team
Contributors:  Sarah Ihden (transcriber), Sarah Ihden (annotator), Sarah Ihden (researcher), Norbert Nagel (transcriber), Norbert Nagel (researcher), Katharina Dreessen (transcriber), Katharina Dreessen (annotator), Katharina Dreessen (researcher), Annemarie Bischoff (annotator), Annemarie Bischoff (researcher), Meike Glawe (transcriber), Meike Glawe (researcher), Verena Kleymann (transcriber), Verena Kleymann (researcher), Elmar Schilling (annotator), Elmar Schilling (researcher), Nadine Wallmeier (annotator), Nadine Wallmeier (researcher), Fabian Barteld (developer), Fabian Barteld (researcher), Timm Lehmberg (developer), Timm Lehmberg (researcher), Katharina Birr (annotator), Carlotta Hübener (transcriber), Paul Kahre (annotator), Vanessa Nasielski (annotator), Anabel Recker (annotator), Lena Schnee (annotator), Henny Sluyter-Gäthje (developer), Sebastian Tews (annotator), Ilka Tran (transcriber), Anna Wollenschläger (transcriber), Lisa Dücker (annotator), Sarah Matthies (transcriber), Hamasa Mirahmadi (transcriber), Eleonore Schmitt (annotator), Christina Eichhorn-Hartmeyer (annotator), Julia Hütter (annotator), Meile-Andrea Schroeder (annotator), Mattias Liebl (transcriber), Katharina Schröder (transcriber)
Project:  Referenzkorpus Mittelniederdeutsch/ Niederrheinisch (1200-1650), Deutsche Forschungsgemeinschaft (DFG)
Languages:  Middle Low German (gml), Low Rhenish (mis)
Size:  40 Text, 277229 tok_anno, 283170 tok_dipl
Annotation types:  Transcription (manual), Part-of-speech (semi-automatic), Lemma (semi-automatic), Morphology (semi-automatic)
Temporal Coverage:  1200/1650
Genre:  various
Modality:  written
References:  Ingrid Schröder (2014) Das Referenzkorpus: Neue Perspektiven für die mittelniederdeutsche Grammatikographie References:  Robert Peters; Norbert Nagel (2014) Das digitale 'Referenzkorpus Mittelniederdeutsch / Niederrheinisch (ReN)'

 </dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/1689</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.1689</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:1689</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.1668</dc:relation>
          <dc:rights>info:eu-repo/semantics/openAccess</dc:rights>
          <dc:rights>https://creativecommons.org/licenses/by/4.0/legalcode</dc:rights>
          <dc:subject>Middle Low German</dc:subject>
          <dc:subject>Low Rhenish</dc:subject>
          <dc:subject>ANNIS</dc:subject>
          <dc:subject>linguistics</dc:subject>
          <dc:subject>diacronic</dc:subject>
          <dc:subject>historical</dc:subject>
          <dc:title>Reference Corpus Middle Low German/Low Rhenish (1200-1650); Referenzkorpus Mittelniederdeutsch/Niederrheinisch (1200-1650)</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:1696</identifier>
        <datestamp>2022-07-15T11:46:05Z</datestamp>
        <setSpec>user-hzsk</setSpec>
        <setSpec>user-uhh</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:contributor>Schnee, Lena</dc:contributor>
          <dc:contributor>Wollenschläger, Anna</dc:contributor>
          <dc:contributor>Mirahmadi, Hamasa</dc:contributor>
          <dc:contributor>Kahre, Paul</dc:contributor>
          <dc:contributor>Sluyter-Gäthje, Henny</dc:contributor>
          <dc:contributor>Tiedemann, Meike</dc:contributor>
          <dc:contributor>Recker, Anabel</dc:contributor>
          <dc:contributor>Schmitt, Eleonore</dc:contributor>
          <dc:contributor>Hübener, Carlotta</dc:contributor>
          <dc:contributor>Ihden, Sarah</dc:contributor>
          <dc:contributor>Schilling, Elmar</dc:contributor>
          <dc:contributor>Birr, Katharina</dc:contributor>
          <dc:contributor>Dücker, Lisa</dc:contributor>
          <dc:contributor>Nasielski, Vanessa</dc:contributor>
          <dc:contributor>Peterberns, Hellen</dc:contributor>
          <dc:contributor>Tran, Ilka</dc:contributor>
          <dc:contributor>Barteld, Fabian</dc:contributor>
          <dc:contributor>Tews, Sebastian</dc:contributor>
          <dc:contributor>Eichhorn-Hartmeyer, Christina</dc:contributor>
          <dc:contributor>Dreessen, Katharina</dc:contributor>
          <dc:contributor>Wallmeier, Nadine</dc:contributor>
          <dc:contributor>Kleymann, Verena</dc:contributor>
          <dc:contributor>Schröder, Katharina</dc:contributor>
          <dc:contributor>Schroeder, Meile-Andrea</dc:contributor>
          <dc:contributor>Nagel, Norbert</dc:contributor>
          <dc:contributor>Sandmann, Lena</dc:contributor>
          <dc:contributor>Matthies, Sarah</dc:contributor>
          <dc:contributor>Sturm, John</dc:contributor>
          <dc:contributor>Lehmberg, Timm</dc:contributor>
          <dc:contributor>Liebl, Mattias</dc:contributor>
          <dc:contributor>Hütter, Julia</dc:contributor>
          <dc:contributor>Bischoff, Annemarie</dc:contributor>
          <dc:creator>ReN-Team</dc:creator>
          <dc:date>2019-05-28</dc:date>
          <dc:description>The project "Reference Corpus Middle Low German/Low Rhenish (1200–1650)", abbreviated as "ReN", is part of the "Corpus of Historical German Texts", which includes the projects "Old German Reference Corpus (750–1050)", the "Reference Corpus Middle High German (1050–1350)" and the "Reference Corpus EarlyModern New High German (1350–1650)". The project "Reference Corpus Middle Low German/Low Rhenish (1200–1650)" deals with a structured selection of Middle Low German and Low Rhenish monuments of speech from 1200 to 1650. This selection is based on the parameters "space", "time" and "field of writing".

The reference corpus of Middle Low German and Low Rhenish texts is based on manuscripts, prints and inscriptions. It is intended to provide an insight into the culture of speech and writing in Middle Low German and Low Rhenish regions. This spectrum of texttypes can be used to trace the linguistic development on the base of diatopic and diacronic subcategorisation. The aim of the project is the publication of diplomatic transcribed, lemmatised and grammatically annotated texts. The processed data – especially on the grammatical level – enables a linguistic analysis of the Middle Low German and Low Rhenish language, which goes far beyond what was possible until now.

Das Referenzkorpus mittelniederdeutscher und niederrheinischer Texte basiert auf Handschriften, Drucken und Inschriften. Es soll einen Einblick in die Sprach- und Textkultur des niederdeutschen und niederrheinischen Raums geben. Die historische Sprachentwicklung soll in ihrer diatopischen und diachronischen Untergliederung anhand des Textsortenspektrums nachgezeichnet werden. Das Ziel des Projektes besteht in der Veröffentlichung diplomatisch transkribierter, lemmatisierter und grammatisch annotierter Texte. Die so bearbeiteten Daten ermöglichen - insbesondere auf grammatischer Ebene - sprachwissenschaftliche Analysen des Mittelniederdeutschen und Niederrheinischen, die weit über das bisher Mögliche hinausgehen.

Ingrid Schröder. 2014. "Das Referenzkorpus: Neue Perspektiven für die mittelniederdeutsche Grammatikographie." Jahrbuch für Germanistische Sprachgeschichte, 5(1), pp. 150-164. http://www.degruyter.com/view/j/jbgsg.2014.5.issue-1/jbgsg-2014-0011/jbgsg-2014-0011.xml .

Robert Peters, Norbert Nagel. 2014. "Das digitale 'Referenzkorpus Mittelniederdeutsch / Niederrheinisch (ReN)'." Jahrbuch für Germanistische Sprachgeschichte, 5(1), pp. 165-175. https://www.degruyter.com/view/j/jbgsg.2014.5.issue-1/jbgsg-2014-0012/jbgsg-2014-0012.xml .

 

CLARIN Metadata summary for Referenzkorpus Mittelniederdeutsch/Niederrheinisch (1200-1650)Reference Corpus Middle Low German/Low Rhenish (1200-1650) (CMDI-based)

Title: Referenzkorpus Mittelniederdeutsch/Niederrheinisch (1200-1650)
Title: Reference Corpus Middle Low German/Low Rhenish (1200-1650)
Description: The reference corpus of Middle Low German and Low Rhenish texts is based on manuscripts, prints and inscriptions. It is intended to provide an insight into the culture of speech and writing in Middle Low German and Low Rhenish regions. This spectrum of texttypes can be used to trace the linguistic development on the base of diatopic and diacronic subcategorisation. The aim of the project is the publication of diplomatic transcribed, lemmatised and grammatically annotated texts. The processed data – especially on the grammatical level – enables a linguistic analysis of the Middle Low German and Low Rhenish language, which goes far beyond what was possible until now.
Description: Das Referenzkorpus mittelniederdeutscher und niederrheinischer Texte basiert auf Handschriften, Drucken und Inschriften. Es soll einen Einblick in die Sprach- und Textkultur des niederdeutschen und niederrheinischen Raums geben. Die historische Sprachentwicklung soll in ihrer diatopischen und diachronischen Untergliederung anhand des Textsortenspektrums nachgezeichnet werden. Das Ziel des Projektes besteht in der Veröffentlichung diplomatisch transkribierter, lemmatisierter und grammatisch annotierter Texte. Die so bearbeiteten Daten ermöglichen - insbesondere auf grammatischer Ebene - sprachwissenschaftliche Analysen des Mittelniederdeutschen und Niederrheinischen, die weit über das bisher Mögliche hinausgehen.
Data owner:  ReN-Team
Contributors:  Sarah Ihden (transcriber), Sarah Ihden (annotator), Sarah Ihden (researcher), Norbert Nagel (transcriber), Norbert Nagel (researcher), Katharina Dreessen (transcriber), Katharina Dreessen (annotator), Katharina Dreessen (researcher), Annemarie Bischoff (annotator), Annemarie Bischoff (researcher), Meike Tiedemann (transcriber), Meike Tiedemann (annotator), Meike Tiedemann (researcher), Verena Kleymann (transcriber), Verena Kleymann (annotator), Verena Kleymann (researcher), Elmar Schilling (annotator), Elmar Schilling (researcher), Nadine Wallmeier (annotator), Nadine Wallmeier (researcher), Fabian Barteld (developer), Fabian Barteld (researcher), Timm Lehmberg (developer), Timm Lehmberg (researcher), Katharina Birr (annotator), Carlotta Hübener (transcriber), Carlotta Hübener (annotator), Paul Kahre (annotator), Vanessa Nasielski (annotator), Anabel Recker (annotator), Lena Schnee (annotator), Henny Sluyter-Gäthje (developer), Lena Sandmann (developer), Hellen Peterberns (annotator), Sebastian Tews (annotator), Ilka Tran (transcriber), Anna Wollenschläger (transcriber), Lisa Dücker (annotator), Sarah Matthies (transcriber), Hamasa Mirahmadi (transcriber), Eleonore Schmitt (annotator), Christina Eichhorn-Hartmeyer (annotator), Julia Hütter (annotator), Meile-Andrea Schroeder (annotator), Mattias Liebl (transcriber), Katharina Schröder (transcriber), John Sturm (annotator)
Project:  Referenzkorpus Mittelniederdeutsch/ Niederrheinisch (1200-1650), Deutsche Forschungsgemeinschaft (DFG)
Languages:  Middle Low German (gml), Low Rhenish (mis)
Size:  146 Text, 1415341 tok_anno, 1450560 tok_dipl
Annotation types:  Transcription (manual), Part-of-speech (semi-automatic), Lemma (semi-automatic), Morphology (semi-automatic)
Temporal Coverage:  1200/1650
References:  Ingrid Schröder (2014) Das Referenzkorpus: Neue Perspektiven für die mittelniederdeutsche Grammatikographie References:  Robert Peters; Norbert Nagel (2014) Das digitale 'Referenzkorpus Mittelniederdeutsch / Niederrheinisch (ReN)'

 </dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/1696</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.1696</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:1696</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.1668</dc:relation>
          <dc:rights>info:eu-repo/semantics/openAccess</dc:rights>
          <dc:rights>https://creativecommons.org/licenses/by/4.0/legalcode</dc:rights>
          <dc:subject>Middle Low German</dc:subject>
          <dc:subject>Low Rhenish</dc:subject>
          <dc:subject>ANNIS</dc:subject>
          <dc:subject>linguistics</dc:subject>
          <dc:subject>diacronic</dc:subject>
          <dc:subject>historical</dc:subject>
          <dc:title>Reference Corpus Middle Low German/Low Rhenish (1200-1650); Referenzkorpus Mittelniederdeutsch/Niederrheinisch (1200-1650)</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:1703</identifier>
        <datestamp>2020-09-29T18:49:13Z</datestamp>
        <setSpec>user-hzsk</setSpec>
        <setSpec>user-uhh</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:contributor>Jacob, Peggy</dc:contributor>
          <dc:contributor>Hartmann, Katharina</dc:contributor>
          <dc:creator>Hartmann, Katharina</dc:creator>
          <dc:date>2015-09-29</dc:date>
          <dc:description>Guruntum sample: sample, status: final, manually transcribed, glossed and translated to English, annotated wrt. morphology, parts of speech, syntax, gramm. function, sem. roles, focus and focus position (e.g. ex situ) in EXMARaLDA.CLARIN Metadata summary for B2 Guruntum (CMDI-based)    	    	    	    		Title: B2 Guruntum    	        	Description: Guruntum sample: sample, status: final, manually transcribed, glossed and translated to English, annotated wrt. morphology, parts of speech, syntax, gramm. function, sem. roles, focus and focus position (e.g. ex situ) in EXMARaLDA.					Publication date: 2015					Data owner: 			Univ.-Prof. Dr. Katharina Hartmann											                		Contributors:                 	Katharina Hartmann (editor), Peggy Jacob (researcher)                			                		            	Project:             	Special Research Centre 632 Information structure, German Research Foundation            				            			            	Keywords:             	focus, Fokus            				            			    	Language:     	Guruntum (grd)    						    					            	Size:             	2171 Token            					            				                		Segmentation units:                 	other                			                		            	Genre:             	discourse            				            			            	Modality:             	written, spoken            				            					</dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/1703</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.1703</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:1703</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.1702</dc:relation>
          <dc:rights>info:eu-repo/semantics/openAccess</dc:rights>
          <dc:rights>https://creativecommons.org/licenses/by-nc/3.0/legalcode</dc:rights>
          <dc:subject>focus</dc:subject>
          <dc:subject>Fokus</dc:subject>
          <dc:subject>linguistics</dc:subject>
          <dc:subject>information structure</dc:subject>
          <dc:subject>Guruntum</dc:subject>
          <dc:title>B2 Guruntum</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:1715</identifier>
        <datestamp>2020-09-29T19:04:28Z</datestamp>
        <setSpec>user-hzsk</setSpec>
        <setSpec>user-uhh</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:contributor>Coniglio, Marco</dc:contributor>
          <dc:contributor>Donhauser, Karin</dc:contributor>
          <dc:contributor>Rasskazova, Oxana</dc:contributor>
          <dc:contributor>Petrova, Svetlana</dc:contributor>
          <dc:contributor>Gehrlein, Anke</dc:contributor>
          <dc:contributor>Schlachter, Eva</dc:contributor>
          <dc:creator>Petrova, Svetlana</dc:creator>
          <dc:date>2015-09-29</dc:date>
          <dc:description>Complete text, status: work in progress, digitalization, translation to English, manually annotated with parts of speech, syntactic category, grammatical function, clause status, numbers of syllables (per constituent), information status, topic/comment, position of constituent in sentence, definiteness, focus/background, focus marker, comments, source (bibliography).CLARIN Metadata summary for B4 Muspilli (CMDI-based)    	    	    	    		Title: B4 Muspilli    	        	Description: Complete text, status: work in progress, digitalization, translation to English, manually annotated with parts of speech, syntactic category, grammatical function, clause status, numbers of syllables (per constituent), information status, topic/comment, position of constituent in sentence, definiteness, focus/background, focus marker, comments, source (bibliography).					Publication date: 2015					Data owner: 			Prof. Dr. Svetlana Petrova											                		Contributors:                 	Svetlana Petrova (editor), Karin Donhauser (editor), Eva Schlachter (annotator), Marco Coniglio (annotator), Oxana Rasskazova (annotator), Anke Gehrlein (annotator)                			                		            	Project:             	Special Research Centre 632 Information structure, German Research Foundation            				            			            	Keywords:             	historical texts, religious texts, information structure            				            			    	Language:     	Old High German (goh)    						    					            	Size:             	909 Token            					            				                		Segmentation units:                 	other                			                		            	Genre:             	historic manuscript            				            			            	Modality:             	written            				            					</dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/1715</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.1715</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:1715</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.1714</dc:relation>
          <dc:rights>info:eu-repo/semantics/openAccess</dc:rights>
          <dc:rights>https://creativecommons.org/licenses/by-nc/3.0/legalcode</dc:rights>
          <dc:subject>historical texts</dc:subject>
          <dc:subject>religious texts</dc:subject>
          <dc:subject>information structure</dc:subject>
          <dc:subject>linguistics</dc:subject>
          <dc:subject>Old High German</dc:subject>
          <dc:title>B4 Muspilli</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <record>
      <header>
        <identifier>oai:fdr.uni-hamburg.de:1733</identifier>
        <datestamp>2020-09-29T21:12:41Z</datestamp>
        <setSpec>user-hzsk</setSpec>
        <setSpec>user-uhh</setSpec>
      </header>
      <metadata>
        <oai_dc:dc xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:oai_dc="http://www.openarchives.org/OAI/2.0/oai_dc/" xsi:schemaLocation="http://www.openarchives.org/OAI/2.0/oai_dc/ http://www.openarchives.org/OAI/2.0/oai_dc.xsd">
          <dc:contributor>House, Juliane</dc:contributor>
          <dc:creator>House, Juliane</dc:creator>
          <dc:date>2011-01-01</dc:date>
          <dc:description>Translation corpora of original texts with translations and comparable texts from the genre external business communication.CLARIN Metadata summary for Covert translation: Business Communication (old) (CMDI-based)    	    	    	    		Title: Covert translation: Business Communication (old)    	        	Description:           Translation corpora of original texts with translations and comparable          texts from the genre external business communication        					Publication date: 2011-01-01					Data owner: 			Juliane House											                		Contributors:                 	Juliane House (compiler)                			                		            	Project:             	K4 "Covert Translation"            				            			            	Keywords:             	translated texts, business communication, parallel corpus, comparable corpus            				            			    	Languages:     	German (deu), English (eng)    						    					            	Size:             	53 texts, 64980 words            					            				                		Segmentation units:                 	orthographic sentence                			                		                		Annotation types:                 	Morphological analysis                			                		            	Temporal Coverage:             	1978/1999            	                        			Spatial Coverage:             		DE            				            			            	Genre:             	discourse            				            			            	Modality:             	written            				            					</dc:description>
          <dc:identifier>https://www.fdr.uni-hamburg.de/record/1733</dc:identifier>
          <dc:identifier>10.25592/uhhfdm.1733</dc:identifier>
          <dc:identifier>oai:fdr.uni-hamburg.de:1733</dc:identifier>
          <dc:relation>doi:10.25592/uhhfdm.1731</dc:relation>
          <dc:rights>info:eu-repo/semantics/restrictedAccess</dc:rights>
          <dc:subject>translated texts</dc:subject>
          <dc:subject>business communication</dc:subject>
          <dc:subject>parallel corpus</dc:subject>
          <dc:subject>comparable corpus</dc:subject>
          <dc:subject>linguistics</dc:subject>
          <dc:subject>German</dc:subject>
          <dc:subject>English</dc:subject>
          <dc:title>Covert translation: Business Communication (old)</dc:title>
          <dc:type>info:eu-repo/semantics/other</dc:type>
          <dc:type>dataset</dc:type>
        </oai_dc:dc>
      </metadata>
    </record>
    <resumptionToken expirationDate="2026-09-12T23:15:02Z" cursor="0" completeListSize="114">.eJyNzLsOgjAYBeB3-Wc0bUEUEgcTjS5c4gXsRJpSgViFUBDF-O52cbbJOdP5ct6ghMjBR9MF8Vwdm9gORshzLWhYIcAnFlwH1hYK_LfGHfjQK9FOylFdwYKb6FjOOha34lI99VizKss5fCxQvK2lzCp9D_ycSH6XJ7ZN-v0tQZQU8-C40g2G8EifkWzCHFH78MKPE252NA1LTjgWaRNHyebnHEM3M3N8MPx7GbrR0CFDhw0dMXT2X7dWS_h8AeburYY.HYdt_g.NBvGYwbwHHi-XnqDIOQGEeQY9JM</resumptionToken>
  </ListRecords>
</OAI-PMH>
