Summer School
Language Science Meets Linguistic Diversity:
The Next Generation


28 August to 1 September, 2028
Berlin, Germany

.

The Endangered Languages Archive (ELAR), is offering an in person summer school in Linguistic Diversity and Language Science from Monday, August 28st through Friday, September 1st, 2028. The training will take place in person in Berlin.

The courses offered are:

  • Capturing Language Use with Mandana Seyfeddinipur (ELDP/ELAR/BBAW)
  • Small scale multilingualism with Jeff Good (Buffalo University)
  • Mobilising documentation data: Internet in a box with Pierpaolo di Carlo (Naples University)
  • Child language acquisition in the field with Birgit Hellwig (Cologne University)
  • Language on the move with Karo Obert (University Texas at Austin)
  • Endangered Materials Knowledge with Mareike Wulff (Humboldt Universitaet)
  • Social networks and kinship with Adam Tallmann (CNRS)
  • Archiving in language documentation with Zak O’Hagen (Berkeley)
  • Introduction to ELAN Christian Döhler (BBAW)
  • Introduction to Flex with with Kelsey Neely (ELDP/BBAW)
  • Lexicography with Flex with Kelsey Neely (ELDP/BBAW)

Information about the Summer School

This five-day intensive training will involve hands-on practice, including homework assignments to be completed in the evenings. The language of instruction is English. Applicants should have demonstrated commitment and/or plans to carry out research on an under-resourced language.

Students may enrol in up to 3 courses. The cost for participation is 250 Euros. Participants will need to arrange their own travel and accommodations in Berlin.

Key dates:

15 November – Applications open

15 January – Applications due

15 February – Notification of acceptance sent

15 May  – Registration payment due

28 August to 1 September – Summer school week

For more information about the training, please contact us at summerschool@langdoc.org

9:30-11:00

Small scale multilingualism with Jeff Good

The course provides theoretical foundations and practical skills for understanding acquisition of oral narratives and for eliciting, analyzing, and evaluating fictional and personal narratives across languages. It introduces the Multilingual Assessment Instrument for Narratives (MAIN) and integrates theory on narrative development with contemporary elicitation methods and systematic scoring procedures. Students gain hands-on experience in assessing oral narratives in different languages and adapting MAIN to new linguistic contexts. The course also highlights applications of narrative assessment in educational and research settings, enabling students to apply their knowledge to both academic and applied purposes.

The course provides theoretical foundations and practical skills for understanding acquisition of oral narratives and for eliciting, analyzing, and evaluating fictional and personal narratives across languages. It introduces the Multilingual Assessment Instrument for Narratives (MAIN) and integrates theory on narrative development with contemporary elicitation methods and systematic scoring procedures. Students gain hands-on experience in assessing oral narratives in different languages and adapting MAIN to new linguistic contexts. The course also highlights applications of narrative assessment in educational and research settings, enabling students to apply their knowledge to both academic and applied purposes.

11:30-13:00

Mobilising documentation data: Internet in a box with Pierpaolo di Carlo

Central America is a fascinating crossroads of linguistic diversity, where two distinct linguistic areas converge: the well-known Mesoamerican Sprachbund, home to the Mayan languages, and the lesser-studied Intermediate Area, which includes the Chibchan languages. These regions differ markedly in their nominal and verbal morphosyntax as well as in their sentence structures, making Central America an ideal setting for typological and areal-linguistic exploration. The course begins with an overview of the key areal features that shape the linguistic landscape of Central America. Building on this foundation, we will explore the characteristic typological patterns of both areas through case studies of selected languages. During the course, participants will focus on a language of their choice and contribute their findings to a collaborative cluster analysis of the Central American linguistic area.

Multiple expoence is defined as the occurrence of multiple realizations of a single morphosemantic feature, bundle of features, or derivational category within a word (Harris 2017). In this course, we will discuss data from a broad variety of typologically diverse languages in order to identify the main types of multiple exponence that have been discussed in the literature. We will then entertain various  treatments of multiple exponence and raise the question of whether it is possible to treat all different types thereof via the same operation.

This course introduces fundamental concepts and practices for the creation of structured data from documentary recordings of under-described and under-resourced languages. Participants will learn how to structure and curate corpora derived from language documentation projects so that they can be reused for comparative analysis, variationist studies, and corpus-informed typology. We will carefully consider the analytic and practical choices documenters must make at each stage of annotation and analysis, from segmenting natural language data to querying a corpus. We will use the transcription and annotation tool ELAN and the interlinear glossing and lexical database creation tool Fieldworks Language Explorer (FLEx). [Note: FLEx is only available for Windows and Linux systems.]

14:00-15:30

Child language acquisition in the field with Birgit Hellwig

Gender is traditionally defined as noun classification that involves syntactic agreement (cf., for example, Hockett 1958, Corbett 1991). Since both criteria can at times be hard to pin down, the recent past saw the emergence of more complex approaches to gender that are to capture also related phenomena of noun classification, notably classifiers, as within Corbett’s (2014) canonical typology or Wälchli and Di Garbo’s (2019: 330-1) “dynamic” characterization. On the basis of a wide range of cross-linguistic data, the course locates gender in its a wider domain of noun classification, provides a unified analytical approach for achieving transparent cross-linguistic comparability of systemic structures, and also examines its diverse diachronic dynamics from simple to so-called “mature” gender. Particular topics to be scrutinized include the nature of agreement class and its partly problematic relation to gender, “overt” gender on nominal controllers, the typology of semantic gender assignment, and the very origin of gender as a grammaticalized system.

All languages can refer to entities or concepts in the shared situation of the participants or in their shared background knowledge, they can introduce new entities into the discourse, and they can pick them up by anaphoric expressions. In the course, I will discuss important theoretical distinctions and known ways that languages use for these purposes, like various types of demonstratives, definite and indefinite articles, pronominal expressions that mark lexical distinctions of their antecedents like gender and number, and their saliency status, as well as prosodic and syntactic features. This includes more subtle phenomena like partitive and associative anaphors and reference to manners and to speech acts. The goal is to enable course participants to identify such more fine-grained distinctions in their documentary work. In particular, we will get familiar with annotation schemes of reference tracking, such as the RefLex Scheme by Riester and Baumann and the RefIND Annotation Guidelines by Schiborr, Schnell and Thiele. 

16:00-17:30

Archiving in language documentation with Zak O'Hagen

While tense is a widespread grammatical category, a third of the languages of the world do not encode tense (based on Grambank data). In this course, we will ask how tenseless languages, or languages without the grammatical category of tense, talk about temporal meanings, such as past, present, and future. We will look both at the broad typology of tenselessness and at individual underdescribed languages. While some languages rely primarily on aspectual categories to talk about time, other languages use mood categories, or leave verbs unmarked for TAM (tense, aspect, mood). We will also discuss potential language universals in this domain, e.g. unmarked verbs favoring past and present interpretations over future meanings cross-linguistically.

This practical course focuses on audiovisual methods for documenting language as it is used in real-life interaction. Participants will gain hands-on experience in planning and recording natural communicative events, learning camera techniques and audio recording suited to field conditions. The course integrates discussion of ethical practices, informed consent, and metadata standards for the creation of accessible, multipurpose records of under-documented languages. Emphasis is placed on producing recordings that illuminate gesture, gaze, and other multimodal aspects of communication, equipping participants to create rich, analyzable records of linguistic and social practice.

Nach oben scrollen