Article In: Register Studies: Online-First Articles
A scale of conceptual orality and literacy
Automatic text categorization in the tradition of ‘Nähe und Distanz’
This content is being prepared for publication; it may be subject to changes.
Abstract
Koch and Oesterreicher’s model of ‘Nähe und Distanz’ (‘Nähe’ = immediacy, conceptual orality; ‘Distanz’ =
distance, conceptual literacy) is widely used in German linguistics. However, there is no statistical foundation for its use in
corpus linguistic analyses, even though it is increasingly being incorporated into empirical corpus linguistics. Theoretically, it
is proposed, among other things, that written texts can be assessed on a scale of conceptual orality and literacy based on
linguistic features, which were later derived by Ágel, V., & Hennig, M. (2006). Theorie
des Nähe- und Distanzsprechens / Praxis des Nähe- und
Distanzsprechens. In V. Ágel & M. Hennig (Eds.), Grammatik
aus Nähe und Distanz: Theorie und Praxis am Beispiel von Nähetexten
1650–2000 (pp. 3–31, 33–74). Tübingen: Niemeyer. .
This article establishes such a scale based on PCA, emphasizing a statistical implementation that seeks to reflect
core theoretical assumptions in practical terms, and combines it with automatic analysis. It critically engages with the
theoretical tradition of ‘Nähe und Distanz’ and, drawing on analytical procedures associated with Multidimensional Analysis,
develops a methodologically robust implementation for corpus-linguistic application. Two corpora of New High German provide the
foundation for both the analysis and the evaluation. When evaluating established features identified by Ágel, V., & Hennig, M. (2006). Theorie
des Nähe- und Distanzsprechens / Praxis des Nähe- und
Distanzsprechens. In V. Ágel & M. Hennig (Eds.), Grammatik
aus Nähe und Distanz: Theorie und Praxis am Beispiel von Nähetexten
1650–2000 (pp. 3–31, 33–74). Tübingen: Niemeyer. , a central finding is that features of conceptual orality and literacy must be
distinguished in order to rank texts in a differentiated way.
The present approach is theory-guided but not confirmatory; rather, it seeks to make a theoretical tradition
fruitful in its function as an explanatory tool for practical analysis. While (1988). Variation
across speech and writing. Cambridge: Cambridge University Press. empirically derives dimensions aimed at describing functional variation across registers, the present
implementation of ‘Nähe und Distanz’ is better understood as a deliberately defined (‘tailored’) dimension, which draws on,
reflects, and implements core assumptions of the theoretical tradition. This leads to a different and more narrowly framed set of
possible applications: not the investigation of conceptual orality and literacy itself, but rather supportive or controlling
functions in the context of corpus compilation and analysis — especially in historical stages.
This is demonstrated using the ‘Deutsches Textarchiv’, the largest collection of New High German texts. Applying
the COL scale allows text categories to be situated relative to each other based on their linguistic properties, reveals
considerable internal variation within seemingly homogeneous categories, and assesses the representativeness of individual text
segments for their source texts.
Article outline
- 1.A scale of conceptual orality and literacy for text categorization
- 1.1‘Nähe und Distanz’: Introductory orientation and comparison with Biber’s Dimension 1
- 1.2Studies within the ‘Nähe und Distanz’ and Multidimensional Analysis traditions
- 1.3Previous research within the ‘Nähe und Distanz’ tradition: A critical review as a basis for the present approach
- 1.4Framework, methods, and corpora: From critique to implementation
- 2.The derivation of a practically applicable scale of conceptual orality and literacy
- 2.1The starting point: Evaluation of the features of conceptual orality by Ágel and Hennig (2006)
- 2.2Automatic analysis of the 5 features of conceptual orality, complementary features of conceptual literacy, and the COL scale in
comparison to other approaches
- 2.2.1Automatic sentence boundary detection
- 2.2.2Automatic analysis of the other features of conceptual orality
- 2.2.3Complementary features of conceptual literacy
- 2.2.4The COL scale
- 2.2.5A comparison to other approaches
- 2.3The COL scale: An application to the DTA
- 3.An initial evaluation: Internal and external features
- 4.Conclusions
- Notes
- Author queries
Literature Sources
References (171)
Ädel, A. (2020). Corpus
compilation. In M. Paquot & S. Gries (Eds.), A
practical handbook of corpus
linguistics (pp. 3–24). New York: Springer.
Ágel, V. (2000). Syntax
des Neuhochdeutschen bis zur Mitte des 20. Jahrhunderts. In W. Besch, A. Betten, O. Reichmann, & S. Sonderegger (Eds.), Sprachgeschichte.
Ein Handbuch zur Geschichte der deutschen Sprache und ihrer Erforschung (= Handbücher zur Sprach- und
Kommunikationswissenschaft, 2.2, 2nd ed., pp.
1855–1903). Berlin: De Gruyter.
Ágel, V., & Hennig, M. (2006). Theorie
des Nähe- und Distanzsprechens / Praxis des Nähe- und
Distanzsprechens. In V. Ágel & M. Hennig (Eds.), Grammatik
aus Nähe und Distanz: Theorie und Praxis am Beispiel von Nähetexten
1650–2000 (pp. 3–31, 33–74). Tübingen: Niemeyer.
Ágel, V. (2007). Was
ist ‚grammatische Aufklärung‘ in einer Schriftkultur? Die Parameter ‚Aggregation‘ und
‚Integration‘. In H. Feilke, C. Knobloch & P. L. Völzing (Eds.), Was
heißt linguistische Aufklärung? Sprachauffassungen zwischen Systemvertrauen und
Benutzerfürsorge (pp. 39–57). Heidelberg: Synchron Wiss.-Verl. der Autoren.
Ágel, V., & Hennig, M. (2007a). Überlegungen
zur Theorie und Praxis des Nähe- und Distanzsprechens. In V. Ágel & M. Hennig (Eds.), Zugänge
zur Grammatik der gesprochenen
Sprache (pp. 179–214). Tübingen: Niemeyer.
(2007b). DFG-Projekt
‘Explizite und elliptische Junktion in der Syntax des Neuhochdeutschen‘.
Forschungsnotiz. Zeitschrift für Germanistische
Linguistik, 351, 185–189.
Ágel, V., & Diegelmann, C. (2010). Theorie
und Praxis der expliziten Junktion. In V. Ágel & M. Hennig (Eds.), Nähe
und Distanz im Kontext variationslinguistischer
Forschung (pp. 345–393). Berlin: De Gruyter.
Ágel, V. (2012). Junktionsprofile
aus Nähe und Distanz: Ein Beitrag zur Vertikalisierung der neuhochdeutschen
Grammatik. In J. Bär & M. Müller (Eds.), Geschichte
der Sprache — Sprache der Geschichte. Probleme und Perspektiven der historischen Sprachwissenschaft des Deutschen, Oskar
Reichmann zum 75.
Geburtstag (pp. 181–206). Berlin: Akademie (LHG 3).
Auer, P. (2000). Online-Syntax
— oder: Was es bedeuten könnte, die Zeitlichkeit der mündlichen Sprache ernst zu nehmen. Sprache und Literatur,
85 (= Themenheft, Die Medialität der Gesprochenen
Sprache), pp. 43–56.
Bernstein, B. (1962/2003). Linguistic
codes, hesitation phenomena and intelligence. In B. Bernstein, Class,
codes and control. Volume I. Theoretical studies towards a sociology of
language (pp. 76–94). London, New York: Routledge.
Biber, D. (1985). Investigating
macroscopic textual variation through multifeature/multidimensional
analyses. Linguistics, 23(2), 337–360.
(1986). Spoken
and written textual dimensions in English: Resolving the contradictory
findings. Language, 62(2), 384–414.
(1996). Investigating
language use through corpus-based analyses of association patterns. International Journal of
Corpus
Linguistics, 1(2), 171–197.
(2001). Dimensions
of variation among eighteenth-century speech-based and written
registers. In S. Conrad & D. Biber (Eds.), Variation
in English: Multi-dimensional
studies (pp. 200–214). London: Longman.
Biber, D., Conrad, S., Reppen, R., Byrd, P., & Helt, M. (2002). Speaking
and writing in the university: A multi-dimensional comparison. TESOL
Quarterly, 361, 9–48.
Biber, D., Conrad, S. & Cortes, V. (2004a). If
you look at …: Lexical bundles in university teaching and textbooks. Applied
Linguistics, 25(3), 371–405.
Biber, D., Conrad, S., Reppen, R., Byrd, P., Helt, M., Clark, V., Cortes, V., Csomay, E., & Urzua, A. (2004b). Representing
language use in the university: Analysis of the TOEFL 2000 spoken and written academic language
corpus. TOEFL Monograph Series. Princeton, NJ: Educational Testing Service.
Biber, D., & Jones, J. (2009). Quantitative
methods in corpus linguistics. In A. Lüdeling & M. Kytö (Eds.), Corpus
linguistics: An international handbook (Vol. 2, Handbücher zur Sprach- und
Kommunikationswissenschaft, 29(2), pp. 1286–1304). Berlin: De Gruyter.
Biber, D., & Gray, B. (2010). Grammatical
complexity in academic English: Linguistic change in
writing. Cambridge: Cambridge University Press.
Biber, D. (2014). Opening:
Multi-dimensional analysis. A personal history. In T. Sardinha & M. Pinto (Eds.), Multi-dimensional
analysis, 25 years on: A tribute to Douglas Biber (Studies in Corpus Linguistics, 60, pp.
xxvii–xxxviii). Amsterdam: Benjamins.
Biber, D., & Egbert, J. (2016). Register
variation on the searchable web: A multi-dimensional analysis. Journal of English
Linguistics, 44(2), 95–137.
Biber, D., Egbert, J., & Keller, D. (2020). Reconceptualizing
register in a continuous situational space. Corpus Linguistics and Linguistic
Theory, 16(3), 581–616.
Biber, D., Egbert, J., Keller, D., & Wizner, S. (2021). Extending
text-linguistic studies of register variation to a continuous situational space: Case studies from the web and natural
conversation. In E. Seoane & D. Biber (Eds.), Corpus-based
approaches to register variation (Studies in Corpus Linguistics, 103, pp.
19–50). Amsterdam: Benjamins.
Biber, D., Larsson, T. & Hancock, G. R. (2024). Dimensions
of text complexity in the spoken and written modes: A comparison of theory-based
models. Journal of English
Linguistics, 52(1), 65–94.
Booth, H., Breitbarth, A., Ecay, A., & Farasyn, M. (2020). A
Penn-style treebank of Middle Low German. In Proceedings of the
Twelfth Language Resources and Evaluation
Conference (pp. 766–775). Marseille, France: European Language Resources Association.
Botha, Y., & van Zyl, M. (2021). Register
and modification in the noun phrase. In E. Seoane & D. Biber (Eds.), Corpus-based
approaches to register variation (Studies in Corpus Linguistics, 103, pp.
179–208). Amsterdam: Benjamins.
Broll, S., & Schneider, R. (2023). Empirische
Verortung konzeptioneller Nähe/Mündlichkeit inner- und außerhalb schriftsprachlicher
Korpora. Journal for Language Technology and Computational
Linguistics, 36(1), 113–150.
Brown, P., & Fraser, C. (1979). Speech
as a marker of situation. In K. Scherer & H. Giles (Eds.), Social
markers in
speech (pp. 33–62). Cambridge: Cambridge University Press.
Callies, M. (2013). Agentivity
as a determinant of lexico-grammatical variation in L2 academic writing. International Journal
of Corpus
Linguistics, 18(3), 357–390.
Chafe, W. L. (1982). Integration
and involvement in speaking, writing, and oral literature. In D. Tannen (Ed.), Spoken
and written language: Exploring orality and
literacy (pp. 35–53). Norwood, NJ: Ablex.
Châu, Q., & Bulté, B. (2023). A
comparison of automated and manual analyses of syntactic complexity in L2 English
writing. International Journal of Corpus
Linguistics, 28(2), 232–262.
Claridge, C. (2008). Historical
corpora. In A. Lüdeling & M. Kytö (Eds.), Corpus
linguistics: An international handbook (Vol. 1, Handbücher zur Sprach- und
Kommunikationswissenschaft, 29(1), pp. 242–259). Berlin: De Gruyter.
Crossley, S., & Louwerse, M. (2007). Multi-dimensional
register classification using bigrams. International Journal of Corpus
Linguistics, 12(4), 453–478.
Culpeper, J., & Kytö, M. (2000). Data
in historical pragmatics: Spoken interaction (re)cast as writing. Journal of Historical
Pragmatics, 1(2), 175–199.
(2010). Early
Modern English dialogues: Spoken interaction as
writing. Cambridge: Cambridge University Press.
Cummins, J. (1979). Cognitive/academic
language proficiency, linguistic interdependence, the optimum age question and some other
matters. Working Papers on
Bilingualism, 191, 198–205.
Cvrček, V., Laubeová, Z., Lukeš, D., Poukarová, P., Řehořková, A., & Zasina, A. (2020). Author
and register as sources of variation: A corpus-based study using elicited texts. International
Journal of Corpus
Linguistics, 25(4), 461–488.
Dahl, Ö. (2004). The
growth and maintenance of linguistic complexity. Amsterdam, Philadelphia: John Benjamins.
Degaetano-Ortlieb, S. (2021). Measuring
informativity: The rise of compounds as informationally dense structures in 20th-century scientific
English. In E. Seoane & D. Biber (Eds.), Corpus-based
approaches to register variation (Studies in Corpus Linguistics,
103, pp. 291–312). Amsterdam: Benjamins.
Denkler, M. & Elspaß, S. (2007). Nähesprachlichkeit
und Regionalsprachlichkeit in historischer Perspektive. Niederdeutsches
Jahrbuch, 1301, 79–108.
Díez-Bedmar, M. B., & Pérez-Paredes, P. (2020). Noun
phrase complexity in young Spanish EFL learners’ writing: Complementing syntactic complexity indices with corpus-driven
analyses. International Journal of Corpus
Linguistics, 25(1), 4–35.
Douglas, F. (2003). The
Scottish corpus of texts and speech: Problems of corpus design. Literary and Linguistic
Computing, 18(1), 23–37.
Egbert, J., Biber, D., Keller, D. & Gracheva, M. (2024). Register
and the dual nature of functional correspondence: Accounting for text-linguistic variation between registers, within
registers, and without registers. Corpus Linguistics and Linguistic
Theory, 20(3), 505–538.
Egbert, J., & Gracheva, M. (2022). Linguistic
variation within registers: Granularity in textual units and situational parameters. Corpus
Linguistics and Linguistic
Theory, 19(1), 115–143.
Egbert, J. & Mahlberg, M. (2020). Fiction
— one register or two? Speech and narration in novels. Register
Studies, 2(1), 72–101.
Ehret, K. & Taboada, M. (2020). Are
online news comments like face-to-face conversation? A multi-dimensional analysis of an emerging
register. Register
Studies, 2(1), 1–36.
Elspaß, S. (2008). Briefe
rheinischer Auswanderer als Quellen einer Regionalsprachgeschichte. Rheinische
Vierteljahrsblätter, 721, 147–165.
(2010a). Klammerstrukturen
in nähesprachlichen Texten des 19. und frühen 20. Jahrhunderts: Ein Plädoyer für die Verknüpfung von historischer und
Gegenwartsgrammatik. In A. Ziegler (Ed.), Historische
Textgrammatik und historische Syntax des Deutschen: Traditionen, Innovationen, Perspektiven (Vol. 2, pp.
1011–1026). Berlin: De Gruyter.
(2010b). Zum
Verhältnis von „Nähegrammatik“ und Regionalsprachlichkeit in historischen
Texten. In V. Ágel & M. Hennig (Eds.), Nähe
und Distanz im Kontext variationslinguistischer
Forschung (pp. 63–84). Berlin, New York: De Gruyter.
Emmrich, V. (2025). GiesKaNe:
Bridging past and present in grammatical theory and practical application.
Emmrich, V., & Hennig, M. (2023). GiesKaNe:
Korpusaufbau zwischen Standard und Innovation. In A. Deppermann, C. Fandrych, M. Kupietz, & T. Schmidt (Eds.), Korpora
in der germanistischen Sprachwissenschaft: Mündlich, schriftlich, multimedial (Jahrbuch des Instituts
für deutsche Sprache, Mannheim, pp. 199–224). Berlin: De Gruyter.
(2025). ‘Nähe
und Distanz’ in der (germanistischen) Sprachgeschichtsforschung: Nutzen, Nutzung,
Neuansatz. Jahrbuch der Gesellschaft für Germanistische
Sprachgeschichte (accepted).
Emmrich, V., Hennig, M., & Meisner, P. (2026a). Attribution
im Neuhochdeutschen: Genitivattribute als Distanzmarker. In V. Ágel (Ed.), Grammatik
des Neuhochdeutschen zwischen Gegenwart und Geschichte (accepted).
(2026b). Pronominale
Autorreferenz in der Wissenschaftssprache des
Neuhochdeutschen. In W. Imo et al. (Eds.), Gattungsspezifik
des Pronomengebrauchs (accepted).
Engel, U. (1974). Syntaktische
Besonderheiten der deutschen Alltagssprache. In Gesprochene Sprache.
Jahrbuch 1972 des Instituts für Deutsche
Sprache (pp. 199–228). Düsseldorf: Schwann.
Erfurt, J. (1996). Sprachwandel
und Schriftlichkeit. In H. Günther & O. Ludwig (Eds.), Schrift und Schriftlichkeit / Writing and its use. Ein interdisziplinäres Handbuch internationaler
Forschung / An interdisciplinary handbook of international
research, 21. Halbbd. (pp. 1387–1404). Berlin: De Gruyter.
Feilke, H. (2011). Literalität
und literale Kompetenz: Kultur, Handlung,
Struktur. Leseforum.ch: Online-Plattform für Literalität, 1(2011). [URL]
(2012b). Schulsprache
— wie Schule Sprache macht. In S. Günthner, W. Imo, D. Meer & J. G. Schneider (Eds.), Kommunikation
und Öffentlichkeit: Sprachwissenschaftliche Potenziale zwischen Empirie und
Norm (pp. 149–175). Berlin, Boston: De Gruyter.
(2012c). Was
sind Textroutinen? Zur Theorie und Methodik des
Forschungsfeldes. In H. Feilke & K. Lehnen (Eds.), Schreib-
und Textroutinen: Theorie, Erwerb und didaktisch-mediale Modellierung (Forum Angewandte Linguistik,
Bd. 52, pp. 1–31). Frankfurt a.M. u. a.: Lang.
Feilke, H., & Hennig, M. (Eds.). (2016). Zur
Karriere von ›Nähe und Distanz‹: Rezeption und Diskussion des
Koch-Oesterreicher-Modells. Berlin: De Gruyter.
Fiehler, R. (2000). Gesprochene
Sprache — gibt’s die? In Gesellschaft Ungarischer Germanisten /
DAAD (Eds.), Jahrbuch der ungarischen Germanistik (Reihe
Germanistik, pp. 93–104). Budapest: Gondolat Kiadói Kör.
(2014). Von
der Mündlichkeit zur Multimodalität … und darüber hinaus. In E. Grundler & C. Spiegel (Eds.), Konzeptionen
des Mündlichen — Wissenschaftliche Perspektiven und didaktische
Konsequenzen (pp. 13–31). Bern: hep Verlag.
Fischer, K. (2007). Komplexität
und semantische Transparenz im Deutschen und
Englischen. Sprachwissenschaft, 32(4), 355–405.
Fischer, H. (2011). Dialektalität
und Nähesprachlichkeit: Eine Anwendung des Nähechecks auf regional markiertes
Sprechen. In B. Ganswindt & C. Purschke (Eds.), Perspektiven
der Variationslinguistik. Beiträge aus dem Forum Sprachvariation (Germanistische Linguistik, 216–217,
pp. 121–147). Hildesheim u. a.: Olms.
Flowerdew, L. (2004). The
argument for using English specialized corpora to understand academic and professional
settings. In U. Connor & T. Upton (Eds.), Discourse
in the
professions (pp. 11–33). Amsterdam: Benjamins.
Geertz, C. (1983). Local
knowledge: Further essays in interpretive anthropology. New York: Basic Books.
Geyken, A., & Gloning, T. (2015). A
living text archive of 15th–19th century German: Corpus strategies, technology,
organization. In J. Gippert & R. Gehrke (Eds.), Corpus
linguistics and interdisciplinary perspectives on language — CLIP (Vol. 5: Historical Corpora:
Challenges and Perspectives. Proceedings of the conference Historical Corpora 2012, pp.
165–180). Tübingen: Narr.
Givón, T. (2009). The
genesis of syntactic complexity: Diachrony, ontogeny, neuro-cognition, evolution. Amsterdam, Philadelphia: John Benjamins.
Gogolin, I., Kaiser, G., Roth, H.-J., Deseniss, A., Hawighorst, B., & Schwarz, I. (2004). Mathematiklernen
im Kontext sprachlich-kultureller Diversität. DFG Go
614/6. Abschlussbericht. Hamburg: Universität Hamburg. [URL]
Gries, S. (2006). Exploring
variability within and between corpora: Some methodological
considerations. Corpora, 1(2), 109–151.
Grieve, J. (2014). Chapter
1.1: A multi-dimensional analysis of regional variation in American
English. In T. B. Sardinha & M. V. Pinto (Eds.), Multi-dimensional
analysis, 25 years on: A tribute to Douglas
Biber (pp. 3–34). Amsterdam: John Benjamins.
Harwood, N. (2005). Nowhere
has anyone attempted … in this article I aim to do just that: A corpus-based study of self-promotional I and We in academic
writing across four disciplines. Journal of
Pragmatics, 37(1), 1207–1231.
Hegedüs, I. (2006). Hans
Ludwig Nehrlich: Erlebnisse eines frommen Handwerkers im späten 17.
Jahrhundert. In V. Ágel & M. Hennig (Eds.), Grammatik
aus Nähe und Distanz: Theorie und Praxis am Beispiel von Nähetexten
1650–2000 (pp. 141–162). Tübingen: Niemeyer.
Hennig, M. (2007). Thesen
zur Erforschung historischer Nähesprachlichkeit. In M. Balaskó & P. Szatmári (Eds.), Sprach-
und Literaturwissenschaftliche Brückenschläge. Vorträge der 13. Jahrestagung der GESUS in Szombathely, 12.–14. Mai
2004 (pp. 13–26). München: Lincom (Edition
Linguistik 59). [URL]
(2009). Nähe
und Distanzierung. Verschriftlichung und Reorganisation des
Nähebereichs. Kassel: University Press. [URL]
Hennig, M., & Niemann, R. (2013). Unpersönliches
Schreiben in der Wissenschaft: Kompetenzunterschiede im interkulturellen
Vergleich. InfoDaF, 401, 622–646.
Hennig, M. (2015). Die
Bundespressekonferenz zwischen Nähe und Distanz. In S. Staffeldt & J. Hagemann (Eds.), Pragmatiktheorien:
Analysen im Vergleich (Stauffenburg-Einführungen, 27, pp.
247–279). Tübingen: Stauffenburg.
Hennig, M., & Meisner, P. (2023). Textmusterbildung
durch nominale Komplexität. In S. Haaf & B. Schuster (Eds.), Historische
Textmuster im Wandel (Reihe Germanistische Linguistik, 331, pp.
327–360). Berlin: De Gruyter.
Hennig, M. & Jacob, J. (2024). Realisiert
ein Dialog im literarischen Drama Mündlichkeit? Linguistische und literarästhetische
Überlegungen. In W. Imo & J. Wesche (Eds.), Sprechen
und Gespräch in historischer Perspektive (LiLi: Studien zu Literaturwissenschaft und Linguistik, Vol.
7, pp. 45–63). Berlin, Heidelberg: J.B. Metzler.
Hiltunen, T. (2021). Exploring
sub-register variation in Victorian newspapers: Evidence from the British Library Newspapers
database. In E. Seoane & D. Biber (Eds.), Corpus-based
approaches to register variation (Studies in Corpus Linguistics, 103, pp.
313–338). Amsterdam: Benjamins.
Hundt, M. (2008). Text
corpora. In A. Lüdeling & M. Kytö (Eds.), Corpus
linguistics: An international handbook (Vol. 1, Handbücher zur Sprach- und Kommunikationswissenschaft,
29(1), pp. 168–186). Berlin: De Gruyter.
Hunston, S. (2008). Collection
strategies and design decisions. In A. Lüdeling & M. Kytö (Eds.), Corpus
linguistics: An international handbook (Vol. 1, Handbücher zur Sprach- und Kommunikationswissenschaft,
29(1), pp. 154–168). Berlin: De Gruyter.
Hsiao, Y., Dawson, N., Bandyopadhyay, N., & Nation, K. (2024). A
corpus-based developmental investigation of linguistic complexity in children’s
writing. Applied Corpus
Linguistics, 4(1).
Imo, W. (2013). Sprache
in Interaktion: Analysemethoden und Untersuchungsfelder. Berlin, Boston: De Gruyter.
(2016). Das
Nähe-Distanz-Modell in der Konversationsanalyse/Interaktionalen Linguistik: Ein Versuch der Skizzierung einer
‚Nicht-Karriere‘. In H. Feilke & M. Hennig (Eds.), Zur
Karriere von ›Nähe und Distanz‹: Rezeption und Diskussion des
Koch-Oesterreicher-Modells (pp. 155–186). Berlin: De Gruyter.
Joulin, A., Grave, E., Bojanowski, P., & Mikolov, T. (2016). Bag
of tricks for efficient text classification.
Jucker, A. (1989). Douglas
Biber, Variation across speech and writing. Cambridge University Press, Cambridge. Journal of
Linguistics, 25(2), 480–484.
(2021). Features
of orality in the language of fiction: A corpus-based investigation. Language and
Literature, 30(4), 341–360.
Kappel, P. (2006a). Augustin
Güntzer: Kleines Biechlin von meinem gantzen Leben. Die Autobiographie eines Elsässer Kannengießers aus dem 17.
Jahrhundert. In V. Ágel & M. Hennig (Eds.), Grammatik
aus Nähe und Distanz: Theorie und Praxis am Beispiel von Nähetexten
1650–2000 (pp. 101–120). Tübingen: Niemeyer.
(2006b). Überlegungen
zur diatopischen Variation in der gesprochenen Sprache. In V. Ágel & M. Hennig (Eds.), Zugänge
zur Grammatik der gesprochenen
Sprache (pp. 215–244). Berlin, Boston: Max Niemeyer Verlag.
Kehrein, R. & Fischer, H. (2016). Nähe,
Distanz und Regionalsprache. In H. Feilke & M. Hennig (Eds.), Zur
Karriere von ›Nähe und Distanz‹: Rezeption und Diskussion des
Koch-Oesterreicher-Modells (pp. 213–258). Berlin, Boston: De Gruyter.
Kleinschmidt-Schinke, K. (2018). Die
an die Schüler/-innen gerichtete Sprache (SgS): Studien zur Veränderung der Lehrer/-innensprache von der Grundschule bis zur
Oberstufe. Berlin, Boston: De Gruyter.
Koch, P., & Oesterreicher, W. (1985). Sprache
der Nähe — Sprache der Distanz: Mündlichkeit und Schriftlichkeit im Spannungsfeld von Sprachtheorie und
Sprachgeschichte. Romanistisches
Jahrbuch, 361, 15–43.
(2007). Schriftlichkeit
und kommunikative Distanz. Zeitschrift für Germanistische
Linguistik, 35(3), 346–375.
(2011). Gesprochene
Sprache in der Romania: Französisch, Italienisch, Spanisch (2nd updated and expanded
ed.). Berlin: De Gruyter.
(2012). Language
of immediacy — language of distance: Orality and literacy from the perspective of language theory and linguistic history
(Transl. of Koch & Oesterreicher 1985 by F. H. Bäuml & U.
Schaefer). In U. Schaefer, C. Lange, B. Weber, & G. Wolf (Eds.), Communicative
spaces: Variation, contact, and change. Papers in honour of Ursula
Schaefer (pp. 441–473). Frankfurt am Main: Lang.
Koester, A. (2022). Building
small specialised corpora. In A. O’Keeffe & M. McCarthy (Eds.), The
Routledge handbook of corpus
linguistics (pp. 48–61). London: Routledge.
Kytö, M. (2019). Register
in historical linguistics. Register
Studies, 1(1), 136–167.
Laippala, V., Kyllönen, R., Egbert, J., Biber, D., & Pyysalo, S. (2019). Toward
multilingual identification of online registers. In Proceedings of
the 22nd Nordic Conference on Computational Linguistics, Turku,
Finland (pp. 292–297). Linköping: Linköping University Electronic Press. [URL]
Landert, D. & Jucker, A. H. (2011). Private
and public in mass media communication: From letters to the editor to online
commentaries. Journal of
Pragmatics, 43(5), 1422–1434.
Lenzhofer, M. (2017). Jugendkommunikation
und Dialekt: Syntax gesprochener Sprache bei Jugendlichen in Osttirol. Berlin, Boston: De Gruyter.
Leska, C. (1965). Vergleichende
Untersuchungen zur Syntax gesprochener und geschriebener deutscher Gegenwartssprache. Beiträge
zur Geschichte der deutschen Sprache und
Literatur, 871, 427–464.
Liimatta, A. (2023). Register
variation across text lengths: Evidence from social media. International Journal of Corpus
Linguistics, 28(2), 202–231.
Lu, X. (2009). Automatic
measurement of syntactic complexity in child language acquisition. International Journal of
Corpus
Linguistics, 14(1), 3–28.
(2010). Automatic
analysis of syntactic complexity in second language writing. International Journal of Corpus
Linguistics, 15(4), 474–496.
Macha, J. (2001). Michael
Zimmer’s diary: Ein deutsches Tagebuch aus dem Amerikanischen Bürgerkrieg. Frankfurt am Main: Lang.
Mair, C. (2018). Erfolgsgeschichte
Korpuslinguistik? In M. Kupietz & T. Schmidt (Eds.), Korpuslinguistik (pp. 5–26). Berlin: De Gruyter.
Mánássy, I. (2006). Bauernleben
im Zeitalter des Dreißigjährigen Krieges: Die Stausebacher Chronik des Caspar Preis 1636–1667 [Bauernleben
I]. In V. Ágel & M. Hennig (Eds.), Grammatik
aus Nähe und Distanz: Theorie und Praxis am Beispiel von Nähetexten
1650–2000 (pp. 77–100). Tübingen: Niemeyer.
Moisl, H. (2015). Cluster
analysis for corpus linguistics (Quantitative Linguistics,
66). Berlin: De Gruyter.
Molnár, P., & Zóka, E. (2006). Wenn
doch dies Elend ein Ende hätte: Ein Briefwechsel aus dem Deutsch-Französischen Krieg
1870/71. In V. Ágel & M. Hennig (Eds.), Grammatik
aus Nähe und Distanz: Theorie und Praxis am Beispiel von Nähetexten
1650–2000 (pp. 279–296). Tübingen: Niemeyer.
Nini, A. (2019). The
Multi-Dimensional Analysis Tagger. In T. Berber Sardinha & M. Pinto (Eds.), Multi-dimensional
analysis: Research methods and current
issues (pp. 67–94). London/New York: Bloomsbury Academic.
Odebrecht, C., Belz, M., Zeldes, A., Lüdeling, A., & Krause, T. (2017). RIDGES
Herbology: Designing a diachronic multi-layer corpus. Language Resources &
Evaluation, 51(3), 695–725.
Ortmann, K., & Dipper, S. (2019). Variation
between different discourse types: Literate vs. oral. In Proceedings
of the Sixth Workshop on NLP for Similar Languages, Varieties and Dialects, Ann Arbor,
Michigan (pp. 64–79). Stroudsburg, PA: Association for Computational Linguistics.
(2020). Automatic
orality identification in historical texts. In Proceedings of the
Twelfth Language Resources and Evaluation Conference, Marseille,
France (pp. 1293–1302). European Language Resources Association. [URL]
(2024). Nähetexte
automatisch erkennen: Entwicklung eines linguistischen Scores für konzeptionelle Mündlichkeit in historischen
Texten. In W. Imo & J. Wesche (Eds.), Sprechen
und Gespräch in historischer Perspektive (LiLi: Studien zu Literaturwissenschaft und Linguistik, 7,
pp.
17–36). Berlin: Metzler.
Osborne, J. (2011). Fluency,
complexity and informativeness in native and non-native speech. International Journal of Corpus
Linguistics, 16(2), 276–298.
Peters, J. (Ed.). (1993). Ein
Söldnerleben im Dreißigjährigen Krieg: Eine Quelle zur
Sozialgeschichte. Berlin: Akademie Verlag.
Petran, F. (2012). Studies
for segmentation of historical texts: Sentences or
chunks? In Proceedings of the Second Workshop on Annotation of
Corpora for Research in the Humanities
(ACRH-2) (pp. 75–86).
Polenz, P. (1999). Deutsche
Sprachgeschichte vom Spätmittelalter bis zur Gegenwart (Vol. 3: 19. und 20.
Jahrhundert). Berlin: De Gruyter.
Raible, W. (2019). Variation
in language: How to characterise types of texts and communication strategies between orality and scripturality. Answers given
by Koch/Oesterreicher and by Biber. International Journal of Language and
Linguistics, 6(2).
Rauzs, O. (2006a). Meister
Johann Dietz: Des Großen Kurfürsten Feldscher und Königlicher
Hofbarbier. In V. Ágel & M. Hennig (Eds.), Grammatik
aus Nähe und Distanz: Theorie und Praxis am Beispiel von Nähetexten
1650–2000 (pp. 163–181). Tübingen: Niemeyer.
(2006b). Ulrich
Bräker: Lebensgeschichte und natürliche Ebentheur des armen Mannes im
Tockenburg. In V. Ágel & M. Hennig (Eds.), Grammatik
aus Nähe und Distanz: Theorie und Praxis am Beispiel von Nähetexten
1650–2000 (pp. 201–219). Tübingen: Niemeyer.
Repo, L., Hashimoto, B., & Laippala, V. (2023). In
search of founding era registers: Automatic modeling of registers from the corpus of Founding Era American
English. Digital Scholarship in the
Humanities, 38(4), 1659–1677.
Riebling, L. (2013). Heuristik
der Bildungssprache. In I. Gogolin, I. Lange, U. Michel & H. H. Reich (Eds.), Herausforderung
Bildungssprache — und wie man sie
meistert (pp. 106–153). Münster: Waxmann.
Rodríguez-Puente, P. (2019). The
English phrasal verb, 1650–present: History, stylistic drifts, and
lexicalisation. Cambridge: Cambridge University Press.
Roth, H.-J., Neumann, U., & Gogolin, I. (2007). Schulversuch
bilinguale Grundschulklassen in Hamburg: Abschlussbericht über die italienisch-deutschen, portugiesisch-deutschen und
spanisch-deutschen Modellklassen. Hamburg. [URL]
Rudnicka, K. (2018). Variation
of sentence length across time and genre: Influence on syntactic usage in
English. In R. Whitt (Ed.), Diachronic
corpora, genre, and language
change (pp. 219–240). Amsterdam: Benjamins.
Sardinha, T., & Pinto, M. (Eds.). (2014). Multi-dimensional
analysis, 25 years on: A tribute to Douglas Biber (Studies in Corpus Linguistics,
60). Amsterdam: Benjamins.
Schaefer, U. (2021). Communicative
distance: The (non-)reception of Koch and Oesterreicher in English-speaking
linguistics. Anglistik, 32(2), 15–42.
Schank, G., & Schoenthal, G. (1976). Gesprochene
Sprache: Eine Einführung in Forschungsansätze und
Analysemethoden. Tübingen: Niemeyer.
Schleppegrell, M. J. (2004). The
language of schooling: A functional linguistics perspective. Mahwah, NJ: Erlbaum.
Schlieben-Lange, B. (1973). Soziolinguistik:
Eine Einführung (3rd revised and expanded ed.). Stuttgart et al.: Kohlhammer.
Schmoller, G., & Kraus, H. (2010). Über
die „Gedanken und Erinnerungen“ von Otto Fürst von Bismarck: Mit einem Nachwort von Hans-Christof Kraus (1st
ed.). Berlin: Duncker & Humblot.
Seoane, E., & Biber, D. (Eds.). (2021). Corpus-based
approaches to register variation (Studies in Corpus Linguistics,
103). Amsterdam: Benjamins.
Sigley, R. (1997). Text
categories and where you can stick them: A crude formality index. International Journal of
Corpus
Linguistics, 2(2), 199–237.
(2006). Corpora
in studies of variation. In K. Brown (Ed.), Encyclopedia
of language & linguistics (2nd
ed., pp. 220–226). Amsterdam: Elsevier.
(2005). Corpus
and text — basic principles. In M. Wynne (Ed.), Developing
linguistic corpora: A guide to good
practice (pp. 1–16). Oxford: Oxbow Books.
Smirnova, E. (2021). The
diachronic development of the verbal bracket construction in
German. In S. Hartmann & A. Quick (Eds.), Yearbook
of the German Cognitive Linguistics
Association, 9(1), 157–176.
Steger, H. (1987). Bilden
gesprochene und geschriebene Sprache eigene
Sprachvarietäten? In H. Aust & T. Lewandowski (Eds.), Wörter: Schätze, Fugen und Fächer des Wissens. Festgabe für Theodor Lewandowski zum
60. Geburtstag (pp. 35–58). Tübingen: Narr.
Tagliavini, C. (1998). Einführung
in die romanische Philologie. Tübingen: UTB für Wissenschaft, Francke.
Vinckel-Roisin, H. (2015). Das
Nachfeld im Deutschen: Theorie und Empirie (Reihe Germanistische Linguistik,
303). Berlin: De Gruyter.
Vetter, F. (2021). Issues
of corpus comparability and register variation in the International Corpus of English: Theories and computer
applications (Doctoral
dissertation, Otto-Friedrich-Universität Bamberg).
Weiss, Z., Lange-Schubert, K., Geist, B. & Meurers, D. (2022). Sprachliche
Komplexität im Unterricht: Eine computerlinguistische Analyse der gesprochenen Sprache von Lehrenden und Lernenden im
naturwissenschaftlichen Unterricht in der Primar- und Sekundarstufe. Zeitschrift für
germanistische
Linguistik, 50(1), 159–201.
Zahn, G. (1991). Beobachtungen
zur Ausklammerung und Nachfeldbesetzung in gesprochenem
Deutsch. Erlangen: Palm und Enke.
Zeman, S. (2010). Tempus
und „Mündlichkeit“ im Mittelhochdeutschen: Zur Interdependenz grammatischer Perspektivensetzung und »Historischer
Mündlichkeit « im mittelhochdeutschen Tempussystem. Berlin, New York: De Gruyter.
(2013a). Mündlichkeit
ist nicht gleich Mündlichkeit: Implikationen für eine Theorie der gesprochenen
Sprache. In J. Hagemann, W. P. Klein & S. Staffeldt (Eds.), Pragmatischer
Standard (pp. 191–206). Tübingen: Stauffenburg.
GiesKaNe, corpus texts: [URL]
Allgemeine Zeitung (1840). Nr. 1.
Augsburg, 1. Januar 1840. In: Deutsches
Textarchiv. [URL]
Bismarck, Otto von. (1898): Gedanken und Erinnerungen.
Bd. 2. Stuttgart, 1898. In: Deutsches
Textarchiv. [URL]
Eckermann, Johann Peter. (1836): Gespräche mit Goethe in
den letzten Jahren seines
Lebens. Bd. 11. Leipzig, 1836. In: Deutsches
Textarchiv. [URL]
Humboldt, Alexander von. (1795): Brief an Samuel Thomas
Soemmerring. Bayreuth, 07.06.1795. In: Deutsches
Textarchiv. [URL]
. (1854): Auszug aus einem Briefe
Alexanders von Humboldt an den Verfasser. In: Berg, Albert:
Physiognomie der tropischen Vegetation
Süd-Americas. Düsseldorf, 1854. In: Deutsches Textarchiv. [URL]
Neue Rheinische Zeitung (1848). Nr.
1. Köln, 1. Juni 1848. In: Deutsches Textarchiv. [URL]
(1848). Nr.
85. Köln, 25. August 1848. Beilage. In: Deutsches
Textarchiv. [URL]
(1849). Nr.
198. Köln, 18. Januar 1849. Beilage. In: Deutsches
Textarchiv. [URL]
[N. N.] (1488): Historia Dracole
Waida. [Nürnberg], 1488. In: Deutsches
Textarchiv. [URL]
(1732): Ausführliche
Nachricht von dem, was alhier zu Halle mit denen Saltzburgischen Emigranten vorgegangen. [s. l.],
1732. In: Deutsches Textarchiv. [URL]
Sanders, Daniel. (1871): Brief
an Adolf Glaßbrenner. Altstrelitz, 26. März 1871. In: Deutsches
Textarchiv. [URL]
. (1879): Brief
an Wilhelm Scherer. Altstrelitz, 7. November 1879. In: Deutsches
Textarchiv. [URL]
. (1884): Brief
an Wilhelm Scherer. Altstrelitz, 7. November 1884. In: Deutsches
Textarchiv. [URL]