References (46)
References
Altman, S. (2023, February 24). Planning for AGI and beyond. OpenAI. [URL]
Amodei, D. (2024, October 24). Machines of loving grace: How AI could transform the world for the better. darioamodei.com. [URL]
Anthropic. (2024, June 8). Claude’s character. Anthropic. [URL]
Bai, Y., Kadavath, S., Kundu, S., Askell, A., Kernion, J., Jones, A., Chen, A., Goldie, A., Mirhoseini, A., McKinnon, C., Chen, C., Olsson, C., Olah, C., Hernandez, D., Drain, D., Ganguli, D., Li, D., Tran-Johnson, E., Perez, E., Kerr, J., Mueller, J., Ladish, J., Landau, J., Ndousse, K., Lukosuite, K., Lovitt, L., Sellitto, M., Elhage, N., Schiefer, N., Mercado, N., DasSarma, N., Lasenby, R., Larson, R., Ringer, S., Johnston, S., Kravec, S., El Showk, S., Fort, S., Lanham, T., Telleen-Lawton, T., Conerly, T., Henighan, T., Hume, T., Bowman, S. R., Hatfield-Dodds, Z., Mann, B., Amodei, D., Joseph, N., McCandlish, S., Brown, T., & Kaplan, J. (2022). Constitutional AI: Harmlessness from AI feedback. arXiv preprint arXiv:2212.08073. [URL]
Balloccu, S., Schmidtová, P., Lango, M., & Dušek, O. (2024). Leak, cheat, repeat: Data contamination and evaluation malpractices in closed-source LLMs. In Y. Graham & M. Purver (Eds.), Proceedings of the 18th conference of the European chapter of the Association for Computational Linguistics (pp. 67–93). Association for Computational Linguistics. Google Scholar logo with link to Google Scholar
Berber Sardinha, T. (2024). AI-generated vs human-authored texts: A multidimensional comparison. Applied Corpus Linguistics, 4(1), 100083. Google Scholar logo with link to Google Scholar
Biber, D. (1988). Variation across speech and writing. Cambridge University Press. Google Scholar logo with link to Google Scholar
(2019). Multi-dimensional analysis: A historical synopsis. In T. Berber Sardinha & M. Veirano Pinto (Eds.), Multi-dimensional analysis: Research methods and current issues (pp. 11–26). Bloomsbury. Google Scholar logo with link to Google Scholar
Biber, D., & Conrad, S. (2009). Register, genre, and style. Cambridge University Press. Google Scholar logo with link to Google Scholar
Bitton, Y., Bitton, E., & Nisan, S. (2025). Detecting stylistic fingerprints of large language models. arXiv:2503.01659v1. Google Scholar logo with link to Google Scholar
Brown, T. B., Mann, B., Ryder, N., Subbiah, M., Kaplan, J., Dhariwal, P., Neelakantan, A., Shyam, P., Sastry, G., Askell, A., Agarwal, S., Herbert-Voss, A., Krueger, G., Henighan, T., Child, R., Ramesh, A., Ziegler, D. M., Wu, J., Winter, C., Hesse, C., Chen, M., Sigler, E., Litwin, M., Gray, S., Chess, B., Clark, J., Berner, C., MCandlish, S., Radford, A., Sutskever, I., & Amodei, D. (2020). Language models are few-shot learners. arXiv:2005.14165v4. [URL]
Common Crawl Foundation. (2025). Common Crawl. [URL]
Cvrček, V., Komrsková, Z., Lukeš, D., Poukarová, P., Řehořková, A., & Zasina, A. J. (2021). From extra- to intratextual characteristics: Charting the space of variation in Czech through MDA. Corpus Linguistics and Linguistic Theory, 17(2), 351–382. Google Scholar logo with link to Google Scholar
Cvrček, V., Laubeová, Z., Lukeš, D., Poukarová, P., Řehořková, A., & Zasina, A. J. (2020). Registry v češtině. NLN.Google Scholar logo with link to Google Scholar
da Silva, T. H., Furtado, V., Furtado, E., Mendes, M., Almeida, V., & Sales, L. (2024). How do illiterate people interact with an intelligent voice assistant? International Journal of Human–Computer Interaction, 40(3), 584–602. Google Scholar logo with link to Google Scholar
Gage, P. (1994). A new algorithm for data compression. The C Users Journal, 12(2), 23–38. Google Scholar logo with link to Google Scholar
Garg, A. (2025, June 10). Google claims AI helping engineers do 10 percent more productive tasks, says AI agents are happening. India Today. [URL]
Goulart, L., Matte, M. L., Mendoza, A., Alvarado, L., & Veloso, I. (2024). AI or student writing? Analyzing the situational and linguistic characteristics of undergraduate student writing and AI-generated assignments. Journal of Second Language Writing, 661, 101160. Google Scholar logo with link to Google Scholar
Johnson, R. L., Pistilli, G., Menédez-González, N., Duran, L. D. D., Panai, E., Kalpokiene, J., & Bertulfo, D. J. (2022). The ghost in the machine has an American accent: Value conflict in GPT-3. arXiv:2203.07785v1. Google Scholar logo with link to Google Scholar
Kubát, M. (2014). Moving window type-token ratio and text length. In G. Altmann, R. Čech, J. Mačutek, & L. Uhlířová (Eds.), Empirical approaches to text and language analysis (pp. 105–113). RAM-Verlag.Google Scholar logo with link to Google Scholar
Kumarage, T., Garland, J., Bhattacharjee, A., Trapeznikov, K., Ruston, S., & Liu, H. (2023). Stylometric detection of AI-generated text in Twitter timelines. arXiv:2303.03697v1. [URL]
Malik, M., Jiang, J., & Chai, K. M. A. (2024). An empirical analysis of the writing styles of persona-assigned LLMs. In Y. Al-Onaizan, M. Bansal, & Y.-N. Chen (Eds.), Proceedings of the 2024 conference on empirical methods in natural language processing (pp. 19369–19388). Association for Computational Linguistics. Google Scholar logo with link to Google Scholar
Mikros, G. (2025). Beyond the surface: Stylometric analysis of GPT-4o’s capacity for literary style imitation. Digital Scholarship in the Humanities, 40(2), 587–600. Google Scholar logo with link to Google Scholar
Mikros, G. K., Koursaris, A., Bilianos, D., & Markopoulos, G. (2023). AI-writing detection using an ensemble of transformers and stylometric features. In M. Montes-y-Gómez, F. Rangel, S. M. Jiménez-Zafra, M. Casavantes, B. Altuna, M. Á. Álvarez-Carmona, G. Bel-Enguix, L. Chiruzzo, I. de la Iglesia, H. J. Escalante, M. Á. García-Cumbreras, J. A. García-Díaz, J. Á. González Barba, R. L. Tamayo, S. Lima, P. Moral, F. Miriam, P. del Arco, & R. Valencia-García (Eds.), Proceedings of the Iberian Languages Evaluation Forum (IberLEF 2023) co-located with the conference of the Spanish Society for Natural Language Processing (SEPLN 2023). CEUR Workshop Proceedings. [URL]
Milička, J., Marklová, A., VanSlambrouck, K., Pospíšilová, E., Šimsová, J., Harvan, S., & Drobil, O. (2024). Large language models are able to downplay their cognitive abilities to fit the persona they simulate. PLOS One, 19(3), e0298522. Google Scholar logo with link to Google Scholar
Milička, J., Marklová, A., & Cvrček, V. (2025a). AI Brown v1. LINDAT/CLARIAH-CZ Digital library at the Institute of Formal and Applied Linguistics (ÚFAL). [URL]
(2025b). AI Koditex v1. LINDAT/CLARIAH-CZ Digital library at the Institute of Formal and Applied Linguistics (ÚFAL), [URL]
(in press). AI Brown and AI Koditex: LLM-generated corpora comparable to human-written corpora of English and Czech texts. Language Resources and Evaluation.
Milička, J., Marklová, A., Drobil, O., & Pospíšilová, E. (2025). Learning to detect AI texts and learning the limits. PloS One, 20(10), e0333007. Google Scholar logo with link to Google Scholar
Ni, S., Kong, X., Li, C., Hu, X., Xu, R., Zhu, J., & Yang, M. (2025). Training on the benchmark is not all you need. Proceedings of the AAAI Conference on Artificial Intelligence, 39(23), 24948–24956. Google Scholar logo with link to Google Scholar
Nini, A. (2019). The multi-dimensional analysis tagger. In T. Berber Sardinha & M. Veirano Pinto (Eds.), Multi-dimensional analysis: Research methods and current issues (pp. 67–94). Bloomsbury. Google Scholar logo with link to Google Scholar
Peeperkorn, M., Kouwenhoven, T., Brown, D., & Jordanous, A. (2024). Is temperature the creativity parameter of large language models? arXiv:2405.00492v1. Google Scholar logo with link to Google Scholar
Przystalski, K., Argasiński, J. K., Grabska-Gradzińska, I., & Ochab, J. K. (2026). Stylometry recognizes human and LLM-generated texts in short samples. Expert Systems with Applications, 2961, 129001. Google Scholar logo with link to Google Scholar
Rao, Z., Mohamed, Y., Liu, S., & Liu, Z. (2025). Two birds with one stone: Multi-task detection and attribution of LLM-generated text. arXiv:2508.14190v1. Google Scholar logo with link to Google Scholar
Reinhart, A., Markey, B., Laudenbach, M., Pantusen, K., Yurko, R., Weinberg, G., & Brown, D. W. (2025). Do LLMs write like humans? Variation in grammatical and rhetorical styles. Proceedings of the National Academy of Sciences, 122(8), e2422455122. Google Scholar logo with link to Google Scholar
Rozado, D. (2024). The political preferences of LLMs. PLOS One, 19(7), e0306621. Google Scholar logo with link to Google Scholar
Rudnicka, K. (2025a, July 9). Each AI chatbot has its own, distinctive writing style just as humans do. Scientific American. [URL]
(2025b). The language of AI tools as idiolects — thus comparable to other idiolects. OSFPREPRINTS. Google Scholar logo with link to Google Scholar
Schut, L., Gal, Y., & Farquhar, S. (2025). Do multilingual LLMs think in English? arXiv:2502.15603v1. Google Scholar logo with link to Google Scholar
Shanahan, M., McDonell, K., & Reynolds, L. (2023). Role play with large language models. Nature, 6231, 493–498. Google Scholar logo with link to Google Scholar
Tully, T., Redfern, J., Das, D., & Xiao, D. (2025, July 13). 2025 Mid-year LLM market update: Foundation model landscape + economics. Menlo Ventures. [URL]
Vaswani, A., Shazeer, N., Parmar, N., Uszkoreit, J., Jones, L., Gomez, A. N., Kaiser, Ł., & Polosukhin, I. (2017). Attention is all you need. In U. von Luxburg, I. Guyon, S. Bengio, H. Wallach, & R. Fergus (Eds.), NIPS’17: Proceedings of the 31st international conference on Neural Information Processing System, (pp. 6000–6010). Association for Computing Machinery. Google Scholar logo with link to Google Scholar
Xu, H., Shi, Z. J., & Shi, M. (2025). Bonding with AI: Investigating the love relationships between humans and AI companions [Master’s thesis]. The Hong Kong University of Science and Technology.
Zasina, J., Lukeš, D., Komrsková, Z., Poukarová, P., & Řehořková, A. (2018). Koditex: A corpus of diversified texts. Institute of the Czech National Corpus, Faculty of Arts, Charles University. [URL]
Zhong, C., Cheng, F., Liu, Q., Jiang, J., Wan, Z., Chu, C., Murawaki, Y., & Kurohashi, S. (2024). Beyond English-centric LLMs: What language do multilingual language models think in? arXiv:2408.10811v1. Google Scholar logo with link to Google Scholar
Mobile Menu Logo with link to supplementary files background Layer 1 prag Twitter_Logo_Blue