Thomas Stephan Juzek

Thomas Stephan Juzek

Assistant Professor of Computational Linguistics
Department of Scientific Computing, Florida State University

Thomas Stephan Juzek

Hi there, and thanks for visiting! Please feel free to call me Tommie. As a computational linguist at Florida State University, I focus on large language model alignment; specifically: the underlying mechanisms driving different language behaviours, AI-associated language change, and the broader societal impacts of language technology. I am based in the Department of Scientific Computing and also hold a courtesy appointment in Modern Languages and Linguistics. I joined FSU in 2022, after a DPhil at the University of Oxford in 2016 and (mostly) industry work in natural language processing.

Much of my current research starts from a simple observation: language models have developed recognisable ways of using language, and those patterns are increasingly interacting with human language use. I am interested in where these behaviours come from, how they differ across models and training methods, and what happens as model-generated language becomes part of the linguistic environment humans learn from and write in.

We are seeing some remarkable shifts in human language use whose timing and pattern give AI a plausible role; establishing what actually drives them is part of my current work. More broadly, language behaviour is an excellent testbed for model behaviour, and language alignment is part of the wider effort to align models with human expectations.

I also teach courses in computational linguistics, and I enjoy mentoring undergraduate and graduate research on language technology and AI. I see my work as being in service of the community and the field, whether through public talks, open tools, or freely shared data. I am always happy to hear from students, and open to media enquiries.

Selected research

2026 · preprint

AI-Associated Lexical Shifts Across 34 Languages

Cross-lingual convergence: emphasize-type verbs surface in 24 of 34 languages.

LREC 2026

Fully Automated Identification of Lexical Alignment & Preference-Stage Shifts

A reusable, curation-free diagnostic for AI lexical overuse: the foundational pipeline.

AIES 2025

Model Misalignment & Language Change

First peer-reviewed evidence of AI-associated language traces in unscripted spoken English.

Interactive tool

AI Word Explorer

Browse AI-overused words by language, register, and model, now including GPT-5.2.

aiwordexplorer.com

For more papers, see the research page or Google Scholar.

In the news

… and many more. See all coverage →