← People
Catherine Arnett
interacted with
EleutherAI
met · ACL 2026 · 2026-07-02
@@linguist_cat
LinkedIn
Google Scholar
Website
Papers in the feed →
Papers · 20
Position: Don't Just "Fix it in Post": A Science of AI Must Study Training Dynamics
arXiv.org
2026-06-03
alphaXiv
arXiv
S2
Weight Tying Biases Token Embeddings Towards the Output Space
Annual Meeting of the Association for Computational Linguistics
2026-03-27
alphaXiv
arXiv
S2
How Open Must Language Models be to Enable Reliable Scientific Inference?
arXiv.org
2026-03-27
alphaXiv
arXiv
S2
CommonLID: Re-evaluating State-of-the-Art Language Identification Performance on Web Data
Annual Meeting of the Association for Computational Linguistics
2026-01-25
alphaXiv
arXiv
S2
Disaggregation Reveals Hidden Training Dynamics: The Case of Agreement Attraction
arXiv.org
2025-10-28
alphaXiv
arXiv
S2
Global PIQA: Evaluating Commonsense Reasoning Across 100+ Languages and Cultures
2025-10-28
alphaXiv
arXiv
S2
Explaining and Mitigating Crosslingual Tokenizer Inequities
Neural Information Processing Systems
2025-10-24
alphaXiv
arXiv
S2
Evaluating Morphological Alignment of Tokenizers in 70 Languages
arXiv.org
2025-07-08
alphaXiv
arXiv
S2
On the Acquisition of Shared Grammatical Representations in Bilingual Language Models
Annual Meeting of the Association for Computational Linguistics
2025-03-05
alphaXiv
arXiv
S2
Why do language models perform worse for morphologically complex languages?
International Conference on Computational Linguistics
2025
S2
Global PIQA: Evaluating Physical Commonsense Reasoning Across 100+ Languages and Cultures
arXiv.org
2025
S2
Why do language models perform worse for morphologically complex languages?
arXiv.org
2024-11-21
alphaXiv
arXiv
S2
Syntax drives default language selection in bilingual connected speech production
Journal of Experimental Psychology. Learning, Memory and Cognition
2024-10-17
S2
Goldfish: Monolingual Language Models for 350 Languages
arXiv.org
2024-08-19
alphaXiv
arXiv
S2
Revenge of the Fallen? Recurrent Models Match Transformers at Predicting Human Language Comprehension Metrics
arXiv.org
2024-04-30
alphaXiv
arXiv
S2
Different Tokenization Schemes Lead to Comparable Performance in Spanish Number Agreement
Special Interest Group on Computational Morphology and Phonology Workshop
2024-03-20
alphaXiv
arXiv
S2
A Bit of a Problem: Measurement Disparities in Dataset Sizes across Languages
SIGUL
2024-03-01
alphaXiv
arXiv
S2
When Is Multilinguality a Curse? Language Modeling for 250 High- and Low-Resource Languages
Conference on Empirical Methods in Natural Language Processing
2023-11-15
alphaXiv
arXiv
S2
Structural Priming Demonstrates Abstract Grammatical Representations in Multilingual Language Models
Conference on Empirical Methods in Natural Language Processing
2023-11-15
alphaXiv
arXiv
S2
Crosslingual Structural Priming and the Pre-Training Dynamics of Bilingual Language Models
arXiv.org
2023-10-11
alphaXiv
arXiv
S2