← People
María Grandury
following
EPFL / SomosNLP
@mariagrandury
Papers in the feed →
Papers · 17
Updating the German Psycholinguistic Word Toolbox with AI-Generated Estimates of Concreteness, Valence, Arousal, Age of Acquisition, and Familiarity
Journal of Cognition
2026-01-08
S2
Apertus: Democratizing Open and Compliant LLMs for Global Language Environments
Annual Meeting of the Association for Computational Linguistics
2026
S2
Learning Vision-Language Alignment in Unified LLMs with 24 Text Tokens per Image
International Workshop on Spoken Language Translation
2026
S2
Measuring what Matters: Construct Validity in Large Language Model Benchmarks
Advances in Neural Information Processing Systems 38
2025-11-03
alphaXiv
arXiv
S2
BabyBabelLM: A Multilingual Benchmark of Developmentally Plausible Training Data
Conference of the European Chapter of the Association for Computational Linguistics
2025-10-11
alphaXiv
arXiv
S2
Spanish is not just one: A dataset of Spanish dialect recognition for LLMs
Data in Brief
2025-09-18
S2
Apertus: Democratizing Open and Compliant LLMs for Global Language Environments
2025-09-17
alphaXiv
arXiv
S2
Adding LLMs to the psycholinguistic norming toolbox: A practical guide to getting the most out of human ratings
Behavior Research Methods
2025-09-17
alphaXiv
arXiv
S2
La Leaderboard: A Large Language Model Leaderboard for Spanish Varieties and Languages of Spain and Latin America
Annual Meeting of the Association for Computational Linguistics
2025-07-01
alphaXiv
arXiv
S2
Psycholinguistic Word Features: a New Approach for the Evaluation of LLMs Alignment with Humans
arXiv.org
2025-05-29
alphaXiv
arXiv
S2
Kaleidoscope: In-language Exams for Massively Multilingual Vision Evaluation
arXiv.org
2025-04-09
alphaXiv
arXiv
S2
It's the same but not the same: Do LLMs distinguish Spanish varieties?
Proces. del Leng. Natural
2025-04-08
alphaXiv
arXiv
S2
Multiple Choice Questions: Reasoning Makes Large Language Models (LLMs) More Self-Confident, specially When They are Wrong
IEEE Intelligent Systems
2025-01-16
alphaXiv
arXiv
S2
Apertus: Democratizing Open and Compliant LLMs for Global Language Environments
arXiv.org
2025
S2
Multiple Choice Questions: Reasoning Makes Large Language Models (LLMs) More Self-Confident Even When They Are Wrong
arXiv.org
2025
S2
Evaluating Large Language Models with Tests of Spanish as a Foreign Language: Pass or Fail?
arXiv.org
2024-09-08
alphaXiv
arXiv
S2
Data in Brief
S2