← People
Sara Hooker
following
Cohere Labs
@sarahookr
Papers in the feed →
Papers · 42
Open-World Evaluations for Measuring Frontier AI Capabilities
arXiv.org
2026-05-19
alphaXiv
arXiv
S2
Tiny Aya: Bridging Scale and Multilingual Depth
arXiv.org
2026-03-12
alphaXiv
arXiv
S2
SimMerge: Learning to Select Merge Operators from Similarity Signals
arXiv.org
2026-01-14
alphaXiv
arXiv
S2
The Art of Asking: Multilingual Prompt Optimization for Synthetic Data
arXiv.org
2025-10-22
alphaXiv
arXiv
S2
The Disparate Impacts of Speculative Decoding
arXiv.org
2025-10-02
alphaXiv
arXiv
S2
Verification Limits Code LLM Training
arXiv.org
2025-09-25
alphaXiv
arXiv
S2
When Life Gives You Samples: The Benefits of Scaling up Inference Compute for Multilingual LLMs
Conference on Empirical Methods in Natural Language Processing
2025-06-25
alphaXiv
arXiv
S2
Language Models can perform Single-Utterance Self-Correction of Perturbed Reasoning
arXiv.org
2025-06-18
alphaXiv
arXiv
S2
Treasure Hunt: Real-time Targeting of the Long Tail using Training-Time Markers
Neural Information Processing Systems
2025-06-17
alphaXiv
arXiv
S2
One Tokenizer To Rule Them All: Emergent Language Plasticity via Multilingual Tokenizers
Annual Meeting of the Association for Computational Linguistics
2025-06-12
alphaXiv
arXiv
S2
The Multilingual Divide and Its Impact on Global AI Safety
arXiv.org
2025-05-27
alphaXiv
arXiv
S2
Aya Vision: Advancing the Frontier of Multilingual Multimodality
arXiv.org
2025-05-13
alphaXiv
arXiv
S2
The Leaderboard Illusion
Neural Information Processing Systems
2025-04-29
alphaXiv
arXiv
S2
Kaleidoscope: In-language Exams for Massively Multilingual Vision Evaluation
arXiv.org
2025-04-09
alphaXiv
arXiv
S2
Command A: An Enterprise-Ready Large Language Model
2025-04-01
alphaXiv
arXiv
S2
MMTEB: Massive Multilingual Text Embedding Benchmark
arXiv.org
2025-02-19
alphaXiv
arXiv
S2
Multilingual Machine Translation with Open Large Language Models at Practical Scale: An Empirical Study
North American Chapter of the Association for Computational Linguistics
2025-02-04
alphaXiv
arXiv
S2
Fairness of Deep Ensembles: On the interplay between per-group task difficulty and under-representation
Conference on Fairness, Accountability and Transparency
2025-01-24
alphaXiv
arXiv
S2
Global MMLU: Understanding and Addressing Cultural and Linguistic Biases in Multilingual Evaluation
Annual Meeting of the Association for Computational Linguistics
2025
S2
RLHF Algorithms Ranked: An Extensive Evaluation Across Diverse Tasks, Rewards, and Hyperparameters
Conference on Empirical Methods in Natural Language Processing
2025
S2
To Code or Not To Code? Exploring Impact of Code in Pre-training
International Conference on Learning Representations
2025
S2
Nexus: Adaptive Upcycling to Efficiently Pretrain Mixture of Experts
Conference on Empirical Methods in Natural Language Processing
2025
S2
Multilingual Arbitration: Optimizing Data Pools to Accelerate Multilingual Progress
Annual Meeting of the Association for Computational Linguistics
2025
S2
Open Problems in Technical AI Governance
Trans. Mach. Learn. Res.
2025
S2
Bridging the Data Provenance Gap Across Text, Speech and Video
arXiv.org
2024-12-19
alphaXiv
arXiv
S2
Aya Expanse: Combining Research Breakthroughs for a New Multilingual Frontier
arXiv.org
2024-12-05
alphaXiv
arXiv
S2
Global MMLU: Understanding and Addressing Cultural and Linguistic Biases in Multilingual Evaluation
arXiv.org
2024-12-04
alphaXiv
arXiv
S2
The Reality of AI and Biorisk
Conference on Fairness, Accountability and Transparency
2024-12-02
alphaXiv
arXiv
S2
INCLUDE: Evaluating Multilingual Language Understanding with Regional Knowledge
International Conference on Learning Representations
2024-11-29
alphaXiv
arXiv
S2
M-RewardBench: Evaluating Reward Models in Multilingual Settings
Annual Meeting of the Association for Computational Linguistics
2024-10-20
alphaXiv
arXiv
S2
Mix Data or Merge Models? Optimizing for Diverse Multi-Task Learning
arXiv.org
2024-10-14
alphaXiv
arXiv
S2
Nexus: Specialization meets Adaptability for Efficiently Training Mixture of Experts
arXiv.org
2024-08-28
alphaXiv
arXiv
S2
Multilingual Arbitrage: Optimizing Data Pools to Accelerate Multilingual Progress
arXiv.org
2024-08-27
alphaXiv
arXiv
S2
To Code, or Not To Code? Exploring Impact of Code in Pre-training
arXiv.org
2024-08-20
alphaXiv
arXiv
S2
The future of open human feedback
Nature Machine Intelligence
2024-08-15
alphaXiv
arXiv
S2
LLM See, LLM Do: Leveraging Active Inheritance to Target Non-Differentiable Objectives
Conference on Empirical Methods in Natural Language Processing
2024
S2
Robust distillation for worst-class performance: on the interplay between teacher and student objectives
Conference on Uncertainty in Artificial Intelligence
2023
S2
The Goldilocks of Pragmatic Understanding: Fine-Tuning Strategy Matters for Implicature Resolution by LLMs
Neural Information Processing Systems
2023
S2
Cross-lingual Transfer Dynamics in BLOOMZ: Insights into Multilingual Generalization
S2
Breaking mBad! Supervised Fine-tuning for Cross-Lingual Detoxification
S2
The Data Provenance Project
S2
Capabilities and risks from frontier AI
S2