← People
Marzieh Fadaee

Marzieh Fadaee following

Cohere Labs
@mziizmPapers in the feed →

Papers · 43
  1. CALIBER: Calibrating Confidence Before and After Reasoning in Language Models
    arXiv.org2026-06-23alphaXiv arXiv S2
  2. AI Exposure Scores: what they measure, what they miss, and what comes next
    arXiv.org2026-06-22alphaXiv arXiv S2
  3. The Culture Funnel: You Can't Align What isn't in the Data
    arXiv.org2026-06-11alphaXiv arXiv S2
  4. Soft-SVeRL: Self-Verified Reinforcement Learning with Soft Rewards
    arXiv.org2026-05-27alphaXiv arXiv S2
  5. Tiny Aya: Bridging Scale and Multilingual Depth
    arXiv.org2026-03-12alphaXiv arXiv S2
  6. CIRCLE: A Framework for Evaluating AI from a Real-World Lens
    arXiv.org2026-02-27alphaXiv arXiv S2
  7. Unlocking Reasoning Capability on Machine Translation in Large Language Models
    arXiv.org2026-02-16alphaXiv arXiv S2
  8. SimMerge: Learning to Select Merge Operators from Similarity Signals
    arXiv.org2026-01-14alphaXiv arXiv S2
  9. The Art of Asking: Multilingual Prompt Optimization for Synthetic Data
    arXiv.org2025-10-22alphaXiv arXiv S2
  10. Making, not Taking, the Best of N
    arXiv.org2025-10-01alphaXiv arXiv S2
  11. Verification Limits Code LLM Training
    arXiv.org2025-09-25alphaXiv arXiv S2
  12. From KMMLU-Redux to Pro: A Professional Korean Benchmark Suite for LLM Evaluation
    Conference on Empirical Methods in Natural Language Processing2025-07-11alphaXiv arXiv S2
  13. NeoBabel: A Multilingual Open Tower for Visual Generation
    arXiv.org2025-07-08alphaXiv arXiv S2
  14. One Tokenizer To Rule Them All: Emergent Language Plasticity via Multilingual Tokenizers
    Annual Meeting of the Association for Computational Linguistics2025-06-12alphaXiv arXiv S2
  15. The State of Multilingual LLM Safety Research: From Measuring the Language Gap to Mitigating It
    Conference on Empirical Methods in Natural Language Processing2025-05-30alphaXiv arXiv S2
  16. The Multilingual Divide and Its Impact on Global AI Safety
    arXiv.org2025-05-27alphaXiv arXiv S2
  17. Reality Check: A New Evaluation Ecosystem Is Necessary to Understand AI's Real World Effects
    arXiv.org2025-05-24alphaXiv arXiv S2
  18. Aya Vision: Advancing the Frontier of Multilingual Multimodality
    arXiv.org2025-05-13alphaXiv arXiv S2
  19. The Leaderboard Illusion
    Neural Information Processing Systems2025-04-29alphaXiv arXiv S2
  20. A Post-trainer's Guide to Multilingual Training Data: Uncovering Cross-lingual Transfer Dynamics
    arXiv.org2025-04-23alphaXiv arXiv S2
  21. Déjà Vu: Multilingual LLM Evaluation through the Lens of Machine Translation Evaluation
    arXiv.org2025-04-16alphaXiv arXiv S2
  22. Kaleidoscope: In-language Exams for Massively Multilingual Vision Evaluation
    arXiv.org2025-04-09alphaXiv arXiv S2
  23. Command A: An Enterprise-Ready Large Language Model
    2025-04-01alphaXiv arXiv S2
  24. From Tools to Teammates: Evaluating LLMs in Multi-Session Coding Interactions
    Annual Meeting of the Association for Computational Linguistics2025-02-19alphaXiv arXiv S2
  25. Multilingual Machine Translation with Open Large Language Models at Practical Scale: An Empirical Study
    North American Chapter of the Association for Computational Linguistics2025-02-04alphaXiv arXiv S2
  26. Towards Best Practices for Open Datasets for LLM Training
    arXiv.org2025-01-14alphaXiv arXiv S2
  27. Global MMLU: Understanding and Addressing Cultural and Linguistic Biases in Multilingual Evaluation
    Annual Meeting of the Association for Computational Linguistics2025S2
  28. Findings of the WMT25 Multilingual Instruction Shared Task: Persistent Hurdles in Reasoning, Generation, and Evaluation
    Conference on Machine Translation2025S2
  29. RLHF Algorithms Ranked: An Extensive Evaluation Across Diverse Tasks, Rewards, and Hyperparameters
    Conference on Empirical Methods in Natural Language Processing2025S2
  30. To Code or Not To Code? Exploring Impact of Code in Pre-training
    International Conference on Learning Representations2025S2
  31. Command-A-Translate: Raising the Bar of Machine Translation with Difficulty Filtering
    Conference on Machine Translation2025S2
  32. Generating Complex Question Decompositions in the Face of Distribution Shifts
    North American Chapter of the Association for Computational Linguistics2025S2
  33. Aya Expanse: Combining Research Breakthroughs for a New Multilingual Frontier
    arXiv.org2024-12-05alphaXiv arXiv S2
  34. Global MMLU: Understanding and Addressing Cultural and Linguistic Biases in Multilingual Evaluation
    arXiv.org2024-12-04alphaXiv arXiv S2
  35. INCLUDE: Evaluating Multilingual Language Understanding with Regional Knowledge
    International Conference on Learning Representations2024-11-29alphaXiv arXiv S2
  36. M-RewardBench: Evaluating Reward Models in Multilingual Settings
    Annual Meeting of the Association for Computational Linguistics2024-10-20alphaXiv arXiv S2
  37. Mix Data or Merge Models? Optimizing for Diverse Multi-Task Learning
    arXiv.org2024-10-14alphaXiv arXiv S2
  38. Diversify and Conquer: Diversity-Centric Data Selection with Iterative Refinement
    arXiv.org2024-09-17alphaXiv arXiv S2
  39. Mathematical modeling of free vibration of star-shaped auxetic rectangular plate
    Archive of applied mechanics (1991)2024-08-21S2
  40. To Code, or Not To Code? Exploring Impact of Code in Pre-training
    arXiv.org2024-08-20alphaXiv arXiv S2
  41. LLM See, LLM Do: Leveraging Active Inheritance to Target Non-Differentiable Objectives
    Conference on Empirical Methods in Natural Language Processing2024S2
  42. Cross-lingual Transfer Dynamics in BLOOMZ: Insights into Multilingual Generalization
    S2
  43. Breaking mBad! Supervised Fine-tuning for Cross-Lingual Detoxification
    S2