← People
Caglar Gulcehre

Caglar Gulcehre following

EPFL
@caglarmlPapers in the feed →

Papers · 35
  1. Context-Aware Toxicity Detection in Game Chat: Domain-Adaptive Pretraining with Match Metadata under Limited Labels
    Games2026-07-17S2
  2. Diffuse AI Control on Fuzzy Tasks
    arXiv.org2026-06-08alphaXiv arXiv S2
  3. BlockGen: Flexible Blockwise Sequence Modeling with Hybrid Samplers
    arXiv.org2026-06-01alphaXiv arXiv S2
  4. Augmenting Attention with Exponentially Decaying Memory Improves Query-Aware KV Sparsity
    arXiv.org2026-05-27alphaXiv arXiv S2
  5. The Future of Facts: Tracing the Factual Generation-Verification Gap
    arXiv.org2026-05-26alphaXiv arXiv S2
  6. Language Modeling with Hyperspherical Flows
    arXiv.org2026-05-11alphaXiv arXiv S2
  7. Sequence Modeling Architectures: Foundations [Special Issue on the Mathematics of Deep Learning]
    IEEE Signal Processing Magazine2026-05-01S2
  8. The Diffusion Duality, Chapter II: Ψ-Samplers and Efficient Curriculum
    arXiv.org2026-02-24alphaXiv arXiv S2
  9. RAT+: Train Dense, Infer Sparse - Recurrence Augmented Attention for Dilated Inference
    arXiv.org2026-02-20alphaXiv arXiv S2
  10. Apertus: Democratizing Open and Compliant LLMs for Global Language Environments
    Annual Meeting of the Association for Computational Linguistics2026S2
  11. Learning Vision-Language Alignment in Unified LLMs with 24 Text Tokens per Image
    International Workshop on Spoken Language Translation2026S2
  12. Loopholing Discrete Diffusion: Deterministic Bypass of the Sampling Wall
    arXiv.org2025-10-22alphaXiv arXiv S2
  13. Adaptive Attacks on Trusted Monitors Subvert AI Control Protocols
    arXiv.org2025-10-10alphaXiv arXiv S2
  14. Apertus: Democratizing Open and Compliant LLMs for Global Language Environments
    2025-09-17alphaXiv arXiv S2
  15. Quantile Reward Policy Optimization: Alignment with Pointwise Regression and Exact Partition Functions
    Neural Information Processing Systems2025-07-10alphaXiv arXiv S2
  16. RAT: Bridging RNN Efficiency and Attention Accuracy via Chunk-based Sequence Modeling
    Advances in Neural Information Processing Systems 382025-07-06alphaXiv arXiv S2
  17. The 2025 PNPL Competition: Speech Detection and Phoneme Classification in the LibriBrain Dataset
    arXiv.org2025-06-11alphaXiv arXiv S2
  18. Control Tax: The Price of Keeping AI in Check
    arXiv.org2025-06-05alphaXiv arXiv S2
  19. Partition Generative Modeling: Masked Modeling Without Masks
    arXiv.org2025-05-24alphaXiv arXiv S2
  20. Algorithm Discovery With LLMs: Evolutionary Search Meets Reinforcement Learning
    arXiv.org2025-04-07alphaXiv arXiv S2
  21. Context-Aware Toxicity Detection in Multiplayer Games: Integrating Domain-Adaptive Pretraining and Match Metadata
    arXiv.org2025-04-02alphaXiv arXiv S2
  22. From Markov to Laplace: How Mamba In-Context Learns Markov Chains
    arXiv.org2025-02-14alphaXiv arXiv S2
  23. Regret-Optimized Portfolio Enhancement through Deep Reinforcement Learning and Future Looking Rewards
    International Conference on AI in Finance2025-02-04alphaXiv arXiv S2
  24. Apertus: Democratizing Open and Compliant LLMs for Global Language Environments
    arXiv.org2025S2
  25. One-Step is Enough: Sparse Autoencoders for Text-to-Image Diffusion Models
    Neural Information Processing Systems2025S2
  26. RAT: Bridging RNN Efficiency and Attention Accuracy in Language Modeling
    arXiv.org2025S2
  27. One-Step is Enough: Sparse Autoencoders for Text-to-Image Diffusion Models
    2024-10-28alphaXiv arXiv S2
  28. Beyond Autoregression: Fast LLMs via Self-Distillation Through Time
    International Conference on Learning Representations2024-10-28alphaXiv arXiv S2
  29. SIKeD: Self-guided Iterative Knowledge Distillation for mathematical reasoning
    Annual Meeting of the Association for Computational Linguistics2024-10-24alphaXiv arXiv S2
  30. The Role of Deep Learning Regularizations on Actors in Offline RL
    arXiv.org2024-09-11alphaXiv arXiv S2
  31. Universality of Linear Recurrences Followed by Non-linear Projections: Finite-Width Guarantees and Benefits of Complex Eigenvalues
    International Conference on Machine Learning2024S2
  32. The Effect of Scheduling and Preemption on the Efficiency of LLM Inference Serving
    arXiv.org2024S2
  33. Fleet of Agents: Coordinated Problem Solving with Large Language Models using Genetic Particle Filtering
    arXiv.org2024S2
  34. Building on Efficient Foundations: Effective Training of LLMs with Structured Feedforward Layers
    Advances in Neural Information Processing Systems 372024S2
  35. Building on Efficient Foundations: Effective Training of LLMs with Structured Feedforward Layers
    Neural Information Processing Systems2024S2