← People
Alisa Liu
following
University of Washington
@alisawuffles
Papers in the feed →
Papers · 19
Nemotron 3 Ultra: Open, Efficient Mixture-of-Experts Hybrid Mamba-Transformer Model for Agentic Reasoning
2026-06-12
alphaXiv
arXiv
S2
Compute Optimal Tokenization
arXiv.org
2026-05-02
alphaXiv
arXiv
S2
Are you going to finish that? A Practical Study of the Partial Token Problem
2026-01-30
alphaXiv
arXiv
S2
When One LLM Drools, Multi-LLM Collaboration Rules
Annual Meeting of the Association for Computational Linguistics
2026
S2
Are you going to finish that? A Practical Study of the Tokenization Boundary Problem
arXiv.org
2026
S2
Olmo 3
2025-12-15
alphaXiv
arXiv
S2
Broken Tokens? Your Language Model can Secretly Handle Non-Canonical Tokenizations
Advances in Neural Information Processing Systems 38
2025-06-23
alphaXiv
arXiv
S2
Sampling from Your Language Model One Byte at a Time
arXiv.org
2025-06-17
alphaXiv
arXiv
S2
LLAMAPIE: Proactive In-Ear Conversation Assistants
Annual Meeting of the Association for Computational Linguistics
2025-05-07
alphaXiv
arXiv
S2
SuperBPE: Space Travel for Language Models
arXiv.org
2025-03-17
alphaXiv
arXiv
S2
When One LLM Drools, Multi-LLM Collaboration Rules
arXiv.org
2025-02-06
alphaXiv
arXiv
S2
Olmo 3
arXiv.org
2025
S2
TÜLU 3: Pushing Frontiers in Open Language Model Post-Training
arXiv.org
2024-11-22
alphaXiv
arXiv
S2
Does Liking Yellow Imply Driving a School Bus? Semantic Leakage in Language Models
North American Chapter of the Association for Computational Linguistics
2024-08-12
alphaXiv
arXiv
S2
Data Mixture Inference Attack: BPE Tokenizers Reveal Training Data Compositions
Neural Information Processing Systems
2024
S2
Data Mixture Inference Attack: BPE Tokenizers Reveal Training Data Compositions
Advances in Neural Information Processing Systems 37
2024
S2
Boyd-Graber . A Good Plan is Hard to Find: Aligning Models with Preferences is Misaligned with What Helps Users . Empirical Methods in Natural Language Processing
S2
Exploring Personalization Shifts in Representation Space of LLMs
S2
Reliability Scales Inversely: Bigger Models Compound Mistakes Faster via a Hidden Auto-Regressive Risk Regime
S2