← People
Zhaofeng Wu
following
MIT
@zhaofeng_wu
Papers in the feed →
Papers · 9
Variable-Width Transformers
arXiv.org
2026-06-16
alphaXiv
arXiv
S2
Implicit Representations of Grammaticality in Language Models
Annual Meeting of the Association for Computational Linguistics
2026-05-06
alphaXiv
arXiv
S2
Parallel-SFT: Improving Zero-Shot Cross-Programming-Language Transfer for Code RL
Annual Meeting of the Association for Computational Linguistics
2026-04-22
alphaXiv
arXiv
S2
Translation or Recitation? Calibrating Evaluation Scores for Machine Translation of Extremely Low-Resource Languages
Annual Meeting of the Association for Computational Linguistics
2026-03-26
alphaXiv
arXiv
S2
reWordBench: Benchmarking and Improving the Robustness of Reward Models with Transformed Inputs
Conference on Empirical Methods in Natural Language Processing
2025-03-14
alphaXiv
arXiv
S2
SelfCite: Self-Supervised Alignment for Context Attribution in Large Language Models
International Conference on Machine Learning
2025-02-13
alphaXiv
arXiv
S2
The Semantic Hub Hypothesis: Language Models Share Semantic Representations Across Languages and Modalities
International Conference on Learning Representations
2024-11-07
alphaXiv
arXiv
S2
Sparkle: Mastering Basic Spatial Capabilities in Vision Language Models Elicits Generalization to Spatial Reasoning
Conference on Empirical Methods in Natural Language Processing
2024-10-21
alphaXiv
arXiv
S2
Ascending the Infinite Ladder: Benchmarking Spatial Deformation Reasoning in Vision-Language Models
S2