← People
Mrinmaya Sachan

Mrinmaya Sachan interacted with

ETH Zurich
@@mrinmayasachanLinkedInGoogle ScholarWebsitePapers in the feed →

Papers · 197
  1. Dynamically Allocating Evaluation Effort for Model Ranking
    2026-08-04alphaXiv arXiv S2
  2. ThinkBooster: A Unified Framework for Seamless Test-Time Scaling of LLM Reasoning
    Annual Meeting of the Association for Computational Linguistics2026-06-05alphaXiv arXiv S2
  3. Tackling the Root of Misinformation by Teaching Laypeople about Logical Fallacies via Socratic Questioning and Critical Argumentation
    Annual Meeting of the Association for Computational Linguistics2026-05-31alphaXiv arXiv S2
  4. Diversity Matters: Revisiting Test-Time Compute in Vision-Language Models
    arXiv.org2026-05-29alphaXiv arXiv S2
  5. Benchmarking and Enhancing Text-to-Image Models for Generating Visual Representations in Early Arithmetic Education
    arXiv.org2026-05-29alphaXiv arXiv S2
  6. Unveiling the Visual Counting Bottleneck in Vision-Language Models
    arXiv.org2026-05-28alphaXiv arXiv S2
  7. Learning to Reason Efficiently with A* Post-Training
    arXiv.org2026-05-23alphaXiv arXiv S2
  8. Simulating Students or Sycophantic Problem Solving? On Misconception Faithfulness of LLM Simulators
    arXiv.org2026-05-12alphaXiv arXiv S2
  9. GeoDial: A Multimodal Conversational Tutoring Dataset for Geometry Problem-Solving with Visual Tutor Turns
    arXiv.org2026-05-08alphaXiv arXiv S2
  10. Efficient Test-Time Inference via Deterministic Exploration of Truncated Decoding Trees
    arXiv.org2026-04-22alphaXiv arXiv S2
  11. Misconception Acquisition Dynamics in Large Language Models
    International Conference on Artificial Intelligence in Education2026-04-01alphaXiv arXiv S2
  12. Can LLMs Model Incorrect Student Reasoning? A Case Study on Distractor Generation
    arXiv.org2026-03-16alphaXiv arXiv S2
  13. Post-Training Language Models for Crosslingual Consistency
    2026-03-04alphaXiv arXiv S2
  14. When AI Benchmarks Plateau: A Systematic Study of Benchmark Saturation
    arXiv.org2026-02-18alphaXiv arXiv S2
  15. Fluid Reasoning Representations
    2026-02-04alphaXiv arXiv S2
  16. Uncovering Hidden Correctness in LLM Causal Reasoning via Symbolic Verification
    Conference of the European Chapter of the Association for Computational Linguistics2026-01-29alphaXiv arXiv S2
  17. Bridging Instead of Replacing Online Coding Communities with AI through Community-Enriched Chatbot Designs
    Proc. ACM Hum. Comput. Interact.2026-01-26alphaXiv arXiv S2
  18. PATS: Personality-Aware Teaching Strategies with Large Language Model Tutors
    Conference of the European Chapter of the Association for Computational Linguistics2026-01-13alphaXiv arXiv S2
  19. Early-Exit and Instant Confidence Translation Quality Estimation
    Conference of the European Chapter of the Association for Computational Linguistics2026S2
  20. Optimizing Language Models for Crosslingual Knowledge Consistency
    arXiv.org2026S2
  21. PaperMentor: A Human-Centered Multi-Agent Writing Tutor for AI Research Papers in Overleaf
    Annual Meeting of the Association for Computational Linguistics2026S2
  22. Efficient Test-Time Scaling of Multi-Step Reasoning by Probing Internal States of Large Language Models
    Annual Meeting of the Association for Computational Linguistics2026S2
  23. Don't Throw Away Your Beams: Improving Consistency-based Uncertainties in LLMs via Beam Search
    arXiv.org2025-12-10alphaXiv arXiv S2
  24. Harmonizing Assistance: Moderating Visual and Textual Aids in AI-Enhanced Textbook Reading with IRead
    International Journal of Artificial Intelligence in Education2025-11-11S2
  25. ReProbe: Efficient Test-Time Scaling of Multi-Step Reasoning by Probing Internal States of Large Language Models
    2025-11-09alphaXiv arXiv S2
  26. On the Emergence of Induction Heads for In-Context Learning
    arXiv.org2025-11-02alphaXiv arXiv S2
  27. Are Language Models Efficient Reasoners? A Perspective from Logic Programming
    Neural Information Processing Systems2025-10-29alphaXiv arXiv S2
  28. Sample Smart, Not Hard: Correctness-First Decoding for Better Reasoning in LLMs
    arXiv.org2025-10-07alphaXiv arXiv S2
  29. Compose and Fuse: Revisiting the Foundational Bottlenecks in Multimodal Reasoning
    arXiv.org2025-09-28alphaXiv arXiv S2
  30. Chimera: Diagnosing Shortcut Learning in Visual-Language Understanding
    arXiv.org2025-09-26alphaXiv arXiv S2
  31. Can Vision-Language Models Solve Visual Math Equations?
    Conference on Empirical Methods in Natural Language Processing2025-09-10alphaXiv arXiv S2
  32. Test of Time: Rethinking Temporal Signal of Benchmark Contamination
    Annual Meeting of the Association for Computational Linguistics2025-08-26alphaXiv arXiv S2
  33. COMET-poly: Machine Translation Metric Grounded in Other Candidates
    Conference on Machine Translation2025-08-25alphaXiv arXiv S2
  34. Lean Meets Theoretical Computer Science: Scalable Synthesis of Theorem Proving Challenges in Formal-Informal Pairs
    arXiv.org2025-08-21alphaXiv arXiv S2
  35. Probing for Arithmetic Errors in Language Models
    Conference on Empirical Methods in Natural Language Processing2025-07-16alphaXiv arXiv S2
  36. Personalized Exercise Recommendation with Semantically-Grounded Knowledge Tracing
    Advances in Neural Information Processing Systems 382025-07-15alphaXiv arXiv S2
  37. Co-DETECT: Collaborative Discovery of Edge Cases in Text Classification
    Conference on Empirical Methods in Natural Language Processing2025-07-07alphaXiv arXiv S2
  38. The Medium Is Not the Message: Deconfounding Document Embeddings via Linear Concept Erasure
    Conference on Empirical Methods in Natural Language Processing2025-07-01alphaXiv arXiv S2
  39. Corrupted by Reasoning: Reasoning Language Models Become Free-Riders in Public Goods Games
    arXiv.org2025-06-29alphaXiv arXiv S2
  40. Can Reasoning Help Large Language Models Capture Human Annotator Disagreement?
    Conference of the European Chapter of the Association for Computational Linguistics2025-06-24alphaXiv arXiv S2
  41. Dense SAE Latents Are Features, Not Bugs
    Neural Information Processing Systems2025-06-18alphaXiv arXiv S2
  42. Improving Large Language Model Safety with Contrastive Representation Learning
    Conference on Empirical Methods in Natural Language Processing2025-06-13alphaXiv arXiv S2
  43. Educators' Perceptions of Large Language Models as Tutors: Comparing Human and AI Tutors in a Blind Text-only Setting
    Workshop on Innovative Use of NLP for Building Educational Applications2025-06-10alphaXiv arXiv S2
  44. Generating Pedagogically Meaningful Visuals for Math Word Problems: A New Benchmark and Analysis of Text-to-Image Models
    Annual Meeting of the Association for Computational Linguistics2025-06-04alphaXiv arXiv S2
  45. Faithfulness-Aware Uncertainty Quantification for Fact-Checking the Output of Retrieval Augmented Generation
    Annual Meeting of the Association for Computational Linguistics2025-05-27alphaXiv arXiv S2
  46. Efficient Hallucination Detection for LLMs Using Uncertainty-Aware Attention Heads
    2025-05-26alphaXiv arXiv S2
  47. SeePhys: Does Seeing Help Thinking? - Benchmarking Vision-Based Physics Reasoning
    Neural Information Processing Systems2025-05-25alphaXiv arXiv S2
  48. From Problem-Solving to Teaching Problem-Solving: Aligning LLMs with Pedagogy using Reinforcement Learning
    Conference on Empirical Methods in Natural Language Processing2025-05-21alphaXiv arXiv S2
  49. LEXam: Benchmarking Legal Reasoning on 340 Law Exams
    arXiv.org2025-05-19alphaXiv arXiv S2
  50. A Head to Predict and a Head to Question: Pre-trained Uncertainty Quantification Heads for Hallucination Detection in LLM Outputs
    Conference on Empirical Methods in Natural Language Processing2025-05-13alphaXiv arXiv S2
  51. Multilingual Performance Biases of Large Language Models in Education
    arXiv.org2025-04-24alphaXiv arXiv S2
  52. MathTutorBench: A Benchmark for Measuring Open-ended Pedagogical Capabilities of LLM Tutors
    Conference on Empirical Methods in Natural Language Processing2025-02-26alphaXiv arXiv S2
  53. Balancing Truthfulness and Informativeness with Uncertainty-Aware Instruction Fine-Tuning
    2025-02-17alphaXiv arXiv S2
  54. Grammar Control in Dialogue Response Generation for Language Learning Chatbots
    North American Chapter of the Association for Computational Linguistics2025-02-11alphaXiv arXiv S2
  55. Investigating the Zone of Proximal Development of Language Models for In-Context Learning
    North American Chapter of the Association for Computational Linguistics2025-02-10alphaXiv arXiv S2
  56. How to Select Datapoints for Efficient Human Evaluation of NLG Models?
    Transactions of the Association for Computational Linguistics2025-01-30alphaXiv arXiv S2
  57. The Medium Is Not the Message: Deconfounding Text Embeddings via Linear Concept Erasure
    arXiv.org2025S2
  58. Beyond Memorization: Reasoning-Driven Synthesis as a Mitigation Strategy Against Benchmark Contamination
    arXiv.org2025S2
  59. Large Language Models for Education: Understanding the Needs of Stakeholders, Current Capabilities and the Path Forward
    Workshop on Innovative Use of NLP for Building Educational Applications2025S2
  60. Can LLMs Effectively Simulate Human Learners? Teachers' Insights from Tutoring LLM Students
    Workshop on Innovative Use of NLP for Building Educational Applications2025S2
  61. Do Vision-Language Models Really Understand Visual Language?
    International Conference on Machine Learning2025S2
  62. Can Large Language Models Capture Human Annotator Disagreements?
    arXiv.org2025S2
  63. Are Large Language Models for Education Reliable Across Languages?
    Workshop on Innovative Use of NLP for Building Educational Applications2025S2
  64. Pointwise Mutual Information as a Performance Gauge for Retrieval-Augmented Generation
    North American Chapter of the Association for Computational Linguistics2024-11-12alphaXiv arXiv S2
  65. SIKeD: Self-guided Iterative Knowledge Distillation for mathematical reasoning
    Annual Meeting of the Association for Computational Linguistics2024-10-24alphaXiv arXiv S2
  66. SMART: Self-learning Meta-strategy Agent for Reasoning Tasks
    arXiv.org2024-10-21alphaXiv arXiv S2
  67. Efficiently Computing Susceptibility to Context in Language Models
    Conference on Empirical Methods in Natural Language Processing2024-10-18alphaXiv arXiv S2
  68. MathGAP: Out-of-Distribution Evaluation on Problems with Arbitrarily Complex Proofs
    International Conference on Learning Representations2024-10-17alphaXiv arXiv S2
  69. LLM-based Cognitive Models of Students with Misconceptions
    arXiv.org2024-10-16alphaXiv arXiv S2
  70. Towards the Pedagogical Steering of Large Language Models for Tutoring: A Case Study with Modeling Productive Failure
    Annual Meeting of the Association for Computational Linguistics2024-10-03alphaXiv arXiv S2
  71. Automated Knowledge Concept Annotation and Question Representation Learning for Knowledge Tracing
    arXiv.org2024-10-02alphaXiv arXiv S2
  72. Do Vision-Language Models Really Understand Visual Language?
    arXiv.org2024-09-30alphaXiv arXiv S2
  73. GPT-4 as a Homework Tutor can Improve Student Engagement and Learning Outcomes
    Annual Meeting of the Association for Computational Linguistics2024-09-24alphaXiv arXiv S2
  74. RETRO-LI: Small-Scale Retrieval Augmented Generation Supporting Noisy Similarity Searches and Domain Shift Generalization
    European Conference on Artificial Intelligence2024-09-12alphaXiv arXiv S2
  75. AI-assisted Automated Short Answer Grading of Handwritten University Level Mathematics Exams
    2024-08-21alphaXiv arXiv S2
  76. Towards Aligning Language Models with Textual Feedback
    Conference on Empirical Methods in Natural Language Processing2024-07-24alphaXiv arXiv S2
  77. How to Engage Your Readers? Generating Guiding Questions to Promote Active Reading
    Annual Meeting of the Association for Computational Linguistics2024-07-19alphaXiv arXiv S2
  78. Stepwise Verification and Remediation of Student Reasoning Errors with Large Language Model Tutors
    Conference on Empirical Methods in Natural Language Processing2024-07-12alphaXiv arXiv S2
  79. Language Model Alignment in Multilingual Trolley Problems
    International Conference on Learning Representations2024-07-02alphaXiv arXiv S2
  80. Confidence Regulation Neurons in Language Models
    Neural Information Processing Systems2024-06-24alphaXiv arXiv S2
  81. DIRAS: Efficient LLM Annotation of Document Relevance for Retrieval Augmented Generation
    North American Chapter of the Association for Computational Linguistics2024-06-20alphaXiv arXiv S2
  82. AI-Assisted Human Evaluation of Machine Translation
    North American Chapter of the Association for Computational Linguistics2024-06-18alphaXiv arXiv S2
  83. Error Span Annotation: A Balanced Approach for Human Evaluation of Machine Translation
    Conference on Machine Translation2024-06-17alphaXiv arXiv S2
  84. What Do Language Models Learn in Context? The Structured Task Hypothesis
    Annual Meeting of the Association for Computational Linguistics2024-06-06alphaXiv arXiv S2
  85. On Affine Homotopy between Language Encoders
    Neural Information Processing Systems2024-06-04alphaXiv arXiv S2
  86. Quriosity: Analyzing Human Questioning Behavior and Causal Inquiry through Curiosity-Driven Queries
    IJCNLP-AACL2024-05-30alphaXiv arXiv S2
  87. Implicit Personalization in Language Models: A Systematic Study
    Conference on Empirical Methods in Natural Language Processing2024-05-23alphaXiv arXiv S2
  88. A Transformer with Stack Attention
    NAACL-HLT2024-05-07alphaXiv arXiv S2
  89. Cooperate or Collapse: Emergence of Sustainable Cooperation in a Society of LLM Agents
    Neural Information Processing Systems2024-04-25alphaXiv arXiv S2
  90. Autoformalizing Natural Language to First-Order Logic: A Case Study in Logical Fallacy Detection
    IJCNLP-AACL2024-04-18alphaXiv arXiv S2
  91. Do LLMs Think Fast and Slow? A Causal Study on Sentiment Analysis
    Conference on Empirical Methods in Natural Language Processing2024-04-17alphaXiv arXiv S2
  92. Slicing, Chatting, and Refining: A Concept-Based Approach for Machine Learning Model Validation with ConceptSlicer
    International Conference on Intelligent User Interfaces2024-03-18S2
  93. Book2Dial: Generating Teacher-Student Interactions from Textbooks for Cost-Effective Development of Educational Chatbots
    Annual Meeting of the Association for Computational Linguistics2024-03-05alphaXiv arXiv S2
  94. Calibrating Large Language Models with Sample Consistency
    AAAI Conference on Artificial Intelligence2024-02-21alphaXiv arXiv S2
  95. Competition of Mechanisms: Tracing How Language Models Handle Facts and Counterfactuals
    Annual Meeting of the Association for Computational Linguistics2024-02-18alphaXiv arXiv S2
  96. AFaCTA: Assisting the Annotation of Factual Claim Detection with Reliable LLM Annotators
    Annual Meeting of the Association for Computational Linguistics2024-02-16alphaXiv arXiv S2
  97. AutoTutor meets Large Language Models: A Language Model Tutor with Rich Pedagogy and Guardrails
    ACM Conference on Learning @ Scale2024-02-14alphaXiv arXiv S2
  98. Do Language Models Exhibit the Same Cognitive Biases in Problem Solving as Human Learners?
    International Conference on Machine Learning2024-01-31alphaXiv arXiv S2
  99. CLadder: Assessing Causal Reasoning in Language Models
    Advances in Neural Information Processing Systems 362023-12-07alphaXiv arXiv S2
  100. RELIC: Investigating Large Language Model Responses using Self-Consistency
    International Conference on Human Factors in Computing Systems2023-11-28alphaXiv arXiv S2
  101. Exploring the Jungle of Bias: Political Bias Attribution in Language Models via Dependency Analysis
    NLP4PI2023-11-15alphaXiv arXiv S2
  102. The ART of LLM Refinement: Ask, Refine, and Trust
    North American Chapter of the Association for Computational Linguistics2023-11-14alphaXiv arXiv S2
  103. CausalCite: A Causal Formulation of Paper Citations
    Annual Meeting of the Association for Computational Linguistics2023-11-05alphaXiv arXiv S2
  104. Towards a Mechanistic Interpretation of Multi-Step Reasoning Capabilities of Language Models
    Conference on Empirical Methods in Natural Language Processing2023-10-23alphaXiv arXiv S2
  105. Let's Synthesize Step by Step: Iterative Dataset Synthesis with Large Language Models by Extrapolating Errors from Small Models
    Conference on Empirical Methods in Natural Language Processing2023-10-20alphaXiv arXiv S2
  106. A Diachronic Perspective on User Trust in AI under Uncertainty
    Conference on Empirical Methods in Natural Language Processing2023-10-20alphaXiv arXiv S2
  107. Agents: An Open-source Framework for Autonomous Language Agents
    arXiv.org2023-09-14alphaXiv arXiv S2
  108. Tokenization and the Noiseless Channel
    Annual Meeting of the Association for Computational Linguistics2023-06-29alphaXiv arXiv S2
  109. A Formal Perspective on Byte-Pair Encoding
    Annual Meeting of the Association for Computational Linguistics2023-06-29alphaXiv arXiv S2
  110. Can Large Language Models Infer Causation from Correlation?
    International Conference on Learning Representations2023-06-09alphaXiv arXiv S2
  111. World Models for Math Story Problems
    Annual Meeting of the Association for Computational Linguistics2023-06-07alphaXiv arXiv S2
  112. Infusing Lattice Symmetry Priors in Attention Mechanisms for Sample-Efficient Abstract Geometric Reasoning
    International Conference on Machine Learning2023-06-05alphaXiv arXiv S2
  113. Adaptive and Personalized Exercise Generation for Online Language Learning
    Annual Meeting of the Association for Computational Linguistics2023-06-04alphaXiv arXiv S2
  114. Membership Inference Attacks against Language Models via Neighbourhood Comparison
    Annual Meeting of the Association for Computational Linguistics2023-05-29alphaXiv arXiv S2
  115. A Mechanistic Interpretation of Arithmetic Reasoning in Language Models using Causal Mediation Analysis
    Conference on Empirical Methods in Natural Language Processing2023-05-24alphaXiv arXiv S2
  116. Linear-Time Modeling of Linguistic Structure: An Order-Theoretic Perspective
    Conference on Empirical Methods in Natural Language Processing2023-05-24alphaXiv arXiv S2
  117. All Roads Lead to Rome? Exploring the Invariance of Transformers' Representations
    arXiv.org2023-05-23alphaXiv arXiv S2
  118. When Does Aggregating Multiple Skills with Multi-Task Learning Work? A Case Study in Financial NLP
    Annual Meeting of the Association for Computational Linguistics2023-05-23alphaXiv arXiv S2
  119. MathDial: A Dialogue Tutoring Dataset with Rich Pedagogical Properties Grounded in Math Reasoning Problems
    Conference on Empirical Methods in Natural Language Processing2023-05-23alphaXiv arXiv S2
  120. RecurrentGPT: Interactive Generation of (Arbitrarily) Long Text
    arXiv.org2023-05-22alphaXiv arXiv S2
  121. Re-visiting Automated Topic Model Evaluation with Large Language Models
    Conference on Empirical Methods in Natural Language Processing2023-05-20alphaXiv arXiv S2
  122. Discourse Centric Evaluation of Machine Translation with a Densely Annotated Parallel Corpus
    arXiv.org2023-05-18alphaXiv arXiv S2
  123. Efficient Prompting via Dynamic In-Context Learning
    arXiv.org2023-05-18alphaXiv arXiv S2
  124. Variational Classification
    Trans. Mach. Learn. Res.2023-05-17alphaXiv arXiv S2
  125. Beyond Good Intentions: Reporting the Research Landscape of NLP for Social Good
    Conference on Empirical Methods in Natural Language Processing2023-05-09alphaXiv arXiv S2
  126. Psychologically-Inspired Causal Prompts
    arXiv.org2023-05-02alphaXiv arXiv S2
  127. Controlled Text Generation with Natural Language Instructions
    International Conference on Machine Learning2023-04-27alphaXiv arXiv S2
  128. Enhancing Textbooks with Visuals from the Web for Improved Learning
    Conference on Empirical Methods in Natural Language Processing2023-04-18alphaXiv arXiv S2
  129. PWESuite: Phonetic Word Embeddings and Tasks They Facilitate
    International Conference on Language Resources and Evaluation2023-04-05alphaXiv arXiv S2
  130. Elastic Weight Removal for Faithful and Abstractive Dialogue Generation
    North American Chapter of the Association for Computational Linguistics2023-03-30alphaXiv arXiv S2
  131. Strategize Before Teaching: A Conversational Tutoring System with Pedagogy Self-Distillation
    Findings2023-02-27alphaXiv arXiv S2
  132. Opportunities and Challenges in Neural Dialog Tutoring
    Conference of the European Chapter of the Association for Computational Linguistics2023-01-24alphaXiv arXiv S2
  133. Poor Man’s Quality Estimation: Predicting Reference-Based MT Metrics Without the Reference
    Conference of the European Chapter of the Association for Computational Linguistics2023-01-21alphaXiv arXiv S2
  134. Distilling Reasoning Capabilities into Smaller Language Models
    Annual Meeting of the Association for Computational Linguistics2022-12-01alphaXiv arXiv S2
  135. Automatic Generation of Socratic Subquestions for Teaching Math Word Problems
    Conference on Empirical Methods in Natural Language Processing2022-11-23alphaXiv arXiv S2
  136. Beyond prompting: Making Pre-trained Language Models Better Zero-shot Learners by Clustering Representations
    Conference on Empirical Methods in Natural Language Processing2022-10-29alphaXiv arXiv S2
  137. Investigating the Role of Centering Theory in the Context of Neural Coreference Resolution Systems
    arXiv.org2022-10-26alphaXiv arXiv S2
  138. A Bilingual Parallel Corpus with Discourse Annotations
    arXiv.org2022-10-26alphaXiv arXiv S2
  139. Autoregressive Structured Prediction with Language Models
    Conference on Empirical Methods in Natural Language Processing2022-10-26alphaXiv arXiv S2
  140. Differentially Private Language Models for Secure Data Sharing
    Conference on Empirical Methods in Natural Language Processing2022-10-25alphaXiv arXiv S2
  141. Adapters for Enhanced Modeling of Multilingual Knowledge and Text
    Conference on Empirical Methods in Natural Language Processing2022-10-24alphaXiv arXiv S2
  142. A Causal Framework to Quantify the Robustness of Mathematical Reasoning with Language Models
    Annual Meeting of the Association for Computational Linguistics2022-10-21alphaXiv arXiv S2
  143. Longtonotes: OntoNotes with Longer Coreference Chains
    Findings2022-10-07alphaXiv arXiv S2
  144. When to Make Exceptions: Exploring Language Models as Accounts of Human Moral Judgment
    Neural Information Processing Systems2022-10-04alphaXiv arXiv S2
  145. Probing via Prompting
    North American Chapter of the Association for Computational Linguistics2022-07-04alphaXiv arXiv S2
  146. A Structured Span Selector
    North American Chapter of the Association for Computational Linguistics2022-05-08alphaXiv arXiv S2
  147. Original or Translated? A Causal Analysis of the Impact of Translationese on Machine Translation Performance
    North American Chapter of the Association for Computational Linguistics2022-05-04alphaXiv arXiv S2
  148. Calibration of Machine Reading Systems at Scale
    Findings2022-03-20alphaXiv arXiv S2
  149. Slangvolution: A Causal Analysis of Semantic Change and Frequency Dynamics in Slang
    Annual Meeting of the Association for Computational Linguistics2022-03-09alphaXiv arXiv S2
  150. Logical Fallacy Detection
    Conference on Empirical Methods in Natural Language Processing2022-02-28alphaXiv arXiv S2
  151. What Has Been Enhanced in my Knowledge-Enhanced Language Model?
    Conference on Empirical Methods in Natural Language Processing2022-02-02alphaXiv arXiv S2
  152. Case-based reasoning for better generalization in textual reinforcement learning
    International Conference on Learning Representations2021-10-16alphaXiv arXiv S2
  153. Learning the Transformer Kernel
    Trans. Mach. Learn. Res.2021-10-15alphaXiv arXiv S2
  154. Causal Direction of Data Collection Matters: Implications of Causal and Anticausal Learning for NLP
    Conference on Empirical Methods in Natural Language Processing2021-10-07alphaXiv arXiv S2
  155. "Let Your Characters Tell Their Story": A Dataset for Character-Centric Narrative Understanding
    Conference on Empirical Methods in Natural Language Processing2021-09-12alphaXiv arXiv S2
  156. Differentiable Subset Pruning of Transformer Heads
    Transactions of the Association for Computational Linguistics2021-08-10alphaXiv arXiv S2
  157. Self-Supervised Contrastive Learning with Adversarial Perturbations for Defending Word Substitution-based Attacks
    NAACL-HLT2021-07-15alphaXiv arXiv S2
  158. How Good Is NLP? A Sober Look at NLP Tasks through the Lens of Social Impact
    Findings2021-06-04alphaXiv arXiv S2
  159. Bird’s Eye: Probing for Linguistic Graph Structures with a Simple Information-Theoretic Approach
    Annual Meeting of the Association for Computational Linguistics2021-05-06alphaXiv arXiv S2
  160. BlonDe: An Automatic Evaluation Metric for Document-level Machine Translation
    North American Chapter of the Association for Computational Linguistics2021-03-22alphaXiv arXiv S2
  161. Clustering Contextualized Representations of Text for Unsupervised Syntax Induction
    arXiv.org2020-10-24S2
  162. Deep Clustering of Text Representations for Supervision-Free Probing of Syntax
    AAAI Conference on Artificial Intelligence2020-10-24alphaXiv arXiv S2
  163. Stronger Transformers for Neural Multi-Hop Question Generation
    arXiv.org2020-10-22alphaXiv arXiv S2
  164. Text-based RL Agents with Commonsense Knowledge: New Challenges, Environments and Baselines
    AAAI Conference on Artificial Intelligence2020-10-08alphaXiv arXiv S2
  165. Knowledge Graph Embedding Compression
    Annual Meeting of the Association for Computational Linguistics2020-07-01S2
  166. Enhancing Text-based Reinforcement Learning Agents with Commonsense Knowledge
    arXiv.org2020-05-02alphaXiv arXiv S2
  167. Towards Literate Artificial Intelligence
    2020-02-26S2
  168. Discourse in Multimedia: A Case Study in Extracting Geometry Knowledge from Textbooks
    Computational Linguistics2020-01-01S2
  169. Discourse in Multimedia: A Case Study in Information Extraction
    arXiv.org2018-11-13alphaXiv arXiv S2
  170. Contextual Parameter Generation for Universal Neural Machine Translation
    Conference on Empirical Methods in Natural Language Processing2018-08-01alphaXiv arXiv S2
  171. Parsing to Programs: A Framework for Situated QA
    Knowledge Discovery and Data Mining2018-07-19S2
  172. Self-Training for Jointly Learning to Ask and Answer Questions
    North American Chapter of the Association for Computational Linguistics2018-06-01S2
  173. Effective Use of Bidirectional Language Modeling for Transfer Learning in Biomedical Named Entity Recognition
    Machine Learning in Health Care2017-11-21alphaXiv arXiv S2
  174. From Textbooks to Knowledge: A Case Study in Harvesting Axiomatic Knowledge from Textbooks to Solve Geometry Problems
    Conference on Empirical Methods in Natural Language Processing2017-09-01S2
  175. Learning to Solve Geometry Problems from Natural Language Demonstrations in Textbooks
    International Workshop on Semantic Evaluation2017-08-01S2
  176. Machine Comprehension using Rich Semantic Representations
    Annual Meeting of the Association for Computational Linguistics2016-08-01S2
  177. Easy Questions First? A Case Study on Curriculum Learning for Question Answering
    Annual Meeting of the Association for Computational Linguistics2016-08-01S2
  178. Grounding Topic Models with Knowledge Bases
    International Joint Conference on Artificial Intelligence2016-07-09S2
  179. Learning Concept Taxonomies from Multi-modal Data
    Annual Meeting of the Association for Computational Linguistics2016-06-29alphaXiv arXiv S2
  180. Science Question Answering using Instructional Materials
    Annual Meeting of the Association for Computational Linguistics2016-02-01alphaXiv arXiv S2
  181. An Active Learning Approach to Coreference Resolution
    International Joint Conference on Artificial Intelligence2015-07-25S2
  182. Learning Answer-Entailing Structures for Machine Comprehension
    Annual Meeting of the Association for Computational Linguistics2015-07-01S2
  183. Spatial compactness meets topical consistency: jointly modeling links and content for community detection
    Web Search and Data Mining2014-02-24S2
  184. A Structured Distributional Semantic Model : Integrating Structure with Semantics
    CVSM@ACL2013-08-05S2
  185. A Structured Distributional Semantic Model for Event Co-reference
    Annual Meeting of the Association for Computational Linguistics2013-08-01S2
  186. Identifying Metaphorical Word Use with Tree Kernels
    2013-06-01S2
  187. Solving electrical networks to incorporate supervision in random walks
    The Web Conference2013-05-13S2
  188. Collective matrix factorization for co-clustering
    The Web Conference2013-05-13S2
  189. Using content and interactions for discovering communities in social networks
    The Web Conference2012-04-16S2
  190. Using Text Reviews for Product Entity Completion
    International Joint Conference on Natural Language Processing2011-11-01S2
  191. Probabilistic model for discovering topic based communities in social networks
    International Conference on Information and Knowledge Management2011-10-24S2
  192. Using Abstract Information and Community Alignment Information for Link Prediction
    2010 Second International Conference on Machine Learning and Computing2010-02-09S2
  193. Towards Understanding Semantic Degeneration in Text-Based Reinforcement Learning
    S2
  194. Causal AI Scientist: Facilitating Causal Data Science with Large Language Models
    S2
  195. AI Poses Risks to Democratic and Social Systems
    S2
  196. Should I Agree with You? Simulating Persuasion and Decision Dynamics in Multi-Agent Moral Dilemmas
    S2
  197. Cooperate or Collapse: Emergence of Sustainability in a Society of LLM Agents
    S2