- Published on
ICLR 2026 — Natural Language Processing
Natural Language Processing
59 papers (0 oral)
LongWriter-Zero: Mastering Ultra-Long Text Generation via Reinforcement Learning
- Link: OpenReview
CounselBench: A Large-Scale Expert Evaluation and Adversarial Benchmarking of Large Language Models in Mental Health Question Answering
- Link: OpenReview
Scalable Multilingual Multimodal Machine Translation with Speech-Text Fusion
- Link: OpenReview
40. Taming Polysemanticity in LLMs: Theory-Grounded Feature Recovery via Sparse Autoencoders
- Topics: LLMs & Foundation Models, Interpretability & Mechanistic Interpretability, Theory & Deep Learning Theory
DAMR: Efficient and Adaptive Context-Aware Knowledge Graph Question Answering with LLM-Guided MCTS
- Link: OpenReview
Text2Arch: A Dataset for Generating Scientific Architecture Diagrams from Natural Language Descriptions
- Link: OpenReview
A.I.R.: Enabling Adaptive, Iterative, and Reasoning-based Frame Selection For Video Question Answering
- Link: OpenReview
Token Alignment Heads: Unveiling Attention's Role in LLM Multilingual Translation
- Link: OpenReview
Personalized Feature Translation for Expression Recognition: An Efficient Source-Free Domain Adaptation Method
- Link: OpenReview
CellAgent: LLM-Driven Multi-Agent Framework for Natural Language-Based Single-Cell Analysis
- Link: OpenReview
Synthesizing High-Quality Visual Question Answering from Medical Documents with Generator-Verifier LMMs
- Link: OpenReview
A Structured, Tagged, and Localized Visual Question Answering Dataset with Full Sentence Answers and Scene Graphs for Chest X-ray Images
- Link: OpenReview
MAD-Logic: Multi-Agent Debate Enhances Symbolic Translation and Reasoning
- Link: OpenReview
Reliable Fine-Grained Evaluation of Natural Language Math Proofs
- Link: OpenReview
Improving Attributed Long-form Question Answering with Intent Awareness
- Link: OpenReview
LatentQA: Teaching LLMs to Decode Activations Into Natural Language
- Link: OpenReview
From Utterance to Vividity: Training Expressive Subtitle Translation LLM via Adaptive Local Preference Optimization
- Link: OpenReview
DiscoX: Benchmarking Discourse-Level Translation in Expert Domains
- Link: OpenReview
FlexiVoice: Enabling Flexible Style Control in Zero-Shot TTS with Natural Language Instructions
- Link: OpenReview
ProofBridge: Auto-Formalization of Natural Language Proofs in Lean via Joint Embeddings
- Link: OpenReview
Meta-Adaptive Prompt Distillation for Few-Shot Visual Question Answering
- Link: OpenReview
Grounding or Guessing? Visual Signals for Detecting Hallucinations in Sign Language Translation
- Link: OpenReview
MoL: Adaptive Mixture-of-Length Reasoning for Efficient Question Answering with Context
- Link: OpenReview
TripleSumm: Adaptive Triple-Modality Fusion for Video Summarization
- Link: OpenReview
Query-Guided Spatial–Temporal–Frequency Interaction for Music Audio–Visual Question Answering
- Link: OpenReview
VERIFY: A Novel Multi-Domain Dataset Grounding LTL in Contextual Natural Language via Provable Intermediate Logic
- Link: OpenReview
Learning to Summarize by Learning to Quiz: Adversarial Agentic Collaboration for Long Document Summarization
- Link: OpenReview
UniSS: Unified Expressive Speech-to-Speech Translation with Your Voice
- Link: OpenReview
Addressing Pitfalls in the Evaluation of Uncertainty Estimation Methods for Natural Language Generation
- Link: OpenReview
Text summarization via global structure awareness
- Link: OpenReview
Retrospective Sparse Attention for Efficient Long-Context Generation
- Link: OpenReview
Topology of Reasoning: Retrieved Cell Complex-Augmented Generation for Textual Graph Question Answering
- Link: OpenReview
What's the plan? Metrics for implicit planning in LLMs and their application to rhyme generation and question answering
- Link: OpenReview
FrugalRAG: Less is More in RL Finetuning for Multi-hop Question Answering
- Link: OpenReview
Enhancing LLMs for Knowledge Base Question Answering by Chain-of-Decomposition
- Link: OpenReview
CogniLoad: A Synthetic Natural Language Reasoning Benchmark With Tunable Length, Intrinsic Difficulty, and Distractor Density
- Link: OpenReview
Adaptive Domain Shift in Diffusion Models for Cross-Modality Image Translation
- Link: OpenReview
MedAraBench: Large-scale Arabic Medical Question Answering Dataset and Benchmark
- Link: OpenReview
Human Uncertainty-Aware Data Selection and Automatic Labeling in Visual Question Answering
- Link: OpenReview
Logit‑KL Flow Matching: Non‑Autoregressive Text Generation via Sampling‑Hybrid Inference
- Link: OpenReview
Universal Multi-Domain Translation via Diffusion Routers
- Link: OpenReview
SC-Arena: A Natural Language Benchmark for Single-Cell Reasoning with Knowledge-Augmented Evaluation
- Link: OpenReview
QuRL: Rubrics As Judge For Open-Ended Question Answering
- Link: OpenReview
Text2Grad: Reinforcement Learning from Natural Language Feedback
- Link: OpenReview
Mathesis: Towards Formal Theorem Proving from Natural Languages
- Link: OpenReview
ASearch: Ambiguity-Aware Question Answering with Reinforcement Learning
- Link: OpenReview
LiveWeb-IE: A Benchmark For Online Web Information Extraction
- Link: OpenReview
Beyond Markovian Drifts: Action-Biased Geometric Walks with Memory for Personalized Summarization
- Link: OpenReview
Natural Language PDDL (NL-PDDL) for Open-world Goal-oriented Commonsense Regression Planning in Embodied AI
- Link: OpenReview
Knowledge Exchange with Confidence: Cost-Effective LLM Integration for Reliable and Efficient Visual Question Answering
- Link: OpenReview
FS-DFM: Fast and Accurate Long Text Generation with Few-Step Diffusion Language Models
- Link: OpenReview
Secret-Protected Evolution for Differentially Private Synthetic Text Generation
- Link: OpenReview
Gumbel Distillation for Parallel Text Generation
- Link: OpenReview
Tight Bounds for Schrodinger Potential Estimation in Unpaired Data Translation
- Link: OpenReview
RobotArena : Scalable Robot Benchmarking via Real-to-Sim Translation
- Link: OpenReview
Learning to Orchestrate Agents in Natural Language with the Conductor
- Link: OpenReview
A High Quality Dataset and Reliable Evaluation for Interleaved Image-Text Generation
- Link: OpenReview
jqBench: a benchmark for reading and editing JSON from natural language and/or examples
- Link: OpenReview