- Published on
ACL 2026 — Generation & Summarization
Generation & Summarization
234 papers Links not yet available — ACL proceedings pending on ACL Anthology.
- No Reader Left Behind: Multi-Agent Summaries Everyone Can Understand
- EASE: Entity-Aware Sub-table Generation for Real-world Multi-table QA
- Parallel Universes, Parallel Languages: A Comprehensive Study on LLM-based Multilingual Counterfactual Example Generation
- PosterForest: Hierarchical Multi-Agent Collaboration for Scientific Poster Generation
- FastV-RAG: Towards Fast and Fine-Grained Video QA with Retrieval-Augmented Generation
- ReasonEmbed: Enhanced Text Embeddings for Reasoning-Intensive Document Retrieval
- ControlAudio: Tackling Text-Guided, Timing-Indicated and Intelligible Audio Generation via Progressive Diffusion Modeling
- Contextual Relevance and Adaptive Sampling for LLM-Based Document Reranking
- MORPHOGEN: A Multilingual Benchmark for Evaluating Gender-Aware Morphological Generation
- CodeHacker: Automated Test Case Generation for Detecting Vulnerabilities in Competitive Programming Solutions
- SDE-SQL: Enhancing Text-to-SQL Generation in Large Language Models via Self-Driven Exploration with SQL Probes
- BracketRank: Large Language Model Document Ranking via Reasoning-based Competitive Elimination
- A Comprehensive Survey of Process Reward Models: Data Generation, Model Construction, and Usage
- CODERL+: Improving Code Generation via Reinforcement with Execution Semantics Alignment
- Saber: Efficient Sampling with Adaptive Acceleration and Backtracking Enhanced Remasking for Diffusion Language Model in Code Generation
- APPSI-139: A Parallel Corpus of English Application Privacy Policy Summarization and Interpretation
- Bringing Real-World Relations into Video Generation with Graph-Structured Knowledge
- SAGE: Sparse Adaptive Guidance for Dependency-Aware Tabular Data Generation
- UniversalRAG: Retrieval-Augmented Generation over Corpora of Diverse Modalities and Granularities
- Looking at Radiology Report Generation through a Causal Lens: A Survey
- Disco-RAG: Discourse-Aware Retrieval-Augmented Generation
- Guaranteeing Knowledge Integration with Joint Decoding for Retrieval-Augmented Generation
- Retrieval as Generation: A Unified Framework with Self-Triggered Information Planning
- CodeFlowBench: A Multi-turn, Iterative Benchmark for Complex Code Generation
- Scaling Beyond Context: A Survey of Multimodal Retrieval-Augmented Generation for Document Understanding
- Open Schrödinger’s Closed Box: Identifying Retrieval Augmented Generation in API-Accessible Large Language Model Services
- AgentOCR: Reimagining Agent History via Optical Self-Compression
- A Theoretically Grounded Approach to Summarizing Conversation Dynamics for Forecasting the Derailment of Online Conversations
- LogicPoison: Logical Attacks on Graph Retrieval-Augmented Generation
- Compete to Complete: Co-opetition Adversarial Learning for Retrieval-Augmented Generation
- What Does LLM Refinement Actually Improve? A Systematic Study on Document-Level Literary Translation
- Measuring Human Contribution in AI-Assisted Content Generation
- ACE-Router: Generalizing History-Aware Routing from MCP Tools to the Agent Web
- UNIKIE-BENCH: Benchmarking Large Multimodal Models for Key Information Extraction in Visual Documents
- StealthGraph: Exposing Domain-Specific Risks in LLMs through Knowledge-Graph-Guided Harmful Prompt Generation
- Efficient Self-Evaluation for Diffusion Language Models via Sequence Regeneration
- Don’t Be Misled by Style: A Style-Adaptive Reranker for Capturing Effective Knowledge in Retrieval-Augmented Generation
- In-depth Research Impact Summarization through Fine-Grained Temporal Citation Analysis
- Is a Document Educational or Just Wikipedia-Style? — Pitfalls of Classifier-Based Quality Filtering
- STEM: Structure-Tracing Evidence Mining for Knowledge Graphs-Driven Retrieval-Augmented Generation
- Arg-LLaDA: Argument Summarization via Large Language Diffusion Models and Sufficiency-Aware Refinement
- arXiv2Table: Toward Realistic Benchmarking and Evaluation for LLM-Based Literature-Review Table Generation
- iTAG: Inverse Design for Natural Text Generation with Accurate Causal Graph Annotations
- LiGen: Active Lipid Generation via a Molecular Language Model
- Investigating More Explainable and Partition-Free Compositionality Estimation for LLMs: A Rule-Generation Perspective
- UniMoE-Audio: Unified Speech and Music Generation with Dynamic-Capacity Mixture-of-Experts
- KCVR: Knowledge-Centric Video Reconstruction for Structured Pedagogical Summarization via Dynamic Graph Planning
- RepoGenesis: Benchmarking End-to-End Microservice Generation from Readme to Repository
- AGSC: Adaptive Granularity and Semantic Clustering for Uncertainty Quantification in Long-text Generation
- TRACE: Traversal Retrieval-Augmented Chain of Evidence for Document Understanding
- Self-Reflective Generation at Test Time
- From Past To Path: Masked History Learning for Next-Item Prediction in Generative Recommendation
- Unified Thinker: A General Reasoning Core for Image Generation
- Visual and Memory–Augmented Soccer Commentary Generation
- Memory-Augmented LLM-based Multi-Agent System for Automated Feature Generation on Tabular Data
- Improving Retrieval-Augmented Generation without Taxonomy-based Error Categorization
- WebCoderBench: Benchmarking Web Application Generation with Comprehensive and Interpretable Evaluation Metrics
- EvoNarrator: Modeling Scientific Evolution for Feasible Hypothesis Generation
- LLM-SLM Collaborative Framework of Idiomatic Expression Generation
- Author-in-the-Loop Response Generation and Evaluation: Integrating Author Expertise and Intent in Responses to Peer Review
- Deriving Character Logic from Storyline as Codified Decision Trees
- SegTune: Structured and Fine-Grained Control for Song Generation
- IUQ: Interrogative Uncertainty Quantification for Long-Form Large Language Model Generation
- ALDEN: Reinforcement Learning for Active Navigation and Evidence Gathering in Long Documents
- Reinforced Informativeness Optimization for Long-Form Retrieval-Augmented Generation
- Illusions of the Gold Standard: A Large-scale Analysis of Human Evaluation Protocols for Long-form Text Generation
- Iterative Dual-Model Alignment for Story Evaluation
- ReFEree: Reference-Free and Fine-Grained Method for Evaluating Factual Consistency in Real-World Code Summarization
- MeanAudio: Fast and Faithful Text-to-Audio Generation with Mean Flows
- ModeX: Evaluator-Free Best-of-N Selection for Open-Ended Generation
- FormulaSPIN: Self-Play Fine-Tuning for Natural Language to Spreadsheet Formula Generation
- SARA: Selective and Adaptive Retrieval-augmented Generation with Context Compression
- SlideAgent: Hierarchical Agentic Framework for Multi-Page Visual Document Understanding
- LAD-RAG: Layout-aware Dynamic RAG for Visually-Rich Document Understanding
- ParaCodex: A Profiling-Guided Autonomous Coding Agent for Reliable Parallel Code Generation and Translation
- MARCH: Multi-Agent Radiology Clinical Hierarchy for CT Report Generation
- Biomedical Question Answering via Multi-Level Summarization on a Local Knowledge Graph
- Narrative License and Model Sycophancy in LLM Summaries of Scientific Work
- ViDoRe V3: A Comprehensive Evaluation of Retrieval Augmented Generation in Complex Real-World Scenarios
- Decoupling Task-Solving and Output Formatting in LLM Generation
- Analyzing and Internalizing Complex Policy Documents for LLM Agents
- Anchor: Branch-Point Data Generation for GUI Agents
- OpenRubrics: Towards Scalable Synthetic Rubric Generation for Reward Modeling and LLM Alignment
- LENS: LLM-Enabled Narrative Synthesis for Mental Health by Aligning Multimodal Sensing with Language Models
- Controlling Distributional Bias in Multi-Round LLM Generation via KL-Optimized Fine-Tuning
- SciFlow-Bench: Evaluating Structure-Aware Scientific Diagram Generation via Inverse Parsing
- LCR-RAG: Enhancing Logical Consistency in Retrieval-Augmented Generation via Neuro-symbolic Reinforcement Learning
- HiKEY: Hierarchical Multimodal Retrieval for Open-Domain Document Question Answering
- Your Reasoning Benchmark May Not Test Reasoning: Revealing Perception Bottleneck in Abstract Reasoning Benchmarks
- Beyond Chunking: Discourse-Aware Hierarchical Retrieval for Long Document Question Answering
- GBV-SQL: Guided Generation and SQL2Text Back-Translation Validation for Multi-Agent Text2SQL
- Adaptive Planning for Multi-Attribute Controllable Summarization with Monte Carlo Tree Search
- Difficulty-Controllable Cloze Question Distractor Generation
- REG: Retrieval via Emotion Similarity for Guiding Empathetic Dialogue Generation
- DisCo_Speech: Controllable Zero-Shot Speech Generation with A Disentangled Speech Codec
- Know the Known and the Unknown: Reasonable Answer Generation with Knowledge-Informed Citations
- Suggest-Verify-Revise: A Three-Stage Document-Level Event Causality Identification with Narrative Consistency
- StoryCoder: Narrative Reformulation for Structured Reasoning in LLM Code Generation
- Discourse Coherence and Response-Guided Context Rewriting for Multi-Party Dialogue Generation
- EvolvR: Self-Evolving Pairwise Reasoning for Story Evaluation to Enhance Generation
- Identity-Robust Language Model Generation via Content Integrity Preservation
- Knowledge Poisoning Attacks on Medical Multi-Modal Retrieval-Augmented Generation
- When Bigger Isn’t Better: A Comprehensive Fairness Evaluation of Political Bias in Multi-News Summarisation
- CAML: A Conflict-Aware Molecular Language Model Merging Framework for Multi-Constraint Molecular Generation
- HiGoE: Hierarchical Graph of Evidence to Enhance Retrieval-Augmented Generation for Long-context Summarization
- Region-Grounded Report Generation for 3D Medical Imaging: A Fine-Grained Dataset and Graph-Enhanced Framework
- DeepGuard: Secure Code Generation via Multi-Layer Semantic Aggregation
- IS-CoT: Breaking the Long-form Generation Collapse via Interleaved Structural Thinking
- MTAVG-Bench: A Diagnostic Benchmark for Multi-Talker Dialogue-Centric Audio-Video Generation
- Preconditioned Test-Time Adaptation for Out-of-Distribution Debiasing in Narrative Generation
- Bridging Internal Consistency and External Alignment: A Causal and Dynamic Interpretability Framework for LLM Generation
- R^3AG: Retriever Routing for Retrieval-Augmented Generation
- ClusterRAG: Cluster-Based Collaborative Filtering for Personalized Retrieval-Augmented Generation
- Aligned Multi-View Scripts for Universal Chart-to-Code Generation
- Breaking Block Boundaries: Anchor-based History-stable Decoding for Diffusion Large Language Models
- Unveiling the Unknown: Open-Set Entity Typing via Two-Stage Generation
- Tailoring Diagnostic Modeling to Individual Learners: Personalized Distractor Generation via MCTS-Guided Reasoning Reconstruction
- Efficient Prior-Guided Reasoning for Robust Retrieval-Augmented Generation under Conflicts
- LitVISTA: A Benchmark for Narrative Orchestration in Literary Text
- MAB-DQA: Addressing Query Aspect Importance in Document Question Answering with Multi-Armed Bandits
- FormalScience: Scalable Human-in-the-Loop Autoformalisation of Science with Agentic Code Generation in Lean
- BiMol-Diff: A Unified Diffusion Framework for Molecular Generation and Captioning
- DEBAR: Mitigating Contextual Bias in Cross-Document Relation Extraction via Dual-Stream Decoupling
- LegalChainReasoner: Grounding Criminal Judicial Opinion Generation via Structured Legal Chains
- InsideOut: Measuring and Mitigating Insider–Outsider Bias in Interview Script Generation
- Attention as Selector: Unlocking VLM Attention for Long Document Page Retrieval
- CharTide: Data-Centric Chart-to-Code Generation via Tri-Perspective Tuning and Inquiry-Driven Evolution
- Agent Newsroom: Efficient Chronological Report Generation via Dynamic Multi-Agent Collaboration
- ChipSeek: Optimizing Verilog Generation via EDA-Integrated Reinforcement Learning
- CSPO: Alleviating Reward Ambiguity for Structured Table-to-LaTeX Generation
- Chart-MRAG: Benchmarking Multimodal Retrieval Augmented Generation on Chart-based Documents
- Revisiting Metric Reliability for Fine-grained Evaluation of Machine Translation and Summarization in Indian Languages
- HAG: Hierarchical Demographic Tree-based Agent Generation for Topic-Adaptive Simulation
- Stable-RAG: Mitigating Retrieval-Permutation-Induced Hallucinations in Retrieval-Augmented Generation
- CIRAG: Construction–Integration Retrieval and Adaptive Generation for Multi-hop Question Answering
- Dynamic Infilling Anchors for Format-Constrained Generation in Diffusion Large Language Models
- Think Parallax: Solving Multi-Hop Problems via Multi-View Knowledge-Graph-Based Retrieval-Augmented Generation
- DocLens: A Tool-Augmented Multi-Agent Framework for Long Visual Document Understanding
- Attention Weights as an Indicator: Analyzing and Improving Document Utilization in Retrieval-Augmented Generation
- LADR: Locality-Aware Dynamic Rescue for Efficient Text-to-Image Generation with Diffusion Large Language Models
- A Multi-Agent Framework for Feature-Constrained Difficulty Control in Reading Comprehension Item Generation
- FusionFlow: Enabling Deep Structural Exploration for Automated Agentic Workflow Generation
- UniSonate: A Unified Model for Speech, Music, and Sound Effect Generation with Text Instructions
- HopWeaver: Cross-Document Synthesis of High-Quality and Authentic Multi-Hop Questions
- VerilogLAVD: LLM-Aided Pattern Generation for Verilog CWE Detection
- Bloom-Eval: A Hierarchical Evaluation Benchmark for Automatic Survey Generation Based on Bloom’s Taxonomy
- GenesisFunc: Multi-Agent Data Generation for Accurate and Generalizable Function-Calling
- FactVerse: A Benchmark for Factual Consistency in Interleaved Image–Text Generation
- Incorporating Temporal Coherence to Cross-Document Event Coreference Resolution
- HiChunk: Evaluating and Enhancing Retrieval Augmented Generation with Hierarchical Chunking
- PersonalityDBench: A Dataset for Personality Disorders - from Modeling to Controlled Generation
- Beyond Explicit Refusals: Soft-Failure Attacks on Retrieval-Augmented Generation
- MavenCoder: Competitive Code Generation via Model Adaptive Planning Strategies and Multi-Perspective Verification Enhancement
- CPR-RAG: Clinical Prior-Regularized Retrieval for Anatomy-Aware 3D CT Report Generation
- ArgGenBench: Benchmarking the Complex Controlled Argument Generation Capability of Large Language Models
- OSCBench: Benchmarking Object State Change in Text-to-Video Generation
- Attribution, Citation, and Quotation: A Survey of Evidence-based Text Generation with Large Language Models
- SwissGov-RSD: A Human-annotated, Cross-lingual Benchmark for Token-level Recognition of Semantic Differences Between Related Documents
- RubricHub: A Comprehensive and Highly Discriminative Rubric Dataset via Automated Coarse-to-Fine Generation
- Frankentext: Stitching random text fragments into long-form narratives
- Decisive: Guiding User Decisions with Optimal Preference Elicitation from Unstructured Documents
- Stress Testing Factual Consistency Metrics for Long-Document Summarization
- LLM-Generated Text May Harm Your Retrieval! A Robust Detection Strategy for Retrieval-Augmented Generation
- EvoSpark: Endogenous Interactive Agent Societies for Unified Long-Horizon Narrative Evolution
- HiddenGuard: Fine-Grained Safe Generation with Specialized Representation Router
- ThreadSumm: Summarization of Nested Discourse Threads Using Tree of Thoughts
- Enhanced Reasoning for Biomedical Document-Level Relation Extraction via a Novel Cascade Language Model Framework
- Multimodal Large Language Models for Multi-Subject In-Context Image Generation
- Synthetic Data Generation for Training Diversified Commonsense Reasoning Models
- MARS²: Scaling Multi-Agent Tree Search via Reinforcement Learning for Code Generation
- LoVeC: Reinforcement Learning for Better Verbalized Confidence in Long-Form Generation
- Anchoring the Cache: Mitigating Contextual Hallucination in KV-Compressed Long-Context Summarization
- Beyond Static Artifacts: An Evolutionary Framework for Synthetic Claim Generation
- ChiKhaPo: A Large-Scale Multilingual Benchmark for Evaluating Lexical Comprehension and Generation in Large Language Models
- Beyond Code Pairs: Dialogue-Based Data Generation for LLM Code Translation
- Fiction Flows: A Replication and Reinterpretation of Narrative Sequentiality
- Evaluating Visual Narrative Coherence in Story Visualization via Diversified Storylines
- CypherSmith: Transforming Text-to-Cypher Generation for LLMs with Synthetic Data
- ATGL: An Adaptive-Threshold Global Loss for Document-level Relation Extraction
- Utility-Oriented Visual Evidence Selection for Multimodal Retrieval-Augmented Generation
- DefGen-Bench: A Benchmark for Chinese Criminal Defence Opinion Generation in LegalAI
- SeDev: Structured Semantic Exploration for LLM-Driven Code Generation
- EVM-QuestBench: An Execution-Grounded Benchmark for Natural-Language Transaction Code Generation
- TexOCR: Advancing Document OCR Models for Compilable Page-to-LaTeX Reconstruction
- If an LLM Were a Character, Would It Know Its Own Story? Evaluating Lifelong Learning in LLMs
- Flow-Based Page Unique Semantic Mapping Architecture for Document Visual Question Answering
- SWAN: Semantic Watermarking with Abstract Meaning Representation
- PaT: Planning-after-Trial for Efficient Test-Time Code Generation
- The Role of Mixed-Language Documents for Multilingual Large Language Model Pretraining
- ES4R: Speech Encoding Based on Prepositive Affective Modeling for Empathetic Response Generation
- IGenBench: Benchmarking the Reliability of Text-to-Infographic Generation
- Enhancing Reinforcement Learning for Radiology Report Generation with Evidence-aware Rewards and Self-correcting Preference Learning
- Grammar as Control: Modular Language Generation for the Long Tail
- LegalGraphRAG: Multi-Agent Graph Retrieval-Augmented Generation for Reliable Legal Reasoning
- Omni-I2C: A Holistic Benchmark for High-Fidelity Image-to-Code Generation
- AdaDPI: Document-level Translation Adaptive Agent via Dynamic Parametric Internalization
- Cross-Examination Framework: A Task-Agnostic Diagnostic for Information Fidelity in Text-to-Text Generation
- Dynamic Generation of Multi LLM Agents Communication Topologies with Graph Diffusion Models
- One-step Nonautoregressive Natural Language Generation with Shortcut Flow Matching Models
- PHOTON: Hierarchical Autoregressive Modeling for Lightspeed and Memory-Efficient Language Generation
- More Aligned, Less Diverse? Analyzing the Grammar and Lexicon of Two Generations of LLMs
- Structure Guided Retrieval-Augmented Generation for Factual Queries
- From Competition to Synergy: Unlocking Reinforcement Learning for Subject-Driven Image Generation
- Deep-Reporter: Deep Research for Grounded Multimodal Long-Form Generation
- Survey Response Generation: Generating Closed-Ended Survey Responses In-Silico with Large Language Models
- Social Story Frames: Contextual Reasoning about Narrative Intent and Reception
- From Nodes to Narratives: Explaining Graph Neural Networks with LLMs and Graph Context
- RealChart2Code: Bridging the Gap in Real-World Chart-to-Code Generation via Multi-Task Evaluation
- HypoEval: Hypothesis-Guided Evaluation for Natural Language Generation
- A Structured Clustering Approach for Inducing Media Narratives
- Simple Agents, Biased Judges: Efficient Multi-Party Dialogue Generation & The Evaluation Gap
- Locate and Explain: Joint Multimodal Emotion Cause Extraction and Summarization in Conversation
- ServImage: An Image Generation and Editing Benchmark from Real-world Commercial Imaging Services
- PodBench: A Comprehensive Benchmark for Instruction-Aware Audio-Oriented Podcast Script Generation
- A Survey of Large Language Models for Text-Guided Molecular Discovery: From Molecule Generation to Optimization
- ReCode: Reinforcing Code Generation with Reasoning-Process Rewards
- CT-FineBench: A Diagnostic Fidelity Benchmark for Fine-Grained Evaluation of CT Report Generation
- MTRouter: Cost-Aware Multi-Turn LLM Routing with History–Model Joint Embeddings
- Automatic and Reliable Evaluation for Academic Caption-to-Figure Generation with LMMs
- SciMDR: Advancing Scientific Multimodal Document Reasoning
- Learning Faster with Better Tokens: Parameter-Efficient Vocabulary Adaptation for Specialized Text Summarization
- SHARP: Self-adaptive Harmful Category-aware Prompt Generation for Black-box Jailbreaking
- Mind’s Eye: A Benchmark of Visual Abstraction, Transformation and Composition for Multimodal LLMs
- Synthia: Scalable Grounded Persona Generation from Social Media Data
- From Recognition to Reasoning: Benchmarking and Enhancing MLLMs on Real-World Receipt Document Understanding
- CEBC: Conformal Evidence-Bounded Control for Low-Hallucination Vision–Language Generation
- DiNO: Disinformation Narrative Observer
- GitChameleon 2.0: Evaluating AI Code Generation Against Python Library Version Incompatibilities
- DREAM-S: Speculative Decoding with Searchable Drafting and Target-Aware Refinement for Multimodal Generation
- Protein-STORY: Semantic Text-Oriented Representation Yields biologically meaningful Protein embeddings
- Self-Guided Alignment: Adaptive Preference Sensing for Multi-Objective Generation
- Typology-Aware Multilingual Morphosyntactic Parsing with Joint Abstract Node Modeling
- MegaRAG: Multimodal Knowledge Graph-Based Retrieval Augmented Generation