- Published on
ICLR 2026 — Diffusion Models & Generative AI
Diffusion Models & Generative AI
672 papers (0 oral)
Half-order Fine-Tuning for Diffusion Model: A Recursive Likelihood Ratio Optimizer
- Link: OpenReview
Invisible Safety Threat: Malicious Finetuning for LLM via Steganography
- Link: OpenReview
Improving Diffusion Models for Class-imbalanced Training Data via Capacity Manipulation
- Link: OpenReview
Let Features Decide Their Own Solvers: Hybrid Feature Caching for Diffusion Transformers
- Link: OpenReview
DiffusionNFT: Online Diffusion Reinforcement with Forward Process
- Link: OpenReview
Universal Inverse Distillation for Matching Models with Real-Data Supervision (No GANs)
- Link: OpenReview
GLASS Flows: Efficient Inference for Reward Alignment of Flow and Diffusion Models
- Link: OpenReview
Neon: Negative Extrapolation From Self-Training Improves Image Generation
- Link: OpenReview
Generative Human Geometry Distribution
- Link: OpenReview
P-GenRM: Personalized Generative Reward Model with Test-time User-based Scaling
- Link: OpenReview
Locality-aware Parallel Decoding for Efficient Autoregressive Image Generation
- Link: OpenReview
SANA-Video: Efficient Video Generation with Block Linear Diffusion Transformer
- Link: OpenReview
Generative Universal Verifier as Multimodal Meta-Reasoner
- Link: OpenReview
Partition Generative Modeling: Masked Modeling Without Masks
- Link: OpenReview
NextStep-1: Toward Autoregressive Image Generation with Continuous Tokens at Scale
- Link: OpenReview
VibeVoice: Expressive Podcast Generation with Next-Token Diffusion
- Link: OpenReview
Exploratory Diffusion Model for Unsupervised Reinforcement Learning
- Link: OpenReview
Pareto-Conditioned Diffusion Models for Offline Multi-Objective Optimization
- Link: OpenReview
Compositional Diffusion with Guided search for Long-Horizon Planning
- Link: OpenReview
Stable Video Infinity: Infinite-Length Video Generation with Error Recycling
- Link: OpenReview
Diffusion Language Model Knows the Answer Before It Decodes
- Link: OpenReview
On the Reasoning Abilities of Masked Diffusion Language Models
- Link: OpenReview
Planner Aware Path Learning in Diffusion Language Models Training
- Link: OpenReview
MotionStream: Real-Time Video Generation with Interactive Motion Controls
- Link: OpenReview
TRACE: Your Diffusion Model is Secretly an Instance Edge Detector
- Link: OpenReview
Quotient-Space Diffusion Models
- Link: OpenReview
The Spacetime of Diffusion Models: An Information Geometry Perspective
- Link: OpenReview
Scaling Atomistic Protein Binder Design with Generative Pretraining and Test-Time Compute
- Link: OpenReview
Structured Flow Autoencoders: Learning Structured Probabilistic Representations with Flow Matching
- Link: OpenReview
Spherical Watermark: Encryption-Free, Lossless Watermarking for Diffusion Models
- Link: OpenReview
Enhancing Generative Auto-bidding with Offline Reward Evaluation and Policy Search
- Link: OpenReview
: Mixed Autoregressive and Diffusion Transformers for Continuous Image Generation
- Link: OpenReview
10. ARINBEV: Bird's-Eye View Layout Estimation with Conditional Autoregressive Model
- Topics: LLMs & Foundation Models
Forget Many, Forget Right: Scalable and Precise Concept Unlearning in Diffusion Models
- Link: OpenReview
Dynamic Classifier-Free Diffusion Guidance via Online Feedback
- Link: OpenReview
I-DRUID: Layout to image generation via instance-disentangled representation and unpaired data
- Link: OpenReview
GDR-learners: Orthogonal Learning of Generative Models for Potential Outcomes
- Link: OpenReview
Error as Signal: Stiffness-Aware Diffusion Sampling via Embedded Runge-Kutta Guidance
- Link: OpenReview
Multi-Scale Diffusion-Guided Graph Learning with Power-Smoothing Random Walk Contrast for Multi-View Clustering
- Link: OpenReview
LapFlow: Laplacian Multi-scale Flow Matching for Generative Modeling
- Link: OpenReview
Long-Text-to-Image Generation via Compositional Prompt Decomposition
- Link: OpenReview
Easier Painting Than Thinking: Can Text-to-Image Models Set the Stage, but Not Direct the Play?
- Link: OpenReview
FlowCast: Trajectory Forecasting for Scalable Zero-Cost Speculative Flow Matching
- Link: OpenReview
GRACE: Generative Representation Learning via Contrastive Policy Optimization
- Link: OpenReview
UltraViCo: Breaking Extrapolation Limits in Video Diffusion Transformers
- Link: OpenReview
Turbo-DDCM: Fast and Flexible Zero-Shot Diffusion-Based Image Compression
- Link: OpenReview
ImagenWorld: Stress-Testing Image Generation Models with Explainable Human Evaluation on Open-ended Real-World Tasks
- Link: OpenReview
CIAR: Interval-based Collaborative Decoding for Image Generation Acceleration
- Link: OpenReview
NeuralOS: Towards Simulating Operating Systems via Neural Generative Models
- Link: OpenReview
FSD-CAP: Fractional Subgraph Diffusion with Class-Aware Propagation for Graph Feature Imputation
- Link: OpenReview
MVCustom: Multi-View Customized Diffusion via Geometric Latent Rendering and Completion
- Link: OpenReview
UME-R1: Exploring Reasoning-Driven Generative Multimodal Embeddings
- Link: OpenReview
MAGREF: Masked Guidance for Any-Reference Video Generation with Subject Disentanglement
- Link: OpenReview
D-AR: Diffusion via Autoregressive Models
- Link: OpenReview
Diffusion Models as Dataset Distillation Priors
- Link: OpenReview
Detecting and Mitigating Memorization in Diffusion Models through Anisotropy of the Log-Probability
- Link: OpenReview
Recover Cell Tensor: Diffusion-Equivalent Tensor Completion for Fluorescence Microscopy Imaging
- Link: OpenReview
Stochastic Self-Guidance for Training-Free Enhancement of Diffusion Models
- Link: OpenReview
Beyond Uniformity: Sample and Frequency Meta Weighting for Post-Training Quantization of Diffusion Models
- Link: OpenReview
Cross-ControlNet: Training-Free Fusion of Multiple Conditions for Text-to-Image Generation
- Link: OpenReview
Gradient-Aligned Calibration for Post-Training Quantization of Diffusion Models
- Link: OpenReview
Point Prompting: Counterfactual Tracking with Video Diffusion Models
- Link: OpenReview
Bridging the Distribution Gap to Harness Pretrained Diffusion Priors for Super-Resolution
- Link: OpenReview
Score-Based Density Estimation from Pairwise Comparisons
- Link: OpenReview
RNE: plug-and-play diffusion inference-time control and energy-based training
- Link: OpenReview
Faster Diffusion Through Temporal Attention Decomposition
- Link: OpenReview
122. Video-As-Prompt: Unified Semantic Control for Video Generation
- Topics: Diffusion Models & Generative AI, Robotics & Control
Lyra: Generative 3D Scene Reconstruction via Video Diffusion Model Self-Distillation
- Link: OpenReview
Enhanced Generative Model Evaluation with Clipped Density and Coverage
- Link: OpenReview
Native Adaptive Solution Expansion for Diffusion-based Combinatorial Optimization
- Link: OpenReview
GHOST: Hallucination-Inducing Image Generation for Multimodal LLMs
- Link: OpenReview
Provably Accelerated Imaging with Restarted Inertia and Score-based Image Priors
- Link: OpenReview
Proximal Diffusion Neural Sampler
- Link: OpenReview
Discrete Diffusion for Reflective Vision-Language-Action Models in Autonomous Driving
- Link: OpenReview
Quant-dLLM: Post-Training Extreme Low-Bit Quantization for Diffusion Large Language Models
- Link: OpenReview
ImageDoctor: Diagnosing Text-to-Image Generation via Grounded Image Reasoning
- Link: OpenReview
gen2seg: Generative Models Enable Generalizable Instance Segmentation
- Link: OpenReview
SceneTransporter: Optimal Transport-Guided Compositional Latent Diffusion for Single-Image Structured 3D Scene Generation
- Link: OpenReview
GenDR: Lighten Generative Detail Restoration
- Link: OpenReview
On the Mechanisms of Collaborative Learning in VAE Recommenders
- Link: OpenReview
Unveiling the Potential of Diffusion Large Language Model in Controllable Generation
- Link: OpenReview
Carré du champ flow matching: better quality-generalisation tradeoff in generative models
- Link: OpenReview
Stage-wise Dynamics of Classifier-Free Guidance in Diffusion Models
- Link: OpenReview
Pixel-Level Residual Diffusion Transformer: Scalable 3D CT Volume Generation
- Link: OpenReview
Learning to Parallel: Accelerating Diffusion Large Language Models via Learnable Parallel Decoding
- Link: OpenReview
Diagnosing and Improving Diffusion Models by Estimating the Optimal Loss Value
- Link: OpenReview
ES-dLLM: Efficient Inference for Diffusion Large Language Models by Early-Skipping
- Link: OpenReview
Riemannian Variational Flow Matching for Material and Protein Design
- Link: OpenReview
Diffusion Blend: Inference-Time Multi-Preference Alignment for Diffusion Models
- Link: OpenReview
Continual Unlearning for Text-to-Image Diffusion Models: A Regularization Perspective
- Link: OpenReview
Improving Reasoning for Diffusion Language Models via Group Diffusion Policy Optimization
- Link: OpenReview
UltraLLaDA: Scaling the Context Length to 128K for Diffusion Large Language Models
- Link: OpenReview
Generalization of Diffusion Models Arises with a Balanced Representation Space
- Link: OpenReview
GGBall: Graph Generative Model on Poincaré Ball
- Link: OpenReview
Watermarking Diffusion Language Models
- Link: OpenReview
Provable Separations between Memorization and Generalization in Diffusion Models
- Link: OpenReview
Routing Matters in MoE: Scaling Diffusion Transformers with Explicit Routing Guidance
- Link: OpenReview
Eliminating VAE for Fast and High-Resolution Generative Detail Restoration
- Link: OpenReview
Assessing Robustness via Score-Based Adversarial Image Generation
- Link: OpenReview
414. Towards Anomaly-Aware Pre-Training and Fine-Tuning for Graph Anomaly Detection
- Topics: LLMs & Foundation Models, Graph Neural Networks, Graphs & Combinatorial
FARI: Robust One-Step Inversion for Watermarking in Diffusion Models
- Link: OpenReview
GeoDiv: Framework for Measuring Geographical Diversity in Text-to-Image Models
- Link: OpenReview
STEDiff: Revealing the Spatial and Temporal Redundancy of Backdoor Attacks in Text-to-Image Diffusion Models
- Link: OpenReview
SDErasure: Concept-Specific Trajectory Shifting for Concept Erasure via Adaptive Diffusion Classifier
- Link: OpenReview
Finite-Time Convergence Analysis of ODE-based Generative Models for Stochastic Interpolants
- Link: OpenReview
Early Signs of Steganographic Capabilities in Frontier LLMs
- Link: OpenReview
Towards Knowledge‑and‑Data‑Driven Organic Reaction Prediction: RAG‑Enhanced and Reasoning‑Powered Hybrid System with LLMs
- Link: OpenReview
Geometric Graph Neural Diffusion for Stable Molecular Dynamics Simulations
- Link: OpenReview
Shrinking Proteins with Diffusion
- Link: OpenReview
A Joint Diffusion Model with Pre-Trained Priors for RNA Sequence-Structure Co-Design
- Link: OpenReview
Online Decision Making with Generative Action Sets
- Link: OpenReview
STABLE: Shift-Tolerant Allocation via Black-Litterman Using Conditional Diffusion Estimates
- Link: OpenReview
WebArbiter: A Generative Reasoning Process Reward Model for Web Agents
- Link: OpenReview
Aurora: Towards Universal Generative Multimodal Time Series Forecasting
- Link: OpenReview
TS-DDAE: A Novel Temporal-Spectral Denoising Diffusion AutoEncoder for Wireless Signal Recognition Model Pre-training
- Link: OpenReview
Flow Caching for Autoregressive Video Generation
- Link: OpenReview
Improving Discrete Diffusion Unmasking Policies Beyond Explicit Reference Policies
- Link: OpenReview
MemGen: Weaving Generative Latent Memory for Self-Evolving Agents
- Link: OpenReview
MAS: Self-Generative, Self-Configuring, Self-Rectifying Multi-Agent Systems
- Link: OpenReview
Accelerating Diffusion Large Language Models with SlowFast Sampling: The Three Golden Principles
- Link: OpenReview
FLUX-Reason-6M & PRISM-Bench: A Million-Scale Text-to-Image Reasoning Dataset and Comprehensive Benchmark
- Link: OpenReview
Stroke3D: Lifting 2D strokes into rigged 3D model via latent diffusion models
- Link: OpenReview
Compose Your Policies! Improving Diffusion-based or Flow-based Robot Policies via Test-time Distribution-level Composition
- Link: OpenReview
Scalable Spatio-Temporal SE(3) Diffusion for Long-Horizon Protein Dynamics
- Link: OpenReview
Flow Along the -Amplitude for Generative Modeling
- Link: OpenReview
Uncovering Semantic Selectivity of Latent Groups in Higher Visual Cortex with Mutual Information-Guided Diffusion
- Link: OpenReview
Uniform Discrete Diffusion with Metric Path for Video Generation
- Link: OpenReview
MoGen: Detailed Neuronal Morphology Generation via Point Cloud Flow Matching
- Link: OpenReview
Think Then Embed: Generative Context Improves Multimodal Embedding
- Link: OpenReview
Efficient Zero-shot Inpainting with Decoupled Diffusion Guidance
- Link: OpenReview
Latent Denoising Makes Good Tokenizers
- Link: OpenReview
MoGA: Mixture-of-Groups Attention for End-to-End Long Video Generation
- Link: OpenReview
NExT-OMNI: Towards Any-to-Any Omnimodal Foundation Models with Discrete Flow Matching
- Link: OpenReview
DenseGRPO: From Sparse to Dense Reward for Flow Matching Model Alignment
- Link: OpenReview
Pretraining Scaling Laws for Generative Evaluations of Language Models
- Link: OpenReview
Ada-Diffuser: Latent-Aware Adaptive Diffusion for Decision-Making
- Link: OpenReview
Much Ado About Noising: Dispelling the Myths of Generative Robotic Control
- Link: OpenReview
Fastcar: Cache Attentive Replay for Fast Auto-Regressive Video Generation on the Edge
- Link: OpenReview
Frame Guidance: Training-Free Guidance for Frame-Level Control in Video Diffusion Models
- Link: OpenReview
WithAnyone: Toward Controllable and ID Consistent Image Generation
- Link: OpenReview
SIGMA-Gen: Structure and Identity Guided Multi-Subject Assembly for Image Generation
- Link: OpenReview
-DPO: Robust Preference Alignment for Diffusion Models via Divergence
- Link: OpenReview
VMDiff: Visual Mixing Diffusion for Limitless Cross-Object Synthesis
- Link: OpenReview
Autoregressive Image Generation with Randomized Parallel Decoding
- Link: OpenReview
Culture in Action: Evaluating Text-to-Image Models through Social Activities
- Link: OpenReview
Any-to-Bokeh: Arbitrary-Subject Video Refocusing with Video Diffusion Model
- Link: OpenReview
Deconstructing Guidance: A Semantic Hierarchy for Precise Diffusion Model Editing
- Link: OpenReview
NoisePrints: Distortion-Free Watermarks for Authorship in Private Diffusion Models
- Link: OpenReview
Trust but Verify: Adaptive Conditioning for Reference-Based Diffusion Super-Resolution via Implicit Reference Correlation Modeling
- Link: OpenReview
Let LLMs Speak Embedding Languages: Generative Text Embeddings via Iterative Contrastive Refinement
- Link: OpenReview
OBS-Diff: Accurate Pruning For Diffusion Models in One-Shot
- Link: OpenReview
SimpleGVR: A Simple Baseline for Latent-Cascaded Generative Video Super-Resolution
- Link: OpenReview
ImageRAG: Dynamic Image Retrieval for Reference-Guided Image Generation
- Link: OpenReview
DeLeaker: Dynamic Inference-Time Reweighting For Semantic Leakage Mitigation in Text-to-Image Models
- Link: OpenReview
FreeAdapt: Unleashing Diffusion Priors for Ultra-High-Definition Image Restoration
- Link: OpenReview
BLADE: Block-Sparse Attention Meets Step Distillation for Efficient Video Generation
- Link: OpenReview
MoSA: Motion-Coherent Human Video Generation via Structure-Appearance Decoupling
- Link: OpenReview
VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video Generation
- Link: OpenReview
Self-Forcing++: Towards Minute-Scale High-Quality Video Generation
- Link: OpenReview
Diffusion Negative Preference Optimization Made Simple
- Link: OpenReview
MIMIC: Mask-Injected Manipulation Video Generation with Interaction Control
- Link: OpenReview
PixNerd: Pixel Neural Field Diffusion
- Link: OpenReview
Temporal Concept Dynamics in Diffusion Models via Prompt-Conditioned Interventions
- Link: OpenReview
Learnable Sparsity for Vision Generative Models
- Link: OpenReview
LiveMoments: Reselected Key Photo Restoration in Live Photos via Reference-guided Diffusion
- Link: OpenReview
Generative Diffusion Prior Distillation for Long-Context Knowledge Transfer
- Link: OpenReview
CREPE: Controlling diffusion with REPlica Exchange
- Link: OpenReview
A Statistical Benchmark for Diffusion-Posterior-Sampling Algorithms
- Link: OpenReview
Distributionally Robust Optimization via Generative Ambiguity Modeling
- Link: OpenReview
DiffusionBlocks: Block-wise Neural Network Training via Diffusion Interpretation
- Link: OpenReview
ParallelBench: Understanding the Trade-offs of Parallel Decoding in Diffusion LLMs
- Link: OpenReview
TEST-TIME SCALING IN DIFFUSION LLMS VIA HIDDEN SEMI-AUTOREGRESSIVE EXPERTS
- Link: OpenReview
Time Is a Feature: Exploiting Temporal Dynamics in Diffusion Language Models
- Link: OpenReview
Variational Autoencoding Discrete Diffusion with Enhanced Dimensional Correlations Modeling
- Link: OpenReview
Your VAR Model is Secretly an Efficient and Explainable Generative Classifier
- Link: OpenReview
Avoid Catastrophic Forgetting with Rank-1 Fisher from Diffusion Models
- Link: OpenReview
GenCape: Structure-Inductive Generative Modeling for Category-Agnostic Pose Estimation
- Link: OpenReview
Taming Score-Based Denoisers in ADMM: A Convergent Plug-and-Play Framework
- Link: OpenReview
FideDiff: Efficient Diffusion Model for High-Fidelity Image Motion Deblurring
- Link: OpenReview
Consistent Text-to-Image Generation via Scene De-Contextualization
- Link: OpenReview
AlignFlow: Improving Flow-based Generative Models with Semi-Discrete Optimal Transport
- Link: OpenReview
Ultra-Fast Language Generation via Discrete Diffusion Divergence Instruct
- Link: OpenReview
Membership Inference Attacks Against Fine-tuned Diffusion Language Models
- Link: OpenReview
Learn to Guide Your Diffusion Model
- Link: OpenReview
DriftLite: Lightweight Drift Control for Inference-Time Scaling of Diffusion Models
- Link: OpenReview
DeRaDiff: Denoising Time Realignment of Diffusion Models
- Link: OpenReview
On the Design of One-step Diffusion via Shortcutting Flow Paths
- Link: OpenReview
Aligning Visual Foundation Encoders to Tokenizers for Diffusion Models
- Link: OpenReview
Noise-Adaptive Diffusion Sampling for Inverse Problems Without Task-Specific Tuning
- Link: OpenReview
Beyond Masks: Efficient, Flexible Diffusion Language Models via Deletion-Insertion Processes
- Link: OpenReview
SCOPED: Score–Curvature Out-of-distribution Proximity Evaluator for Diffusion
- Link: OpenReview
Scaling Laws for Diffusion Transformers
- Link: OpenReview
Principled RL for Diffusion LLMs Emerges from a Sequence-Level Perspective
- Link: OpenReview
GAS: Improving Discretization of Diffusion ODEs via Generalized Adversarial Solver
- Link: OpenReview
Texture Vector-Quantization and Reconstruction Aware Prediction for Generative Super-Resolution
- Link: OpenReview
ReFusion: A Diffusion Large Language Model with Parallel Autoregressive Decoding
- Link: OpenReview
ConfHit: Conformal Generative Design with Oracle-Free Guarantees
- Link: OpenReview
Edit-Based Flow Matching for Temporal Point Processes
- Link: OpenReview
Fine-Tuning Diffusion Models via Intermediate Distribution Shaping
- Link: OpenReview
AEGIS: Adversarial Target-Guided Retention-Data-Free Robust Concept Erasure from Diffusion Models
- Link: OpenReview
TrajFlow: Nation-wide Pseudo GPS Trajectory Generation with Flow Matching Models
- Link: OpenReview
Flow Matching with Semidiscrete Couplings
- Link: OpenReview
Forward-Learned Discrete Diffusion: Learning how to noise to denoise faster
- Link: OpenReview
HoloPart: Generative 3D Part Amodal Segmentation
- Link: OpenReview
Projected Coupled Diffusion for Test-Time Constrained Joint Generation
- Link: OpenReview
AdaBlock-dLLM: Semantic-Aware Diffusion LLM Inference via Adaptive Block Size
- Link: OpenReview
Low-Rank Few-Shot Node Classification by Node-Level Graph Diffusion
- Link: OpenReview
GenCtrl -- A Formal Controllability Toolkit for Generative Models
- Link: OpenReview
The Devil behind the mask: An emergent safety vulnerability of Diffusion LLMs
- Link: OpenReview
Constrained Diffusion for Protein Design with Hard Structural Constraints
- Link: OpenReview
MarS-FM: Generative Modeling of Molecular Dynamics via Markov State Models
- Link: OpenReview
Iterative Distillation for Reward-Guided Fine-Tuning of Diffusion Models in Biomolecular Design
- Link: OpenReview
BioMD: All-atom Generative Model for Biomolecular Dynamics Simulation
- Link: OpenReview
Enhancing Diffusion-Based Sampling with Molecular Collective Variables
- Link: OpenReview
Mirror Flow Matching with Heavy-Tailed Priors for Generative Modeling on Convex Domains
- Link: OpenReview
Flow Matching Policy Gradients
- Link: OpenReview
Improving 2D Diffusion Models for 3D Medical Imaging with Inter‑Slice Consistent Stochasticity
- Link: OpenReview
Beyond Penalization: Diffusion-based Out-of-Distribution Detection and Selective Regularization in Offline Reinforcement Learning
- Link: OpenReview
Contractive Diffusion Policies
- Link: OpenReview
Accelerating Diffusion Planners in Offline RL via Reward-Aware Consistency Trajectory Distillation
- Link: OpenReview
Block-wise Adaptive Caching for Accelerating Diffusion Policy
- Link: OpenReview
Masked Generative Policy for Robotic Control
- Link: OpenReview
Hierarchical Entity-centric Reinforcement Learning with Factored Subgoal Diffusion
- Link: OpenReview
HybridVLA: Collaborative Diffusion and Autoregression in a Unified Vision-Language-Action Model
- Link: OpenReview
Unified Diffusion VLA: Vision-Language-Action Model via Joint Discrete Denosing Diffusion Process
- Link: OpenReview
MMPD: Diverse Time Series Forecasting via Multi-Mode Patch Diffusion Loss
- Link: OpenReview
ICDiffAD: Implicit Conditioning Diffusion Model for Time Series Anomaly Detection
- Link: OpenReview
Revisiting Long-context Modeling from Context Denoising Perspective
- Link: OpenReview
Gauge Flow Matching: Efficient Constrained Generative Modeling over General Convex Set and Beyond
- Link: OpenReview
Reinforcing Diffusion Models by Direct Group Preference Optimization
- Link: OpenReview
Token-Based Audio Inpainting via Discrete Diffusion
- Link: OpenReview
SenseFlow: Scaling Distribution Matching for Flow-based Text-to-Image Distillation
- Link: OpenReview
Phantom-Data: Towards a General Subject-Consistent Video Generation Dataset
- Link: OpenReview
Ctrl-World: A Controllable Generative World Model for Robot Manipulation
- Link: OpenReview
A Noise is Worth Diffusion Guidance
- Link: OpenReview
AC-Sampler: Accelerate and Correct Diffusion Sampling with Metropolis-Hastings Algorithm
- Link: OpenReview
NarrLV: Towards a Comprehensive Narrative-Centric Evaluation for Long Video Generation
- Link: OpenReview
AlignSep: Temporally-Aligned Video-Queried Sound Separation with Flow Matching
- Link: OpenReview
LaTo: Landmark-tokenized Diffusion Transformer for Fine-grained Human Face Editing
- Link: OpenReview
Relational Feature Caching for Accelerating Diffusion Transformers
- Link: OpenReview
Syncphony: Synchronized Audio-to-Video Generation with Diffusion Transformers
- Link: OpenReview
NewtonGen: Physics-consistent and Controllable Text-to-Video Generation via Neural Newtonian Dynamics
- Link: OpenReview
VMoBA: Mixture-of-Block Attention for Video Diffusion Models
- Link: OpenReview
Dual-IPO: Dual-Iterative Preference Optimization for Text-to-Video Generation
- Link: OpenReview
Latent Diffusion Model without Variational Autoencoder
- Link: OpenReview
TINKER: Diffusion's Gift to 3D--Multi-View Consistent Editing From Sparse Inputs without Per-Scene Optimization
- Link: OpenReview
Towards Better Optimization For Listwise Preference in Diffusion Models
- Link: OpenReview
DiffSDA: Unsupervised Diffusion Sequential Disentanglement Across Modalities
- Link: OpenReview
Everything in Its Place: Benchmarking Spatial Intelligence of Text-to-Image Models
- Link: OpenReview
Learning Patient-Specific Disease Dynamics With Latent Flow Matching For Longitudinal Imaging Generation
- Link: OpenReview
Scale-wise Distillation of Diffusion Models
- Link: OpenReview
CreatiDesign: A Unified Multi-Conditional Diffusion Transformer for Creative Graphic Design
- Link: OpenReview
Light-X: Generative 4D Video Rendering with Camera and Illumination Control
- Link: OpenReview
Purrception: Variational Flow Matching for Vector-Quantized Image Generation
- Link: OpenReview
Time-to-Move: Training-Free Motion-Controlled Video Generation via Dual-Clock Denoising
- Link: OpenReview
Realtime Video Frame Interpolation using One-Step Diffusion Sampling
- Link: OpenReview
Localized Concept Erasure in Text-to-Image Diffusion Models via High-Level Representation Misdirection
- Link: OpenReview
ERTACache: Error Rectification and Timesteps Adjustment for Efficient Diffusion
- Link: OpenReview
PQGAN: Product-Quantised Image Representation for High-Quality Image Synthesis
- Link: OpenReview
STORK: Faster Diffusion and Flow Matching Sampling by Resolving both Stiffness and Structure-Dependence
- Link: OpenReview
CineTrans: Learning to Generate Videos with Cinematic Transitions via Masked Diffusion Models
- Link: OpenReview
3D Scene Prompting for Scene-Consistent Camera-Controllable Video Generation
- Link: OpenReview
Target-Aware Video Diffusion Models
- Link: OpenReview
Model Already Knows the Best Noise: Bayesian Active Noise Selection via Attention in Video Diffusion Model
- Link: OpenReview
DiffPBR: Point-Based Rendering via Spatial-Aware Residual Diffusion
- Link: OpenReview
Discrete Guidance Matching: Exact Guidance for Discrete Flow Matching
- Link: OpenReview
G4Splat: Geometry-Guided Gaussian Splatting with Generative Prior
- Link: OpenReview
Condition Matters in Full-head 3D GANs
- Link: OpenReview
Clustering by Denoising: Latent plug-and-play diffusion for single-cell embeddings
- Link: OpenReview
Lavida-O: Elastic Large Masked Diffusion Models for Unified Multimodal Understanding and Generation
- Link: OpenReview
DoFlow: Flow-based Generative Models for Interventional and Counterfactual Forecasting on Time Series
- Link: OpenReview
On Discriminative vs. Generative classifiers: Rethinking MLLMs for Action Understanding
- Link: OpenReview
Beyond Text-to-Image: Liberating Generation with a Unified Discrete Diffusion Model
- Link: OpenReview
Automatic Image-Level Morphological Trait Annotation for Organismal Images
- Link: OpenReview
Generative Bayesian Optimization: Generative Models as Acquisition Functions
- Link: OpenReview
Constrained Decoding of Diffusion LLMs with Context-Free Grammars
- Link: OpenReview
Cross-Timestep: 3D Diffusion Model with Trans-temporal Memory LSTM and Adaptive Priori Decoding Strategy for Medical Segmentation
- Link: OpenReview
Beyond Fixed: Training-Free Variable-Length Denoising for Diffusion Large Language Models
- Link: OpenReview
FSOD-VFM: Few-Shot Object Detection with Vision Foundation Models and Graph Diffusion
- Link: OpenReview
Divergence-Free Neural Networks with Application to Image Denoising
- Link: OpenReview
TangoFlux: Super Fast and Faithful Text to Audio Generation with Flow Matching and Clap-Ranked Preference Optimization
- Link: OpenReview
Adaptive Moments are Surprisingly Effective for Plug-and-Play Diffusion Sampling
- Link: OpenReview
Improving Classifier-Free Guidance in Masked Diffusion: Low-Dim Theoretical Insights with High-Dim Impact
- Link: OpenReview
Image Can Bring Your Memory Back: A Novel Multi-Modal Guided Attack against Image Generation Model Unlearning
- Link: OpenReview
SparseD: Sparse Attention for Diffusion Language Models
- Link: OpenReview
SoFlow: Solution Flow Models for One-Step Generative Modeling
- Link: OpenReview
TEN-DM: Topology-Enhanced Diffusion Model for Spatio-Temporal Event Prediction
- Link: OpenReview
DiffInk: Glyph- and Style-Aware Latent Diffusion Transformer for Text to Online Handwriting Generation
- Link: OpenReview
A Study of Posterior Stability in Time-Series Latent Diffusion
- Link: OpenReview
The Diffusion Duality, Chapter II: -Samplers and Efficient Curriculum
- Link: OpenReview
Generative Modeling from Black-Box Corruptions via Self-Consistent Stochastic Interpolants
- Link: OpenReview
FlashDLM: Accelerating Diffusion Language Model Inference via Efficient KV Caching and Guided Diffusion
- Link: OpenReview
Stopping Computation for Converged Tokens in Masked Diffusion-LM Decoding
- Link: OpenReview
Sample Reward Soups: Query-efficient Multi-Reward Guidance for Text-to-Image Diffusion Models
- Link: OpenReview
Soft-Masked Diffusion Language Models
- Link: OpenReview
Diverse Text-to-Image Generation via Contrastive Noise Optimization
- Link: OpenReview
Composition of Pretrained Diffusion Models: A Logic-Based Calculus
- Link: OpenReview
Antithetic Noise in Diffusion Models
- Link: OpenReview
Diffusion Fine-Tuning via Reparameterized Policy Gradient of the Soft Q-Function
- Link: OpenReview
Dual-Solver: A Generalized ODE Solver for Diffusion Models with Dual Prediction
- Link: OpenReview
Patronus: Interpretable Diffusion Models with Prototypes
- Link: OpenReview
Generative Value Conflicts Reveal LLM Priorities
- Link: OpenReview
Denoising Neural Reranker for Recommender Systems
- Link: OpenReview
Learning Flexible Forward Trajectories for Masked Molecular Diffusion
- Link: OpenReview
Global and Local Topology-Aware Graph Generation via Dual Conditioning Diffusion
- Link: OpenReview
Controllable diffusion-based generation for multi-channel biological data
- Link: OpenReview
Doloris: Dual Conditional Diffusion Implicit Bridges with Sparsity Masking Strategy for Unpaired Single-Cell Perturbation Estimation
- Link: OpenReview
scDFM: Distributional Flow Matching Model for Robust Single-Cell Perturbation Prediction
- Link: OpenReview
Physics vs Distributions: Pareto Optimal Flow Matching with Physics Constraints
- Link: OpenReview
Incomplete Data, Complete Dynamics: A Diffusion Approach
- Link: OpenReview
Vid2World: Crafting Video Diffusion Models to Interactive World Models
- Link: OpenReview
Improving Human-AI Coordination through Online Adversarial Training and Generative Models
- Link: OpenReview
VITA: Vision-to-Action Flow Matching Policy
- Link: OpenReview
LaDiR: Latent Diffusion Enhances LLMs for Text Reasoning
- Link: OpenReview
Diffusion LLMs Can Do Faster-Than-AR Inference via Discrete Diffusion Forcing
- Link: OpenReview
Planned Diffusion
- Link: OpenReview
ProteinAE: Protein Diffusion Autoencoders for Structure Encoding
- Link: OpenReview
Massive Activations are the Key to Local Detail Synthesis in Diffusion Transformers
- Link: OpenReview
NI Sampling: Accelerating Discrete Diffusion Sampling by Token Order Optimization
- Link: OpenReview
Scaling Speech Tokenizers with Diffusion Autoencoders
- Link: OpenReview
Inpainting-Guided Policy Optimization for Diffusion Large Language Models
- Link: OpenReview
HiCache: A Plug-in Scaled-Hermite Upgrade for Taylor-Style Cache-then-Forecast Diffusion Acceleration
- Link: OpenReview
Representation Alignment for Diffusion Transformers without External Components
- Link: OpenReview
PCPO: Proportionate Credit Policy Optimization for Preference Alignment of Image Generation Models
- Link: OpenReview
CL-DPS: A Contrastive Learning Approach to Blind Nonlinear Inverse Problem Solving via Diffusion Posterior Sampling
- Link: OpenReview
Scaling Behavior of Discrete Diffusion Language Models
- Link: OpenReview
Functional MRI Time Series Generation via Wavelet-Based Image Transform and Spectral Flow Matching for Brain Disorder Identification
- Link: OpenReview
Moving Beyond Diffusion: Hierarchy-to-Hierarchy Autoregression for fMRI-to-Image Reconstruction
- Link: OpenReview
Quantization-Aware Diffusion Models For Maximum Likelihood Training
- Link: OpenReview
BridgeDrive: Diffusion Bridge Policy for Closed-Loop Trajectory Planning in Autonomous Driving
- Link: OpenReview
Large Scale Diffusion Distillation via Score-Regularized Continuous-Time Consistency
- Link: OpenReview
Diffusion Transformers with Representation Autoencoders
- Link: OpenReview
Generative Adversarial Post-Training Mitigates Reward Hacking in Live Human-AI Music Interaction
- Link: OpenReview
Astraea: A Token-wise Acceleration Framework for Video Diffusion Transformers
- Link: OpenReview
ZeroGR: A Generalizable and Scalable Framework for Zero-Shot Generative Retrieval
- Link: OpenReview
Conditionally Whitened Generative Models for Probabilistic Time Series Forecasting
- Link: OpenReview
GenSR: Symbolic regression based on equation generative space
- Link: OpenReview
Multi-agent Coordination via Flow Matching
- Link: OpenReview
ConsisDrive: Identity-Preserving Driving World Models for Video Generation by Instance Mask
- Link: OpenReview
Generative Blocks World: Moving Things Around in Pictures
- Link: OpenReview
JavisDiT: Joint Audio-Video Diffusion Transformer with Hierarchical Spatio-Temporal Prior Synchronization
- Link: OpenReview
SeedVR2: One-Step Video Restoration via Diffusion Adversarial Post-Training
- Link: OpenReview
TTOM: Test-Time Optimization and Memorization for Compositional Video Generation
- Link: OpenReview
FastFlow: Accelerating The Generative Flow Matching Models with Bandit Inference
- Link: OpenReview
SIGMark: Scalable In-Generation Watermark with Blind Extraction for Video Diffusion
- Link: OpenReview
FilMaster: Bridging Cinematic Principles and Generative AI for Automated Film Generation
- Link: OpenReview
JavisDiT++: Unified Modeling and Optimization for Joint Audio-Video Generation
- Link: OpenReview
Object Fidelity Diffusion for Remote Sensing Image Generation
- Link: OpenReview
MMaDA-Parallel: Multimodal Large Diffusion Language Models for Thinking-Aware Editing and Generation
- Link: OpenReview
MILR: Improving Multimodal Image Generation via Test-Time Latent Reasoning
- Link: OpenReview
EchoMotion: Unified Human Video and Motion Generation via Dual-Modality Diffusion Transformer
- Link: OpenReview
RefAny3D: 3D Asset-Referenced Diffusion Models for Image Generation
- Link: OpenReview
Many-for-Many: Unify the Training of Multiple Video and Image Generation and Manipulation Tasks
- Link: OpenReview
AdaViewPlanner: Adapting Video Diffusion Models for Viewpoint Planning in 4D Scenes
- Link: OpenReview
Asynchronous Denoising Diffusion Models for Aligning Text-to-Image Generation
- Link: OpenReview
Streaming Autoregressive Video Generation via Diagonal Distillation
- Link: OpenReview
DiffVax: Optimization-Free Image Immunization Against Diffusion-Based Editing
- Link: OpenReview
Towards One-step Causal Video Generation via Adversarial Self-Distillation
- Link: OpenReview
Learning Video Generation for Robotic Manipulation with Collaborative Trajectory Control
- Link: OpenReview
Condition Errors Refinement in Autoregressive Image Generation with Diffusion Loss
- Link: OpenReview
MeanCache: From Instantaneous to Average Velocity for Accelerating Flow Matching Inference
- Link: OpenReview
BindWeave: Subject-Consistent Video Generation via Cross-Modal Integration
- Link: OpenReview
MoCa: Modeling Object Consistency for 3D Camera Control in Video Generation
- Link: OpenReview
ACCORD: Alleviating Concept Coupling through Dependence Regularization for Text-to-Image Diffusion Personalization
- Link: OpenReview
A Minimum Variance Path Principle for Accurate and Stable Score-Based Density Ratio Estimation
- Link: OpenReview
Source-Guided Flow Matching
- Link: OpenReview
From Prediction to Perfection: Introducing Refinement to Autoregressive Image Generation
- Link: OpenReview
Diffusion and Flow-based Copulas: Forgetting and Remembering Dependencies
- Link: OpenReview
There and Back Again: On the relation between Noise and Image Inversions in Diffusion Models
- Link: OpenReview
Rethinking Global Text Conditioning in Diffusion Transformers
- Link: OpenReview
QuantSparse: Comprehensively Compressing Video Diffusion Transformer with Model Quantization and Attention Sparsification
- Link: OpenReview
Preserve and Personalize: Personalized Text-to-Image Diffusion Models without Distributional Drift
- Link: OpenReview
Interp3D: Correspondence-aware Interpolation for Generative Textured 3D Morphing
- Link: OpenReview
Autoregressive Models Rival Diffusion Models at ANY-ORDER Generation
- Link: OpenReview
Gen-DFL: Decision-Focused Generative Learning for Robust Decision Making
- Link: OpenReview
ViMo: A Generative Visual GUI World Model for App Agents
- Link: OpenReview
Bridging Generalization Gap of Heterogeneous Federated Clients Using Generative Models
- Link: OpenReview
H2OFlow: Grounding Human-Object Affordances with 3D Generative Models and Dense Diffused Flows
- Link: OpenReview
Translate Policy to Language: Flow Matching Generated Rewards for LLM Explanations
- Link: OpenReview
Stochastic Self-Organization in Multi-Agent Systems
- Link: OpenReview
Spatial CAPTCHA: Generatively Benchmarking Spatial Reasoning for Human-Machine Differentiation
- Link: OpenReview
Hierarchy Decoding: A Training-free Parallel Decoding Strategy for Diffusion Large Language Models
- Link: OpenReview
KernelFusion: Zero-Shot Blind Super-Resolution via Patch Diffusion
- Link: OpenReview
DiffuDETR: Rethinking Detection Transformers with Denoising Diffusion Process
- Link: OpenReview
Massive Memorization with Hundreds of Trillions of Parameters for Sequential Transducer Generative Recommenders
- Link: OpenReview
DPad: Efficient Diffusion Language Models with Suffix Dropout
- Link: OpenReview
GoR: A Unified and Extensible Generative Framework for Ordinal Regression
- Link: OpenReview
FAST‑DIPS: Adjoint‑Free Analytic Steps and Hard‑Constrained Likelihood Correction for Diffusion‑Prior Inverse Problems
- Link: OpenReview
NeRV-Diffusion: Diffuse Implicit Neural Representation for Video Synthesis
- Link: OpenReview
Measurement Score-Based Diffusion Model
- Link: OpenReview
Entering the Era of Discrete Diffusion Models: A Benchmark for Schrödinger Bridges and Entropic Optimal Transport
- Link: OpenReview
VSF: Simple, Efficient, and Effective Negative Guidance in Few-Step Image Generation Models By Value Sign Flip
- Link: OpenReview
Secure Inference for Diffusion Models via Unconditional Scores
- Link: OpenReview
Mitigating Semantic Collapse in Generative Personalization with Test-Time Embedding Adjustment
- Link: OpenReview
STEER AWAY FROM MODE COLLISIONS: IMPROVING COMPOSITION IN DIFFUSION MODELS
- Link: OpenReview
wd1: Weighted Policy Optimization for Reasoning in Diffusion Language Models
- Link: OpenReview
Steering Diffusion Models Towards Credible Content Recommendation
- Link: OpenReview
A Probabilistic Hard Concept Bottleneck for Steerable Generative Models
- Link: OpenReview
Continuously Augmented Discrete Diffusion model for Categorical Generative Modeling
- Link: OpenReview
Propaganda AI: An Analysis of Semantic Divergence in Large Language Models
- Link: OpenReview
Harpoon: Generalised Manifold Guidance for Conditional Tabular Diffusion
- Link: OpenReview
HOG-Diff: Higher-Order Guided Diffusion for Graph Generation
- Link: OpenReview
Co-occurring Associated REtained concepts in Diffusion Unlearning
- Link: OpenReview
Video-GPT via Next Clip Diffusion
- Link: OpenReview
Purifying Generative LLMs from Backdoors without Prior Knowledge or Clean Reference
- Link: OpenReview
Contrastive Diffusion Guidance for Spatial Inverse Problems
- Link: OpenReview
Topological Flow Matching
- Link: OpenReview
Bures-Wasserstein Flow Matching for Graph Generation
- Link: OpenReview
Diffusion & Adversarial Schrödinger Bridges via Iterative Proportional Markovian Fitting
- Link: OpenReview
VFScale: Intrinsic Reasoning through Verifier-Free Test-time Scalable Diffusion Model
- Link: OpenReview
A Unification of Discrete, Gaussian, and Simplicial Diffusion
- Link: OpenReview
Discovering and Steering Interpretable Concepts in Large Generative Music Models
- Link: OpenReview
Learning Energy-Based Generative Models via Potential Flow: A Variational Principle Approach to Probability Density Homotopy Matching
- Link: OpenReview
3112. Localizing Task Recognition and Task Learning in In-Context Learning via Attention Head Analysis
- Topics: Other / Unclassified
Improving Black-Box Generative Attacks via Generator Semantic Consistency
- Link: OpenReview
TRACEDET: HALLUCINATION DETECTION FROM THE DECODING TRACE OF DIFFUSION LARGE LANGUAGE MODELS
- Link: OpenReview
Gradient Descent Dynamics of Rank-One Matrix Denoising
- Link: OpenReview
Continual Low-Rank Adapters for LLM-based Generative Recommender Systems
- Link: OpenReview
FragFM: Hierarchical Framework for Efficient Molecule Generation via Fragment-Level Discrete Flow Matching
- Link: OpenReview
h-MINT: Modeling Pocket-Ligand Binding with Hierarchical Molecular Interaction Network
- Link: OpenReview
OXtal: An All-Atom Diffusion Model for Organic Crystal Structure Prediction
- Link: OpenReview
MolEditRL: Structure-Preserving Molecular Editing via Discrete Diffusion and Reinforcement Learning
- Link: OpenReview
Horizon Imagination: Efficient On-Policy Rollout in Diffusion World Models
- Link: OpenReview
Latent-to-Data Cascaded Diffusion Models for Unconditional Time Series Generation
- Link: OpenReview
PINFDiT: Energy-Based Physics-Informed Diffusion Transformers for General-purpose Time Series Tasks
- Link: OpenReview
GAR: Generative Adversarial Reinforcement Learning for Formal Theorem Proving
- Link: OpenReview
Generative Adversarial Reasoner: Enhancing LLM Reasoning with Adversarial Reinforcement Learning
- Link: OpenReview
SpeechOp: Inference-Time Task Composition for Generative Speech Processing
- Link: OpenReview
Weak-to-Strong Diffusion with Reflection
- Link: OpenReview
Computational Bottlenecks for Denoising Diffusions
- Link: OpenReview
MATRIX: Mask Track Alignment for Interaction-aware Video Generation
- Link: OpenReview
Discrete Diffusion Trajectory Alignment via Stepwise Decomposition
- Link: OpenReview
Market Games for Generative Models: Equilibria, Welfare, and Strategic Entry
- Link: OpenReview
TAVAE: A VAE with Adaptable Priors Explains Contextual Modulation in the Visual Cortex
- Link: OpenReview
Discrete Diffusion for Bundle Construction
- Link: OpenReview
Automatic Stage Lighting Control: Is it a Rule-Driven Process or Generative Task?
- Link: OpenReview
WavefrontDiffusion: Dynamic Decoding Schedule for Improved Reasoning
- Link: OpenReview
Any-Order Flexible Length Masked Diffusion
- Link: OpenReview
Guaranteed Simply Connected Mesh Reconstruction from an Unorganized Point Cloud
- Link: OpenReview
Adaptive Domain Shift in Diffusion Models for Cross-Modality Image Translation
- Link: OpenReview
EA3D: Event-Augmented 3D Diffusion for Generalizable Novel View Synthesis
- Link: OpenReview
SpikeGen: Decoupled “Rods and Cones” Visual Representation Processing with Latent Generative Framework
- Link: OpenReview
La-Proteina: Atomistic Protein Generation via Partially Latent Flow Matching
- Link: OpenReview
Astra: General Interactive World Model with Autoregressive Denoising
- Link: OpenReview
GenCompositor: Generative Video Compositing with Diffusion Transformer
- Link: OpenReview
Mixture of Contexts for Long Video Generation
- Link: OpenReview
DistillKac: Few-Step Image Generation via Damped Wave Equations
- Link: OpenReview
Composition of Memory Experts for Diffusion World Models
- Link: OpenReview
Score-based Greedy Search for Structure Identification of Partially Observed Causal Models
- Link: OpenReview
Consistent Noisy Latent Rewards for Trajectory Preference Optimization in Diffusion Models
- Link: OpenReview
AttriCtrl: A Generalizable Framework for Controlling Semantic Attribute Intensity in Diffusion Models
- Link: OpenReview
VisualPrompter: Semantic-Aware Prompt Optimization with Visual Feedback for Text-to-Image Synthesis
- Link: OpenReview
Generative View Stitching
- Link: OpenReview
Arbitrary Generative Video Interpolation
- Link: OpenReview
Vivid-VR: Distilling Concepts from Text-to-Video Diffusion Transformer for Photorealistic Video Restoration
- Link: OpenReview
Mitigating Noise Shift in Denoising Generative Models with Noise Awareness Guidance
- Link: OpenReview
Concept-TRAK: Understanding how diffusion models learn concepts through concept attribution
- Link: OpenReview
BranchGRPO: Stable and Efficient GRPO with Structured Branching in Diffusion Models
- Link: OpenReview
MoAlign: Motion-Centric Representation Alignment for Video Diffusion Models
- Link: OpenReview
Verification of the Implicit World Model in a Generative Model via Adversarial Sequences
- Link: OpenReview
RePrompt: Reasoning-Augmented Reprompting for Text-to-Image Generation via Reinforcement Learning
- Link: OpenReview
EasyTune: Efficient Step-Aware Fine-Tuning for Diffusion-Based Motion Generation
- Link: OpenReview
SteinsGate: Adding Causality to Diffusions for Long Video Generation via Path Integral
- Link: OpenReview
Soft-Di[M]O: Improving One-Step Discrete Image Generation with Soft Embeddings
- Link: OpenReview
Directional Textual Inversion for Personalized Text-to-Image Generation
- Link: OpenReview
LikePhys: Evaluating Intuitive Physics Understanding in Video Diffusion Models via Likelihood Preference
- Link: OpenReview
BWCache: Accelerating Video Diffusion Transformers through Block-Wise Caching
- Link: OpenReview
Real-Time Motion-Controllable Autoregressive Video Diffusion
- Link: OpenReview
ToonComposer: Streamlining Cartoon Production with Generative Post-Keyframing
- Link: OpenReview
Generalized Compressed Sensing for Image Reconstruction with Diffusion Probabilistic Models
- Link: OpenReview
3760. In-Context Learning of Temporal Point Processes with Foundation Inference Models
- Topics: Efficiency & Compression
Alternating Diffusion for Proximal Sampling with Zeroth Order Queries
- Link: OpenReview
ReSplat: Degradation-agnostic Feed-forward Gaussian Splatting via Self-guided Residual Diffusion
- Link: OpenReview
Diffusion-DFL: Decision-focused Diffusion Models for Stochastic Optimization
- Link: OpenReview
A Sharp KL Convergence Analysis for Diffusion Models under Minimal Assumptions
- Link: OpenReview
Strictly Constrained Generative Modeling via Split Augmented Langevin Sampling
- Link: OpenReview
Test-Time Scaling with Reflective Generative Model
- Link: OpenReview
NatADiff: Adversarial Boundary Guidance for Natural Adversarial Diffusion
- Link: OpenReview
Logit‑KL Flow Matching: Non‑Autoregressive Text Generation via Sampling‑Hybrid Inference
- Link: OpenReview
EnvSocial-Diff: A Diffusion-Based Crowd Simulation Model with Environmental Conditioning and Individual-Group Interaction
- Link: OpenReview
TEDM: Time Series Forecasting with Elucidated Diffusion Models
- Link: OpenReview
What Exactly Does Guidance Do in Masked Discrete Diffusion Models
- Link: OpenReview
DeepWeightFlow: Re-Basined Flow Matching for Generating Neural Network Weights
- Link: OpenReview
Text-Aware Image Restoration with Diffusion Models
- Link: OpenReview
Quartet of Diffusions: Structure-Aware Point Cloud Generation through Part and Symmetry Guidance
- Link: OpenReview
On the Interpolation Effect of Score Smoothing in Diffusion Models
- Link: OpenReview
Universal Multi-Domain Translation via Diffusion Routers
- Link: OpenReview
Energy-oriented Diffusion Bridge for Image Restoration with Foundational Diffusion Models
- Link: OpenReview
Bringing Stability to Diffusion: Decomposing and Reducing Variance of Training Masked Diffusion Models
- Link: OpenReview
There is No VAE: End-to-End Pixel-Space Generative Modeling via Self-Supervised Pre-Training
- Link: OpenReview
Dual-Path Condition Alignment for Diffusion Transformers
- Link: OpenReview
Foundational Automatic Evaluators: Scaling Multi-Task Generative Evaluator Training for Reasoning-Centric Domains
- Link: OpenReview
Delay Flow Matching
- Link: OpenReview
SPREAD: Sampling-based Pareto front Refinement via Efficient Adaptive Diffusion
- Link: OpenReview
Generalised Flow Maps for Few-Step Generative Modelling on Riemannian Manifolds
- Link: OpenReview
CheckMate! Watermarking Graph Diffusion Models in Polynomial Time
- Link: OpenReview
Foresight Diffusion: Improving Sampling Consistency in Predictive Diffusion Models
- Link: OpenReview
Why Adversarially Train Diffusion Models?
- Link: OpenReview
Toward Safer Diffusion Language Models: Discovery and Mitigation of Priming Vulnerability
- Link: OpenReview
Are Deep Speech Denoising Models Robust to Adversarial Noise?
- Link: OpenReview
Uncovering Conceptual Blindspots in Generative Image Models Using Sparse Autoencoders
- Link: OpenReview
On Powerful Ways to Generate: Autoregression, Diffusion, and Beyond
- Link: OpenReview
PepTri: Tri-Guided All-Atom Diffusion for Peptide Design via Physics, Evolution, and Mutual Information
- Link: OpenReview
PoseX: AI Defeats Physics-based Methods on Protein Ligand Cross-Docking
- Link: OpenReview
SAIR: Enabling Deep Learning for Protein-Ligand Interactions with a Synthetic Structural Dataset
- Link: OpenReview
Interpolation-Based Conditioning of Flow Matching Models for Bioisosteric Ligand Design
- Link: OpenReview
A Hidden Semantic Bottleneck in Conditional Embeddings of Diffusion Transformers
- Link: OpenReview
FlowCast: Advancing Precipitation Nowcasting with Conditional Flow Matching
- Link: OpenReview
SigmaDock: Untwisting Molecular Docking with Fragment-Based SE(3) Diffusion
- Link: OpenReview
RIDER: 3D RNA Inverse Design with Reinforcement Learning-Guided Diffusion
- Link: OpenReview
An Optimal Diffusion Approach to Quadratic Rate-Distortion Problems: New Solution and Approximation Methods
- Link: OpenReview
Inference-Time Scaling of Discrete Diffusion Models via Importance Weighting and Optimal Proposal Design
- Link: OpenReview
GAS: Enhancing Reward-Cost Balance of Generative Model-assisted Offline Safe RL
- Link: OpenReview
One-Step Flow Q-Learning: Addressing the Diffusion Policy Bottleneck in Offline Reinforcement Learning
- Link: OpenReview
GenCP: Towards Generative Modeling Paradigm of Coupled physics
- Link: OpenReview
Bridging Successor Measure and Online Policy Learning with Flow Matching-Based Representations
- Link: OpenReview
Grounding Generative Planners in Verifiable Logic: A Hybrid Architecture for Trustworthy Embodied AI
- Link: OpenReview
Abstracting Robot Manipulation Skills via Mixture-of-Experts Diffusion Policies
- Link: OpenReview
DAK-UCB: Diversity-Aware Prompt Routing for LLMs and Generative Models
- Link: OpenReview
Compositional Visual Planning via Inference-Time Diffusion Scaling
- Link: OpenReview
Flow Matching with Injected Noise for Offline-to-Online Reinforcement Learning
- Link: OpenReview
Confident and Adaptive Generative Speech Recognition via Risk Control
- Link: OpenReview
LongLive: Real-time Interactive Long Video Generation
- Link: OpenReview
Smarter Not Harder: Generative Process Evaluation with Intrinsic-Signal Driving and Ability‑Adaptive Reward Shaping
- Link: OpenReview
Revolutionizing Reinforcement Learning Framework for Diffusion Large Language Models
- Link: OpenReview
SLA: Beyond Sparsity in Diffusion Transformers via Fine-Tunable Sparse–Linear Attention
- Link: OpenReview
SpatialHand: Generative Object Manipulation from 3D Prespective
- Link: OpenReview
Improved Adversarial Diffusion Compression for Real-World Video Super-Resolution
- Link: OpenReview
Dynamic-dLLM: Dynamic Cache-Budget and Adaptive Parallel Decoding for Training-Free Acceleration of Diffusion LLM
- Link: OpenReview
BideDPO: Conditional Image Generation with Simultaneous Text and Condition Alignment
- Link: OpenReview
Diffusion Language Models are Provably Optimal Parallel Samplers
- Link: OpenReview
DiffuCoder: Understanding and Improving Masked Diffusion Models for Code Generation
- Link: OpenReview
Pusa V1.0: Unlocking Temporal Control in Pretrained Video Diffusion Models via Vectorized Timestep Adaptation
- Link: OpenReview
Fast-dLLM: Training-free Acceleration of Diffusion LLM by Enabling KV Cache and Parallel Decoding
- Link: OpenReview
Deforming Videos to Masks: Flow Matching for Referring Video Segmentation
- Link: OpenReview
Reconciling Visual Perception and Generation in Diffusion Models
- Link: OpenReview
SAIL: Self-Amplified Iterative Learning for Diffusion Model Alignment with Minimal Human Feedback
- Link: OpenReview
RAPID: Tri-Level Reinforced Acceleration Policies for Diffusion Transformer
- Link: OpenReview
LumosX: Relate Any Identities with Their Attributes for Personalized Video Generation
- Link: OpenReview
AttTok: Marrying Attribute Tokens with Generative Pre-trained Vision-Language Models towards Medical Image Understanding
- Link: OpenReview
Dichotomous Diffusion Policy Optimization
- Link: OpenReview
Rolling Forcing: Autoregressive Long Video Diffusion in Real Time
- Link: OpenReview
CoDA: From Text-to-Image Diffusion Models to Training-Free Dataset Distillation
- Link: OpenReview
Frozen Priors, Fluid Forecasts: Prequential Uncertainty for Low-Data Deployment with Pretrained Generative Models
- Link: OpenReview
Evolutionary Caching to Accelerate Your Off-the-Shelf Diffusion Model
- Link: OpenReview
ReactID: Synchronizing Realistic Actions and Identity in Personalized Video Generation
- Link: OpenReview
Lumos-1: On Autoregressive Video Generation with Discrete Diffusion from a Unified Model Perspective
- Link: OpenReview
Plug-and-Play Fidelity Optimization for Diffusion Transformer Acceleration via Cumulative Error Minimization
- Link: OpenReview
DiCache: Let Diffusion Model Determine Its Own Cache
- Link: OpenReview
Anchor Frame Bridging for Coherent First-Last Frame Video Generation
- Link: OpenReview
Group Critical-token Policy Optimization for Autoregressive Image Generation
- Link: OpenReview
HiGS: History-Guided Sampling for Plug-and-Play Enhancement of Diffusion Models
- Link: OpenReview
SPEED: Scalable, Precise, and Efficient Concept Erasure for Diffusion Models
- Link: OpenReview
SPRINT: Sparse-Dense Residual Fusion for Efficient Diffusion Transformers
- Link: OpenReview
QVGen: Pushing the Limit of Quantized Video Generative Models
- Link: OpenReview
Arbitrary-Shaped Image Generation via Spherical Neural Field Diffusion
- Link: OpenReview
Geometry Forcing: Marrying Video Diffusion and 3D Representation for Consistent World Modeling
- Link: OpenReview
TS-Attn: Temporal-wise Separable Attention for Multi-Event Video Generation
- Link: OpenReview
LazyDrag: Enabling Stable Drag-Based Editing on Multi-Modal Diffusion Transformers via Explicit Correspondence
- Link: OpenReview
DiffSparse: Accelerating Diffusion Transformers with Learned Token Sparsity
- Link: OpenReview
WILD-Diffusion: A WDRO Inspired Training Method for Diffusion Models under Limited Data
- Link: OpenReview
Controllable Video Generation with Provable Disentanglement
- Link: OpenReview
From Broad Exploration to Stable Synthesis: Entropy-Guided Optimization for Autoregressive Image Generation
- Link: OpenReview
Training-Free Text-Guided Color Editing with Multi-Modal Diffusion Transformer
- Link: OpenReview
Disentanglement of Variations with Multimodal Generative Modeling
- Link: OpenReview
PI-Light: Physics-Inspired Diffusion for Full-Image Relighting
- Link: OpenReview
Factuality Matters: When Image Generation and Editing Meet Structured Visuals
- Link: OpenReview
Motion Prior Distillation in Time Reversal Sampling for Generative Inbetweening
- Link: OpenReview
TPDiff: Temporal Pyramid Video Diffusion Model
- Link: OpenReview
PreciseCache: Precise Feature Caching for Efficient and High-fidelity Video Generation
- Link: OpenReview
Test-Time Iterative Error Correction for Efficient Diffusion Models
- Link: OpenReview
CoDi: Subject-Consistent and Pose-Diverse Text-to-Image Generation
- Link: OpenReview
Latent Wavelet Diffusion For Ultra High-Resolution Image Synthesis
- Link: OpenReview
Disentangled Hierarchical VAE for 3D Human-Human Interaction Generation
- Link: OpenReview
TreeGRPO: Tree-Advantage GRPO for Online RL Post-Training of Diffusion Models
- Link: OpenReview
DSA: Efficient Inference For Video Generation Models via Distributed Sparse Attention
- Link: OpenReview
Geometric Image Editing via Effects-Sensitive In-Context Inpainting with Diffusion Transformers
- Link: OpenReview
Scaling Sequence-to-Sequence Generative Neural Rendering
- Link: OpenReview
Query-Aware Flow Diffusion for Graph-Based RAG with Retrieval Guarantees
- Link: OpenReview
dCache: Accelerating Diffusion-Based LLMs via Dual Adaptive Caching
- Link: OpenReview
Consolidating Reinforcement Learning for Multimodal Discrete Diffusion Models
- Link: OpenReview
Attention Is All You Need for KV Cache in Diffusion LLMs
- Link: OpenReview
Loopholing Discrete Diffusion: Deterministic Bypass of the Sampling Wall
- Link: OpenReview
FS-DFM: Fast and Accurate Long Text Generation with Few-Step Diffusion Language Models
- Link: OpenReview
Self-Speculative Masked Diffusions
- Link: OpenReview
Information Estimation with Discrete Diffusion
- Link: OpenReview
KinemaDiff: Towards Diffusion for Coherent and Physically Plausible Human Motion Prediction
- Link: OpenReview
LucidFlux: Caption-Free Universal Image Restoration via a Large-Scale Diffusion Transformer
- Link: OpenReview
ProtoKV: Long-context Knowledges Are Already Well-Organized Before Your Query
- Link: OpenReview
Rainbow Padding: Mitigating Early Termination in Instruction-Tuned Diffusion LLMs
- Link: OpenReview
One step further with Monte-Carlo sampler to guide diffusion better
- Link: OpenReview
Shortcut Diffusion Training with Cumulative Consistency Loss: An Optimal Control View
- Link: OpenReview
Animating the Uncaptured: Humanoid Mesh Animation with Video Diffusion Models
- Link: OpenReview
Diffusion Alignment as Variational Expectation-Maximization
- Link: OpenReview
DM4CT: Benchmarking Diffusion Models for Computed Tomography Reconstruction
- Link: OpenReview
Parallel Sampling from Masked Diffusion Models via Conditional Independence Testing
- Link: OpenReview
Neodragon: Mobile Video Generation Using Diffusion Transformer
- Link: OpenReview
Improved Object-Centric Diffusion Learning with Registers and Contrastive Alignment
- Link: OpenReview
High-dimensional Mean-Field Games by Particle-based Flow Matching
- Link: OpenReview
CardioComposer: Leveraging Differentiable Geometry for Compositional Control of Anatomical Diffusion Models
- Link: OpenReview
Multiplicative Diffusion Models: Beyond Gaussian Latents
- Link: OpenReview
EigenScore: OOD Detection using Posterior Covariance in Diffusion Models
- Link: OpenReview
SERUM: Simple, Efficient, Robust, and Unifying Marking for Diffusion-based Image Generation
- Link: OpenReview
Guidance Watermarking for Diffusion Models
- Link: OpenReview
DiffuGuard: How Intrinsic Safety is Lost and Found in Diffusion Large Language Models
- Link: OpenReview
A2D: Any-Order, Any-Step Safety Alignment for Diffusion Language Models
- Link: OpenReview
Do We Need All the Synthetic Data? Targeted Image Augmentation via Diffusion Models
- Link: OpenReview
GDGB: A Benchmark for Generative Dynamic Text-Attributed Graph Learning
- Link: OpenReview
MambaVoiceCloning: Efficient and Expressive Text-to-Speech via State-Space Modeling and Diffusion Control
- Link: OpenReview
Pallatom-Ligand: an All-Atom Diffusion Model for Designing Ligand-Binding Proteins
- Link: OpenReview
Graph Diffusion Transformers are In-Context Molecular Designers
- Link: OpenReview
Polynomial Convergence of Riemannian Diffusion Models
- Link: OpenReview
SafeFlowMatcher: Safe and Fast Planning using Flow Matching with Control Barrier Functions
- Link: OpenReview
HDP: Triply‑Hierarchical Diffusion Policy for Visuomotor Learning
- Link: OpenReview
Demystifying Robot Diffusion Policies: Action Memorization and a Simple Lookup Table Alternative
- Link: OpenReview
DrivingGen: A Comprehensive Benchmark for Generative Video World Models in Autonomous Driving
- Link: OpenReview
DecompGAIL: Learning Realistic Traffic Behaviors with Decomposed Multi-Agent Generative Adversarial Imitation Learning
- Link: OpenReview
Geometry-aware 4D Video Generation for Robot Manipulation
- Link: OpenReview
Time-Gated Multi-Scale Flow Matching for Time-Series Imputation
- Link: OpenReview
GCGNet: Graph-Consistent Generative Network for Time Series Forecasting with Exogenous Variables
- Link: OpenReview
SE-Diff: Simulator and Experience Enhanced Diffusion Model for Comprehensive ECG Generation
- Link: OpenReview
GlobeDiff: State Diffusion Process for Partial Observability in Multi-Agent System
- Link: OpenReview
KL-Regularized Reinforcement Learning for Generative Modelling is Designed to Mode Collapse
- Link: OpenReview
What Generative Search Engines Like and How to Optimize Web Content Cooperatively
- Link: OpenReview
DreamOn: Diffusion Language Models For Code Infilling Beyond Fixed-size Canvas
- Link: OpenReview
CoCoDiff: Correspondence-Consistent Diffusion Model for Fine-grained Style Transfer
- Link: OpenReview
Flow2GAN: Hybrid Flow Matching and GAN with Multi-Resolution Network for Few-step High-Fidelity Audio Generation
- Link: OpenReview
StepORLM: A Self-Evolving Framework With Generative Process Supervision For Operations Research Language Models
- Link: OpenReview
SPG: Sandwiched Policy Gradient for Masked Diffusion Language Models
- Link: OpenReview
Guidance Matters: Rethinking the Evaluation Pitfall for Text-to-Image Generation
- Link: OpenReview
Score Distillation Beyond Acceleration: Generative Modeling from Corrupted Data
- Link: OpenReview
MindPilot: Closed-loop Visual Stimulation Optimization for Brain Modulation with EEG-guided Diffusion
- Link: OpenReview
Multi-Subspace Multi-Modal Modeling for Diffusion Models: Estimation, Convergence and Mixture of Experts
- Link: OpenReview
iFusion: Integrating Dynamic Interest Streams via Diffusion Model for Click-Through Rate Prediction
- Link: OpenReview
Don't Settle Too Early: Self-Reflective Remasking for Diffusion Language Models
- Link: OpenReview
Multi-Marginal Flow Matching with Adversarially Learnt Interpolants
- Link: OpenReview
Zero-shot Human Pose Estimation using Diffusion-based Inverse solvers
- Link: OpenReview
Fast-dLLM v2: Efficient Block-Diffusion LLM
- Link: OpenReview
Landing with the Score: Riemannian Optimization through Denoising
- Link: OpenReview
SpaceControl: Introducing Test-Time Spatial Control to 3D Generative Modeling
- Link: OpenReview
Interleaving Reasoning for Better Text-to-Image Generation
- Link: OpenReview
Inference-time scaling of diffusion models through classical search
- Link: OpenReview
CryoNet.Refine: A One-step Diffusion Model for Rapid Refinement of Structural Models with Cryo-EM Density Map Restraints
- Link: OpenReview
TrustGen: A Platform of Dynamic Benchmarking on the Trustworthiness of Generative Foundation Models
- Link: OpenReview
Implicit Regularisation in Diffusion Models: An Algorithm-Dependent Generalisation Analysis
- Link: OpenReview
Step-Aware Residual-Guided Diffusion for EEG Spatial Super-Resolution
- Link: OpenReview