- Published on
ICLR 2026 — Theory & Deep Learning Theory
Theory & Deep Learning Theory
256 papers (0 oral)
Improving Diffusion Models for Class-imbalanced Training Data via Capacity Manipulation
- Link: OpenReview
Compactness and Consistency: A Conjoint Framework for Deep Graph Clustering
- Link: OpenReview
The Shape of Adversarial Influence: Characterizing LLM Latent Spaces with Persistent Homology
- Link: OpenReview
TileLang: Bridge Programmability and Performance in Modern Neural Kernels
- Link: OpenReview
Depth Anything 3: Recovering the Visual Space from Any Views
- Link: OpenReview
Probabilistic Kernel Function for Fast Angle Testing
- Link: OpenReview
On the Generalization Capacities of MLLMs for Spatial Intelligence
- Link: OpenReview
Vid-LLM: A Compact Video-based 3D Multimodal LLM with Reconstruction–Reasoning Synergy
- Link: OpenReview
Exploring Synthesizable Chemical Space with Iterative Pathway Refinements
- Link: OpenReview
Efficient Resource-Constrained Training of Transformers via Subspace Optimization
- Link: OpenReview
Mamba-3: Improved Sequence Modeling using State Space Principles
- Link: OpenReview
Navigating the Latent Space Dynamics of Neural Models
- Link: OpenReview
To Infinity and Beyond: Tool-Use Unlocks Length Generalization in State Space Models
- Link: OpenReview
Quotient-Space Diffusion Models
- Link: OpenReview
The Spacetime of Diffusion Models: An Information Geometry Perspective
- Link: OpenReview
Discount Model Search for Quality Diversity Optimization in High-Dimensional Measure Spaces
- Link: OpenReview
Block-Sample MAC-Bayes Generalization Bounds
- Link: OpenReview
2. Enhancing Communication Compression via Discrepancy-aware Calibration for Federated Learning
- Topics: Efficiency & Compression
Overshoot and Shrinkage in Classifier-Free Guidance: From Theory to Practice
- Link: OpenReview
KDP: Simplifying Representation Dynamics in Kernel Space
- Link: OpenReview
MHLA: Restoring Expressivity of Linear Attention via Token-Level Multi-Head
- Link: OpenReview
Scalable Random Wavelet Features: Efficient Non-Stationary Kernel Approximation with Convergence Guarantees
- Link: OpenReview
Temporal Test-Time Adaptation with State-Space Models
- Link: OpenReview
131. ShapeGen4D: Towards High Quality 4D Shape Generation from Videos
- Topics: Other / Unclassified
Null-Space Filtering for Data-Free Continual Model Merging: Preserving Stability, Promoting Plasticity
- Link: OpenReview
DISK: Differentiable Sparse Kernel Complex for Efficient Spatially-Variant Convolution
- Link: OpenReview
Sculpting Subspaces: Constrained Full Fine-Tuning in LLMs for Continual Learning
- Link: OpenReview
Theory of Space: Can Foundation Models Construct Spatial Beliefs through Active Exploration?
- Link: OpenReview
Unveiling the Cognitive Compass: Theory-of-Mind–Guided Multimodal Emotion Reasoning
- Link: OpenReview
Achieving low-bit Muon through subspace preservation and grid quantization
- Link: OpenReview
Towards Reliable Detection of Empty Space: Conditional Marked Point Processes for Object Detection
- Link: OpenReview
xLSTM Scaling Laws: Competitive Performance with Linear Time-Complexity
- Link: OpenReview
LaplacianFormer:Rethinking Linear Attention with Laplacian Kernel
- Link: OpenReview
Vulcan: Crafting Compact Class-Specific Vision Transformers For Edge Intelligence
- Link: OpenReview
Rethinking Expressivity and Degradation-Awareness in Attention for All-in-One Blind Image Restoration
- Link: OpenReview
PonderLM: Pretraining Language Models to Ponder in Continuous Space
- Link: OpenReview
Next-ToBE: Probabilistic Next Token-Bag Exploitation for Activating Anticipatory Capacity in LLMs
- Link: OpenReview
Generalization of Diffusion Models Arises with a Balanced Representation Space
- Link: OpenReview
Provable Separations between Memorization and Generalization in Diffusion Models
- Link: OpenReview
Mitigating the Safety Alignment Tax with Null-Space Constrained Policy Optimization
- Link: OpenReview
Theory of Scaling Laws for In-Context Regression: Depth, Width, Context and Time
- Link: OpenReview
SSDi8: Accurate and Efficient 8-bit Quantization for State Space Duality
- Link: OpenReview
Dual-Kernel Adapter: Expanding Spatial Horizons for Data-Constrained Medical Image Analysis
- Link: OpenReview
Locally Subspace-Informed Neural Operators for Efficient Multiscale PDE Solving
- Link: OpenReview
SPACeR: Self-Play Anchoring with Centralized Reference Models
- Link: OpenReview
Weight-Space Linear Recurrent Neural Networks
- Link: OpenReview
Learning Admissible Heuristics for A*: Theory and Practice
- Link: OpenReview
Operator Theory-Driven Autoformulation of MDPs for Control of Queueing Systems
- Link: OpenReview
Unraveling the Complexity of Memory in RL Agents: an Approach for Classification and Evaluation
- Link: OpenReview
Tucker-FNO: Tensor Tucker-Fourier Neural Operator and its Universal Approximation Theory
- Link: OpenReview
Graph Representational Learning: When Does More Expressivity Hurt Generalization?
- Link: OpenReview
Weight Space Representation Learning on Diverse NeRF Architectures
- Link: OpenReview
Subspace Kernel Learning on Tensor Sequences
- Link: OpenReview
The Intricate Dance of Prompt Complexity, Quality, Diversity and Consistency in T2I Models
- Link: OpenReview
Separable Neural Networks: Approximation Theory, NTK Regime, and Preconditioned Gradient Descent
- Link: OpenReview
Supporting High-Stakes Decision Making Through Interactive Preference Elicitation in the Latent Space
- Link: OpenReview
PAC-Bayes bounds for cumulative loss in Continual Learning
- Link: OpenReview
MaskPro: Linear-Space Probabilistic Learning for Strict (N:M)-Sparsity on LLMs
- Link: OpenReview
Memory-Free Continual Learning with Null Space Adaptation for Zero-Shot Vision-Language Models
- Link: OpenReview
Symmetry-Aware Bayesian Optimization via Max Kernels
- Link: OpenReview
Improving LLM-based Global Optimization with Search Space Partitioning
- Link: OpenReview
SpaCE-Eval: A Benchmark for Real-World Multi-Modal Reasoning
- Link: OpenReview
STARK: Strategic Team of Agents for Refining Kernels
- Link: OpenReview
Capacity-Aware Inference: Mitigating the Straggler Effect in Mixture of Experts
- Link: OpenReview
MaRS: Memory-Adaptive Routing for Reliable Capacity Expansion and Knowledge Retention
- Link: OpenReview
TyphoonMLA: A Mixed Naive-Absorb MLA Kernel For Shared Prefix
- Link: OpenReview
Chimera: State Space Models Beyond Sequences
- Link: OpenReview
1171. Enhancing Vision Transformers for Object Detection via Context-Aware Token Selection and Packing
- Topics: LLMs & Foundation Models, Computer Vision, Theory & Deep Learning Theory
ActivationReasoning: Logical Reasoning in Latent Activation Spaces
- Link: OpenReview
PE-SGD: Differentially Private Deep Learning via Evolution of Gradient Subspace for Text
- Link: OpenReview
LS-Merge: Merging Language Models in Latent Space
- Link: OpenReview
Obfuscated Activations Bypass LLM Latent-Space Defenses
- Link: OpenReview
Unpacking Human Preference for LLMs: Demographically Aware Evaluation with the HUMAINE Framework
- Link: OpenReview
Dual-Space Smoothness for Robust and Balanced LLM Unlearning
- Link: OpenReview
AIRE-Prune: Asymptotic Impulse-Response Energy for State Pruning in State Space Models
- Link: OpenReview
Enabling True Global Perception in State Space Models for Visual Tasks
- Link: OpenReview
ConvT3: Structured State Kernels for Convolutional State Space Models
- Link: OpenReview
Do We Really Need Permutations? Impact of Model Width on Linear Mode Connectivity
- Link: OpenReview
Prune-then-Quantize or Quantize-then-Prune? Understanding the Impact of Compression Order in Joint Model Compression
- Link: OpenReview
Understanding In-Context Learning on Structured Manifolds: Bridging Attention to Kernel Methods
- Link: OpenReview
PiCa: Parameter-Efficient Fine-Tuning with Column Space Projection
- Link: OpenReview
Out of the Shadows: Exploring a Latent Space for Neural Network Verification
- Link: OpenReview
Formal Mechanistic Interpretability: Automated Circuit Discovery with Provable Guarantees
- Link: OpenReview
Beyond Ensembles: Simulating All-Atom Protein Dynamics in a Learned Latent Space
- Link: OpenReview
Concepts' Information Bottleneck Models
- Link: OpenReview
Orbital Transformers for Predicting Wavefunctions in Time-Dependent Density Functional Theory
- Link: OpenReview
Complexity Analysis of Normalizing Constant Estimation: from Jarzynski Equality to Annealed Importance Sampling and beyond
- Link: OpenReview
The Sample Complexity of Online Reinforcement Learning: A Multi-model Perspective
- Link: OpenReview
From atom to space: A region-based readout function for spatial properties of materials
- Link: OpenReview
Sample Complexity and Representation Ability of Test-time Scaling Paradigms
- Link: OpenReview
Deep Learning for Subspace Regression
- Link: OpenReview
From Parameters to Behaviors: Unsupervised Compression of the Policy Space
- Link: OpenReview
Distributions as Actions: A Unified Framework for Diverse Action Spaces
- Link: OpenReview
COSA: Context-aware Output-Space Adapter for Test-Time Adaptation in Time Series Forecasting
- Link: OpenReview
Provable and Practical In-Context Policy Optimization for Self-Improvement
- Link: OpenReview
FlowNIB: An Information Bottleneck Analysis of Bidirectional vs. Unidirectional Language Models
- Link: OpenReview
Entropy-Monitored Kernelized Token Distillation for Audio-Visual Compression
- Link: OpenReview
Unifying Formal Explanations: A Complexity-Theoretic Perspective
- Link: OpenReview
Meta-Learning Theory-Informed Inductive Biases using Deep Kernel Gaussian Processes
- Link: OpenReview
Reasoning in Space via Grounding in the World
- Link: OpenReview
Unified Vision–Language Modeling via Concept Space Alignment
- Link: OpenReview
PACE: Pretrained Audio Continual Learning
- Link: OpenReview
Let's Explore Step by Step: Generating Provable Formal Statements with Deductive Exploration
- Link: OpenReview
Flow Straight and Fast in Hilbert Space: Functional Rectified Flow
- Link: OpenReview
Achieving Olympia-Level Geometry Large Language Model Agent via Complexity Boosting Reinforcement Learning
- Link: OpenReview
Domain Expansion: A Latent Space Construction Framework for Multi-Task Learning
- Link: OpenReview
A Single Architecture for Representing Invariance Under Any Space Group
- Link: OpenReview
Knowledge Fusion of Large Language Models via Modular SkillPacks
- Link: OpenReview
Bilevel Optimization with Lower-Level Uniform Convexity: Theory and Algorithm
- Link: OpenReview
Tighter Performance Theory of FedExProx
- Link: OpenReview
Point-Focused Attention Meets Context-Scan State Space: Robust Biological Visual Perception for Point Cloud Representation
- Link: OpenReview
BioTamperNet: Affinity-Guided State-Space Model Detecting Tampered Biomedical Images
- Link: OpenReview
EnsembleSHAP: Faithful and Certifiably Robust Attribution for Random Subspace Method
- Link: OpenReview
Improving Classifier-Free Guidance in Masked Diffusion: Low-Dim Theoretical Insights with High-Dim Impact
- Link: OpenReview
Exploring the Design Space of Transition Matching
- Link: OpenReview
Paradigm Shift of GNN Explainer from Label Space to Prototypical Representation Space
- Link: OpenReview
VERIFY: A Novel Multi-Domain Dataset Grounding LTL in Contextual Natural Language via Provable Intermediate Logic
- Link: OpenReview
Two failure modes of deep transformers and how to avoid them: a unified theory of signal propagation at initialisation
- Link: OpenReview
Exploring Mode Connectivity in Krylov Subspace for Domain Generalization
- Link: OpenReview
Training-Free Determination of Network Width via Neural Tangent Kernel
- Link: OpenReview
Deep Hierarchical Learning with Nested Subspace Networks for Large Language Models
- Link: OpenReview
A Statistical Theory of Overfitting for Imbalanced Classification
- Link: OpenReview
FlexLinearAttention: Compiling a Unified Abstraction into Scalable Kernels for Linear Attention
- Link: OpenReview
The Curious Case of In-Training Compression of State Space Models
- Link: OpenReview
Topology and geometry of the learning space of ReLU networks: connectivity and singularities
- Link: OpenReview
Active Learning for Decision Trees with Provable Guarantees
- Link: OpenReview
Characterizing and Optimizing the Spatial Kernel of Multi Resolution Hash Encodings
- Link: OpenReview
Learning Molecular Chirality via Chiral Determinant Kernels
- Link: OpenReview
How hard is learning to cut? Trade-offs and sample complexity
- Link: OpenReview
CMT-Benchmark: A Benchmark for Condensed Matter Theory Built by Expert Researchers
- Link: OpenReview
Geometry of Uncertainty: Learning Metric Spaces for Multimodal State Estimation in RL
- Link: OpenReview
Dancing in Chains: Strategic Persuasion in Academic Rebuttal via Theory of Mind
- Link: OpenReview
Laplacian Kernelized Bandit
- Link: OpenReview
Measuring Audio's Impact on Correctness: Audio-Contribution-Aware Post-Training of Large Audio Language Models
- Link: OpenReview
-Reasoner: LLM Reasoning via Test-Time Gradient Descent in Latent Space
- Link: OpenReview
Scaling Linear Attention Capacity with Sparse State Expansion
- Link: OpenReview
Characterizing Human Semantic Navigation in Concept Production as Trajectories in Embedding Space
- Link: OpenReview
QuestA: Expanding Reasoning Capacity in LLMs via Question Augmentation
- Link: OpenReview
HEAPr: Hessian-based Efficient Atomic Expert Pruning in Output Space
- Link: OpenReview
GT-Space: Enhancing Heterogeneous Collaborative Perception with Ground Truth Feature Space
- Link: OpenReview
Sharp asymptotic theory for Q-learning with learning rate and its generalization
- Link: OpenReview
Trion: FFT-based Dynamic Subspace Selection for Low-Rank Adaptive Optimization of LLMs
- Link: OpenReview
GenSR: Symbolic regression based on equation generative space
- Link: OpenReview
Symmetric Space Learning for Combinatorial Generalization
- Link: OpenReview
Rethinking LLM-as-a-Judge: Representation-as-a-Judge with Small Language Models via Semantic Capacity Asymmetry
- Link: OpenReview
HARP: Hallucination Detection via Reasoning Subspace Projection
- Link: OpenReview
UniCon: Unified Framework for Efficient Contrastive Alignment via Kernels
- Link: OpenReview
Evaluating Cross-Modal Reasoning Ability and Problem Characteristics with Multimodal Item Response Theory
- Link: OpenReview
Calibrated Information Bottleneck for Trusted Multi-modal Clustering
- Link: OpenReview
Training Dynamics Impact Post-Training Quantization Robustness
- Link: OpenReview
Understanding the Mixture-of-Experts with Nadaraya-Watson Kernel
- Link: OpenReview
Learning Global Hypothesis Space for Enhancing Synergistic Reasoning Chain
- Link: OpenReview
KernelFusion: Zero-Shot Blind Super-Resolution via Patch Diffusion
- Link: OpenReview
FSA: An Alternative Efficient Implementation of Native Sparse Attention Kernel
- Link: OpenReview
Exploring State-Space Models for Data-Specific Neural Representations
- Link: OpenReview
LoRAGen: Structure-Aware Weight Space Learning for LoRA Generation
- Link: OpenReview
gLSTM: Mitigating Over-Squashing by Increasing Storage Capacity
- Link: OpenReview
On the trade-off between expressivity and privacy in graph representation learning
- Link: OpenReview
Minimax Sample Complexity of Graph Neural Networks: Lower Bounds and Structural Effects
- Link: OpenReview
Bridging Input Feature Spaces Towards Graph Foundation Models
- Link: OpenReview
A universal compression theory for lottery ticket hypothesis and neural scaling laws
- Link: OpenReview
Jailbreaking the Matrix: Nullspace Steering for Controlled Model Subversion
- Link: OpenReview
A Rich Knowledge Space for Scalable Deepfake Detection
- Link: OpenReview
Safety Subspaces are Not Linearly Distinct: A Fine-Tuning Case Study
- Link: OpenReview
On the Ability of Deep Networks to Learn Symmetries from Data – A Neural Kernel Theory
- Link: OpenReview
3190. Reusing Pre-Training Data at Test Time is a Compute Multiplier
- Topics: LLMs & Foundation Models
SUIT: Knowledge Editing with Subspace-Aware Key-Value Mappings
- Link: OpenReview
SGD-Based Knowledge Distillation with Bayesian Teachers: Theory and Guidelines
- Link: OpenReview
Sampling Complexity of TD and PPO in RKHS
- Link: OpenReview
MASS: MoErging through Adaptive Subspace Selection
- Link: OpenReview
Emotions Where Art Thou: Understanding and Characterizing the Emotional Latent Space of Large Language Models
- Link: OpenReview
Convergence of an actor-critic gradient flow for entropy regularised MDPs in general spaces
- Link: OpenReview
Improving and Accelerating Offline RL in Large Discrete Action Spaces with Structured Policy Initialization
- Link: OpenReview
Learning linear state-space models with sparse system matrices
- Link: OpenReview
Kevin: Multi-Turn RL for Generating CUDA Kernels
- Link: OpenReview
FormalML: A Benchmark for Evaluating Formal Subgoal Completion in Machine Learning Theory
- Link: OpenReview
PerFit: Exploring Personalization Shifts in Representation Space of LLMs
- Link: OpenReview
Accelerating Eigenvalue Dataset Generation via Chebyshev Subspace Filter
- Link: OpenReview
A Biologically Plausible Dense Associative Memory with Exponential Capacity
- Link: OpenReview
Breaking the Total Variance Barrier: Sharp Sample Complexity for Linear Heteroscedastic Bandits with Fixed Action Set
- Link: OpenReview
SpaCE-10: A Comprehensive Benchmark for Multimodal Large Language Models in Compositional Spatial Intelligence
- Link: OpenReview
Understanding and Improving Continuous LLM Adversarial Training via In-context Learning Theory
- Link: OpenReview
Almost Bayesian: Dynamics of SGD Through Singular Learning Theory
- Link: OpenReview
Adaptive Thinking: Large Language Models Know When to Think in Latent Space
- Link: OpenReview
AlphaSteer: Learning Refusal Steering with Principled Null-Space Constraint
- Link: OpenReview
Elastic Optimal Transport: Theory, Application, and Empirical Evaluation
- Link: OpenReview
Efficient Orthogonal Fine-Tuning with Principal Subspace Adaptation
- Link: OpenReview
Boosted Trees on a Diet: Compact Models for Resource-Constrained Devices
- Link: OpenReview
GenFusion: Feed-forward Human Performance Capture via Progressive Canonical Space Updates
- Link: OpenReview
Sharpness-Aware Minimization in Logit Space Efficiently Enhances Direct Preference Optimization
- Link: OpenReview
Long-Context Attention Benchmark: From Kernel Efficiency to Distributed Context Parallelism
- Link: OpenReview
Ringleader ASGD: The First Asynchronous SGD with Optimal Time Complexity under Data Heterogeneity
- Link: OpenReview
Breaking the Correlation Plateau: On the Optimization and Capacity Limits of Attention-Based Regressors
- Link: OpenReview
VideoAnchor: Reinforcing Subspace-Structured Visual Cues for Coherent Visual-Spatial Reasoning
- Link: OpenReview
KV Cache Transform Coding for Compact Storage in LLM Inference
- Link: OpenReview
Pixel3DMM: Versatile Screen-Space Priors for Single-Image 3D Face Reconstruction
- Link: OpenReview
There is No VAE: End-to-End Pixel-Space Generative Modeling via Self-Supervised Pre-Training
- Link: OpenReview
Adapting Self-Supervised Representations as a Latent Space for Efficient Generation
- Link: OpenReview
On the Impact of the Utility in Semivalue-based Data Valuation
- Link: OpenReview
PolySHAP: Extending KernelSHAP with Interaction-Informed Polynomial Regression
- Link: OpenReview
Some Neural Networks Inherently Preserve Subspace Clustering Structure
- Link: OpenReview
Near-Optimal Sample Complexity Bounds for Constrained Average-Reward MDPs
- Link: OpenReview
Mitigating the Curse of Detail: Scaling Arguments for Feature Learning and Sample Complexity
- Link: OpenReview
Non-Clashing Teaching in Graphs: Algorithms, Complexity, and Bounds
- Link: OpenReview
The Geometry of Reasoning: Flowing Logics in Representation Space
- Link: OpenReview
Towards a Transferable Acceleration Method for Density Functional Theory
- Link: OpenReview
Revisiting Nonstationary Kernel Design for Multi-Output Gaussian Processes
- Link: OpenReview
Unveiling the Mechanism of Continuous Representation Full-Waveform Inversion: A Wave Based Neural Tangent Kernel Framework
- Link: OpenReview
Complexity- and Statistics-Guided Anomaly Detection in Time Series Foundation Models
- Link: OpenReview
AutoDrive-R²: Incentivizing Reasoning and Self-Reflection Capacity for VLA Model in Autonomous Driving
- Link: OpenReview
Predicting Kernel Regression Learning Curves from Only Raw Data Statistics
- Link: OpenReview
SplitLoRA: Balancing Stability and Plasticity in Continual Learning Through Gradient Space Splitting
- Link: OpenReview
The Hot Mess of AI: How Does Misalignment Scale With Model Intelligence and Task Complexity?
- Link: OpenReview
Continuous Space-Time Video Super-Resolution with 3D Fourier Fields
- Link: OpenReview
Aligning Collaborative View Recovery and Tensorial Subspace Learning via Latent Representation for Incomplete Multi-View Clustering
- Link: OpenReview
Compositional Generalization through Gradient Search in Nonparametric Latent Space
- Link: OpenReview
Trajectory-aware Shifted State Space Models for Online Video Super-Resolution
- Link: OpenReview
Controllable Video Generation with Provable Disentanglement
- Link: OpenReview
Beyond the Heatmap: A Rigorous Evaluation of Component Impact in MCTS-Based TSP Solvers
- Link: OpenReview
From Sequential to Parallel: Reformulating Dynamic Programming as GPU Kernels for Large-Scale Stochastic Combinatorial Optimization
- Link: OpenReview
Sequential Information Bottleneck Fusion: Towards Robust and Generalizable Multi-Modal Brain Tumor Segmentation
- Link: OpenReview
ReLaSH: Reconstructing Joint Latent Spaces for Efficient Generation of Synthetic Hypergraphs with Hyperlink Attributes
- Link: OpenReview
On the Universality and Complexity of GNN for Solving Second-order Cone Programs
- Link: OpenReview
A Guardrail for Safety Preservation: When Safety-Sensitive Subspace Meets Harmful-Resistant Null-Space
- Link: OpenReview
Decomposing Representation Space into Interpretable Subspaces with Unsupervised Learning
- Link: OpenReview
PACEbench: A Framework for Evaluating Practical AI Cyber-Exploitation Capabilities
- Link: OpenReview
MambaVoiceCloning: Efficient and Expressive Text-to-Speech via State-Space Modeling and Diffusion Control
- Link: OpenReview
A Genetic Algorithm for Navigating Synthesizable Molecular Spaces
- Link: OpenReview
A Theoretical Analysis of Mamba’s Training Dynamics: Filtering Relevant Features for Generalization in State Space Models
- Link: OpenReview
Understanding the Dynamics of Forgetting and Generalization in Continual Learning via the Neural Tangent Kernel
- Link: OpenReview
Bridging Kolmogorov Complexity and Deep Learning: Asymptotically Optimal Description Length Objectives for Transformers
- Link: OpenReview
PCF Learned Sort: a Learning Augmented Sort Algorithm with O(nloglogn) Expected Complexity
- Link: OpenReview
5058. PatchDNA: A Flexible and Biologically-Informed Alternative to Tokenization for DNA
- Topics: Other / Unclassified
Path Channels and Plan Extension Kernels: a Mechanistic Description of Planning in a Sokoban RNN
- Link: OpenReview
On the Expressiveness of State Space Models via Temporal Logics
- Link: OpenReview
Efficient Estimation of Kernel Surrogate Models for Task Attribution
- Link: OpenReview
Action Chunking and Data Augmentation Yield Exponential Improvements in Behavior Cloning for Continuous Spaces
- Link: OpenReview
DriveMamba: Task-Centric Scalable State Space Model for Efficient End-to-End Autonomous Driving
- Link: OpenReview
Policy Newton Algorithm in Reproducing Kernel Hilbert Space
- Link: OpenReview
OrthAlign: Orthogonal Subspace Decomposition for Non-Interfering Multi-Objective Alignment
- Link: OpenReview
Theory-Grounded Evaluation of Human-Like Fallacy Patterns in LLM Reasoning
- Link: OpenReview
Multi-Subspace Multi-Modal Modeling for Diffusion Models: Estimation, Convergence and Mixture of Experts
- Link: OpenReview
SpaceControl: Introducing Test-Time Spatial Control to 3D Generative Modeling
- Link: OpenReview
Tractability via Low Dimensionality: The Parameterized Complexity of Training Quantized Neural Networks
- Link: OpenReview
Why Less is More (Sometimes): A Theory of Data Curation
- Link: OpenReview