← All conferences

ICLR 2026 Accepted Papers

The full list of 5,356 papers accepted at ICLR 2026 (International Conference on Learning Representations). Click any title for details, similar papers, and links to the original source. You can also search these papers by meaning, not just keywords.

Poster: 5,118Oral: 223ICLR 2026 ConditionalPoster: 14ICLR 2026 ConditionalOral: 1
  1. UniCon: Unified Framework for Efficient Contrastive Alignment via KernelsPoster
  2. UniEdit-Flow: Unleashing Inversion and Editing in the Era of Flow ModelsPoster
  3. UniF$^2$ace: A $\underline{Uni}$fied $\underline{F}$ine-grained $\underline{Face}$ Understanding and Generation ModelPoster
  4. UniFlow: A Unified Pixel Flow Tokenizer for Visual Understanding and GenerationPoster
  5. UniHM: Unified Dexterous Hand Manipulation with Vision Language ModelPoster
  6. UniHand: A Unified Model for Diverse Controlled 4D Hand Motion ModelingPoster
  7. UniLiP: Adapting CLIP for Unified Multimodal Understanding, Generation and EditingPoster
  8. UniOD: A Universal Model for Outlier Detection across Diverse DomainsPoster
  9. UniQL: Unified Quantization and Low-rank Compression for Adaptive Edge LLMsPoster
  10. UniRestorer: Universal Image Restoration via Adaptively Estimating Image Degradation at Proper GranularityPoster
  11. UniSS: Unified Expressive Speech-to-Speech Translation with Your VoicePoster
  12. UniSplat: Unified Spatio-Temporal Fusion via 3D Latent Scaffolds for Dynamic Driving Scene ReconstructionPoster
  13. UniTrack: Differentiable Graph Representation Learning for Multi-Object TrackingPoster
  14. UniUGG: Unified 3D Understanding and Generation via Geometric-Semantic EncodingPoster
  15. Unified 3D Scene Understanding Through Physical World ModelingPoster
  16. Unified Analyses for Hierarchical Federated Learning: Topology Selection under Data HeterogeneityPoster
  17. Unified Biomolecular Trajectory Generation via Pretrained Variational BridgePoster
  18. Unified Brain Surface and Volume RegistrationPoster
  19. Unified Diffusion VLA: Vision-Language-Action Model via Joint Discrete Diffusion Diffusion ProcessPoster
  20. Unified In-Context Video EditingPoster
  21. Unified Multi-Modal Interactive and Reactive 3D Motion Generation via Rectified FlowPoster
  22. Unified Privacy Guarantees for Decentralized Learning via Matrix FactorizationPoster
  23. Unified Vision-Language-Action ModelPoster
  24. Unified Vision–Language Modeling via Concept Space AlignmentPoster
  25. Unified and Efficient Multi-view Clustering from Probabilistic PerspectivePoster
  26. Uniform Discrete Diffusion with Metric Path for Video GenerationPoster
  27. Unifying Diffusion and Autoregression for Generalizable Vision-Language-Action ModelPoster
  28. Unifying Formal Explanations: A Complexity-Theoretic PerspectivePoster
  29. Unifying Stable Optimization and Reference Regularization in RLHFPoster
  30. Universal Beta SplattingPoster
  31. Universal Inverse Distillation for Matching Models with Real-Data Supervision (No GANs)Oral
  32. Universal Model Routing for Efficient LLM InferencePoster
  33. Universal Multi-Domain Translation via Diffusion RoutersPoster
  34. Universal Properties of Activation Sparsity in Modern Large Language ModelsPoster
  35. Universal Value-Function UncertaintiesPoster
  36. Unlearning Evaluation through Subset Statistical IndependencePoster
  37. Unlearning Isn't Invisible: Detecting Unlearning Traces in LLMs from Model OutputsPoster
  38. Unlearning during Training: Domain-Specific Gradient Ascent for Domain GeneralizationPoster
  39. Unleashing Guidance Without Classifiers for Human-Object Interaction AnimationPoster
  40. Unleashing LLMs in Bayesian Optimization: Preference-Guided Framework for Scientific DiscoveryPoster
  41. Unleashing Perception-Time Scaling to Multimodal Reasoning ModelsPoster
  42. Unleashing Scientific Reasoning for Bio-experimental Protocol Generation via Structured Component-based Reward MechanismPoster
  43. Unlocking Full Efficiency of Token Filtering in Large Language Model TrainingPoster
  44. Unlocking Long-Horizon Agentic Search with Large-Scale End-to-End RLPoster
  45. Unlocking the Essence of Beauty: Advanced Aesthetic Reasoning with Relative-Absolute Policy OptimizationPoster
  46. Unlocking the Potential of Weighting Methods in Federated Learning Through Communication CompressionPoster
  47. Unlocking the Power of Co-Occurrence in CLIP: A DualPrompt-Driven Method for Training-Free Zero-Shot Multi-Label ClassificationPoster
  48. Unlocking the Power of Multi-Agent LLM for Reasoning: From Lazy Agents to DeliberationPoster
  49. Unlocking the Value of Text: Event-Driven Reasoning and Multi-Level Alignment for Time Series ForecastingPoster
  50. Unmasking Backdoors: An Explainable Defense via Gradient-Attention Anomaly Scoring for Pre-trained Language ModelsPoster
  51. Unmute the Patch Tokens: Rethinking Probing in Multi-Label Audio ClassificationPoster
  52. Unpacking Human Preference for LLMs: Demographically Aware Evaluation with the HUMAINE FrameworkPoster
  53. Unraveling the Complexity of Memory in RL Agents: an Approach for Classification and EvaluationPoster
  54. Unsupervised Learning of Efficient Exploration: Pre-training Adaptive Policies via Self-Imposed GoalsPoster
  55. Unsupervised Representation Learning - an Invariant Risk Minimization PerspectivePoster
  56. Unsupervised Representation Learning for 3D Mesh Parameterization with Semantic and Visibility ObjectivesPoster
  57. Untraceable DeepFakes via Traceable Fingerprint EliminationPoster
  58. Unveiling Downstream Performance Scaling of LLMs: A Clustering-Based PerspectivePoster
  59. Unveiling Perceptual Artifacts: A Fine-Grained Benchmark for Interpretable AI-Generated Image DetectionPoster
  60. Unveiling Super Experts in Mixture-of-Experts Large Language ModelsPoster
  61. Unveiling the Cognitive Compass: Theory-of-Mind–Guided Multimodal Emotion ReasoningPoster
  62. Unveiling the Mechanism of Continuous Representation Full-Waveform Inversion: A Wave Based Neural Tangent Kernel FrameworkPoster
  63. Unveiling the Potential of Diffusion Large Language Model in Controllable GenerationPoster
  64. Urban Socio-Semantic Segmentation with Vision-Language ReasoningICLR 2026 ConditionalPoster
  65. UrbanFeel:A Comprehensive Benchmark for Temporal and Perceptual Understanding of City Scenes through Human PerspectivePoster
  66. UrbanGS: Efficient and Scalable Architecture for Geometrically Accurate Large-Scene ReconstructionPoster
  67. UrbanGraph: Physics-Informed Spatio-Temporal Dynamic Heterogeneous Graphs for Urban Microclimate PredictionPoster
  68. UrbanVerse: Scaling Urban Simulation by Watching City-Tour VideosPoster
  69. Use the Online Network If You Can: Towards Fast and Stable Reinforcement LearningPoster
  70. Using Reinforcement Learning to Train Large Language Models to Explain Human DecisionsPoster
  71. Using cognitive models to reveal value trade-offs in language modelsPoster
  72. Using maximal information auxiliary variables to improve synthetic data generation based on TabPFN foundation modelsPoster
  73. V2P-Bench: Evaluating Video-Language Understanding with Visual Prompts for Better Human-Model InteractionPoster
  74. VADv2: End-to-End Autonomous Driving via Probabilistic PlanningPoster
  75. VARestorer: One-Step VAR Distillation for Real-World Image Super-ResolutionPoster
  76. VCWorld: A Biological World Model for Virtual Cell SimulationPoster
  77. VEAttack: Downstream-agnostic Vision Encoder Attack against Large Vision Language ModelsPoster
  78. VER: Vision Expert Transformer for Robot Learning via Foundation Distillation and Dynamic RoutingPoster
  79. VERIFY: A Novel Multi-Domain Dataset Grounding LTL in Contextual Natural Language via Provable Intermediate LogicPoster
  80. VERINA: Benchmarking Verifiable Code GenerationPoster
  81. VFScale: Intrinsic Reasoning through Verifier-Free Test-time Scalable Diffusion ModelPoster
  82. VGR: Visual Grounded ReasoningPoster
  83. VINCIE: Unlocking In-context Image Editing from VideoPoster
  84. VIRTUE: Visual-Interactive Text-Image Universal EmbedderPoster
  85. VITA: Vision-to-Action Flow Matching PolicyPoster
  86. VITA: Zero-Shot Value Functions via Test-Time Adaptation of Vision–Language ModelsPoster
  87. VL-JEPA: Joint Embedding Predictive Architecture for Vision-languagePoster
  88. VLBiMan: Vision-Language Anchored One-Shot Demonstration Enables Generalizable Bimanual Robotic ManipulationPoster
  89. VLM-Guided Adaptive Negative Prompting for Creative GenerationPoster
  90. VLM-SubtleBench: How Far Are VLMs from Human-Level Subtle Comparative Reasoning?Poster
  91. VLM4VLA: Revisiting Vision-Language-Models in Vision-Language-Action ModelsPoster
  92. VLMgineer: Vision-Language Models as Robotic ToolsmithsPoster
  93. VLSU: Mapping the Limits of Joint Multimodal Understanding for AI SafetyPoster
  94. VMDiff: Visual Mixing Diffusion for Limitless Cross-Object SynthesisPoster
  95. VMoBA: Mixture-of-Block Attention for Video Diffusion ModelsPoster
  96. VOGUE: Unified Understanding, Generation, and Editing for VideosPoster
  97. VPI-Bench: Visual Prompt Injection Attacks for Computer-Use AgentsPoster
  98. VQ-Transplant: Efficient VQ-Module Integration for Pre-trained Visual TokenizersPoster
  99. VSF: Simple, Efficient, and Effective Negative Guidance in Few-Step Image Generation Models By Value Sign FlipPoster
  100. VTool-R1: VLMs Learn to Think with Images via Reinforcement Learning on Multimodal Tool UsePoster
  101. VUDG: A Dataset for Video Understanding Domain GeneralizationPoster
  102. Value FlowsPoster
  103. Value Gradient Flow: Behavior-Regularized RL without RegularizationPoster
  104. Value Matching: Scalable and Gradient-Free Reward-Guided Flow AdaptationPoster
  105. Variance-Dependent Regret Lower Bounds for Contextual BanditsPoster
  106. Variation in Verification: Understanding Verification Dynamics in Large Language ModelsPoster
  107. Variation-aware Flexible 3D Gaussian EditingPoster
  108. Variational Autoencoding Discrete Diffusion with Enhanced Dimensional Correlations ModelingPoster
  109. Variational Deep Learning via Implicit RegularizationPoster
  110. Variational Inference for Cyclic LearningPoster
  111. Variational Reasoning for Language ModelsPoster
  112. VaseVQA-3D: Benchmarking 3D VLMs on Ancient Greek PotteryPoster
  113. VenusX: Unlocking Fine-Grained Functional Understanding of ProteinsPoster
  114. VeriCoT: Neuro-symbolic Chain-of-Thought Validation via Logical Consistency ChecksPoster
  115. VeriEquivBench: An Equivalence Score for Ground-Truth-Free Evaluation of Formally Verifiable CodePoster
  116. VeriRole: Verifiable Role-Awareness through Hint-Guided Reinforcement LearningPoster
  117. VeriTrail: Closed-Domain Hallucination Detection with TraceabilityPoster
  118. Verification and Co-Alignment via Heterogeneous Consistency for Preference-Aligned LLM AnnotationsPoster
  119. Verification of the Implicit World Model in a Generative Model via Adversarial SequencesPoster
  120. Verifier-free Test-Time Sampling for Vision Language Action ModelsPoster
  121. VerifyBench: Benchmarking Reference-based Reward Systems for Large Language ModelsPoster
  122. Verifying Chain-of-Thought Reasoning via Its Computational GraphOral
  123. Veritas: Generalizable Deepfake Detection via Pattern-Aware ReasoningOral
  124. ViMo: A Generative Visual GUI World Model for App AgentsPoster
  125. ViPER: Empowering the Self-Evolution of Visual Perception Abilities in Vision-Language ModelsPoster
  126. ViPO: Visual Preference Optimization at ScalePoster
  127. ViPRA: Video Prediction for Robot ActionsPoster
  128. ViTSP: A Vision Language Models Guided Framework for Large-Scale Traveling Salesman ProblemsPoster
  129. VibeVoice: Expressive Podcast Generation with Next-Token DiffusionOral
  130. Vid-LLM: A Compact Video-based 3D Multimodal LLM with Reconstruction–Reasoning SynergyOral
  131. Vid2World: Crafting Video Diffusion Models to Interactive World ModelsPoster
  132. VidBridge-R1: Bridging QA and Captioning for RL-based Video Understanding Models with Intermediate Proxy TasksPoster
  133. VidGuard-R1: AI-Generated Video Detection and Explanation via Reasoning MLLMs and RLPoster
  134. Video Scene Segmentation with Genre and Duration SignalsPoster
  135. Video Unlearning via Low-Rank Refusal VectorPoster
  136. Video-As-Prompt: Unified Semantic Control for Video GenerationPoster
  137. Video-GPT via Next Clip DiffusionPoster
  138. Video-KTR: Reinforcing Video Reasoning via Key Token AttributionPoster
  139. Video-LevelGauge: Investigating Contextual Positional Bias in Video Language Models.Poster
  140. Video-STAR: Reinforcing Open-Vocabulary Action Recognition with ToolsPoster
  141. VideoAgentTrek: Computer-Use Pretraining from Unlabeled VideosPoster
  142. VideoAnchor: Reinforcing Subspace-Structured Visual Cues for Coherent Visual-Spatial ReasoningPoster
  143. VideoChat-Flash: Hierarchical Compression for Long-Context Video ModelingPoster
  144. VideoJudge: Bootstrapping Enables Scalable Supervision of MLLM-as-a-Judge for Video UnderstandingPoster
  145. VideoMathQA: Benchmarking Mathematical Reasoning via Multimodal Understanding in VideoPoster
  146. VideoMind: A Chain-of-LoRA Agent for Temporal-Grounded Video ReasoningPoster
  147. VideoNSA: Native Sparse Attention Scales Video UnderstandingPoster
  148. VideoPhy-2: A Challenging Action-Centric Physical Commonsense Evaluation in Video GenerationPoster
  149. VideoReasonBench: Can MLLMs Perform Vision-Centric Complex Video Reasoning?Poster
  150. VideoZoomer: Reinforcement-Learned Temporal Focusing for Long Video ReasoningPoster
  151. Virne: A Comprehensive Benchmark for RL-based Network Resource Allocation in NFVPoster
  152. Virtual Community: An Open World for Humans, Robots, and SocietyPoster
  153. VisCoder2: Building Multi-Language Visualization Coding AgentsPoster
  154. VisCodex: Unified Multimodal Code Generation via Merging Vision and Coding ModelsPoster
  155. VisJudge-Bench: Aesthetics and Quality Assessment of VisualizationsPoster
  156. VisioMath: Benchmarking Figure-based Mathematical Reasoning in LMMsPoster
  157. Vision Language Models are BiasedPoster
  158. Vision-Language-Action Instruction Tuning: From Understanding to ManipulationPoster
  159. Vision-R1: Incentivizing Reasoning Capability in Multimodal Large Language ModelsPoster
  160. Vision-Zero: Scalable VLM Self-Improvement via Strategic Gamified Self-PlayPoster
  161. VisionLaw: Inferring Interpretable Intrinsic Dynamics from Visual Observations via Bilevel OptimizationPoster
  162. VisionReasoner: Unified Reasoning-Integrated Visual Perception via Reinforcement LearningPoster
  163. VisionTrim: Unified Vision Token Compression for Training-Free MLLM AccelerationPoster
  164. VisuLogic: A Benchmark for Evaluating Visual Reasoning in Multi-modal Large Language ModelsPoster
  165. VisuRiddles: Fine-grained Perception is a Primary Bottleneck for Multimodal Large Language Models in Abstract Visual ReasoningPoster
  166. Visual Autoregressive Modeling for Instruction-Guided Image EditingPoster
  167. Visual Jigsaw Post-Training Improves MLLMsPoster
  168. Visual Multi-Agent System: Mitigating Hallucination Snowballing via Visual FlowPoster
  169. Visual Planning: Let's Think Only with ImagesOral
  170. Visual Prompt-Agnostic EvolutionPoster
  171. Visual Self-Refine: A Pixel-Guided Paradigm for Accurate Chart ParsingPoster
  172. Visual symbolic mechanisms: Emergent symbol processing in Vision Language ModelsOral
  173. VisualPRM400K: An Effective Dataset for Training Multimodal Process Reward ModelsPoster
  174. VisualPrompter: Semantic-Aware Prompt Optimization with Visual Feedback for Text-to-Image SynthesisPoster
  175. VitaBench: Benchmarking LLM Agents with Versatile Interactive Tasks in Real-world ApplicationsPoster
  176. Vivid-VR: Distilling Concepts from Text-to-Video Diffusion Transformer for Photorealistic Video RestorationPoster
  177. Vlaser: Vision-Language-Action Model with Synergistic Embodied ReasoningPoster
  178. VoG: Enhancing LLM Reasoning through Stepwise Verification on Knowledge GraphsPoster
  179. VoMP: Predicting Volumetric Mechanical Property FieldsPoster
  180. VowelPrompt: Hearing Speech Emotions from Text via Vowel-level Prosodic AugmentationPoster
  181. VoxPrivacy: A Benchmark for Evaluating Interactional Privacy of Speech Language ModelsPoster
  182. Vulcan: Crafting Compact Class-Specific Vision Transformers For Edge IntelligencePoster
  183. W-EDIT: A Wavelet-Based Frequency-Aware Framework for Text-Driven Image EditingPoster
  184. WAFT: Warping-Alone Field Transforms for Optical FlowOral
  185. WALT: Web Agents that Learn ToolsPoster
  186. WARC-Bench: Web Archive based Benchmark for GUI Subtask ExecutionsPoster
  187. WARP: Weight Teleportation for Attack-Resilient Unlearning ProtocolsPoster
  188. WATS: Wavelet-Aware Temperature Scaling for Reliable Graph Neural NetworksPoster
  189. WAVE: Learning Unified & Versatile Audio-Visual Embeddings with Multimodal LLMOral
  190. WFR-FM: Simulation-Free Dynamic Unbalanced Optimal TransportPoster
  191. WILD-Diffusion: A WDRO Inspired Training Method for Diffusion Models under Limited DataPoster
  192. WIMFRIS: WIndow Mamba Fusion and Parameter Efficient Tuning for Referring Image SegmentationPoster
  193. WIMLE: Uncertainty‑Aware World Models with IMLE for Sample‑Efficient Continuous ControlPoster
  194. WINA: Weight Informed Neuron Activation for Accelerating Large Language Model InferencePoster
  195. WMPO: World Model-based Policy Optimization for Vision-Language-Action ModelsPoster
  196. WOW-Seg: A Word-free Open World Segmentation ModelPoster
  197. WRING Out The Bias: A Rotation-Based Alternative To Projection DebiasingPoster
  198. WSM: Decay-Free Learning Rate Schedule via Checkpoint Merging for LLM Pre-trainingOral
  199. WSVD: Weighted Low-Rank Approximation for Fast and Efficient Execution of Low-Precision Vision-Language ModelsPoster
  200. Watch the Weights: Unsupervised monitoring and control of fine-tuned LLMsPoster
  201. Watch your steps: Dormant Adversarial Behaviors that Activate upon LLM FinetuningICLR 2026 ConditionalOral
  202. WaterDrum: Watermark-based Data-centric Unlearning MetricPoster
  203. Watermark-based Attribution of AI-Generated ContentPoster
  204. Watermarking Diffusion Language ModelsPoster
  205. WavePolyp: Video Polyp Segmentation via Hierarchical Wavelet-Based Feature Aggregation and Inter-Frame Divergence PerceptionPoster
  206. WavefrontDiffusion: Dynamic Decoding Schedule for Improved ReasoningPoster
  207. Wavelet Predictive Representations for Non-Stationary Reinforcement LearningPoster
  208. We-Math 2.0: A Versatile MathBook System for Incentivizing Visual Mathematical ReasoningPoster
  209. WeTok: Powerful Discrete Tokenization for High-Fidelity Visual ReconstructionPoster
  210. Weak Correlations as the Underlying Principle for Linearization of Gradient-Based Learning SystemsPoster
  211. Weak-to-Strong DiffusionPoster
  212. Weak-to-Strong Generalization with Failure TrajectoriesPoster
  213. WearVox: An Egocentric Multichannel Voice Assistant Benchmark for WearablesPoster
  214. Web-CogReasoner: Towards Knowledge-Induced Cognitive Reasoning for Web AgentsPoster
  215. WebArbiter: A Generative Reasoning Process Reward Model for Web AgentsPoster
  216. WebDS: An End-to-End Benchmark for Web-based Data SciencePoster
  217. WebDevJudge: Evaluating (M)LLMs as Critiques for Web Development QualityOral
  218. WebFactory: Automated Compression of Foundational Language Intelligence into Grounded Web AgentsPoster
  219. WebGen-Agent: Enhancing Interactive Website Generation with Multi-Level Feedback and Step-Level Reinforcement LearningPoster
  220. WebSailor-V2: Bridging the Chasm to Proprietary Agents via Synthetic Data and Scalable Reinforcement LearningPoster
  221. WebSeer: Training Deeper Search Agents through Reinforcement Learning with Self-ReflectionPoster
  222. WebShaper: Agentically Data Synthesizing via Information-Seeking FormalizationPoster
  223. WebWatcher: Breaking New Frontiers of Vision-Language Deep Research AgentPoster
  224. WebWeaver: Structuring Web-Scale Evidence with Dynamic Outlines for Open-Ended Deep ResearchPoster
  225. Webscale-RL: Automated Data Pipeline for Scaling RL Data to Pretraining LevelsPoster
  226. Weight Decay may matter more than µP for Learning Rate Transfer in PracticePoster
  227. Weight Space Representation Learning on Diverse NeRF ArchitecturesPoster
  228. Weight-Space Linear Recurrent Neural NetworksPoster
  229. Welfarist Formulations for Diverse Similarity SearchPoster
  230. What "Not" to Detect: Negation-Aware VLMs via Structured Reasoning and Token MergingPoster
  231. What Do Large Language Models Know About Opinions?Poster
  232. What Exactly Does Guidance Do in Masked Discrete Diffusion ModelsPoster
  233. What Generative Search Engines Like and How to Optimize Web Content CooperativelyPoster
  234. What Happens Next? Anticipating Future Motion by Generating Point TrajectoriesPoster
  235. What Layers When: Learning to Skip Compute in LLMs with Residual GatesPoster
  236. What Matters for Batch Online Reinforcement Learning in Robotics?Poster
  237. What Scales in Cross-Entropy Scaling Law?Poster
  238. What happens when generative AI models train recursively on each others' outputs?Poster
  239. What matters for Representation Alignment: Global Information or Spatial Structure?Poster
  240. What's In My Human Feedback? Learning Interpretable Descriptions of Preference DataOral
  241. What's the plan? Metrics for implicit planning in LLMs and their application to rhyme generationPoster
  242. Whatever Remains Must Be True: Filtering Drives Reasoning in LLMs, Shaping DiversityPoster
  243. When Agents “Misremember” Collectively: Exploring the Mandela Effect in LLM-based Multi-Agent SystemsPoster
  244. When Bias Helps Learning: Bridging Initial Prejudice and TrainabilityPoster
  245. When Data is the Algorithm: A Systematic Study and Curation of Preference Optimization DatasetsPoster
  246. When Does Divide and Conquer Work for Long Context LLM? A Noise Decomposition FrameworkPoster
  247. When Flatness Does (Not) Guarantee Adversarial RobustnessPoster
  248. When Foundation Models are One-Liners: Limitations and Future Directions for Time Series Anomaly DetectionPoster
  249. When Greedy Wins: Emergent Exploitation Bias in Meta-Bandit LLM TrainingPoster
  250. When Is Diversity Rewarded in Cooperative Multi-Agent Learning?Poster
  251. When LLMs get significantly worse: A statistical approach to detect model degradationsPoster
  252. When Language Models Lose Their Mind: The Consequences of Brain MisalignmentPoster
  253. When Large Multimodal Models Confront Evolving Knowledge: Challenges and ExplorationsPoster
  254. When MLLMs Meets Compression Distortion: A Coding Paradigm Tailored to MLLMsPoster
  255. When Machine Learning Gets Personal: Evaluating Prediction and ExplanationPoster
  256. When More is Less: Understanding Chain-of-Thought Length in LLMsPoster
  257. When Priors Backfire: On the Vulnerability of Unlearnable Examples to PretrainingPoster
  258. When Reasoning Meets Compression: Understanding the Effects of LLMs Compression on Large Reasoning ModelsPoster
  259. When Scores Learn Geometry: Rate Separations under the Manifold HypothesisPoster
  260. When Shift Happens - Confounding Is to BlamePoster
  261. When Silence Is Golden: Can LLMs Learn to Abstain in Temporal QA and Beyond?Poster
  262. When Style Breaks Safety: Defending LLMs Against Superficial Style AlignmentPoster
  263. When Thinking Backfires: Mechanistic Insights into Reason-induced MisalignmentPoster
  264. When Weak LLMs Speak with Confidence, Preference Alignment Gets StrongerPoster
  265. When a Robot is More Capable than a Human: Learning from Constrained DemonstratorsPoster
  266. When and Where to Reset Matters for Long-Term Test-Time AdaptationPoster
  267. When to Ensemble: Identifying Token-Level Points for Stable and Fast LLM EnsemblingPoster
  268. When to Retrain after Drift: A Data-Only Test of Post-Drift Data Size SufficiencyPoster
  269. When to use Graphs in RAG: A Comprehensive Analysis for Graph Retrieval-Augmented GenerationPoster
  270. When would Vision-Proprioception Policies Fail in Robotic Manipulation?Poster
  271. Where Did It Go Wrong? Attributing Undesirable LLM Behaviors via Representation Gradient TracingPoster
  272. Where Did This Sentence Come From? Tracing Provenance in LLM Reasoning DistillationPoster
  273. Who Matters Matters: Agent-Specific Conservative Offline MARLPoster
  274. WholeBodyVLA: Towards Unified Latent VLA for Whole-body Loco-manipulation ControlPoster
  275. Why Adversarially Train Diffusion Models?Poster
  276. Why Ask One When You Can Ask $k$? Learning-to-Defer to the Top-$k$ ExpertsPoster
  277. Why Attention Patterns Exist: A Unifying Temporal Perspective AnalysisPoster
  278. Why DPO is a Misspecified Estimator and How to Fix ItOral
  279. Why Do Unlearnable Examples Work: A Novel Perspective of Mutual InformationPoster
  280. Why High-rank Neural Networks Generalize?: An Algebraic Framework with RKHSsPoster
  281. Why Keep Your Doubts to Yourself? Trading Visual Uncertainties in Multi-Agent Bandit SystemsPoster
  282. Why Less is More (Sometimes): A Theory of Data CurationPoster
  283. Why Low-Precision Transformer Training Fails: An Analysis on Flash AttentionOral
  284. Why Prototypes Collapse: Diagnosing and Preventing Partial Collapse in Prototypical Self-Supervised LearningPoster
  285. Why Reinforcement Fine-Tuning Enables MLLMs Preserve Prior Knowledge Better: A Data PerspectivePoster
  286. Why We Need New Benchmarks for Local Intrinsic Dimension EstimationPoster
  287. Why is Your Language Model a Poor Implicit Reward Model?Poster
  288. Wide-In, Narrow-Out: Revokable Decoding for Efficient and Effective DLLMsPoster
  289. WideSearch: Benchmarking Agentic Broad Info-SeekingPoster
  290. Wiki-R1: Incentivizing Multimodal Reasoning for Knowledge-based VQA via Data and Sampling CurriculumPoster
  291. Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmasPoster
  292. WinT3R: Window-Based Streaming Reconstruction with Camera Token PoolPoster
  293. Winter Soldier: Backdooring Language Models at Pre-Training with Indirect Data PoisoningPoster
  294. WithAnyone: Toward Controllable and ID Consistent Image GenerationICLR 2026 ConditionalPoster
  295. WoW!: World Models in a Closed-Loop WorldOral
  296. World2Minecraft: Occupancy-Driven simulated scenes ConstructionPoster
  297. WorldEdit: Towards Open-World Image Editing with a Knowledge-Informed BenchmarkPoster
  298. WorldGym: World Model as An Environment for Policy EvaluationPoster
  299. WorldSense: Evaluating Real-world Omnimodal Understanding for Multimodal LLMsPoster
  300. WorldSplat: Gaussian-Centric Feed-Forward 4D Scene Generation for Autonomous DrivingPoster
  301. WorldTree: Towards 4D Dynamic Worlds from Monocular Video using Tree-ChainsPoster
  302. X-VLA: Soft-Prompted Transformer as Scalable Cross-Embodiment Vision-Language-Action ModelPoster
  303. XIL: Cross-Expanding Incremental LearningPoster
  304. XModBench: Benchmarking Cross-Modal Capabilities and Consistency in Omni-Language ModelsPoster
  305. XQC: Well-conditioned Optimization Accelerates Deep Reinforcement LearningPoster
  306. YoNoSplat: You Only Need One Model for Feedforward 3D Gaussian SplattingPoster
  307. You Point, I Learn: Online Adaptation of Interactive Segmentation Models for Handling Distribution Shifts in Medical ImagingPoster
  308. Your Agent May Misevolve: Emergent Risks in Self-evolving LLM AgentsPoster
  309. Your Language Model Secretly Contains Personality SubnetworksPoster
  310. Your Models Have Thought Enough: Training Large Reasoning Models to Stop OverthinkingPoster
  311. Your VAR Model is Secretly an Efficient and Explainable Generative ClassifierPoster
  312. Youtu-GraphRAG: Vertically Unified Agents for Graph Retrieval-Augmented Complex ReasoningPoster
  313. YuE: Scaling Open Foundation Models for Long-Form Music GenerationPoster
  314. ZIP-RC: Zero-overhead Inference-time Prediction of Reward and Cost for Adaptive and Interpretable GenerationPoster
  315. Zebra-CoT: A Dataset for Interleaved Vision-Language ReasoningPoster
  316. Zephyrus: An Agentic Framework for Weather SciencePoster
  317. Zero-Sacrifice Persistent-Robustness Adversarial Defense for Pre-Trained EncodersPoster
  318. Zero-Shot Adaptation of Behavioral Foundation Models to Unseen DynamicsPoster
  319. Zero-shot Forecasting by Simulation AlonePoster
  320. Zero-shot HOI Detection with MLLM-based Detector-agnostic Interaction RecognitionPoster
  321. Zero-shot Human Pose Estimation using Diffusion-based Inverse solversPoster
  322. ZeroGR: A Generalizable and Scalable Framework for Zero-Shot Generative RetrievalPoster
  323. ZeroSiam: An Efficient Siamese for Test-Time Entropy Optimization without CollapsePoster
  324. ZeroTuning: Unlocking the Initial Token's Power to Enhance Large Language Models Without TrainingPoster
  325. Zeros can be Informative: Masked Binary U-Net for Image Segmentation on Tensor CoresPoster
  326. ``Noisier'’ Noise Contrastive Estimation is (Almost) Maximum LikelihoodPoster
  327. cadrille: Multi-modal CAD Reconstruction with Reinforcement LearningOral
  328. d$^2$Cache: Accelerating Diffusion-Based LLMs via Dual Adaptive CachingPoster
  329. dParallel: Learnable Parallel Decoding for dLLMsPoster
  330. e3: Learning to Explore Enables Extrapolation of Test-Time Compute for LLMsPoster
  331. f-INE: A Hypothesis Testing Framework for Estimating Influence under Training RandomnessPoster
  332. floq: Training Critics via Flow-Matching for Scaling Compute in Value-Based RLPoster
  333. gLSTM: Mitigating Over-Squashing by Increasing Storage CapacityPoster
  334. gen2seg: Generative Models Enable Generalizable Instance SegmentationPoster
  335. h-MINT: Modeling Pocket-Ligand Binding with Hierarchical Molecular Interaction NetworkPoster
  336. iFusion: Integrating Dynamic Interest Streams via Diffusion Model for Click-Through Rate PredictionPoster
  337. iLLaVA: An Image is Worth Fewer Than 1/3 Input Tokens in Large Multimodal ModelsPoster
  338. jqBench: a benchmark for reading and editing JSON from natural language and/or examplesPoster
  339. lmgame-Bench: How Good are LLMs at Playing Games?Poster
  340. mCLM: A Modular Chemical Language Model that Generates Functional and Makeable MoleculesOral
  341. mR3: Multilingual Rubric-Agnostic Reward Reasoning ModelsPoster
  342. pFedMMA: Personalized Federated Fine-Tuning with Multi-Modal Adapter for Vision-Language ModelsPoster
  343. pi-Flow: Policy-Based Few-Step Generation via Imitation DistillationPoster
  344. pySpatial: Generating 3D Visual Programs for Zero-Shot Spatial ReasoningPoster
  345. reAR: Rethinking Visual Autoregressive Models via Token-wise Consistency RegularizationPoster
  346. scDFM: Distributional Flow Matching Model for Robust Single-Cell Perturbation PredictionPoster
  347. sleep2vec: Unified Cross-Modal Alignment for Heterogeneous Nocturnal BiosignalsPoster
  348. ssToken: Self-modulated and Semantic-aware Token Selection for LLM Fine-tuningPoster
  349. station2radar: query‑conditioned gaussian splatting for precipitation fieldPoster
  350. t-SNE Exaggerates Clusters, ProvablyPoster
  351. vAttention: Verified Sparse Attention via SamplingPoster
  352. vCache: Verified Semantic Prompt CachingPoster
  353. villa-X: Enhancing Latent Action Modeling in Vision-Language-Action ModelsPoster
  354. wd1: Weighted Policy Optimization for Reasoning in Diffusion Language ModelsPoster
  355. xLSTM Scaling Laws: Competitive Performance with Linear Time-ComplexityPoster
  356. xRFM: Accurate, scalable, and interpretable feature learning models for tabular dataPoster

Looking for submission deadlines instead? See the conference deadline calendar.