← All conferences

ICML 2025 Accepted Papers

The full list of 3,333 papers accepted at ICML 2025 (International Conference on Machine Learning). Click any title for details, similar papers, and links to the original source. You can also search these papers by meaning, not just keywords.

Poster: 2,988Spotlight: 225Oral: 120
  1. TMetaNet: Topological Meta-Learning Framework for Dynamic Link PredictionPoster
  2. TOPLOC: A Locality Sensitive Hashing Scheme for Trustless Verifiable InferencePoster
  3. TRACE Back from the Future: A Probabilistic Reasoning Approach to Controllable Language GenerationPoster
  4. TRUST-VLM: Thorough Red-Teaming for Uncovering Safety Threats in Vision-Language ModelsPoster
  5. TS-SNN: Temporal Shift Module for Spiking Neural NetworksPoster
  6. TSP: A Two-Sided Smoothed Primal-Dual Method for Nonconvex Bilevel OptimizationPoster
  7. TTFSFormer: A TTFS-based Lossless Conversion of Spiking TransformerPoster
  8. TUMTraf VideoQA: Dataset and Benchmark for Unified Spatio-Temporal Video Understanding in Traffic ScenesPoster
  9. TabFSBench: Tabular Benchmark for Feature Shifts in Open EnvironmentsPoster
  10. TabNAT: A Continuous-Discrete Joint Generative Framework for Tabular DataPoster
  11. TabSDS: a Lightweight, Fully Non-Parametric, and Model Free Approach for Generating Synthetic Tabular DataPoster
  12. Tackling Dimensional Collapse toward Comprehensive Universal Domain AdaptationPoster
  13. Tackling View-Dependent Semantics in 3D Language Gaussian SplattingPoster
  14. Taming Diffusion for Dataset Distillation with High RepresentativenessPoster
  15. Taming Knowledge Conflicts in Language ModelsSpotlight
  16. Targeted Low-rank Refinement: Enhancing Sparse Language Models with PrecisionPoster
  17. Targeted Unlearning with Single Layer Unlearning GradientPoster
  18. Targeted control of fast prototyping through domain-specific interfacePoster
  19. Task Generalization with Autoregressive Compositional Structure: Can Learning from $D$ Tasks Generalize to $D^T$ Tasks?Poster
  20. Task-Agnostic Pre-training and Task-Guided Fine-tuning for Versatile Diffusion PlannerPoster
  21. Task-Aware Virtual Training: Enhancing Generalization in Meta-Reinforcement Learning for Out-of-Distribution TasksPoster
  22. Task-Gated Multi-Expert Collaboration Network for Degraded Multi-Modal Image FusionPoster
  23. TeDS: Joint Learning of Diachronic and Synchronic Perspectives in Quaternion Space for Temporal Knowledge Graph CompletionPoster
  24. TeLoGraF: Temporal Logic Planning via Graph-encoded Flow MatchingPoster
  25. Teaching Physical Awareness to LLMs through SoundsPoster
  26. Telling Peer Direct Effects from Indirect Effects in Observational Network DataPoster
  27. Temporal Difference FlowsOral
  28. Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement LearningPoster
  29. Temporal Misalignment in ANN-SNN Conversion and its Mitigation via Probabilistic Spiking NeuronsPoster
  30. Temporal Query Network for Efficient Multivariate Time Series ForecastingPoster
  31. Tensor Decomposition Based Memory-Efficient Incremental LearningPoster
  32. Tensor Product Neural Networks for Functional ANOVA ModelPoster
  33. Tensor-Var: Efficient Four-Dimensional Variational Data AssimilationPoster
  34. Tensorized Multi-View Multi-Label Classification via Laplace Tensor RankPoster
  35. Test-Time Adaptation for Online Vision-Language Navigation with Feedback-based Reinforcement LearningPoster
  36. Test-Time Adaptation with Binary FeedbackPoster
  37. Test-Time Canonicalization by Foundation Models for Robust PerceptionPoster
  38. Test-Time Graph Neural Dataset Search With Generative ProjectionPoster
  39. Test-Time Learning for Large Language ModelsPoster
  40. Test-Time Multimodal Backdoor Detection by Contrastive PromptingPoster
  41. Test-Time Selective Adaptation for Uni-Modal Distribution Shift in Multi-Modal DataPoster
  42. Test-time Adaptation on Graphs via Adaptive Subgraph-based Selection and Regularized PrototypesPoster
  43. Test-time Adapted Reinforcement Learning with Action Entropy RegularizationPoster
  44. Test-time Correlation AlignmentPoster
  45. Testing Conditional Mean Independence Using Generative Neural NetworksPoster
  46. Testing the Limits of Fine-Tuning for Improving Visual Cognition in Vision Language ModelsPoster
  47. Text-to-LoRA: Instant Transformer AdaptionPoster
  48. TextCenGen: Attention-Guided Text-Centric Background Adaptation for Text-to-Image GenerationPoster
  49. Textural or Textual: How Vision-Language Models Read Text in ImagesPoster
  50. The Batch Complexity of Bandit Pure ExplorationPoster
  51. The Berkeley Function Calling Leaderboard (BFCL): From Tool Use to Agentic Evaluation of Large Language ModelsPoster
  52. The Butterfly Effect: Neural Network Training Trajectories Are Highly Sensitive to Initial ConditionsPoster
  53. The Canary’s Echo: Auditing Privacy Risks of LLM-Generated Synthetic TextPoster
  54. The Case for Learned Provenance-based System Behavior BaselinePoster
  55. The Complexity of Learning Sparse Superposed Features with FeedbackPoster
  56. The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement LearningPoster
  57. The Diffusion DualityPoster
  58. The Elicitation Game: Evaluating Capability Elicitation TechniquesPoster
  59. The Emperor's New Clothes in Benchmarking? A Rigorous Examination of Mitigation Strategies for LLM Benchmark Data ContaminationPoster
  60. The Empirical Mean is Minimax Optimal for Local Glivenko-CantelliPoster
  61. The Energy Loss Phenomenon in RLHF: A New Perspective on Mitigating Reward HackingPoster
  62. The Four Color Theorem for Cell Instance SegmentationPoster
  63. The Generalized Skew Spectrum of GraphsPoster
  64. The Geometry of Refusal in Large Language Models: Concept Cones and Representational IndependencePoster
  65. The Global Convergence Time of Stochastic Gradient Descent in Non-Convex Landscapes: Sharp Estimates via Large DeviationsPoster
  66. The Harder Path: Last Iterate Convergence for Uncoupled Learning in Zero-Sum Games with Bandit FeedbackPoster
  67. The Hidden Dimensions of LLM Alignment: A Multi-Dimensional Analysis of Orthogonal Safety DirectionsPoster
  68. The Hidden Joules: Evaluating the Energy Consumption of Vision Backbones for Progress Towards More Efficient Model InferencePoster
  69. The Hidden Life of Tokens: Reducing Hallucination of Large Vision-Language Models Via Visual Information SteeringPoster
  70. The Illusion of Role Separation: Hidden Shortcuts in LLM Role Learning (and How to Fix Them)Poster
  71. The Impact of On-Policy Parallelized Data Collection on Deep Reinforcement Learning NetworksPoster
  72. The Importance of Being Lazy: Scaling Limits of Continual LearningPoster
  73. The Jailbreak Tax: How Useful are Your Jailbreak Outputs?Spotlight
  74. The Limits of Predicting Agents from BehaviourPoster
  75. The Limits of Tractable MarginalizationPoster
  76. The Lock-in Hypothesis: Stagnation by AlgorithmPoster
  77. The Logical Implication Steering Method for Conditional Interventions on Transformer GenerationPoster
  78. The Missing Alignment Link of In-context Learning on SequencesPoster
  79. The Noisy Laplacian: a Threshold Phenomenon for Non-Linear Dimension ReductionPoster
  80. The Number of Trials Matters in Infinite-Horizon General-Utility Markov Decision ProcessesSpotlight
  81. The Panaceas for Improving Low-Rank Decomposition in Communication-Efficient Federated LearningPoster
  82. The Polynomial Stein Discrepancy for Assessing Moment ConvergencePoster
  83. The Power of Random Features and the Limits of Distribution-Free Gradient DescentPoster
  84. The Price of Freedom: Exploring Expressivity and Runtime Tradeoffs in Equivariant Tensor ProductsPoster
  85. The Price of Linear Time: Error Analysis of Structured Kernel InterpolationPoster
  86. The Relationship Between No-Regret Learning and Online Conformal PredictionPoster
  87. The Ripple Effect: On Unforeseen Complications of Backdoor AttacksPoster
  88. The Role of Sparsity for Length Generalization in LLMsPoster
  89. The Sample Complexity of Online Strategic Decision Making with Information Asymmetry and Knowledge TransportabilityPoster
  90. The Sparse-Plus-Low-Rank Quasi-Newton Method for Entropic-Regularized Optimal TransportPoster
  91. The Surprising Agreement Between Convex Optimization Theory and Learning-Rate Scheduling for Large Model TrainingPoster
  92. The Synergy of LLMs & RL Unlocks Offline Learning of Generalizable Language-Conditioned Policies with Low-fidelity DataSpotlight
  93. The Underlying Universal Statistical Structure of Natural DatasetsPoster
  94. The Value of Prediction in Identifying the Worst-OffOral
  95. The impact of uncertainty on regularized learning in gamesPoster
  96. The underlying structures of self-attention: symmetry, directionality, and emergent dynamics in Transformer trainingPoster
  97. Theoretical Limitations of Ensembles in the Age of OverparameterizationOral
  98. Theoretical Performance Guarantees for Partial Domain Adaptation via Partial Optimal TransportPoster
  99. Theoretically Unmasking Inference Attacks Against LDP-Protected Clients in Federated Vision ModelsPoster
  100. Thickness-aware E(3)-Equivariant 3D Mesh Neural NetworksPoster
  101. Think Twice, Act Once: A Co-Evolution Framework of LLM and RL for Large-Scale Decision MakingPoster
  102. Three-Dimensional Trajectory Prediction with 3DMoTraj DatasetPoster
  103. Tight and Fast Bounds for Multi-Label LearningPoster
  104. Tightening Causal Bounds via Covariate-Aware Optimal TransportPoster
  105. Tilted Sharpness-Aware MinimizationPoster
  106. Time Series Representations with Hard-Coded InvariancesPoster
  107. Time to Spike? Understanding the Representational Power of Spiking Neural Networks in Discrete TimePoster
  108. Time-Aware World Model for Adaptive Prediction and ControlPoster
  109. TimeBase: The Power of Minimalism in Efficient Long-term Time Series ForecastingSpotlight
  110. TimeDART: A Diffusion Autoregressive Transformer for Self-Supervised Time Series RepresentationPoster
  111. TimePoint: Accelerated Time Series Alignment via Self-Supervised Keypoint and Descriptor LearningPoster
  112. TimePro: Efficient Multivariate Long-term Time Series Forecasting with Variable- and Time-Aware Hyper-statePoster
  113. TimeStacker: A Novel Framework with Multilevel Observation for Capturing Nonstationary Patterns in Time Series ForecastingPoster
  114. TimeStep Master: Asymmetrical Mixture of Timestep LoRA Experts for Versatile and Efficient Diffusion Models in VisionPoster
  115. TinyMIG: Transferring Generalization from Vision Foundation Models to Single-Domain Medical ImagingPoster
  116. To Each Metric Its Decoding: Post-Hoc Optimal Decision Rules of Probabilistic Hierarchical ClassifiersPoster
  117. To Steer or Not to Steer? Mechanistic Error Reduction with Abstention for Language ModelsPoster
  118. ToMA: Token Merge with Attention for Diffusion ModelsPoster
  119. Token Cleaning: Fine-Grained Data Selection for LLM Supervised Fine-TuningPoster
  120. Token Coordinated Prompt Attention is Needed for Visual PromptingPoster
  121. Token Signature: Predicting Chain-of-Thought Gains with Token Decoding Feature in Large Language ModelsPoster
  122. TokenSwift: Lossless Acceleration of Ultra Long Sequence GenerationPoster
  123. Tokenized Bandit for LLM Decoding and AlignmentPoster
  124. TopInG: Topologically Interpretable Graph Learning via Persistent Rationale FiltrationPoster
  125. Topological Signatures of Adversaries in Multimodal AlignmentsPoster
  126. Topology-Aware Dynamic Reweighting for Distribution Shifts on GraphPoster
  127. Topology-aware Neural Flux Prediction Guided by PhysicsPoster
  128. Toward Data-centric Directed Graph Learning: An Entropy-driven ApproachPoster
  129. Towards Better-than-2 Approximation for Constrained Correlation ClusteringSpotlight
  130. Towards Efficient Online Tuning of VLM Agents via Counterfactual Soft Reinforcement LearningPoster
  131. Towards Escaping from Class Dependency Modeling for Multi-Dimensional ClassificationPoster
  132. Towards Global-level Mechanistic Interpretability: A Perspective of Modular Circuits of Large Language ModelsPoster
  133. Towards Graph Foundation Models: Learning Generalities Across Graphs via Task-TreesPoster
  134. Towards Learning to Complete Anything in LidarPoster
  135. Towards Lifelong Model Editing via Simulating Ideal EditorPoster
  136. Towards Memorization Estimation: Fast, Formal and FreePoster
  137. Towards Practical Defect-Focused Automated Code ReviewSpotlight
  138. Towards Rationale-Answer Alignment of LVLMs via Self-Rationale CalibrationPoster
  139. Towards Robust Influence Functions with Flat Validation MinimaPoster
  140. Towards Robustness and Explainability of Automatic Algorithm SelectionSpotlight
  141. Towards Theoretical Understanding of Sequential Decision Making with Preference FeedbackPoster
  142. Towards Trustworthy Federated Learning with Untrusted ParticipantsPoster
  143. Towards Understanding Catastrophic Forgetting in Two-layer Convolutional Neural NetworksPoster
  144. Towards Understanding Fine-Tuning Mechanisms of LLMs via Circuit AnalysisPoster
  145. Towards Understanding Gradient Dynamics of the Sliced-Wasserstein Distance via Critical Point AnalysisPoster
  146. Towards Understanding Parametric Generalized Category Discovery on GraphsPoster
  147. Towards Universal Offline Black-Box Optimization via Learning Language Model EmbeddingsPoster
  148. Towards a Formal Theory of Representational CompositionalityPoster
  149. Towards a General Time Series Forecasting Model with Unified Representation and Adaptive TransferPoster
  150. Towards a Mechanistic Explanation of Diffusion Model GeneralizationSpotlight
  151. Towards a Unified Framework of Clustering-based Anomaly DetectionPoster
  152. Towards an Explainable Comparison and Alignment of Feature EmbeddingsPoster
  153. Towards the Causal Complete Cause of Multi-Modal Representation LearningPoster
  154. Towards the Efficient Inference by Incorporating Automated Computational Phenotypes under Covariate ShiftPoster
  155. TraceGrad: a Framework Learning Expressive SO(3)-equivariant Non-linear Representations for Electronic-Structure Hamiltonian PredictionPoster
  156. Tracking Most Significant Shifts in Infinite-Armed BanditsPoster
  157. Tracking The Best Expert PrivatelyPoster
  158. Tractable Transformers for Flexible Conditional GenerationPoster
  159. Training Diffusion-based Generative Models with Limited DataPoster
  160. Training Flexible Models of Genetic Variant Effects from Functional Annotations using Accelerated Linear AlgebraPoster
  161. Training High Performance Spiking Neural Network by Temporal Model CalibrationPoster
  162. Training a Generally Curious AgentOral
  163. Trajectory World Models for Heterogeneous EnvironmentsPoster
  164. TransPL: VQ-Code Transition Matrices for Pseudo-Labeling of Time Series Unsupervised Domain AdaptationPoster
  165. Transfer Learning for Nonparametric Contextual Dynamic PricingPoster
  166. Transfer Q-Learning with Composite MDP StructuresPoster
  167. Transformative or Conservative? Conservation laws for ResNets and TransformersOral
  168. Transformer-Based Spatial-Temporal Counterfactual Outcomes EstimationPoster
  169. Tree-Sliced Wasserstein Distance with Nonlinear ProjectionPoster
  170. Tree-Sliced Wasserstein Distance: A Geometric PerspectivePoster
  171. TreeLoRA: Efficient Continual Learning via Layer-Wise LoRAs Guided by a Hierarchical Gradient-Similarity TreePoster
  172. Triple-Optimistic Learning for Stochastic Contextual Bandits with General ConstraintsPoster
  173. Trust-Region Twisted Policy ImprovementPoster
  174. Trusted Multi-View Classification with Expert Knowledge ConstraintsSpotlight
  175. Trustworthy Machine Learning through Data-Specific IndistinguishabilityPoster
  176. TruthFlow: Truthful LLM Generation via Representation Flow CorrectionPoster
  177. TtBA: Two-third Bridge Approach for Decision-Based Adversarial AttackPoster
  178. TuCo: Measuring the Contribution of Fine-Tuning to Individual Responses of LLMsPoster
  179. Tuning Sequential Monte Carlo Samplers via Greedy Incremental Divergence MinimizationPoster
  180. Two Tickets are Better than One: Fair and Accurate Hiring Under Strategic LLM ManipulationsPoster
  181. TypyBench: Evaluating LLM Type Inference for Untyped Python RepositoriesPoster
  182. UDora: A Unified Red Teaming Framework against LLM Agents by Dynamically Hijacking Their Own ReasoningPoster
  183. UI-Vision: A Desktop-centric GUI Benchmark for Visual Perception and InteractionPoster
  184. Ultra Lowrate Image Compression with Semantic Residual Coding and Compression-aware DiffusionPoster
  185. UltraTWD: Optimizing Ultrametric Trees for Tree-Wasserstein DistancePoster
  186. UnHiPPO: Uncertainty-aware Initialization for State Space ModelsPoster
  187. Unbiased Evaluation of Large Language Models from a Causal PerspectivePoster
  188. Unbiased Recommender Learning from Implicit Feedback via Weakly Supervised LearningPoster
  189. UncertainSAM: Fast and Efficient Uncertainty Quantification of the Segment Anything ModelPoster
  190. Uncertainty Estimation for Heterophilic Graphs Through the Lens of Information TheoryPoster
  191. Uncertainty-Based Extensible Codebook for Discrete Federated Learning in Heterogeneous Data SilosPoster
  192. Unconstrained Robust Online Convex OptimizationPoster
  193. Underestimated Privacy Risks for Minority Populations in Large Language Model UnlearningPoster
  194. Understanding Bias Reinforcement in LLM Agents DebatePoster
  195. Understanding Complexity in VideoQA via Visual Program GenerationPoster
  196. Understanding Fixed Predictions via Confined RegionsPoster
  197. Understanding Input Selectivity in Mamba: Impact on Approximation Power, Memorization, and Associative Recall CapacityPoster
  198. Understanding Model Reprogramming for CLIP via Decoupling Visual PromptsPoster
  199. Understanding Nonlinear Implicit Bias via Region Counts in Input SpacePoster
  200. Understanding Overadaptation in Supervised Fine-Tuning: The Role of Ensemble MethodsPoster
  201. Understanding Sharpness Dynamics in NN Training with a Minimalist Example: The Effects of Dataset Difficulty, Depth, Stochasticity, and MorePoster
  202. Understanding Synthetic Context Extension via Retrieval HeadsPoster
  203. Understanding and Improving Length Generalization in Recurrent ModelsPoster
  204. Understanding and Mitigating Memorization in Diffusion Models for Tabular DataPoster
  205. Understanding and Mitigating Memorization in Generative Models via Sharpness of Probability LandscapesSpotlight
  206. Understanding the Forgetting of (Replay-based) Continual Learning via Feature Learning: Angle MattersPoster
  207. Understanding the Kronecker Matrix-Vector Complexity of Linear AlgebraPoster
  208. Understanding the Logic of Direct Preference Alignment through LogicPoster
  209. Understanding the Statistical Accuracy-Communication Trade-off in Personalized Federated Learning with Minimax GuaranteesPoster
  210. Understanding the Unfairness in Network QuantizationPoster
  211. Understanding the difficulties of posterior predictive estimationPoster
  212. UniDB: A Unified Diffusion Bridge Framework via Stochastic Optimal ControlSpotlight
  213. UniMC: Taming Diffusion Transformer for Unified Keypoint-Guided Multi-Class Image GenerationPoster
  214. UniMate: A Unified Model for Mechanical Metamaterial Generation, Property Prediction, and Condition ConfirmationPoster
  215. UniMoMo: Unified Generative Modeling of 3D Molecules for De Novo Binder DesignPoster
  216. UniSim: A Unified Simulator for Time-Coarsened Dynamics of BiomoleculesPoster
  217. Unifews: You Need Fewer Operations for Efficient Graph Neural NetworksPoster
  218. Unified Analysis of Continuous Weak Features Learning with Applications to Learning from Missing DataPoster
  219. Unified K-Means Clustering with Label-Guided Manifold LearningPoster
  220. Unified Screening for Multiple DiseasesPoster
  221. Uniform Mean Estimation for Heavy-Tailed Distributions via Median-of-MeansPoster
  222. Unifying Knowledge from Diverse Datasets to Enhance Spatial-Temporal Modeling: A Granularity-Adaptive Geographical Embedding ApproachPoster
  223. Unisolver: PDE-Conditional Transformers Towards Universal Neural PDE SolversPoster
  224. Unisoma: A Unified Transformer-based Solver for Multi-Solid SystemsPoster
  225. Universal Approximation of Mean-Field Models via TransformersPoster
  226. Universal Biological Sequence Reranking for Improved De Novo Peptide SequencingPoster
  227. Universal Neural Optimal TransportPoster
  228. Unlocking Post-hoc Dataset Inference with Synthetic DataPoster
  229. Unlocking the Capabilities of Large Vision-Language Models for Generalizable and Explainable Deepfake DetectionPoster
  230. Unlocking the Power of Rehearsal in Continual Learning: A Theoretical PerspectivePoster
  231. Unlocking the Power of SAM 2 for Few-Shot SegmentationPoster
  232. Unnatural Languages Are Not Bugs but Features for LLMsPoster
  233. Unpaired Point Cloud Completion via Unbalanced Optimal TransportPoster
  234. Unraveling the Interplay between Carryover Effects and Reward Autocorrelations in Switchback ExperimentsPoster
  235. Unsupervised Learning for Class Distribution MismatchPoster
  236. Unveiling Markov heads in Pretrained Language Models for Offline Reinforcement LearningPoster
  237. Upcycling Text-to-Image Diffusion Models for Multi-Task CapabilitiesPoster
  238. Update Your Transformer to the Latest Release: Re-Basin of Task VectorsPoster
  239. VCT: Training Consistency Models with Variational Noise CouplingPoster
  240. VIP: Vision Instructed Pre-training for Robotic ManipulationPoster
  241. VTGaussian-SLAM: RGBD SLAM for Large Scale Scenes with Splatting View-Tied 3D GaussiansPoster
  242. Validating Mechanistic Interpretations: An Axiomatic ApproachPoster
  243. Value-Based Deep RL Scales PredictablyPoster
  244. Variance as a Catalyst: Efficient and Transferable Semantic Erasure Adversarial Attack for Customized Diffusion ModelsPoster
  245. Variance-Reduced Forward-Reflected-Backward Splitting Methods for Nonmonotone Generalized EquationsPoster
  246. Variational Control for Guidance in Diffusion ModelsPoster
  247. Variational Counterfactual Intervention Planning to Achieve Target OutcomesPoster
  248. Variational Learning of Fractional PosteriorsPoster
  249. Variational Phylogenetic Inference with Products over BipartitionsPoster
  250. Vector Grimoire: Codebook-based Shape Generation under Raster Image SupervisionPoster
  251. VerbalTS: Generating Time Series from TextsPoster
  252. Verification Learning: Make Unsupervised Neuro-Symbolic System FeasiblePoster
  253. Video-Enhanced Offline Reinforcement Learning: A Model-Based ApproachPoster
  254. VinePPO: Refining Credit Assignment in RL Training of LLMsPoster
  255. Vision Graph Prompting via Semantic Low-Rank DecompositionPoster
  256. Vision-Language Model Selection and Reuse for Downstream AdaptationPoster
  257. Vision-Language Models Create Cross-Modal Task RepresentationsPoster
  258. Visual Abstraction: A Plug-and-Play Approach for Text-Visual RetrievalPoster
  259. Visual Attention Never Fades: Selective Progressive Attention ReCalibration for Detailed Image Captioning in Multimodal Large Language ModelsPoster
  260. Visual Generation Without GuidancePoster
  261. Visual Graph Arena: Evaluating Visual Conceptualization of Vision and Multimodal Large Language ModelsPoster
  262. Visual and Domain Knowledge for Professional-level Graph-of-Thought Medical ReasoningSpotlight
  263. Volume Optimality in Conformal Prediction with Structured Prediction SetsPoster
  264. Volume-Aware Distance for Robust Similarity LearningPoster
  265. Voronoi-grid-based Pareto Front Learning and Its Application to Collaborative Federated LearningPoster
  266. Vulnerability-Aware Alignment: Mitigating Uneven Forgetting in Harmful Fine-TuningPoster
  267. WATCH: Adaptive Monitoring for AI Deployments via Weighted-Conformal MartingalesPoster
  268. WAVE: Weighted Autoregressive Varying Gate for Time Series ForecastingPoster
  269. WGFormer: An SE(3)-Transformer Driven by Wasserstein Gradient Flows for Molecular Ground-State Conformation PredictionPoster
  270. WILTing Trees: Interpreting the Distance Between MPNN EmbeddingsPoster
  271. WMarkGPT: Watermarked Image Understanding via Multimodal Large Language ModelsPoster
  272. WOMD-Reasoning: A Large-Scale Dataset for Interaction Reasoning in DrivingPoster
  273. Wait-Less Offline Tuning and Re-solving for Online Decision MakingPoster
  274. Wasserstein Policy OptimizationPoster
  275. Watch Out Your Album! On the Inadvertent Privacy Memorization in Multi-Modal Large Language ModelsPoster
  276. WeGeFT: Weight‑Generative Fine‑Tuning for Multi‑Faceted Efficient Adaptation of Large ModelsPoster
  277. Weak-to-Strong Generalization Even in Random Feature Networks, ProvablyPoster
  278. Weak-to-Strong Jailbreaking on Large Language ModelsPoster
  279. Weakly Supervised Anomaly Detection via Dual-Tailed KernelPoster
  280. Weakly-Supervised Contrastive Learning for Imprecise Class LabelsSpotlight
  281. Weight matrices compression based on PDB model in deep neural networksPoster
  282. Weisfeiler and Leman Go Gambling: Why Expressive Lottery Tickets WinPoster
  283. What Do Learning Dynamics Reveal About Generalization in LLM Mathematical Reasoning?Poster
  284. What Has a Foundation Model Found? Inductive Bias Reveals World ModelsPoster
  285. What Limits Bidirectional Model's Generative Capabilities? A Uni-Bi-Directional Mixture-of-Expert Method For Bidirectional Fine-tuningPoster
  286. What Limits Virtual Agent Application? OmniBench: A Scalable Multi-Dimensional Benchmark for Essential Virtual Agent CapabilitiesOral
  287. What Makes In-context Learning Effective for Mathematical ReasoningPoster
  288. What Makes a Good Feedforward Computational Graph?Poster
  289. What can large language models do for sustainable food?Poster
  290. What makes an Ensemble (Un) Interpretable?Poster
  291. When Bad Data Leads to Good ModelsPoster
  292. When Can Proxies Improve the Sample Complexity of Preference Learning?Poster
  293. When Data-Free Knowledge Distillation Meets Non-Transferable Teacher: Escaping Out-of-Distribution Trap is All You NeedPoster
  294. When Diffusion Models Memorize: Inductive Biases in Probability Flow of Minimum-Norm Shallow Neural NetsPoster
  295. When Do LLMs Help With Node Classification? A Comprehensive AnalysisPoster
  296. When Dynamic Data Selection Meets Data Augmentation: Achieving Enhanced Training AccelerationPoster
  297. When Every Millisecond Counts: Real-Time Anomaly Detection via the Multimodal Asynchronous Hybrid NetworkSpotlight
  298. When Maximum Entropy Misleads Policy OptimizationPoster
  299. When Model Knowledge meets Diffusion Model: Diffusion-assisted Data-free Image Synthesis with Alignment of Domain and ClassPoster
  300. When Will It Fail?: Anomaly to Prompt for Forecasting Future Anomalies in Time SeriesPoster
  301. When and How Does CLIP Enable Domain and Compositional Generalization?Spotlight
  302. When can in-context learning generalize out of task distribution?Poster
  303. When do neural networks learn world models?Poster
  304. When to Forget? Complexity Trade-offs in Machine UnlearningPoster
  305. When to retrain a machine learning modelPoster
  306. When, Where and Why to Average Weights?Poster
  307. Whitened CLIP as a Likelihood Surrogate of Images and CaptionsPoster
  308. Whoever Started the interference Should End It: Guiding Data-Free Model Merging via Task VectorsPoster
  309. Widening the Network Mitigates the Impact of Data Heterogeneity on FedAvgPoster
  310. WikiBigEdit: Understanding the Limits of Lifelong Knowledge Editing in LLMsPoster
  311. WildChat-50M: A Deep Dive Into the Role of Synthetic Data in Post-TrainingPoster
  312. Winner-takes-all for Multivariate Probabilistic Time Series ForecastingPoster
  313. Wolfpack Adversarial Attack for Robust Multi-Agent Reinforcement LearningPoster
  314. World Model Implanting for Test-time Adaptation of Embodied AgentsPoster
  315. Wrapped Gaussian on the manifold of Symmetric Positive Definite MatricesPoster
  316. WyckoffDiff -- A Generative Diffusion Model for Crystal SymmetryPoster
  317. X-Transfer Attacks: Towards Super Transferable Adversarial Attacks on CLIPPoster
  318. XAttention: Block Sparse Attention with Antidiagonal ScoringPoster
  319. You Always Recognize Me (YARM): Robust Texture Synthesis Against Multi-View CorruptionPoster
  320. Zero Shot Generalization of Vision-Based RL Without Data AugmentationPoster
  321. Zero-Shot Adaptation of Parameter-Efficient Fine-Tuning in Diffusion ModelsPoster
  322. Zero-Shot Cyclic Peptide Design via Composable Geometric ConstraintsPoster
  323. Zero-Shot Offline Imitation Learning via Optimal TransportPoster
  324. ZipAR: Parallel Autoregressive Image Generation through Spatial LocalityPoster
  325. am-ELO: A Stable Framework for Arena-based LLM EvaluationSpotlight
  326. any4: Learned 4-bit Numeric Representation for LLMsPoster
  327. e-GAI: e-value-based Generalized $\alpha$-Investing for Online False Discovery Rate ControlPoster
  328. iDPA: Instance Decoupled Prompt Attention for Incremental Medical Object DetectionPoster
  329. iN2V: Bringing Transductive Node Embeddings to Inductive GraphsPoster
  330. polybasic Speculative Decoding Through a Theoretical PerspectivePoster
  331. scSSL-Bench: Benchmarking Self-Supervised Learning for Single-Cell DataSpotlight
  332. sciLaMA: A Single-Cell Representation Learning Framework to Leverage Prior Knowledge from Large Language ModelsPoster
  333. unMORE: Unsupervised Multi-Object Segmentation via Center-Boundary ReasoningPoster

Looking for submission deadlines instead? See the conference deadline calendar.