ICML 2026 Accepted Papers
The full list of 6,634 papers accepted at ICML 2026 (International Conference on Machine Learning). Click any title for details, similar papers, and links to the original source. You can also search these papers by meaning, not just keywords.
Poster: 6,060Spotlight: 406Oral: 168
- Robust-U1: Can MLLMs Self-Recover Corrupted Visual Content for Robust Understanding?Poster
- Robustifying Vision-Language Models via Test-Time Prompt AdaptationPoster
- Robustness of Mixtures of Experts to Feature NoisePoster
- Role-Level Inductive Bias for Cross-Task Generalization in Multi-Agent Reinforcement LearningPoster
- Romberg-Extrapolated Zeroth-Order Gradient Estimator: Higher-Order Bias Reduction with Preserved Leading Directional VariancePoster
- Root Cause Analysis of Failures in Microservices via Bayesian Root Cause DiscoverySpotlight
- Rotary Position Encodings for GraphsSpotlight
- Rotation-Invariant Spherical Watermarking via Third-Order SO(3) Representation CouplingPoster
- RouteFinder: Towards Foundation Models for Vehicle Routing ProblemsPoster
- RouterInterp: Understanding Superposed Specialisation in Mixture of Experts RoutingPoster
- Routing and Reasoned Evaluation with Large Language ModelsPoster
- Row-stochastic matrices can provably outperform doubly stochastic matrices in decentralized learningPoster
- RuCL: Stratified Rubric-Based Curriculum Learning for Multimodal Large Language Model ReasoningPoster
- Rubric Curriculum RL: Exploiting the Generation-Verification Gap in Creative WritingPoster
- RubricRobustness: A Simple Framework for Evaluating the Robustness of Rubrics-Based BenchmarksPoster
- Rule2DRC: Benchmarking LLM Agents for DRC Script Synthesis with Execution-Guided Test GenerationPoster
- RulePlanner: All-in-One Reinforcement Learner for Unifying Design Rules in 3D FloorplanningPoster
- Rényi Diffusion ModelsPoster
- S$^3$GNN: Efficient Global Mixing and Local Message Passing for Long-Range Graph LearningSpotlight
- S-Quant: Rethinking Weight Quantization with Seed-Based GenerationPoster
- S2GS: Streaming Semantic Gaussian Splatting for Online Scene Understanding and ReconstructionPoster
- S3Audio: Towards Streaming Synchronized Spatial Audio Generation via Autoregressive Diffusion TransformerPoster
- SABER: Continual Learning with Representation Conflict ManagementPoster
- SAD-Flower: Flow Matching for Safe, Admissible, and Dynamically Consistent PlanningPoster
- SAEmnesia: Erasing Concepts in Diffusion Models with Supervised Sparse AutoencodersPoster
- SAEs-BrainMap: Unveiling the Emergence of Specialized Concepts in Deep Models via Brain AlignmentPoster
- SAGE-NAS: Synergizing LLM-Based Semantic Agent with Graph-Based Evaluator for Neural Architecture SearchPoster
- SAGE: A Dataflow-Native Framework for Modular, Controllable, and Transparent LLM-Augmented ReasoningPoster
- SAGE: Shaping Anchors for Guided Exploration in RLVR of LLMsPoster
- SALAAD: Sparse And Low-Rank Adaptation via ADMM for Large Language Model InferencePoster
- SALE : Low-bit Estimation for Efficient Sparse Attention in Long-context LLM PrefillingPoster
- SALSA-V: Shortcut-Augmented Long-form Synchronized Audio from VideosPoster
- SAM Audio: Segment Anything in AudioPoster
- SAME: Stabilized Mixture-of-Experts for Multimodal Continual Instruction TuningPoster
- SAMT: Generating Structured Avatar Meshes and Textures from a Single ImagePoster
- SAOT: Self-Supervised Continual Graph Learning with Structure-Aware Optimal TransportPoster
- SAQNN: Spectral Adaptive Quantum Neural Network as a Universal ApproximatorPoster
- SARL: Structure-Aligned Reinforcement Learning for Bridging the Perception-Action Gap in AirspacePoster
- SARSteer: Safeguarding Large Audio Language Models via Safe-Ablated Refusal SteeringPoster
- SC$^{2}$-WM: A Self-Correcting World Model with Closed-Loop Feedback for Vision-and-Language Navigation in Continuous EnvironmentsPoster
- SC-FAGC: Size Constrained Fast Anchor-based Graph ClusteringPoster
- SCALE: Self-uncertainty Conditioned Adaptive Looking and Execution for Vision-Language-Action ModelsSpotlight
- SCNS: Continual Personalization of Diffusion Models via Submodular Concept Neuron SelectionPoster
- SCOPE and SCION: Benchmark and Method for Ontology Induction and Fusion from TextPoster
- SCOPE: Evolving Symbolic World for Planning in Open-Ended EnvironmentsPoster
- SCOPE: Selective Conformal Optimized Pairwise LLM JudgingPoster
- SCORE: A Unified Framework for Overshoot Refund in Online FDR ControlPoster
- SCOUT: Active Information Foraging for Long-Text Understanding with Decoupled Epistemic StatesPoster
- SCOUT: Cyclic Causal Discovery Under Soft Interventions with Unknown TargetsPoster
- SCRWKV: Ultra-Compact Structure-Calibrated Vision-RWKV for Topological Crack SegmentationPoster
- SCalDA: Semantics-Calibrated and Diffusion-Enhanced Data AugmentationPoster
- SCoA: Revisiting Domain Generalized Object Detection with Style-Conditioned AdaptationPoster
- SD-MoE: Spectral Decomposition for Effective Expert SpecializationPoster
- SDiD:Shared diffusion prior for efficient distributed stereo image compressionPoster
- SE(3)-Equivariant Flow Matching with Gaussian Process Priors for Geometric Trajectory PredictionPoster
- SE(n)-Invariant Flow Matching: A General Framework with Application to Object ReassemblyPoster
- SE-GA: Memory-Augmented Self-Evolution for GUI AgentsPoster
- SE3Set: Harnessing Equivariant Hypergraph Neural Networks for Molecular Representation LearningPoster
- SEAgent: Self-Evolving Computer Use Agent with Autonomous Learning from ExperiencePoster
- SEDRAS: Symbolically Evaluated Deep Research And SciencePoster
- SEER: Transformer-based Robust Time Series Forecasting via Automated Patch Enhancement and ReplacementPoster
- SEM-CTRL: Semantically Controlled DecodingPoster
- SEMA: a Scalable and Efficient Mamba like Attention via Token Localization and AveragingPoster
- SEMIR: Semantic Minor-Induced Representation Learning on Graphs for Visual SegmentationPoster
- SENDAI: A Hierarchical Sparse-measurement, EfficieNt Data AssImilation FrameworkPoster
- SEPS: Semantic-Enhanced Patch Slimming Framework for Fine-Grained Cross-Modal AlignmentPoster
- SERA: Soft-Verified Efficient Repository AgentsPoster
- SF-Mamba: Rethinking State Space Model for VisionPoster
- SFCLTA: Spectral Fusion Contrastive Learning with Topology-Adaptive Graph AugmentationPoster
- SFedPO: Streaming Federated Learning with a Prediction Oracle under Temporal ShiftsPoster
- SG2Loc: Sequential Visual Localization on 3D Scene GraphsPoster
- SGERA: Stein-Guided ECG-Report Alignment for ECG Representation LearningPoster
- SGMD: Score Gradient Matching Distillation for Few-Step Video Diffusion DistillationPoster
- SHAP-Guided Kernel Actor-Critic for Explainable Reinforcement LearningPoster
- SHARP-Q: Spectral Hessian Alignment and Rectification for Post-training QuantizationPoster
- SHERPA: Fine-tuning Segment Anything Models with Task-relevant GuidancePoster
- SHINE: A Scalable In-Context Hypernetwork for Mapping Context to LoRA in a Single PassPoster
- SI-IGCL: Subject Invariance-aware Inverse Graph Contrastive Learning for Psychiatric Disorder IdentificationPoster
- SIGMA-PPG: Statistical-prior Informed Generative Masking Architecture for PPG Foundation ModelPoster
- SIKA-GP: Accelerating Gaussian Process Inference with Sparse Inducing Kernel Approximations for Bayesian Deep LearningPoster
- SIMPC: Learning Self-Induced Mirror-Point Consistency for Unsupervised Point Cloud DenoisingPoster
- SIMoE: A Probabilistic Framework for Cardinality-Constrained Routing in Mixture-of-ExpertsPoster
- SINQ: Sinkhorn-Normalized Quantization for Calibration-Free Low-Precision LLM WeightsPoster
- SIPO: Stabilized and Improved Preference Optimization for Aligning Diffusion ModelsPoster
- SJD-SV: Speculative Jacobi Decoding with Semantics Verification for Autoregressive Image GenerationPoster
- SKETCH: Semantic Key-Point Conditioning for Long-Horizon Vessel Trajectory PredictionPoster
- SL-VC: A Benchmark and Automated Framework for Separation Logic Verification Condition ProvingPoster
- SLAE: Strictly Local All-atom Environment for Protein RepresentationPoster
- SLAP: The Semantic Least Action Principle for Variational Video-Language ModelingPoster
- SLAT: Segment-Level Adaptive Trimming for Efficient CoT ReasoningPoster
- SLIM: Secure and Efficient Inference for Large Language Models on Untrusted Devices via TEEsPoster
- SLIP-RS: Structured-Attribute Language-Image Pre-Training for Remote Sensing Object DetectionPoster
- SLQ: Bridging Modalities via Shared Latent Queries for Retrieval with Frozen MLLMsPoster
- SMAC: Score-Matched Actor-Critics for Robust Offline-to-Online TransferPoster
- SMART: Scalable Mesh‑free Aerodynamic Simulations from Raw Geometries using a Transformer‑based Surrogate ModelPoster
- SMD: Multi-view Safety-Critical Driving Video Generation in the Real-world DomainPoster
- SMILE: Extended Deep Submodular Function-Based Instruction and In-context Learning Demonstration SelectionPoster
- SMM Transformer: Leveraging Spiking Neural Networks for Multimodal TasksPoster
- SOLAR for Offline MARL: Plateau-Triggered Potential Shaping under World-Model UncertaintyPoster
- SOLAR: Self-supervised Joint Learning for Symmetric Multimodal RetrievalPoster
- SONAR: Spectral‑Contrastive Audio Residuals for Generalizable Deepfake DetectionPoster
- SOPE: Situation-Aware and Statistically Indistinguishable Privacy Exfiltration for MCP-enabled AgentsPoster
- SORA: Free Second Order Attacks in Fast Adversarial TrainingPoster
- SOTAlign: Semi-Supervised Alignment of Unimodal Vision and Language Models via Optimal TransportPoster
- SP-Mind: An Autonomous Reasoning Agent for Spatial Proteomics AnalysisPoster
- SPA: A Simple but Tough-to-Beat Baseline for Knowledge InjectionPoster
- SPADA: A Verifiable Test-Driven Agent for Controllable Parametric CAD Assembly GenerationPoster
- SPAR: Support-Preserving Action RectificationPoster
- SPARC: Separating Perception And Reasoning Circuits for Test-time Scaling of VLMsPoster
- SPARD: Defending Harmful Fine-Tuning Attack via Safety Projection with Relevance–Diversity Data SelectionPoster
- SPARKLING: Balancing Signal Preservation and Symmetry Breaking for Width-Progressive LearningPoster
- SPARe: Stacked Parallelism with Adaptive Reordering for Fault-Tolerant LLM Pretraining Systems with 100k+ GPUsPoster
- SPATIA: Multimodal Generation and Prediction of Spatial Cell PhenotypesPoster
- SPEAR: A Unified SSL Framework for Learning Speech and Audio RepresentationsPoster
- SPEED-Bench: A Unified and Diverse Benchmark for Speculative DecodingPoster
- SPEED: Sharpened-Teacher Distillation for Parallel Decoding of Diffusion Language ModelsPoster
- SPHERE: Mitigating the Loss of Spectral Plasticity in Mixture-of-Experts for Deep Reinforcement LearningPoster
- SPLIT-VLM: Salience-Guided Partitioning towards Local Coverage for Importance-Aware Token Dropping in Vision-Language ModelsPoster
- SPR-RAFT: Parameter-Efficient Regression-Aware Fine-Tuning for Biomedical LLM RegressionPoster
- SPR: A Structured Prompt Refinement Network for Modality MissingPoster
- SPUR: Scale-Partitioned Uncertainty Rectification for Robust UAV-on-UAV InterceptionPoster
- SRPO: Self-Reflective Policy Optimization for Long-Horizon ReasoningPoster
- SSA: Sparse Sparse Attention by Aligning Full and Sparse Attention Outputs in Feature SpacePoster
- SSDCN: Spatial-Spectral Dual-Clustering-based Network for Hyperspectral Image Super-resolutionPoster
- SSL4RL: Revisiting Self-supervised Learning as Intrinsic Reward for Visual-Language ReasoningPoster
- SSR-Merge: Subspace Signal Routing for Training-Free LoRA Merging in Diffusion ModelsPoster
- SS‑TPT: Stability and Suitability-Guided Test-Time Prompt Tuning for Adversarially Robust Vision-Language ModelsPoster
- ST-TGExplainer: Disentangling Stability and Transition Patterns for Temporal GNN InterpretabilityPoster
- ST-Veto: Spatio-Temporal Token Veto for Diffusion MLLMs via Taylor Prediction and Visual GroundingPoster
- STABLE: Simulation-Ready Tabletop Layout Generation via a Semantics–Physics Dual SystemPoster
- STABLEVAL: Disagreement-Aware and Stable Evaluation of AI SystemsPoster
- STAND: Self-Aware Precondition Induction for Interactive Task LearningPoster
- STAR-KV: Low-Rank KV Cache Compression via Soft Thresholding for Adaptive Rank ControlSpotlight
- STAR-VAE: Structured Topology-Aware Regularization for Audio Reconstruction and GenerationPoster
- STAR: Rethinking MoE Routing as Structure-Aware Subspace LearningPoster
- STARCaster: Spatio-Temporal AutoRegressive Video Diffusion for Identity- and View-Aware Talking PortraitsPoster
- STARE: Step-wise Temporal Alignment and Red-teaming Engine for Multi-modal Toxicity AttackPoster
- STD-Former: Image-Conditioned Texture Dictionary Encoding with Sparse Topological Supervision for Texture RecognitionPoster
- STEP: Warm-Started Visuomotor Policies with Spatiotemporal Consistency PredictionPoster
- STFlow: Data-Coupled Flow Matching for Geometric Trajectory SimulationPoster
- STLA: Spatiotemporal Lookahead Alignment for Post-Training QuantizationPoster
- STORM: Segment, Track, and Object Re-Localization from a Single ImagePoster
- STRIDE: Post-Training LLMs to Reason and Refine Bio-Sequences via Edit TrajectoriesPoster
- STT-LLM: Structural-Temporal Tokenization for Adapting LLMs to Longitudinal Clinical ProfilesPoster
- SURF: Separation via Unsupervised Remixing FlowPoster
- SURGE: Surrogate Gradient Adaptation in Binary Neural NetworksPoster
- SURGE:Unbiased Data Assimilation for Diffusion Model via Particle FilteringPoster
- SVD as a Fast Interpretability Method for TransformersSpotlight
- SVL: Empowering Spiking Neural Networks for Efficient 3D Open-World UnderstandingSpotlight
- SVL: Goal-Conditioned Reinforcement Learning as Survival LearningPoster
- SVRG and Beyond via Posterior CorrectionOral
- SWE-ABS: Adversarial Benchmark Strengthening Exposes Inflated Success Rates on Test-based BenchmarkPoster
- SWE-Bench Pro: Can AI Agents Solve Long-Horizon Software Engineering Tasks?Poster
- SWE-Compass: Towards Unified Evaluation of Agentic Coding Abilities for Large Language ModelsPoster
- SWE-MiniSandbox: Container-Free Reinforcement Learning for Building Software Engineering AgentsPoster
- SWE-Perf: Can Language Models Optimize Code Performance on Real-World Repositories?Poster
- SWE-fficiency: Can Language Models Optimize Real-World Repositories on Real Workloads?Poster
- SWE-rebench V2: Language-Agnostic SWE Task Collection at ScalePoster
- SWING: Unlocking Implicit Graph Representations for Graph Random FeaturesSpotlight
- SaTeen: Learning Structural Alignment for Continual Test-Time AdaptationPoster
- Safe Autoregressive Image Generation with Iterative Self-Improving CodebooksPoster
- Safe In-Context Reinforcement LearningPoster
- Safe Reinforcement Learning with Preference-based Constraint InferencePoster
- Safe and Scalable Web Agent Learning via Recreated WebsitesPoster
- SafeCompass: Dynamic Chain-of-Thought Steering via Inference-Time Safety SignalsPoster
- SafeDec: Constrained Decoding for Safe Autoregressive Generalist Robot Navigation PoliciesPoster
- SafeHarbor: Defining Precise Decision Boundaries via Hierarchical Memory-Augmented Guardrail for LLM Agent SafetyPoster
- SafeLab: An Interactive High-Fidelity Benchmark for Embodied Safety in Scientific RoboticsPoster
- SafeSci: Safety Evaluation of Large Language Models in Science Domains and BeyondPoster
- SafeSearch: Automated Red-Teaming of LLM-Based Search AgentsPoster
- SafeSeek: Universal Attribution of Safety Circuits in Language ModelsPoster
- SafeSpec: Fast and Safe LLM via Dynamic Reflective SamplingPoster
- Safeguarded Stochastic Polyak Step Sizes for Non-smooth Optimization: Robust Performance Without Small (Sub)GradientsPoster
- Safety Alignment of LMs via Non-cooperative GamesSpotlight
- Safety Anchor: Defending Harmful Fine-tuning via Geometric BottlenecksPoster
- Safety Game: Inference-Time Alignment of Black-Box LLMs via Constrained OptimizationPoster
- Safety Generalization Under Distribution Shift in Safe Reinforcement Learning: A Diabetes TestbedPoster
- Safety Recovery in Reasoning Models Is Only a Few Early Steering Steps AwayPoster
- Safety-Efficacy Trade Off: Robustness against Data-PoisoningPoster
- Saliency-Aware Model MergingPoster
- Salus: Strategic Diagnostic Testing for Complex Diagnosis via Multi-Agent Reinforcement LearningPoster
- Same Graph Cross-Task Transfer in GNNs: Protocols and PredictorsPoster
- Same Question, Different Lies: Cross-Context Consistency (C³) for Black-Box Sandbagging DetectionPoster
- Sample Complexity Bounds for Robust Mean Estimation with Mean-Shift ContaminationPoster
- Sample Efficient Full-Finetuning of Generative Control PoliciesPoster
- Sample Margin-Aware Recalibration of Temperature ScalingPoster
- Sample from What You See: Visuomotor Policy Learning via Diffusion Bridge with Observation-Embedded Stochastic Differential EquationPoster
- Sample-Efficient Diffusion-based Reinforcement Learning with Critic GuidancePoster
- Sampled hard labels from sparse targets mislead rotation invariant algorithmsPoster
- Sampling and Identity-Testing Without Approximate Tensorization of EntropyPoster
- Sampling from Your Language Model One Byte at a TimePoster
- Saving Foundation Flow-Matching Priors for Inverse ProblemsPoster
- ScDiVa: Masked Discrete Diffusion for Joint Modeling of Single-Cell Identity and ExpressionPoster
- ScaLoRA: Optimally Scaled Low-Rank Adaptation for Efficient High-Rank Fine-TuningPoster
- Scalable Bayesian Inference for Nonlinear Conservation LawsPoster
- Scalable Bayesian Semi-supervised Clustering with Feature Selection and Adaptive Constraint WeightingPoster
- Scalable Event Cloud Network for Event-based ClassificationOral
- Scalable GANs with TransformersPoster
- Scalable Kronecker-Factored Fisher Approximation for Neural Network Parameter SensitivityPoster
- Scalable Medical Multimodal Fusion via Symmetric Consistency ModelingPoster
- Scalable Option Learning in High-Throughput EnvironmentsSpotlight
- Scalable Power Sampling: Unlocking Efficient, Training-Free Reasoning for LLMs via Distribution SharpeningPoster
- Scalable RF Simulation in Generative 4D WorldsPoster
- Scalable Sampling via Generalized Fixed-Point Diffusion MatchingPoster
- Scalable Simulation-Based Model Inference with Test-Time Complexity ControlPoster
- Scalable Single-Cell Gene Expression Generation with Latent Diffusion ModelsPoster
- Scalable Topology-Preserving Graph Coarsening: Concepts and AlgorithmsPoster
- Scalable Traffic Signal Control with Shared Policy FrameworkPoster
- Scalable Training of 3D Gaussian Splatting via Out-of-Core OptimizationSpotlight
- Scalable and Differentiable Point-Cloud Registration Using Maximum Mean DiscrepancyPoster
- Scalable and General Whole-Body Control for Cross-Humanoid LocomotionPoster
- Scalable and Interpretable Representation Alignment with Ordinal SimilarityPoster
- Scalable and Stable Estimation of Amari $\alpha$-Divergence using Random Fourier FeaturesPoster
- Scale-Aware Domain Harmonization for Domain Adaptation Person SearchPoster
- ScaleEnv: Scaling Environment Synthesis from Scratch for Generalist Interactive Tool-Use Agent TrainingPoster
- ScaleErasure: Inference-Time Minimal Intervention for Precise Concept Erasure in Next-Scale Autoregressive Image GenerationPoster
- ScaleMoE: Mixture-of-Experts for Scalable Continuous Control in Actor-Critic Reinforcement LearningSpotlight
- ScaleSim: Serving Large-Scale Multi-Agent Simulation with Invocation Distance-Based Memory ManagementPoster
- Scaling Agentic Verifier for Competitive CodingPoster
- Scaling Beyond Masked Diffusion Language ModelsPoster
- Scaling Continual Learning with Bi-Level Routing Mixture-of-ExpertsPoster
- Scaling Generative Verifiers For Natural Language Mathematical Proof Verification And SelectionPoster
- Scaling Inference-Time Computation via Opponent Simulation: Enabling Online Strategic Adaptation in Repeated NegotiationPoster
- Scaling Law for Quantization-Aware TrainingSpotlight
- Scaling Laws and Architectural Frontiers in Metagenomic Foundation ModelsPoster
- Scaling Laws for Precision in High-Dimensional Linear RegressionPoster
- Scaling Laws in Model Fine-tuning for Audio DeepFake DetectionPoster
- Scaling Laws of Global Weather ModelsPoster
- Scaling Long-Horizon Agent via Context FoldingPoster
- Scaling Multi-Agent Environment Co-Design with Diffusion ModelsPoster
- Scaling Prompt Synthesis for Large Language Model ReasoningPoster
- Scaling Real-World Robot Policy Evaluation via Discrete Diffusion World ModelSpotlight
- Scaling Small Agents Through Strategy AuctionsPoster
- Scaling Transformers for End-to-End Discrete Audio TokenizationPoster
- Scaling Unsupervised Multi-Source Federated Domain Adaptation through Group-Wise Discrepancy MinimizationPoster
- Scaling Vision Transformers for Functional MRI with Flat MapsPoster
- Scaling by Diversified Experience for Vision-Language-Action ModelsPoster
- Scaling the Prior: Size-Consistent Geometric Diffusion for 3D Molecular GenerationPoster
- Scaling the Scaling Logic: Agentic Meta-Synthesis of Logic ReasoningPoster
- Scaling up Multi-Turn Off-Policy RL and Multi-Agent Tree Search for LLM Step-ProversPoster
- Scaling, Benchmarking, and Reasoning of Vision-Language Agents for Mobile GUI NavigationPoster
- Scaling-Aware Adapter for Structure-Grounded LLM ReasoningPoster
- ScalingAR: Scaling Confidence for Autoregressive Image GenerationPoster
- Scam2Prompt: A Scalable Framework for Auditing Malicious Scam Endpoints in Production LLMsPoster
- Scene Graph Thinking: Reinforcing Structured Visual Reasoning for Multimodal Large Language ModelsPoster
- SceneDirector: Bridging Explicit Geometry and Generative Priors for Unified Driving Scene EditingPoster
- ScenePilot: Controllable Boundary-Driven Critical Scenario Generation for Autonomous DrivingPoster
- SceneSmith: Agentic Generation of Simulation-Ready Indoor ScenesSpotlight
- Scheduling LLM Inference with Uncertainty-Aware Output Length PredictionsPoster
- Scheduling Thoughts: Learning the Order of Thought in Diffusion Language ModelsPoster
- Schema-Guided World Modeling for Understanding Hierarchical Visual DynamicsPoster
- Schur-A*: Layer-wise Optimal Expert Pruning for Sparse MoEs via Schur-Complement Guided A* SearchPoster
- SciAgentGym: Benchmarking Multi-Step Scientific Tool-Use in LLM AgentsPoster
- SciNet: Evaluating AI Agents in Relation-Aware Scientific Literature RetrievalPoster
- SciPredict: Can LLMs Predict the Outcomes of Scientific Experiments in Natural Sciences?Poster
- SciVideoBench: Benchmarking Scientific Video Reasoning in Large Multimodal ModelsPoster
- Scientific logicality enriched methodology for LLM reasoning: A practice in physicsPoster
- Score Based Error Correcting Code DecoderPoster
- Score-Repellent Monte Carlo: Toward Efficient Non-Markovian Sampler with Constant Memory in General State SpacesSpotlight
- ScoreMatchingRiesz: Score Matching for Debiased Machine Learning and Policy Path EstimationPoster
- ScoreMix: Synthetic Data Generation by Score Composition in Diffusion Models Improves Face RecognitionPoster
- Scout Before You Attend: Sketch-and-Walk Sparse Attention for Efficient LLM InferencePoster
- Search Space Synthesis for Parametric FunctionsPoster
- Search for Truth from Reasoning: A Dynamic Representation Editing Framework for Steering LLM TrajectoriesPoster
- Search or Accelerate: Confidence-Switched Position Beam Search for Diffusion Language ModelsPoster
- Search-R2: Enhancing Search-Integrated Reasoning via Actor-Refiner CollaborationPoster
- SecCodePRM: A Process Reward Model for Code SecurityPoster
- Second-Order Bilevel Optimization with Accelerated Convergence RatesPoster
- Second-Order Smooth Planning with Optimal-Transport Bellman SmoothingOral
- Secure Multi-agent Reinforcement Learning for Service Systems with Affinity and Byzantine Nodes: Stability Analysis and Protection DesignPoster
- Securing Multimodal AI through Internal Information DecompositionSpotlight
- Security–Fidelity Tradeoffs: No Universal Defense Against Prompt InjectionSpotlight
- See First, Reason Later: Mutual Information-Guided Reinforcement Learning for Vision-Language ModelsPoster
- See More, Forecast Better and Faster: Enhancing Time Series Foundation Models via Inference-Time Plug-and-Play DownsamplingPoster
- See What Matters: Differentiable Grid Sample Pruning for Generalizable Vision-Language-Action ModelPoster
- See the Emotion: A Facial Emoji Proxy Modeling for EEG Emotion RecognitionPoster
- See, Act, Adapt: Active Perception for Unsupervised Cross-Domain Visual Adaptation via Personalized VLM-Guided AgentPoster
- Seeing Realism from Simulation: Efficient Video Transfer for Vision-Language-Action Data AugmentationPoster
- Seeing Symbols, Missing Structure: A Real-World Handwritten Mathematical Expression Recognition Benchmark for Large ModelsPoster
- Seeing Without Understanding: Disentangling Perception, Reasoning, and Simulation in VLM GameplayPoster
- Seeing is Solving: Unlocking Efficient Multimodal RL via View AlignmentPoster
- Seeing is Understanding: Unlocking Causal Attention into Modality-Mutual Attention for Multimodal LLMsPoster
- Seeing the Unseen: Physics-as-Representation for Generalizable Gaze PerceptionPoster
- Seeing to Generalize: How Visual Data Corrects Binding ShortcutsPoster
- Seeking Commonality, Preserving Specificity: A Spectral-Aware Hierarchical Framework for Cross-City Road Representation LearningPoster
- Seg-ReSearch: Segmentation with Interleaved Reasoning and External SearchPoster
- SegPVSG: Panoptic Video Scene Graph Generation via Temporal Focusing and Generative AugmentationPoster
- Segment Anything with Robust Uncertainty-Accuracy CorrelationPoster
- Segment-Aligned Policy Optimization for Multi-Modal ReasoningPoster
- Segment-driven Structural Induction and Semantic Alignment for Heterogeneous Tabular RepresentationPoster
- Segmentation From Attention: Training-Free Layer Selection and One-Shot Tuning for Segmentation in VLMsPoster
- SeisMark: A Large-Scale Open Benchmark for Robust 3D Seismic Fault DetectionPoster
- Seizure-Semiology-Suite($S^3$): A Clinically Multimodal Dataset, Benchmark, and Models for Seizure Semiology UnderstandingSpotlight
- Select to Think: Unlocking SLM Potential with Local SufficiencyPoster
- Selecting Samples on Graphs: A Unified Dataset Pruning Framework for Lossless Training AccelerationPoster
- Selective Concept Bottleneck Models Without Predefined ConceptsPoster
- Selective Coupling of Decoupled Informative Regions: Masked Attention Alignment for Data-Free Quantization of Vision TransformersPoster
- Selective Deferred Routing: Enabling Cost-Efficient Collaboration between Local SLMs and Remote LLMsPoster
- Selective Disclosure Watermarking for Large Language ModelsPoster
- Self Optimizing Language ModelsPoster
- Self-Augmenting Retrieval for Diffusion Language ModelsPoster
- Self-Calibrated Consistency can Fight Back for Adversarial Robustness in Vision-Language ModelsPoster
- Self-Captioning Multimodal Interaction Tuning: Amplifying Exploitable Redundancies for Robust Vision Language ModelsPoster
- Self-CriTeach: LLM Self-Teaching and Self-Critiquing for Improving Robotic Planning via Automated Domain GenerationPoster
- Self-Distillation Enables Continual LearningSpotlight
- Self-Distilled Reasoner: On-Policy Self-Distillation for Large Language ModelsPoster
- Self-Evolving LLM Agents under Offline Data SupportPoster
- Self-Guidance: Enhancing Neural Codecs via Decoder Manifold AlignmentPoster
- Self-Prompting Diffusion Transformer for Open-Vocabulary Scene Text Edit via In-Context LearningPoster
- Self-Prophetic Decoding to Unlock Visual Search in LVLMsPoster
- Self-Refining Video SamplingPoster
- Self-Soupervision: Cooking Model Soups without LabelsSpotlight
- Self-Supervised Dynamical System Representations for Physiological Time-SeriesPoster
- Self-Supervised Flow Matching for Scalable Multi-Modal SynthesisPoster
- Self-Supervised Foundation Model for Calcium-imaging Population DynamicsPoster
- Self-Supervised Learning as Discrete CommunicationPoster
- Self-Supervised Weight Templates for Scalable Vision Model InitializationPoster
- Self-correcting for Debiasing Large Language ModelsPoster
- Self-supervised Hierarchical Visual Reasoning with World ModelPoster
- SelfJudge: Faster Speculative Decoding via Self-Supervised Judge VerificationPoster
- Selling Data as a Digital Good with Scaling ValuationsPoster
- Sem-Detect: Semantic Level Detection of AI Generated Peer-ReviewsPoster
- SemBind: Binding Diffusion Watermarks to Semantics Against Black-Box Forgery AttacksPoster
- SemRep: Code Transformation with Semantics-Preserving RepresentationsPoster
- Semantic Cache Distillation: Efficient State Transfer via Reuse and Selective PatchingPoster
- Semantic Editing with Coupled Stochastic Differential EquationsPoster
- Semantic Granularity Navigation in Image EditingPoster
- Semantic Impact–Driven Visual Scheduling in Vision-Language ModelsPoster
- Semantic Integrity Matters: Benchmarking and Preserving High-Density Reasoning in KV Cache CompressionPoster
- Semantic Robustness Certification for Vision-Language ModelsPoster
- Semantic Router: On the Feasibility of Hijacking MLLMs via a Single Adversarial PerturbationPoster
- Semantic Tube Prediction: Beating LLM Data Efficiency with JEPAPoster
- Semantic-Aware Motion Encoding for Topology-Agnostic Character AnimationPoster
- Semantic-Enriched Latent Visual ReasoningPoster
- Semantic-level Backdoor Attack against Text-to-Image Diffusion ModelsPoster
- SemanticNVS: Improving Semantic Scene Understanding in Generative Novel View SynthesisPoster
- Semi-LAR: Semi-supervised Contrastive Learning with Linear Attention for Removal of Nighttime FlaresPoster
- Semi-Supervised Gaze Estimation via Disentangled Subspace Contrastive LearningPoster
- Semi-Supervised Learning for Molecular Graphs via Ensemble ConsensusPoster
- Semi-Supervised Learning with Noisy Covariates: Generalization Bounds and Distribution RegressionPoster
- Semi-Supervised Neural Super-Resolution for Mesh-Based SimulationsPoster
- Semi-Supervised Noise Adaptation: Transferring Knowledge from Noise DomainPoster
- Semi-knockoffs: a model-agnostic conditional independence testing method with finite-sample guaranteesPoster
- Sentinel-VLA: A Metacognitive VLA Model with Active Status Monitoring for Dynamic Reasoning and Error RecoveryPoster
- Separating representation from reconstruction enables scalable text encodersPoster
- Sequential Group Composition: A Window into the Mechanics of Deep LearningPoster
- Sequential Kernel-based Conditional Independence Testing via Adaptive BettingPoster
- Server-Proximal Aggregation for Federated Domain-Incremental Learning under Partial Participation: Task-Uniform Convergence and Backward TransferPoster
- Set Diffusion: Interpolating Token Orderings between Autoregression and Diffusion for Fast and Flexible DecodingPoster
- Set-Coupled Guidance: Set-Level Coordination in Diffusion-Based Dataset DistillationPoster
- Set-Preserving Calibration from Conformal P-Values to E-ValuesPoster
- SetPO: Set-Level Policy Optimization for Diversity-Preserving LLM ReasoningPoster
- ShapCCS: Shapley-Driven Client Coreset Selection in Federated LearningPoster
- Shape of Thought: Progressive Object Assembly via Visual Chain-of-ThoughtPoster
- Shapley Neuron Values for Continual Learning: Which Neurons Matter Most?Poster
- Shapley Regularized Neural Granger CausalityPoster
- Shared Semantics, Divergent Mechanisms: Unsupervised Feature Discovery by Aligning Semantics and MechanismsSpotlight
- Sharp Concentration Bounds for Vector Bundle-Valued Statistics on ManifoldsPoster
- Sharp Inequalities between Total Variation and Hellinger Distances for Gaussian MixturesSpotlight
- Sharp description of local minima in the loss landscape of high-dimensional two-layer ReLU neural networksPoster
- Sharp empirical Bernstein inequalities for the variance of bounded random variablesPoster
- Sharper Generalization Guarantees for Asynchronous SGD: Beyond Lipschitzness, Smoothness and Data HomogeneityPoster
- Sharpness-Aware Minimization Can Hallucinate MinimizersPoster
- Sharpness-Aware Pretraining Mitigates Catastrophic ForgettingPoster
- Sheaf Neural Networks on SPD Manifolds: Second-Order Geometric Representation LearningPoster
- Shift-Dependent Asymmetry: Orthogonal Inverse Low-Rank Adaptation for Federated Medical SegmentationPoster
- Shifting the Breaking Point of Flow Matching for Multi-Instance EditingPoster
- Short Chains, Deep Thoughts: Balancing Reasoning Efficiency and Intra-Segment Capability via Split-Merge OptimizationPoster
- Shortcut-Resistant CAM Distillation for Long-Tailed RecognitionPoster
- Should I Have Expressed a Different Intent? Counterfactual Generation for LLM-Based Autonomous ControlPoster
- Show, Don't Tell: Morphing Latent Reasoning into Image GenerationPoster
- Shrinking the Variance: Shrinkage Baselines for Reinforcement Learning with Verifiable RewardsPoster
- Shuffle the Context: RoPE-Perturbed Self-Distillation for Long-Context AdaptationPoster
- Shuffling-Aware Optimization for Private Vector Mean EstimationPoster
- SiameseNorm: Breaking the Barrier to Reconciling Pre/Post-NormPoster
- Sign Lock-In: Randomly Initialized Weight Signs Persist and Bottleneck Sub-Bit Model CompressionPoster
- Signal Strength Estimation in Logistic Regression Using Data SplittingPoster
- Signature-Informed Transformer for Asset AllocationPoster
- SilentWood: Efficient Private Inference Over Gradient Boosting Decision ForestsPoster
- SimGFM: Simplifying Discrete Flow Matching for Graph GenerationPoster
- Similarity Is Not Logic: Factored Inference for Dual-Encoder Vision-Language ModelsPoster
- Simple Algorithms for Bad Triangle Transversals with Applications to Correlation ClusteringSpotlight
- Simple Policy Gradients for Reasoning with Diffusion Language ModelsPoster
- Simple Unbiased Derivative Free Inference-Time Scaling for Diffusion Models via Sequential Monte Carlo on Path MeasuresPoster
- Simple yet Effective: Low-Rank Spatial Attention for Neural OperatorsPoster
- SimpleGPT: Improving GPT via A Simple Normalization StrategyPoster
- SimpleMem: Efficient Lifelong Memory for LLM AgentsPoster
- SimulCost: A Cost-Aware Benchmark and Toolkit for Automating Physics Simulations with LLMsPoster
- Simultaneous Confidence Bounds for Aggregated Effects via Exact Subset OptimizationPoster
- Simultaneous Multi-objective Alignment Across Verifiable and Non-verifiable RewardsPoster
- Simultaneous Speech-to-Speech Translation Without Aligned DataOral
- Single-Head Attention in High Dimensions: A Theory of Generalization, Weights Spectra, and Scaling LawsSpotlight
- Single-Rollout Hidden-State Dynamics for Training-Free RLVR Data SelectionPoster
- Singular Bayesian Neural NetworksPoster
- Singular Proxies for Adaptive Caching in Diffusion Language ModelsPoster
- Singular Vectors of Attention Heads Align with FeaturesPoster
- Singularity-aware Optimization via Randomized Geometric Probing: Towards Stable Non-smooth OptimizationPoster
- Sinkhorn Treatment EffectsPoster
- Size Transferability of Graph Convolutional Networks across Sparsity: A Generalized Graphon PerspectivePoster
- SkelHCC: A Hyperbolic CLIP-Driven Cache Adaptation Framework for Skeleton-based One-Shot Action RecognitionPoster
- Sketch-Based Low-Rank Model Merging with Shared Circulant TransformsPoster
- Skewness-Robust Causal Discovery in Location-Scale Noise ModelsPoster
- Ski Rental with Distributional Predictions of Unknown QualityPoster
- Skill Neologisms: Towards Skill-based Continual LearningSpotlight
- SkillNet: Hierarchical Skill Modeling for Compositional Generalization in Vision-Language Action ModelsPoster
- SkillTrojan: Backdoor Attacks on Skill-Based Agent SystemsPoster
- Skip a Layer or Loop It? Learning Program-of-Layers in LLMsOral
- Skip-It? Theoretical Conditions for Layer Skipping in Vision–Language ModelsPoster
- Skipping the Zeros in Diffusion Models for Sparse Data GenerationPoster
- SlaClip: Gradient Norm Slacks can be Indicator for Adaptive Clipping in DP-SGDSpotlight
- Slash the Sink: Sharpening Structural Attention Inside LLMsPoster
- SleepLM: Natural-Language Intelligence for Human SleepSpotlight
- SleepMaMi: A Universal Sleep Foundation Model for Integrating Macro- and Micro-structuresPoster
- SlerpFlow: Spherical Trajectory Correction for Rectified Flow InversionPoster
- SliceFine: The Universal Winning-Slice Hypothesis for Pretrained NetworksPoster
- SlideSparse: Fast and Flexible (2N-2):2N Structured SparsityPoster
- Small Agent Group is the Future of Digital HealthPoster
- Small Generalizable Prompt Predictive Models Can Steer Efficient RL Post-Training of Large Reasoning ModelsPoster
- Smaller Models are Natural Explorers for Policy-Level Diversity in GRPOPoster
- SmartThinker: Progressive Chain-of-Thought Length Calibration for Efficient Large Language Model ReasoningPoster
- Smooth Dynamic Cutoffs for Machine Learning Interatomic PotentialsPoster
- Smooth Multi-Policy Causal Effect Estimation in Longitudinal SettingsPoster
- SmoothSpike: Spiking Transformer with Learnable Hadamard TransformationSpotlight
- Smoothie: Smoothing Diffusion on Token Embeddings for Text GenerationPoster
- Smoothing Slot Attention Iterations and RecurrencesPoster
- Smoothness Errors in Dynamics Models and How to Avoid ThemPoster
- SoMA: A Real-to-Sim Neural Simulator for Robotic Soft-Body ManipulationPoster
- Sobolev Regularized Score Difference Estimation in Diffusion ModelsPoster
- Social Hippocampus Memory LearningPoster
- SoftBinary Coding: A New Information-Theoretic Paradigm for Neural Compression via Fast Channel SimulationPoster
- SoftJAX & SoftTorch: Empowering Automatic Differentiation Libraries with Informative GradientsOral
- SoftMoE: Soft Differentiable Routing for Mixture-of-Experts in LLMsPoster
- Softmax as Linear Attention in the Large-Prompt Regime: a Measure-based PerspectivePoster
- Softplus Attention with Re-weighting Boosts Length Extrapolation in Large Language ModelsPoster
- Softsignum: Smooth Your Signum For Better Heterogeneity HandlingPoster
- Solver-in-the-Loop: MDP-Based Benchmarks for Self-Correction and Behavioral Rationality in Operations ResearchPoster
- Solving Imperfect-Recall Games via Sum-of-Squares OptimizationPoster
- Solving Inverse Problems with Flow-based Models via Model Predictive ControlPoster
- Solving Physics Olympiad via Reinforcement Learning on Physics SimulatorsPoster
- Solving Positive Linear Programs with Differential PrivacyPoster
- Solving Spatial-Spectral Fusion with Latent Spectral OperatorsPoster
- Solving Stochastic Variational Inequalities without the Bounded Variance AssumptionPoster
- Solving Time-Dependent Differential Equations with Physical Dynamical SystemsOral
- Solving the Offline and Online Min-Max Problem of Non-smooth Submodular-Concave Functions: A Zeroth-Order ApproachPoster
- Sonar-TS: Search-Then-Verify Natural Language Querying for Time Series DatabasesPoster
- SonicMaster: Towards Controllable All-in-One Music Restoration and MasteringPoster
- SorryDB: Can AI Provers Complete Real-World Lean Theorems?Poster
- Source-Free Open-World RF Fingerprint IdentificationPoster
- SpaCeFormer: Space-Curve Transformer for Open-Vocabulary 3D Instance Segmentation without ProposalsPoster
- SpaEF: Spatially Resolved Transcriptomics Data Element-Wise Denoising Framework Powered by Large ModelsPoster
- SpaceVista: All-Scale Visual Spatial Reasoning from mm to kmPoster
- SpanNorm: Reconciling Training Stability and Performance in Deep TransformersPoster
- Sparks of Cooperative Reasoning: LLMs as Strategic Hanabi AgentsPoster
- Sparse ActionGen: Accelerating Diffusion Policy with Real-time PruningPoster
- Sparse Autoencoders are Topic ModelsPoster
- Sparse Autoencoders for Interpretable Emotion Control in Text-to-SpeechPoster
- Sparse Bayesian Deep Functional Learning with Structured Region SelectionPoster
- Sparse Models, Sparse Safety: Unsafe Routes in Mixture-of-Experts LLMsPoster
- Sparse Regression with $\ell_0$ Constraints for $\alpha$-Mixing Time Series: Algorithms and GuaranteesPoster
- Sparse Relaxed-Lasso Steering: Automatic Sparse-Autoencoder Feature Selection for Precise Image EditingPoster
- Sparse Tokens Suffice: Jailbreaking Audio Language Models via Token-Aware Gradient OptimizationPoster
- Sparse Topology-Aware Pairwise Scoring for Large-Scale Multi-Agent Reinforcement LearningPoster
- Sparse and Faithful Local Explanations with Piecewise Linear SurrogatesPoster
- Sparse but Wrong: Incorrect L0 Leads to Incorrect Features in Sparse AutoencodersPoster
- SparseInfer: Accelerating Large Language Model Inference with Semantics-Inspired Adaptive Sparse ActivationPoster
- SparseOpt: Addressing Normalization-induced Gradient Skew in Sparse TrainingPoster
- SparseSSM: Efficient Selective Structured State Space Models Can Be Pruned in One-ShotPoster
- Sparser Block-Sparse Attention via Token PermutationPoster
- Sparser, Faster, Lighter Transformer Language ModelsPoster
- Spatial Conformal Inference through Localized Quantile RegressionPoster
- Spatial Deconfounder: Interference-Aware Deconfounding for Spatial Causal InferencePoster
- Spatial Memory for Out-of-Vision Manipulation in Vision-Language-ActionPoster
- Spatial Priors via Space Filling Curves for Small and Limited Data Vision TransformersPoster
- Spatial-Aware Reduction Framework: Towards Efficient and Faithful Visual State Space ModelsPoster
- SpatialJB: How Text Distribution Art Becomes The "Jailbreak Key" for LLM GuardrailsPoster
- SpatialReward: Bridging the Perception Gap in Online RL for Image Editing via Explicit Spatial ReasoningPoster
- Spatially-Adaptive Gradient Re-parameterization for 3D Large Kernel OptimizationPoster
- Spatially-Regularized Entropy for Discriminative Token Merging in Fine-Grained Re-IdentificationPoster
- Spatio-Temporal LLM: Reasoning about Environments and ActionsPoster
- SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language ModelsOral
- Spatiotemporal Imputation with Graph-Informed Flow MatchingPoster
- SpecExit: Accelerating Large Reasoning Model via Speculative ExitPoster
- SpecForge: A Flexible and Efficient Open-Source Training Framework for Speculative DecodingPoster
- SpecMD: A Comprehensive Study On Speculative Expert PrefetchingPoster
- SpecPL: Disentangling Spectral Granularity for Prompt LearningPoster
- SpecPrune-VLA: Accelerating Vision-Language-Action Models via Action-Aware Self-Speculative PruningPoster
- Spectra: Rethinking Optimizers for LLMs Under Spectral AnisotropyPoster
- Spectral Bridge Variational Inference: Dynamic LoRA via Bures-Wasserstein Gradient FlowsPoster
- Spectral Collapse Drives Loss of Plasticity in Deep Continual LearningPoster
- Spectral Evolution Search: Efficient Inference-Time Scaling for Reward-Aligned Image GenerationPoster
- Spectral Flow Matching: Stabilizing Stochastic GFlowNets via Frequency-Domain RegularizationPoster
- Spectral Gradient Descent Mitigates Anisotropy-Driven Misalignment: A Case Study in Phase RetrievalPoster
- Spectral Guidance for Flexible and Efficient Control of Diffusion ModelsPoster
- Spectral Heat Flow for Conservative Token Condensation in Vision-Language ModelsPoster
- Spectral Imbalance Causes Forgetting in Low-Rank Continual AdaptationPoster
- Spectral Reach: Understanding Neural Scaling through Kernel Alignment DynamicsPoster
- Spectral-Informed Neural Networks Outperform Spectral methods in High-dimensional PDEsSpotlight
- Spectral-Progressive Thought Flow for Lightweight Multimodal ReasoningPoster
- Spectral–Spatial Mixing with Morphology-Aware Adaptive Loss for Medical Image Segmentation.Poster
- Speculative Coupled Decoding for Training-Free Lossless Acceleration of Autoregressive Visual GenerationPoster
- Speculative Safety Honeypot: Toward Proactive Defense Against Multi-turn Agent AttacksPoster
- Speculative Sampling For Faster Molecular DynamicsPoster
- Speech-Audio Compositional Attacks on Multimodal LLMs and Their Defense with SALMONN-GuardPoster
- SpeedCP: Fast Kernel-based Conditional Conformal PredictionPoster
- SpeedVFI: One-step Diffusion for Efficient Video Frame InterpolationPoster
- Speedup Patch: Learning a Plug-and-Play Policy to Accelerate Embodied ManipulationPoster
- Spherical Procrustes Alignment for Reliable Medical Audio DiagnosisPoster
- Spherical SO(3) Equivariant Local AttentionPoster
- Spherical Steering: Geometry-Aware Activation Rotation for Language ModelsPoster
- SphericalDreamer: Generating Navigable Immersive 3D Worlds with Panorama FusionPoster
- Spik4lite: Refactoring Neuromorphic Sparsity for Efficient Spiking Neural Networks on Commodity Edge DevicesPoster
- Spike Camera Autofocus via Frequency-Domain Spectral-Centroid MigrationPoster
- Spike-HTR: Spiking Neural Transformer for Handwritten Text RecognitionPoster
- SpikeCLR: Self-Supervised Contrastive Learning for Visual Representations with Spiking Neural NetworksPoster
- SpikeNet: Sparse Spike-Driven Mask Vector Transformer for Energy-Efficient and Stable Spiking Point Cloud ProcessingPoster
- SpikeVLA: Vision-Language-Action Models with Spiking Neural NetworksPoster
- Spiked-CFR: Causal Representation Learning from LLMs via Wasserstein Projection PursuitPoster
- SpikingLM: Towards Fully Spiking Language ModelPoster
- Spiral RoPE: Rotate Your Rotary Positional Embeddings in the 2D PlanePoster
- SplAttN: Bridging 2D and 3D with Gaussian Soft Splatting and Attention for Point Cloud CompletionSpotlight
- Split Group Knockoffs: Controlling False Discovery Rate in Transformational Group SparsityPoster
- Split Personality Training: Revealing Latent Knowledge Through Alternate PersonalitiesPoster
- Sponge Tool Attack: Stealthy Denial-of-Efficiency against Tool-Augmented Agentic ReasoningPoster
- SpreadsheetArena: Decomposing Preference in LLM Generation of Spreadsheet WorkbooksPoster
- Spurious Correlation Learning in Preference Optimization: Mechanisms, Consequences, and Mitigation via Tie TrainingPoster
- Spurious Rewards Paradox: Mechanistically Understanding How RLVR Activates Memorization Shortcuts in LLMsPoster
- Spurious Rewards: Rethinking Training Signals in RLVRPoster
- Squeezing More from the Stream : Learning Representation Online for Streaming Reinforcement LearningPoster
- Stability Analysis of Sharpness-Aware MinimizationPoster
- Stability and Generalization of Nonconvex Optimization with Heavy-Tailed NoisePoster
- Stability beyond bounded differences: sharp generalization bounds under finite $L_p$ momentsPoster
- Stability-Aware Feature Design for Robust Watermark Detection in Machine-Generated TextPoster
- Stabilized Supralinear Networks Learn to Switch Coding Strategies Balancing Cost and PerformancePoster
- Stabilizing In-Context Multi-Source Domain Adaptation for Biomedical Images Through ControlsPoster
- Stabilizing MoE Reinforcement Learning by Aligning Training and Inference RoutersPoster
- Stabilizing Native Low-Rank LLM PretrainingPoster
- Stabilizing PPO via Latent-Space Regularization and KDE-Driven ExplorationPoster
- Stabilizing Recurrent Dynamics for Test-Time Scalable Latent Reasoning in Looped Language ModelsPoster
- Stabilizing Reinforcement Learning for Diffusion Language ModelsPoster
- Stabilizing the Q-Gradient Field for Policy Smoothness in Actor-Critic MethodsOral
- Stable Asynchrony: Variance-Controlled Off-Policy RL for LLMsPoster
- Stable Deep Reinforcement Learning via Isotropic Gaussian RepresentationsSpotlight
- Stable Localized Conformal Prediction via TransductionPoster
- Stable Spectral Copula Alignment for Robust Multimodal LearningPoster
- Stable Velocity: A Variance Perspective on Flow MatchingPoster
- Stable-GFlowNet: Toward Diverse and Robust LLM Red-Teaming via Contrastive Trajectory BalanceSpotlight
- StableI2I: Spotting Unintended Changes in Image-to-Image TransitionPoster
- StableVLA: Towards Robust Vision-Language-Action Models without Extra DataPoster
- Stage-wise Distortion–Perception Traversal in Zero-shot Inverse Problems with Diffusion ModelsPoster
- Star Elastic: Many-in-One Reasoning LLMs with Efficient Budget ControlPoster
- StarEmbed: Benchmarking Time Series Foundation Models on Astronomical Observations of Variable StarsPoster
- State Space Model with Continuous Limit of HiPPO Matrix: Eigenvalue Analysis and Explicit Solution FormulaPoster
- State-Dependent Safety Failures in Multi-Turn Language Model InteractionPoster
- Stationary MMD PointsPoster
- Statistical Consistency and Generalization of Contrastive Representation LearningPoster
- Statistical Early Stopping for Reasoning ModelsPoster
- Statistical Impossibility and Possibility of Aligning LLMs with Human Preferences: From Condorcet Paradox to Nash EquilibriumPoster
- Statistical Learning Theory in Lean 4: Empirical Processes from ScratchPoster
- Statistical-Computational Trade-offs for Recursive Adaptive Partitioning EstimatorsPoster
- Statistically Optimal Scaling for Token Merging in TransformersPoster
- Statistically Undetectable Backdoors in Deep Neural NetworksPoster
- Steady-State Behavior of Constant-Stepsize Stochastic Approximation: Gaussian Approximation and Tail BoundsPoster
- Steal the Patch Size: Adversarially Manipulate Vision Language ModelsPoster
- Steer Like the LLM: Activation Steering that Mimics PromptingSpotlight
- Steer Where It Matters: Token-Level Visual-Sensitivity Steering for LVLMs Hallucination MitigationPoster
- Steering Beyond the Support: Adversarial Training on Unsupervised Jailbroken Activation SimulationPoster
- Steering Large Language Models through the DMTA Cycle: Structure-Based Drug Design via Knowledge-Driven Bi-Level Thompson SamplingPoster
- Steering Out-of-Distribution Generalization with Concept Ablation Fine-TuningPoster
- Steering at the Source: Style Modulation Heads for Robust Persona ControlPoster
- SteeringSafety: Benchmarking Representation Steering in LLMs Across Safety PerspectivesPoster
- Stein Diffusion Guidance: Training-Free Posterior Correction for Sampling Beyond High-Density RegionsPoster
- Stem: Rethinking Causal Information Flow in Sparse AttentionPoster
- Step-Level Sparse Autoencoder for Reasoning Process InterpretationPoster
- Step-Size Stability in Stochastic Optimization: A Theoretical PerspectivePoster
- Step-resolved data attribution for looped transformersPoster
- StepCodeReasoner: Aligning Code Reasoning with Stepwise Execution Traces via Reinforcement LearningPoster
- StethoLM: Audio Language Model for Cardiopulmonary Analysis Across Clinical TasksPoster
- StitchCUDA: An Automated Multi-Agents End-to-End GPU Programing Framework with Rubric-based Agentic Reinforcement LearningPoster
- Stochastic Gradient Methods under Heavy-Tailed Noises in Weakly Convex OptimizationPoster
- Stochastic Gradient Variational Inference with Price's Gradient Estimator from Bures-Wasserstein to Parameter SpacePoster
- Stochastic Lifting for Generating Trajectories of Stochastic Physical SystemsPoster
- Stochastic Linear Bandits with Parameter NoisePoster
- Stochastic Minimum-Cost Reach-Avoid Reinforcement LearningPoster
- Stochastic Neural Ray Tracing for Radio Frequency Channel ModelingPoster
- Stochastic Order Learning: An Approach to Rank Estimation Using Noisy DataPoster
- Stochastic Sparse Attention for Memory-Bound InferencePoster
- Stop Training for the Worst: Progressive Unmasking Accelerates Masked Diffusion TrainingPoster
- Stop When Further Reasoning Won’t Help: Attention-State Adaptive Generation in Reasoning ModelsSpotlight
- Stop the Flip-Flop: Context-Preserving Verification for Fast Revocable Diffusion DecodingPoster
- StormInsight: Hierarchical Environmental Forcing and Vertical Coupling for Weather System EvolutionPoster
- Strat-Reasoner: Reinforcing Strategic Reasoning of LLMs in Multi-Agent GamesPoster
- Strategic Candidacy in Generative AI ArenasPoster
- Strategic Navigation or Stochastic Search? How Agents and Humans Reason Over Document CollectionsOral
- Strategy Executability in Mathematical Reasoning: Leveraging Human–Model Differences for Effective GuidancePoster
- Strategy-Aware Optimization Modeling with Reasoning LLMsPoster
- Stratified GRPO: Handling Structural Heterogeneity in Reinforcement Learning of LLM Search AgentsPoster
- Stream RAG: Instant and Accurate Spoken Dialogue Systems with Streaming Tool UsagePoster
- StreamFlow: Theory, Algorithm, and Implementation for High-Efficiency Rectified Flow GenerationPoster
- Streaming Covariate Balancing via Discrepancy-Based Feature CoresetsPoster
- Streaming Sliced Optimal TransportPoster
- StretchTime: Adaptive Time Series Forecasting via Symplectic AttentionPoster
- Stronger Benchmarks for Prediction as a Service with ConstraintsPoster
- Stronger Semantic Encoders Can Harm Relighting Performance: A Probe of Visual Priors via Augmented Latent IntrinsicsPoster
- StructMAR: Structure-Aware Masked Autoregression for Explicit Layout Alignment in Text-to-Image GenerationPoster
- StructMamPose: From Sequential Perception to Structural Reasoning for 3D Human Pose EstimationPoster
- Structurally Aligned Subtask-Level Memory for Software Engineering AgentsPoster
- Structure Abstraction and Generalization in a Hippocampus-Entorhinal Inspired World ModelPoster
- Structure Enables Effective Self-Localization of Errors in LLMsPoster
- Structure-Aware Consistency Priors for Shape from Polarization in Complex MediaPoster
- Structure-Aware Riemannian Flow Matching for Registration and Fusion of Hyperspectral and Multispectral ImagesPoster
- Structure-Centric Graph Foundation Model via Geometric BasesPoster
- Structure-Induced Information for Rerooting Levin Tree SearchPoster
- Structure-Preserving Learning Improves Geometry Generalization in Neural PDEsPoster
- Structure-aware Granular-Ball based Information Bottleneck for Multi-modal ClusteringPoster
- Structured 4D Latent World Model for Robot PlanningPoster
- Structured Diffusion Bridges: Inductive Bias for Denoising Diffusion BridgesPoster
- Structured Expert Routing with Multi-View Task Priors for Offline Meta-Reinforcement LearningPoster
- Structured Multi-modal Graph Disentanglement for Psychiatric DiagnosisPoster
- Structured Multi-step Jailbreaking under a Hamiltonian Generative FormulationPoster
- Structured Progressive Knowledge Activation for LLM-Driven Neural Architecture SearchPoster
- Student-Centered Distillation Narrows the Agentic Gap Between Small and Large LLMsPoster
- StyleDistillation: A New Insight of Image Style Enables Personalized Aesthetic ManipulationPoster
- SuCo: Sufficiency-guided Continuous Adaptive ReasoningPoster
- Subgroup Discovery with the Cox ModelPoster
- Subliminal Effects in Your Data: A General Mechanism via Log-LinearityPoster
- Submodular Optimization for Minimal Augmentation in Robust Language Model AlignmentPoster
- Subspace-Aware Feature Reshaping for Open-Set Graph Class-Incremental LearningPoster
- SubspacePath Pruner: Inference-time Pruning via Probe-based Representation–Parameter CouplingPoster
- Success-Conditioning as Policy Improvement: The Optimization Problem Solved by Imitating SuccessPoster
- Sufficiency is Relative: Evaluating LLM Explanations under Model-Induced Input DistributionsPoster
- SuperHype: Hypergraph Generation via Graph-Superposition DecompositionPoster
- Supervise Less, See More: Training-free Nuclear Instance Segmentation with Prototype-Guided PromptingPoster
- Supervised Classification Heads as Semantic Prototypes: Unlocking Vision-Language Alignment via Weight RecyclingPoster
- Supervised Graph Contrastive Learning for Gene Regulatory NetworksPoster
- Supervised Guidance Training for Infinite-Dimensional Diffusion ModelsPoster
- Support-Proximity Augmented Diffusion Estimation for Offline Black-Box OptimizationPoster
- Suppress and Diversify: Refining Robust Pathways for Corruption RobustnessSpotlight
- Surgery: Mitigating Harmful Fine-Tuning for Large Language Models via Attention SinkSpotlight
- SurrogateSHAP: Training-Free Contributor Attribution for Text-to-Image (T2I) ModelsPoster
- SurvDiff: A Diffusion Model for Generating Synthetic Data in Survival AnalysisSpotlight
- Swift-SVD: Theoretical Optimality Meets Practical Efficiency in Low-Rank LLM CompressionPoster
- SwiftPFN: Revisiting Row-Wise Attention–Only Tabular Foundation Models with Adaptive Early ExitSpotlight
- SwitchCraft: Programmatic Design of State-Switching ProteinsPoster
- Swordsman: Entropy-Driven Adaptive Block Partition for Efficient Diffusion Language ModelsPoster
- SyMerge: From Non-Interference to Synergistic Merging via Single-Layer AdaptationPoster
- Sycophancy Towards Researchers Drives Performative MisalignmentSpotlight
- SymSpectra: Symmetric Information Bottleneck Framework for Molecular Structure Recognition under Imbalanced SettingsPoster
- Symbal: Detecting Systematic Misalignments in Model-Generated CaptionsPoster
- Symbiosis-Inspired Knowledge Distillation for Incremental Object DetectionPoster
- Symbol-Equivariant Recurrent Reasoning ModelsPoster
- Symbolic Mixture-of-Experts: Adaptive Skill-based Routing for Heterogeneous ReasoningPoster
- Symmetries in PAC-Bayesian LearningPoster
- Symmetries in language statistics shape the geometry of model representationsSpotlight
- Symmetry Reveals the In-Context Classifier: Transformers Implement Mean-Shift DynamicsSpotlight
- SynGR: Unleashing the Potential of Cross-Modal Synergy for Generative RecommendationPoster
- SynLaD: Latent Diffusion for Generating Synthesizable Molecules Conditioned on 3D Pharmacophore ProfilesPoster
- SynerMedGen: Synergizing Medical Multimodal Understanding with Generation via Task AlignmentPoster
- Synergistic Intra- and Cross-Layer Regularization Losses for MoE Expert SpecializationPoster
- Synergistic Space-Vision Processing for Predicate InferencePoster
- Syntax vs. Semantics: How Transformers Learn Deep DependenciesPoster
- Synthesizable Molecular Generation via Soft-constrained GFlowNets with Rich Chemical PriorsPoster
- Synthesizing Multimodal Geometry Datasets from Scratch and Enabling Visual Alignment via Plotting CodePoster
- Synthesizing world models for bilevel planningPoster
- Systematic Failures in Collective Reasoning under Distributed Information in Multi-Agent LLMsPoster
- T$^2$PO: Uncertainty-Guided Exploration Control for Stable Multi-Turn Agentic Reinforcement LearningSpotlight
- T-Edit: Triple-Branch Diffusion Anchoring for Consistent EditingPoster
- T-GINEE: A Tensor-Based Multi-Graph Representation LearningPoster
- T-POP: Test-Time Personalization with Online Preference FeedbackPoster
- T-measure: A Topology-Consistent Metric for Binary SegmentationPoster
- T2AV-Compass: Towards Unified Evaluation for Text-to-Audio-Video GenerationPoster
- TABX: A High-Throughput Sandbox Battle Simulator for Multi-Agent Reinforcement LearningPoster
- TACTIC: Task-Aware Sparse Coordination Graphs for Multi-Task Multi-agent Reinforcement LearningPoster
- TAG: Tangential Amplifying Guidance for Hallucination-Resistant SamplingPoster
- TAGRPO: Boosting GRPO on Image-to-Video Generation with Direct Trajectory AlignmentPoster
- TAMPO: Task- and Model-Aware Automatic Prompt Optimization for Robust and Controllable Auto-Routing in LLM-based SystemsPoster
- TCAP: Tri-Component Attention Profiling for Unsupervised Backdoor Detection in MLLM Fine-TuningPoster
- TD-VAD: Breaking Visual Dependence in Video Anomaly Detection with Text-Driven LearningPoster
- TD3B: Transition-Directed Discrete Diffusion for Allosteric Binder GenerationSpotlight
- TDM-R1: Reinforcing Few-Step Diffusion Models with Non-Differentiable RewardPoster
- TEAM: Temporal–Spatial Consistency Guided Expert Activation for MoE Diffusion Language Model AccelerationPoster
- TEFormer: Structured Bidirectional Temporal Enhancement Modeling in Spiking TransformersPoster
- TF-FACE: Time-Frequency Fusion Learning via Frequency-Domain Adaptive and Controllable Enhancement for Trajectory PredictionPoster
- TFRBench: A Reasoning Benchmark for Evaluating Forecasting SystemsPoster
- TFTF: Training-Free Targeted Flow for Conditional SamplingPoster
- TG-RAG: A Retrieval-Augmented Framework for Reasoning Guidance in Specialized DomainsOral
- TGPO: Efficient Policy Optimization through Sequence Anchor and Information GatingPoster
- TGV-KV: Text-Grounded KV Eviction for Vision-Language ModelsPoster
- THETA: Threshold-Based Exclusive Batching for Memory-Bandwidth-Constrained LLM InferencePoster
- TIC-VLA: A Think-in-Control Vision-Language-Action Model for Robot Navigation in Dynamic EnvironmentsPoster
- TIME: Tensor-Factorized Mixture-of-Experts with Intrinsic Routing for Lifelong Multimodal Knowledge EditingPoster
- TIMI: Training-Free Image-to-3D Multi-Instance Generation with Spatial FidelityPoster
- TINNs: Time-Induced Neural Networks for Solving Time-Dependent PDEsPoster
- TMD-Bench: A Multi-Level Evaluation Paradigm for Music–Dance Co-GenerationPoster
- TMS: Trajectory-Mixed Supervision for Reward-Free, On-Policy SFTPoster
- TN-SHAP-G: Graph-Structured Tensor Network Surrogates for Shapley Values and InteractionsPoster
- TOM-SWE: User Mental Modeling For Software Engineering AgentsPoster
- TPGDiff : Hierarchical Triple-Prior Guided Diffusion for Image RestorationPoster
- TPV: Parameter Perturbations Through the Lens of Test Prediction VariancePoster
- TQL: Scaling Q-Functions with Transformers by Preventing Attention CollapsePoster
- TRACE: Toulmin-based Reasoning Assessment through Constructive Elements for LLM CoT EvaluationPoster
- TRACE: Trajectory Recovery for Continuous Mechanism Evolution in Causal Representation LearningPoster
- TRACER: Persistent Regularization for Robust Multimodal FinetuningPoster
- TRACER: Trajectory Risk Aggregation for Critical Episodes in Agentic ReasoningPoster
- TRAP: Hijacking VLA CoT-Reasoning via Adversarial PatchesPoster
- TRIM: Token-wise Attention-Derived Saliency for Data-Efficient Instruction TuningPoster
- TRIP-Bench: A Benchmark for Long-Horizon Interactive Agents in Real-World ScenariosPoster
- TSFAdv: Frequency-Guided Black-Box Adversarial Attacks on Time Series ForecastingPoster
- TSMGen: Target-Specific Molecule Generation via Higher-Order Structural Dependencies and Context-Aware Bidirectional FusionPoster
- TSP with predictionsPoster
- TSRBench: A Comprehensive Multi-task Multi-modal Time Series Reasoning Benchmark for Generalist ModelsPoster
- TT-Sparse: Learning Sparse Rule Models with Differentiable Truth TablesPoster
- TUR-DPO: Topology- and Uncertainty-Aware Direct Preference OptimizationPoster
- TVDRNet: Text-driven Viewpoint Optimization via Differentiable Rendering for 3D Reasoning SegmentationPoster
- TVI-CoT: Text-Visual Interleaved Chain-of-Thought Reasoning for Multimodal UnderstandingPoster
- TWLA: Breaking the Barrier to W1.58A4 Post-Training Quantization for LLMsPoster
- TaRO: Temporal-Aware Reasoning Optimization for Video Temporal GroundingPoster
- TabICooL: A better, faster, scalable, and open tabular foundation modelPoster
- TabMGP: Martingale Posterior with TabPFNPoster
- TabPack: Efficient Hyperparameter Ensembles for Tabular Deep LearningPoster
- Tabero: Learning Gentle Manipulation with Closed-Loop Force Feedback from Vision, Touch, and LanguagePoster
- TabularBERT: Binning-Based Self-Supervised Learning for Tabular RepresentationPoster
- Tackling Fake Forgetting through Uncertainty QuantificationPoster
- Tackling Length Inflation Without Trade-offs: Group Relative Reward Rescaling for Reinforcement LearningPoster
- TadABench-1M: A Large-Scale Wet-Lab Protein Benchmark For Rigorous OOD EvaluationPoster
- Tail Annealing for Heavy-Tailed Flow MatchingPoster
- Tailoring Strictly Proper Scoring Rules for Downstream Tasks: An Application to Causal InferencePoster
- Tailoring the Training: Difficulty-Aware Learning Strategy Allocation for Large Language ModelsPoster
- Taking the GP Out of the LoopPoster
- Talk, Judge, Cooperate: Gossip-Driven Indirect Reciprocity in Self-Interested LLM AgentsPoster
- Taming I2V models for Image HOI Editing: A Cognitive Benchmark and Agentic Self-Correcting FrameworkPoster
- Taming Stochastic Gradient Descent: Almost Sure Convergence and Saddle-Point Avoidance under $(L_{0},L_{1})$-SmoothnessPoster
- Taming the Aleatoric Impulse in Off-Policy Reinforcement LearningPoster
- Taming the Loss Landscape of PINNs with Noisy Feynman–Kac Supervision: Operator Preconditioning and Non-Asymptotic Error BoundsPoster
- Taming the Recent-Data Bias: Towards Robust Time Series Forecasting with Global ContextPoster
- TapSampling: Inference-Time Sampling with a Task-Progress-Understanding Verifier for Robotic ManipulationPoster
- TarGATE: Target-Aware Data Selection via Token-Attenuation GatesPoster
- Target-Agnostic Calibration under Distribution Shift with Frequency-Aware Gradient RectificationPoster
- Target-Aware Bandit Allocation for Scalable Surrogate Optimization in Chemical SpacePoster
- Target-Driven Policy Optimization for Sequential Counterfactual Outcome ControlPoster
- Target-Oriented Pretraining Data Selection via Neuron-Activated GraphPoster
- Task-Aware Exploration via a Predictive Bisimulation MetricPoster
- Task-Aware Mechanism: Hybrid MoE Vision Tower Towards Holistic Video UnderstandingPoster
- Task-Aware Preference Calibration for Direct Preference OptimizationPoster
- Task-Aware Structured Memory for Dynamic Multi-modal In-Context LearningPoster
- Task-Awareness Improves LLM Generations and UncertaintyPoster
- Task-Driven Subspace Decomposition for Knowledge Sharing and Isolation in LoRA-based Continual LearningPoster
- Task-and-Model-Aware Fractal-Consistency for Efficient LLM ReasoningPoster
- TaskLoom: Weaving Knowledge Across Tasks in World ModelsPoster
- Teaching Agents to Ask Effective Clarification QuestionsPoster
- Teaching Models to Teach Themselves: Reasoning at the Edge of LearnabilitySpotlight
- Teaching Molecular Dynamics to a Non-Autoregressive Ionic Transport PredictorPoster
- TeamTR: Trust-Region Fine-Tuning for Multi-Agent LLM CoordinationPoster
- TeamWork: Multivariate Time Series Anomaly Detection via Asymmetric Role-aware Channel ModelingPoster
- TelecomTS: A Multi-Modal Observability Dataset for Time Series and Language AnalysisPoster
- Telescope: Improving Zero Shot Detection of LLM Generated Content By Measuring Token Repetition ProbabilityPoster
- Temper-Then-Tilt: Principled Unlearning for Generative Models through Tempering and Classifier GuidancePoster
- Tempora: Characterising the Time-Contingent Utility of Online Test-Time AdaptationPoster
- Temporal Context Reinstatement Drives Episodic-Like Order Memory in Long-Context Language ModelsPoster
- Temporal Difference Calibration in Sequential Tasks: Application to Vision-Language-Action ModelsPoster
- Temporal Difference Learning for Diffusion ModelsPoster
- Temporal Difference Learning with Compressed Updates: Error-Feedback meets Reinforcement LearningPoster
- Temporal Preference Optimization for Unsupervised RetrievalPoster
- Temporal Score Rescaling for Temperature Sampling in Diffusion and Flow ModelsPoster
- Temporal Self-Rewarding Language Models: Decoupling Chosen-Rejected via Past-FuturePoster
- Temporal Straightening for Latent PlanningPoster
- Temporal Weighted Encoding: Towards Maximal-Capacity Spike Coding for ANN–SNN ConversionPoster
- Temporal-Emerged Prompting for Segment Anything in Multiframe Infrared Small Target DetectionPoster
- Temporal-aware Flow Matching for Video Generation with Temporally Coherent MotionPoster
- Terminal Dimension Reduction for Time Series with ApplicationsPoster
- Test-Time Anchoring for Discrete Diffusion Posterior SamplingPoster
- Test-Time Debiasing with Probabilistic Prompts via Wasserstein Distance in Vision-Language ModelsPoster
- Test-Time Detoxification without Training or Learning AnythingPoster
- Test-Time Graph Search for Goal-Conditioned Reinforcement LearningPoster
- Test-Time Guidance for Flow-Based Generative Models via Parallel Tempering on Source DistributionsPoster
- Test-Time Learning of Causal Structure from Interventional DataPoster
- Test-Time Reinforcement Learning for Flow MatchingPoster
- Test-Time Training Is Secretly Linear AttentionPoster
- Test-time Generalization for Physics through Neural Operator SplittingPoster
- Test-time Offline Reinforcement Learning on Goal-related ExperiencePoster
- TestExplora: Benchmarking LLMs for Proactive Bug Discovery via Repository-Level Test GenerationPoster
- Testing For Distribution Shifts with Conditional Conformal Test MartingalesPoster
- TetraJet-v2: Accurate NVFP4 Training for Large Language Models with Oscillation Suppression and Outlier ControlSpotlight
- TexEditor: Structure-Preserving Text-Driven texture EditingPoster
- Text Before Vision: Staged Knowledge Injection Matters for Agentic RLVR in Ultra-High-Resolution Remote Sensing UnderstandingPoster
- Text Generation as Continuous Latent Dynamics via Reinforcement LearningPoster
- Text Has CurvaturePoster
- Text-Conditional JEPA for Learning Semantically Rich Visual RepresentationsPoster
- Text-Driven Fusion for Infrared and Visible Images: Achieving Image Scene Adaptation on Hyperbolic SpacePoster
- TextAtlas5M: A Large-Scale Dataset for Long Text Image GenerationPoster
- TextME: Bridging Unseen Modalities Through Text DescriptionsPoster
- TextMesh4D: Zero-shot Text-to-4D Mesh GenerationPoster
- TextResNet: Decoupling and Routing Optimization Signals in Compound AI Systems via Deep Residual TuningPoster
- Textual Stochastic Gradient Descent: Discrete Optimization of External Memory for Reasoning Language AgentsPoster
- Textual Supervision Enhances Geospatial Representations in Vision-Language ModelsPoster
- The (Marginal) Value of a Search Ad: An Online Causal Framework for Repeated Second-price AuctionsPoster
- The ACE Protocol: Operationalizing Language Model Activations for Better Calibration and UtilityPoster
- The Abstraction Gap in Vision-Language Causal ReasoningPoster
- The Accumulation of Score Estimation Error in Diffusion ModelsPoster
- The Art of Interrogation: Consistency Amplifies Factuality in Spatial ReasoningPoster
- The Assistant Axis: Situating and Stabilizing the Default Persona of Language ModelsSpotlight
- The Bridge-Garden Dilemma in LLM Distillation: Why Mixing Hard and Soft Labels WorksPoster
- The Catastrophic Failure of *the* k-Means Algorithm in High Dimensions, and How Hartigan's Algorithm Avoids ItPoster
- The Choice of Normalization Influences Shrinkage in Regularized RegressionPoster
- The Consistency Trap in LLMs: Generator-Evaluator Agreement and Vulnerability to MistakesPoster
- The Convergent Representation of Vision-Language Contrastive Learning: Geometry, Modality Gap and Shared Space AlignmentPoster
- The Cost of Information: Phase Transitions in Contextual Bandits with Paid ObservationsPoster
- The Cost of Learning under Multiple Change PointsPoster
- The Crowded Embedding Space: A Mean-Field Mechanism for Emergent Marginalization in Retrieval-Augmented AgentsPoster
- The Cylindrical Representation Hypothesis for Language Model SteeringPoster
- The Decrypto Benchmark for Multi-Agent Reasoning and Theory of MindPoster
- The Deterministic Horizon: When Extended Reasoning Fails and Tool Delegation Becomes NecessaryPoster
- The Devil is in the Condition Numbers: Why is GLU Better than non-GLU Structure?Poster
- The Devil is in the Spectrum: Mitigating Representation Collapse in LLMs via Topologically Regularized Side-PathPoster
- The Differences Between Direct Alignment Algorithms are a BlurPoster
- The Double Dilemma in Multi-Task Radiology Report Generation: A Gradient Dynamics Analysis and SolutionPoster
- The Double-Edged Nature of the Rashomon Set for Trustworthy Machine LearningSpotlight
- The Efficiency Gap in Byte ModelingPoster
- The Entropic Signature of Class Speciation in Diffusion ModelsPoster
- The Expert Strikes Back: Interpreting Mixture-of-Experts Language Models at Expert LevelPoster
- The Expressive Power of Low Precision Softmax Transformers with (Summarized) Chain-of-ThoughtPoster
- The Expressivity Limits of TransformersOral
- The Extra Tokens Matter: Disentangled Representation Learning with Vision TransformersPoster
- The Fairness Hierarchy: A viewpoint from causal inferencePoster
- The First Drop of Ink: Nonlinear Impact of Misleading Information in Long-Context ReasoningPoster
- The Fisher Dimension: Instance-Dependent Complexity for Causal DiscoveryPoster
- The Flexibility Trap: Rethinking the Value of Arbitrary Order in Diffusion Language ModelsOral
- The Forgetting-Retention Dilemma: Certified Unlearning Theory in Continual LearningPoster
- The Generalization Spectrum: A Chromatographic Approach to Evaluating Learning AlgorithmsPoster
- The Geometric Mechanics of Contrastive Representation Learning: Alignment Potentials, Entropic Dispersion, and Cross-Modal DivergencePoster
- The Geometric Origin of Grokking: Accelerating Generalization via Active Structural ReorganizationPoster
- The Geometric Reasoner: Manifold-Informed Latent Foresight Search for Long-Context ReasoningPoster
- The Geometry of Narrow Fine-Tuning Degradation: Trajectory Lock-in and Spectral BifurcationPoster
- The Geometry of Projection Heads: Conditioning, Invariance, and CollapsePoster
- The Geometry of Reasoning: Self-Evaluation via Layerwise Trajectory EvolutionPoster
- The Geometry of Representational Failures in Vision Language ModelsPoster
- The Geometry of Sequential Learning: Lie-Bracket Prediction of Transfer OrderPoster
- The Geometry of Updates: Fisher Alignment at Vocabulary ScalePoster
- The Heterogeneous Safety Impacts of Benign Multilingual Fine-TuningPoster
- The Hidden Risk: Membership Inference Attacks on Multimodal Federated Learning via Modality ImbalancePoster
- The Hippocampal Place Field Gradient: A Bio-inspired Framework Building Multiscale Representation for Better Sample EfficiencyPoster
- The Ideal Expression Is Not a Local Optimum: A Revisit of EQL with Zero-Point ConstraintsPoster
- The Illusion of Generalization: Instruction-Following, Task Bias and Contamination in Tabular Language Model EvaluationPoster
- The Implicit Bias of Adam and Muon on Smooth Homogeneous Neural NetworksPoster
- The Implicit Bias of Depth: From Neural Collapse to Softmax CodesPoster
- The Implicit Bias of Steepest Descent with Mini-batch Stochastic GradientPoster
- The Information Geometry of Softmax: Probing and SteeringPoster
- The Interplay Between Interpolation and Aggregation in Regression: Optimal Sample ComplexityPoster
- The Invisible Lottery: How Subtle Cues Steer Algorithm Choice in LLM Code GenerationPoster
- The Label Horizon Paradox: Rethinking Supervision Targets in Financial ForecastingPoster
- The Labyrinth and the Thread: Rethinking Regularizations in Sequential Knowledge Editing for Large Language ModelsPoster
- The Latent Color Subspace: Emergent Order in High-Dimensional ChaosPoster
- The Latent Guardian: Defending Collaborative Perception via Feature-Level Consistency VerificationPoster
- The Lie We Tell: Correcting the Euclidean Fallacy in Vision Language Action Policies via Score Matching on Tangent SpacePoster
- The Loss Is Not Enough: Sampling Conditions and Inductive Bias in Contrastive Representation LearningPoster
- The Obfuscation Atlas: Mapping Where Honesty Emerges in RLVR with Deception ProbesOral
- The Optimal Sample Complexity of Linear ContractsPoster
- The Optimal Token Baseline: Variance Reduction for Long-Horizon LLM-RLPoster
- The Oversight Game: Learning to Cooperatively Balance an AI Agent's Safety and AutonomyPoster
- The Pareto-optimal Trade-off between Regret and Statistical Inference in Linear Stochastic Bandits under Safety ConstraintsPoster
- The Perception–Physics Paradox: Probing Scientific Alignment with TC-AtlasPoster
- The Personality Illusion: Revealing Dissociation Between Self-Reports & Behavior in LLMsPoster
- The Power of Power Law: Asymmetry Enables Compositional ReasoningSpotlight
- The Quality-Utility Paradox: Why High-Reward Data Impairs Small Model ReasoningPoster
- The Realignment Problem: When Right becomes Wrong in LLMsPoster
- The Relative Instability of Model Comparison with Cross-validationSpotlight
- The Role of Target Update Frequencies in Q-LearningPoster
- The Safety-Aware Denoiser for Text Diffusion ModelsPoster
- The Secret Engine Behind RLHF: It's Contarstive Learning All AlongPoster
- The Shadow Price of Reasoning: Economic Perspective on Optimal Budget Allocation for LLMsPoster
- The Shape of Addition: Geometric Structures of Arithmetic in Large Language ModelsPoster
- The Sign Estimator: Preference Modeling for LLM Alignment under HeterogeneityPoster
- The Signal is in the Steps: Local Scoring for Reasoning Data SelectionOral
- The Silent Thought: Modeling Internal Cognition in Full-Duplex Spoken Dialogue Models via Latent ReasoningPoster
- The Stability of Singular Distribution: A Spectral Perspective on the Two-Phase Dynamics of Language Model Pre-trainingPoster
- The Structural Origin of Attention Sink: Variance Discrepancy, Super Neurons, and Dimension DisparityPoster
- The Surprising Difficulty of Search in Model-Based Reinforcement LearningPoster
- The Tell-Tale Norm: $\ell_2$ Magnitude as a Signal for Reasoning Dynamics in Large Language ModelsSpotlight
- The Theory and Practice of MAP Inference over Non-Convex ConstraintsPoster
- The Trojan Knowledge: Bypassing Commercial LLM Guardrails via Harmless Prompt Weaving and Adaptive Tree SearchPoster
- The Truth Lies Somewhere in the Middle (of the Generated Tokens)Poster
- The Truth Stays in the Family: Enhancing Contextual Truthfulness via Inherited Heads in Model LineagesPoster
- The Two-Hump Problem: Bridging the Difficulty Gap in Mathematical Reinforcement LearningPoster
- The Unlearnability Phenomenon in RLVR for Language ModelsPoster
- The Value Function Semi-Algebraic Set in Partially Observable Markov Decision ProcessesPoster
- The Value of Variance: Mitigating Debate Collapse in Multi-Agent Systems via Uncertainty-Driven Policy OptimizationSpotlight
- The Velocity Deficit: Initial Energy Injection for Flow MatchingPoster
- The Viscosity of Logic: Phase Transitions and Hysteresis in DPO AlignmentPoster
- The benefits of full data shuffle, now with optimal I/O cost: $k$-wise independence and matrix transposition to the rescuePoster
- The cost of commitment in option-based hierarchical RLPoster
- The data manifold under the microscopePoster
- The impact of LoRA on Oversmoothing $\colon$ Understanding Catastrophic Forgetting in Mean-Field Attention DynamicsPoster
- The surprising strength of weak classifiers for validating neural posterior estimatesPoster
- Theoretical Analysis of Sparse Optimization with Reparameterization, Weight Decay, and Adaptive Learning RatePoster
- Theoretical Challenges in Learning for Branch-and-CutPoster
- Theoretical Characterization of Generalization in Knowledge DistillationPoster
- Theoretical Guarantees for One-Shot Magnitude Pruning and Compute-Adaptive Early ExitPoster
- Theoretical Investigation on Inductive Bias of Isolation ForestPoster
- Theoretical Perspectives on Data Quality and Synergistic Effects in Pre- and Post-Training Reasoning ModelsPoster
- Theory of Continual Learning Against Data Poisoning AttacksPoster
- Theory of Minimal Weight Perturbations in Deep Networks and its Applications for Low-Rank Activated Backdoor AttacksPoster
- ThetaEvolve: Test-time Learning on Open ProblemsPoster
- Think Deep, Not Just Long: Measuring LLM Reasoning Effort via Deep-Thinking TokensPoster
- Think Fast and Slow: Step-Level Cognitive Depth Adaptation for LLM AgentsPoster
- Think Less, Act Early: Reinforced Latent Reasoning with Early Exit in Vision-Language-Action ModelsPoster
- Think Twice Before You Act: Enhancing Agent Behavioral Safety with Thought CorrectionPoster
- Think Twice Before You Act: Protecting LLM Agents Against Tool Description Poisoning via Isolated PlanningPoster
- Think in Cloud, Look at Edges: Semantic-Driven Query Decomposition for Efficient Video ReasoningSpotlight
- Think in Latent, Explain in Language: Self-Explainable Latent ReasoningPoster
- Think-Then-Generate: Reasoning-Aware Text-to-Image Diffusion with LLM EncodersPoster
- Think-at-Hard: Selective Latent Iterations to Improve Reasoning Language ModelsPoster
- Thinking in Flow: A Dissipative Stabilization Operator for Robust Autoregressive ReasoningSpotlight
- Thinking in Latent Space: Progressive Multimodal Simplification for Visual ReasoningPoster
- Thinking in Scales: Accelerating Gigapixel Pathology Image Analysis via Adaptive Continuous ReasoningPoster
- Thinking in Structures: Evaluating Spatial Intelligence through Reasoning on Constrained ManifoldsPoster
- Thinking with Geometry: Active Geometry Integration for Spatial ReasoningPoster
- Thinned Mean Field Langevin DynamicsPoster
- This State Looks Like That: Self-Interpretable Reinforcement Learning Agents using Prototype Soft Actor-CriticPoster
- ThoughtFold: Folding Reasoning Chains via Introspective Preference LearningPoster
- Thoughtbubbles: an Unsupervised Method for Parallel Thinking in Latent SpacePoster
- ThreadWeaver: Adaptive Threading for Efficient Parallel Reasoning in Language ModelsOral
- Threat2Traffic: Multi-Agent Environment Synthesis for Malware Traffic Generation from Threat IntelligencePoster
- Three Years of r/ChatGPT: Societal Impact Evaluations from Social Media DataPoster
- Threshold-Guided Optimization for Visual Generative ModelsPoster
- Through the Stealth Lens: Attention-Aware Defenses Against Poisoning in RAGPoster
- ThunderAgent: A Fast, Simple, and Program-Aware Agentic Inference SystemSpotlight
- TiME: Test-Time Mixture-of-Experts Routing via Asymmetric CO-Optimal Transport for Continual Test-Time AdaptationPoster
- TiMi: Empower Time Series Transformers with Multimodal Mixture of ExpertsPoster
- Tight Margin-Based Generalization Bounds for Voting Classifiers over Finite Hypothesis SetsPoster
- Tight Stability Bounds for Robust Distributed Learning: Byzantine Failures Hurt Generalization More than Data PoisoningPoster
- Tightening the Score Matching Gap for Diffusion ModelsPoster
- Tighter Regret Lower Bound for Gaussian Process Bandits with Squared Exponential Kernel in HyperspherePoster
- TileQ: Efficient Low-Rank Quantization of Mixture-of-Experts with 2D TilingPoster
- TileSparse: Arithmetic-Intensity-Aware Sparse Attention for Compute-Bound LLM DecodingPoster
- Tilt Matching for Scalable Sampling and Fine-TuningSpotlight
- Time Series Reasoning via Process-Verifiable Thinking Data Synthesis and Scheduling for Tailored LLM ReasoningPoster
- Time Series, Vision, and Language: Exploring the Limits of Alignment in Contrastive Representation SpacesPoster
- Time series saliency maps: Explaining models across multiple domainsSpotlight
- Time-CoT: Hierarchical Reasoning with Temporal Semantic Codes for Multivariate Time Series ClassificationPoster
- Time-Conditioned Foreseeing: An EHR-Specific Foundation Model for Irregular Dynamics and Calendrical TimePoster
- Time-Consistent Robust Multi-Objective Reinforcement Learning via a Bellman–Isaacs Weight-Adversary RecursionPoster
- Time-PEFT: Temporal and Multichannel Complexity-Based Fine-Tuning for Time-Series Foundation ModelsPoster
- Time-Series Decomposition as a standalone Task: A Mechanism-Driven Diagnostic BenchmarkPoster
- Time-series forecasting through the lens of dynamicsPoster
- TimeAutoDiff: A Unified Framework for Generation, Imputation, Forecasting, and Time-Varying Metadata Conditioning of Heterogeneous Time Series Tabular DataPoster
- TimeGuard: Channel-wise Pool Training for Backdoor Defense in Time Series ForecastingPoster
- TimeLAVA: Learning-Agnostic Valuation for Time Series DataPoster
- TimeMRA: LLM-Empowered Time Series Forecasting via Multi-Scale Retrieval-Augmented RepresentationsPoster
- TimeOmni-VL: Unified Models for Time Series Understanding and GenerationPoster
- TimeRewarder: Learning Dense Reward from Passive Videos via Frame-wise Temporal DistanceSpotlight
- TimeSAE: Sparse Decoding for Faithful Explanations of Black-Box Time Series ModelsPoster
- TimeSeed: Effective Time Series Forecasting with Sparse Endogenous VariablesPoster
- TimeSpot: Benchmarking Geo-Temporal Understanding in Vision–Language Models in Real-World SettingsPoster
- Timestep Rescheduling in Diffusion InversionPoster
- Tiny Brains, Giant Impact: Uncovering the Keystone Neurons of LLM with Just a Few PromptsPoster
- To Grok Grokking: Provable Grokking in Ridge RegressionOral
- ToMAP: Training Opponent-Aware LLM Persuaders with Theory of MindPoster
- ToMoE: Converting Dense Large Language Models to Mixture-of-Experts through Dynamic Structural PruningPoster
- ToaSt: Token Channel Selection and Structured Pruning for Efficient ViTPoster
- TodoEvolve: Learning to Architect Agent Planning SystemsPoster
- TokSuite: Measuring the Impact of Tokenizer Choice on Language Model BehaviorOral
- Token Sample Complexity of AttentionPoster
- Token Sparse Attention: Efficient Long-Context Inference with Interleaved Token SelectionPoster
- Token-Efficient Change Detection in LLM APIsPoster
- Token-Free Hierarchical Indexing for RAG beyond LLM-based SummarizationPoster
- Token-Level LLM Collaboration via FusionRoutePoster
- Token-Sparse Medical Multimodal Reasoning via Dual-Stream Reinforcement LearningPoster
- TokenDrop: Token-Level Importance-Aware Backward Propagation Skipping for Efficient LLM Fine-TuningPoster
- TokenRatio: Principled Token-Level Preference Optimization via Ratio MatchingPoster
- TokenSwap: Backdoor Attack on the Compositional Understanding of Large Vision-Language ModelsPoster
- Tokenised Flow Matching for Hierarchical Simulation Based InferencePoster
- ToolOrchestra: Elevating Intelligence via Efficient Model and Tool OrchestrationPoster
- TopAdapter: Topology-Aware Prompt Tuning for Efficient Point Cloud UnderstandingPoster
- TopBench: A Benchmark for Implicit Prediction and Reasoning over Tabular Question AnsweringPoster
- TopoDistill: Distilling Global System Topology for Causal Discovery in Multivariate Time SeriesPoster
- Topological Active Inference for Task DisambiguationPoster
- Topology-Aware Contrastive Learning: Regulating Representation Connectivity via Persistent HomologyPoster
- Topology-Preserving Neural Operator Learning via Hodge DecompositionPoster
- Torus Graphs for Large Scale Neural Phase AnalysisPoster
- Toward Calibrated Mixture-of-Experts Under Distribution ShiftPoster
- Toward Culturally Aligned LLMs through Ontology-Guided Multi-Agent ReasoningPoster
- Toward Cybersecurity-Expert Small Language ModelsPoster
- Toward Effective Multimodal Graph Foundation Model: A Divide-and-Conquer Based ApproachPoster
- Toward Identifiable Sparse AutoencodersPoster
- Toward More Reliable Agent Evaluation: A Component-Based Benchmark Auditing PipelinePoster
- Toward Robust Multilingual Adaptation of LLMs for Low-Resource LanguagesPoster
- Toward Safe Quantization-Aware Fine-tuning: Understanding and Mitigating Safety Alignment DegradationPoster
- Toward Scalable and Valid Conditional Independence Testing with Spectral RepresentationsPoster
- Toward Stable Value Alignment: Introducing Independent Modules for Consistent Value GuidanceSpotlight
- Toward Structural Multimodal Representations: Specialization, Selection, and Sparsification via Mixture-of-ExpertsPoster
- Toward Subspace-Perturbed Trajectory-Aware Backdoor Attacks in Deep Reinforcement LearningPoster
- Toward Training Superintelligent Software Agents through Self-Play SWE-RLPoster
- Toward Understanding Adversarial Distillation: Why Robust Teachers FailPoster
- Towards A Generative Protein Evolution Machine with DPLM-EvoPoster
- Towards Achieving Optimal Strong Regret and Constraint Violation via Computational Efficient Model-free RLPoster
- Towards Atoms of Large Language ModelsPoster
- Towards Cold-Start Drafting and Continual Refining: A Value-Driven Memory Approach with Application to NPU Kernel SynthesisPoster
- Towards Complete Multi-Agent Coordination Policy Learning via Denoising Maximum Entropy OptimizationPoster
- Towards Completeness in Causal Discovery from Soft Interventions with Known TargetsPoster
- Towards Context-Invariant Safety Alignment for Large Language ModelsPoster
- Towards Disentangled Preference Optimization DynamicsPoster
ICML accepted papers in other years
Looking for submission deadlines instead? See the conference deadline calendar.