AAAI 2025 Accepted Papers
The full list of 3,478 papers accepted at AAAI 2025 (Association for the Advancement of Artificial Intelligence Annual Conference on Artificial Intelligence). Click any title for details, similar papers, and links to the original source. You can also search these papers by meaning, not just keywords.
Technical: 3,478
- STEM-LTS: Integrating Semantic-Temporal Dynamics in LLM-driven Time Series AnalysisTechnical
- STGC-NeRF: Spatial-Temporal Geometric Consistency for LiDAR Neural Radiance Fields in Dynamic ScenesTechnical
- STLC-KG:A Social Text Steganalysis Method Combining Large-Scale Language Models and Common-Sense Knowledge GraphsTechnical
- STraj: Self-training for Bridging the Cross-Geography Gap in Trajectory PredictionTechnical
- SVGBuilder: Component-Based Colored SVG Generation with Text-Guided Autoregressive TransformersTechnical
- SVRMamba: Slice-to-Volume Reconstruction from Multiple MRI Stacks with Slice Sequence Guided MambaTechnical
- SVTformer: Spatial-View-Temporal Transformer for Multi-View 3D Human Pose EstimationTechnical
- SVasP: Self-Versatility Adversarial Style Perturbation for Cross-Domain Few-Shot LearningTechnical
- SWAMamba: A Sliding Window Attention Mamba Framework for Predicting Translation Elongation RatesTechnical
- SYNAPSE: SYmbolic Neural-Aided Preference Synthesis EngineTechnical
- Safe Online Convex Optimization with Heavy-Tailed Observation NoisesTechnical
- Salient Frequency-aware Exemplar Compression for Resource-constrained Online Continual LearningTechnical
- Sample Complexity of Linear Regression Models for Opinion Formation in NetworksTechnical
- Sample-aware Adaptive Structured Pruning for Large Language ModelsTechnical
- Scaffold-BPE: Enhancing Byte Pair Encoding for Large Language Models with Simple and Effective Scaffold Token RemovalTechnical
- Scalable Acceleration for Classification-Based Derivative-Free OptimizationTechnical
- Scalable Decentralized Algorithms for Online Personalized Mean EstimationTechnical
- Scalable Federated One-Step Multi-View Clustering with Tensorized RegularizationTechnical
- Scalable Hierarchical Deep Neural Network for Time Series Analysis in Wearable Sensor-based Human Activity RecognitionTechnical
- Scalable Knowledge Refactoring Using Constrained OptimisationTechnical
- Scalable One-Pass Incomplete Multi-View Clustering by Aligning AnchorsTechnical
- Scalable Solutions for Decision-Making Systems Using Explainable Policy RepresentationsTechnical
- Scalable Trajectory-User Linking with Dual-Stream Representation NetworksTechnical
- Scalable Vision-Language Understanding and GenerationTechnical
- Scalable and Efficient Probabilistic Inference for Bayesian Deep Learning and Generative ModelingTechnical
- Scalable and Trustworthy Learning in Heterogeneous NetworksTechnical
- Scalable, Sustainable, Generalizable, and Responsible AI for Public SectorTechnical
- ScaleMatch: Multi-scale Consistency Enhancement for Semi-supervised Semantic SegmentationTechnical
- Scaling Diffusion Mamba with Bidirectional SSMs for Efficient 3D Shape GenerationTechnical
- Scaling Effects on Latent Representation Edits in GPT Models (Student Abstract)Technical
- Scaling Trends for Data Poisoning in LLMsTechnical
- ScamNet: Toward Explainable Large Language Model-Based Fraudulent Shopping Website DetectionTechnical
- Scenario-Based Robust Optimization of Tree StructuresTechnical
- Scene Graph-Grounded Image GenerationTechnical
- ScholarGEC: Enhancing Controllability of Large Language Model for Chinese Academic Grammatical Error CorrectionTechnical
- ScoreNet: Consistency-driven Framework with Multi-side Information Fusion for Session-based RecommendationTechnical
- ScreenMark: Watermarking Arbitrary Visual Content on ScreenTechnical
- SdalsNet: Self-Distilled Attention Localization and Shift Network for Unsupervised Camouflaged Object DetectionTechnical
- Search Strategy Generation for Branch and Bound Using Genetic ProgrammingTechnical
- Searching for Unfairness in Algorithms’ Outputs: Novel Tests and InsightsTechnical
- Searching for and Avoiding Hidden Sets Using Queries with Local FeedbackTechnical
- See Through Their Minds: Learning Transferable Brain Decoding Models from Cross-Subject fMRITechnical
- SeeDiff: Off-the-Shelf Seeded Mask Generation from Diffusion ModelsTechnical
- Seeing Beyond Noise: Joint Graph Structure Evaluation and Denoising for Multimodal RecommendationTechnical
- Seg2Box: 3D Object Detection by Point-Wise Semantics SupervisionTechnical
- Self-Attentive Spatio-Temporal Calibration for Precise Intermediate Layer Matching in ANN-to-SNN DistillationTechnical
- Self-Correcting Robot Manipulation via Gaussian-Splatted ForesightTechnical
- Self-Explainable Graph Transformer for Link Sign PredictionTechnical
- Self-Prompting Analogical Reasoning for UAV Object DetectionTechnical
- Self-Supervised Collaborative Information Bottleneck for Text Readability AssessmentTechnical
- Self-attention-based Diffusion Model for Time-series Imputation in Partial Blackout ScenariosTechnical
- Self-supervised Trusted Contrastive Multi-view Clustering with Uncertainty RefinedTechnical
- SemStereo: Semantic-Constrained Stereo Matching Network for Remote SensingTechnical
- Semantic Ambiguity Modeling and Propagation for Fine-Grained Visual Cross View Geo-LocalizationTechnical
- Semantic Enhanced Heterogeneous Hypergraph Network for Collaborative FilteringTechnical
- Semantic Segmentation on Raindrop Degraded Images Using Two-Stage Dual Teacher-Student LearningTechnical
- Semantic-guided Masked Mutual Learning for Multi-modal Brain Tumor Segmentation with Arbitrary Missing ModalitiesTechnical
- Semi-IIN: Semi-Supervised Intra-Inter Modal Interaction Learning Network for Multimodal Sentiment AnalysisTechnical
- Semi-Markovian Planning to Coordinate Aerial and Maritime Medical Evacuation PlatformsTechnical
- Semi-Supervised Clustering Framework for Fine-grained Scene Graph GenerationTechnical
- Semi-Supervised Multi-View Multi-Label Learning with View-Specific Transformer and Enhanced Pseudo-LabelTechnical
- Semi-Supervised Multimodal Classification Through Learning from Modal and Strategic ComplementaritiesTechnical
- Semi-Supervised Online Cross-Modal HashingTechnical
- Semi-supervised 3D Semantic Scene Completion with 2D Vision Foundation Model GuidanceTechnical
- Semi-supervised Infrared Small Target Detection with Thermodynamic-Inspired Uneven Perturbation and Confidence AdaptationTechnical
- SemiDFL: A Semi-Supervised Paradigm for Decentralized Federated LearningTechnical
- Sentence-level Aggregation of Lexical Metrics Correlates Stronger with Human Judgements than Corpus-level AggregationTechnical
- Separating the Wheat from the Chaff: Spatio-Temporal Transformer with View-interweaved Attention for Photon-Efficient Depth SensingTechnical
- Sequence Accumulation and Beyond: Infinite Context Length on Single GPU and Large ClustersTechnical
- Sequence Complementor: Complementing Transformers for Time Series Forecasting with Learnable SequencesTechnical
- Sequential Joint Dependency Aware Human Pose Estimation with State Space ModelTechnical
- Sequential Order Adjustment of Action Decisions for Multi-Agent Transformer (Student Abstract)Technical
- Sequential Preference Optimization: Multi-Dimensional Preference Alignment with Implicit Reward ModelingTechnical
- Set-Valued Sensitivity Analysis of Deep Neural NetworksTechnical
- Shaping AI Interest in Rural Middle Schools with Unplugged Learning: Gender Differences and Teacher InsightsTechnical
- Sharper Error Bounds in Late Fusion Multi-view Clustering with Eigenvalue Proportion OptimizationTechnical
- ShotVL: Human-Centric Highlight Frame Retrieval via Language QueriesTechnical
- Sim4Rec: Data-Free Model Extraction Attack on Sequential RecommendationTechnical
- Sim911: Towards Effective and Equitable 9-1-1 Dispatcher Training with an LLM-Enabled SimulationTechnical
- SimProF: A Simple Probabilistic Framework for Unsupervised Domain AdaptationTechnical
- SimRP: Syntactic and Semantic Similarity Retrieval Prompting Enhances Aspect Sentiment Quad PredictionTechnical
- Similar Modality Enhancement and Action Consistency Learning for Weakly Supervised Temporal Action LocalizationTechnical
- Simplifying Control Mechanism in Text-to-Image Diffusion ModelsTechnical
- Single-Loop Federated Actor-Critic across Heterogeneous EnvironmentsTechnical
- Single-View Graph Contrastive Learning with Soft Neighborhood AwarenessTechnical
- Single-view Image to Novel-view Generation for Hand-Object InteractionsTechnical
- Singular Value Scaling: Efficient Generative Model Compression via Pruned Weights RefinementTechnical
- Situation Calculus Temporally Lifted Abstractions for Generalized PlanningTechnical
- Skill Disentanglement in Reproducing Kernel Hilbert SpaceTechnical
- Skip Mamba Diffusion for Monocular 3D Semantic Scene CompletionTechnical
- SkipPool: Improved Sparse Hierarchical Graph Pooling with Differentiable ExplorationTechnical
- Slice-and-Pack: Tailoring Deep Models for Customized RequirementsTechnical
- Small Language Model Makes an Effective Long Text ExtractorTechnical
- Smart Motor: A Low-Cost Hardware and Software Toolkit for Introducing Supervised Machine Learning to Elementary School StudentsTechnical
- SoLA: Leveraging Soft Activation Sparsity and Low-Rank Decomposition for Large Language Model CompressionTechnical
- Social Recommendation via Graph-Level Counterfactual AugmentationTechnical
- SocialSim: Towards Socialized Simulation of Emotional Support ConversationTechnical
- Solving Epistemic Logic Programs Using Generate-and-Test with PropagationTechnical
- Solving Higher-Order Quantified Boolean Satisfiability via Higher-Order Model CheckingTechnical
- Solving Multiagent Path Finding on Highly Centralized NetworksTechnical
- SongEditor: Adapting Zero-Shot Song Generation Language Model as a Multi-Task EditorTechnical
- SongSong: A Time Phonograph for Chinese SongCi Music from Thousand of Years AwayTechnical
- Sound Over-Approximation of Equational Reasoning with Variable-Preserving Rules Parameterized by Derivation DepthTechnical
- SoundBrush: Sound as a Brush for Visual Scene EditingTechnical
- Sp3ctralMamba: Physics-Driven Joint State Space Model for Hyperspectral Image ReconstructionTechnical
- Sparse Transfer Learning Accelerates and Enhances Certified Robustness: A Comprehensive StudyTechnical
- Spatial Annealing for Efficient Few-shot Neural RenderingTechnical
- Spatial Clustering of Citizen Science Data Improves Downstream Species Distribution ModelsTechnical
- Spatial-Temporal Heterogenous Graph Contrastive Learning for Microservice Workload PredictionTechnical
- Spatial-Temporal Knowledge Distillation for Takeaway RecommendationTechnical
- Spatiotemporal-Aware Neural Fields for Dynamic CT ReconstructionTechnical
- Spatiotemporal-aware Trend-Seasonality Decomposition Network for Traffic Flow ForecastingTechnical
- SpeHeaTal: A Cluster-Enhanced Segmentation Method for Sperm Morphology AnalysisTechnical
- Specifying What You Know or Not for Multi-Label Class-Incremental LearningTechnical
- Spectra of Cardinality Queries over Description Logic Knowledge BasesTechnical
- Speech Recognition Meets Large Language Model: Benchmarking, Models, and ExplorationTechnical
- Speed Master: Quick or Slow Play to Attack Speaker RecognitionTechnical
- SpikeGS: Reconstruct 3D Scene Captured by a Fast-Moving Bio-Inspired CameraTechnical
- SpikingYOLOX: Improved YOLOX Object Detection with Fast Fourier Convolution and Spiking Neural NetworksTechnical
- Spin: Diffusion-based Semantic Image Painting Through Independent Information InjectionTechnical
- SpotDiff: Spatial Gene Expression Imputation Diffusion with Single-Cell RNA Sequencing Data IntegrationTechnical
- Spurious Feature Eraser: Stabilizing Test-Time Adaptation for Vision-Language Foundation ModelTechnical
- SrSv: Integrating Sequential Rollouts with Sequential Value Estimation for Multi-agent Reinforcement LearningTechnical
- Stability and Generalization of Zeroth-Order Decentralized Stochastic Gradient Descent with Changing TopologyTechnical
- Stability-based Generalization Analysis of Randomized Coordinate Descent for Pairwise LearningTechnical
- State Encodings for GNN-Based Lifted PlannersTechnical
- State-Based Disassembly PlanningTechnical
- Statistical Model-driven Similarity Hashing: Bridging Modalities for Efficient Unsupervised RetrievalTechnical
- Stitch, Contrast, and Segment: Learning a Human Action Segmentation Model Using Trimmed Skeleton VideosTechnical
- Stop Diverse OOD Attacks: Knowledge Ensemble for Reliable DefenseTechnical
- Strategic Manipulation in Temporal Voting with Undesirable Candidates (Student Abstract)Technical
- Strategic Network Creation for Enabling Greedy RoutingTechnical
- Stream Aligner: Efficient Sentence-Level Alignment via Distribution InductionTechnical
- Stress-Testing of Multimodal Models in Medical Image-Based Report GenerationTechnical
- Strong Empowered and Aligned Weak Mastered Annotation for Weak-to-Strong GeneralizationTechnical
- StructSR: Refuse Spurious Details in Real-World Image Super-ResolutionTechnical
- Structural Entropy Guided Unsupervised Graph Out-Of-Distribution DetectionTechnical
- Structural Pruning via Spatial-aware Information Redundancy for Semantic SegmentationTechnical
- Structure Balance and Gradient Matching-Based Signed Graph CondensationTechnical
- Structure-Adaptive Multi-View Graph Clustering for Remote Sensing DataTechnical
- Structured Document Generation for Industrial EquipmentTechnical
- Structured IB: Improving Information Bottleneck with Structured Feature LearningTechnical
- Style Nursing with Spatial and Semantic Guidance for Zero-Shot Traffic Scene Style TransferTechnical
- SuBiTO: Synopsis-based Training Optimization for Continuous Real-Time Neural Learning over Big Streaming DataTechnical
- Sub-Interest-Aware Representation Uniformity for Recommender SystemTechnical
- Subgraph Aggregation for Out-of-Distribution Generalization on GraphsTechnical
- Subgraph Invariant Learning Towards Large-Scale Graph Node ClassificationTechnical
- Suboptimal Search with Dynamic Distribution of SuboptimalityTechnical
- SummPilot: Bridging Efficiency and Customization for Interactive Summarization SystemTechnical
- Super-Class Guided Transformer for Zero-Shot Attribute ClassificationTechnical
- Supervised Score-Based Modeling by Gradient BoostingTechnical
- Support Vector-based Estimation of Multilinear Games for Feature Selection and ExplanationTechnical
- Supporting AI Literacy Teaching Through the Development of Assessments for Classroom UseTechnical
- Supportive Negatives Spectral Augmentation for Source-Free Cross-Domain SegmentationTechnical
- SwiftTry: Fast and Consistent Video Virtual Try-On with Diffusion ModelsTechnical
- Symbolic Functional Decomposition: A Reconfiguration ApproachTechnical
- Symbolic Neural Ordinary Differential EquationsTechnical
- SymmCompletion: High-Fidelity and High-Consistency Point Cloud Completion with Symmetry GuidanceTechnical
- SyncNoise: Geometrically Consistent Noise Prediction for Instruction-based 3D EditingTechnical
- Synchronization in Learning in Periodic Zero-Sum Games Triggers Divergence from Nash EquilibriumTechnical
- Synergy of GFlowNet and Protein Language Model Makes a Diverse Antibody DesignerTechnical
- S²DN: Learning to Denoise Unconvincing Knowledge for Inductive Knowledge Graph CompletionTechnical
- S²MILE: Semantic-and-Structure-Aware Music-Driven Lyric GenerationTechnical
- T-MDML: Triplet-based Multiple Distance Metric Learning for Multi-Instance Multi-Label Classification with Label CorrelationTechnical
- TAIL-MIL: Time-Aware and Instance-Learnable Multiple Instance Learning for Multivariate Time Series Anomaly DetectionTechnical
- TAMER: Tree-Aware Transformer for Handwritten Mathematical Expression RecognitionTechnical
- TB-HSU: Hierarchical 3D Scene Understanding with Contextual AffordancesTechnical
- TC-Diffuser: Bi-Condition Multi-Modal Diffusion for Tropical Cyclone ForecastingTechnical
- TC-LLaVA: Rethinking the Transfer of LLava from Image to Video Understanding with Temporal ConsiderationsTechnical
- TCAM-Diff: Triplane-Aware Cross-Attention Medical Diffusion ModelTechnical
- TGBFormer: Transformer-GraphFormer Blender Network for Video Object DetectionTechnical
- TGFormer: Transformer with Track Query Group for Multi-Object TrackingTechnical
- TGLsta: Low-resource Textual Graph Learning with Semantic and Topological Awareness via LLMsTechnical
- THESAURUS: Contrastive Graph Clustering by Swapping Fused Gromov-Wasserstein CouplingsTechnical
- THGNets: Constrained Temporal Hypergraphs and Graph Neural Networks in Hyperbolic Space for Information Diffusion PredictionTechnical
- TIME-FS: Joint Learning of Tensorial Incomplete Multi-View Unsupervised Feature Selection and Missing-View ImputationTechnical
- TNCSE: Tensor Norm Constraints for Unsupervised Contrastive Learning of Sentence EmbeddingsTechnical
- TRACI: A Data-centric Approach for Multi-Domain Generalization on GraphsTechnical
- TRAIL: Trust-Aware Client Scheduling for Semi-Decentralized Federated LearningTechnical
- TSDF-Based Efficient Motion-Compensated Temporal Interpolation for 3D Dynamic SequencesTechnical
- TSGAN: Temporal Social Graph Attention Network for Aggressive Behavior ForecastingTechnical
- TSVC: Tripartite Learning with Semantic Variation Consistency for Robust Image-Text RetrievalTechnical
- TTA-FedDG: Leveraging Test-Time Adaptation to Address Federated Domain GeneralizationTechnical
- TTE: Two Tokens Are Enough to Improve Parameter-Efficient TuningTechnical
- Tab-Shapley: Identifying Top-k Tabular Data Quality InsightsTechnical
- TabGLM: Tabular Graph Language Model for Learning Transferable Representations Through Multi-Modal Consistency MinimizationTechnical
- Tackling Intertwined Data and Device Heterogeneities in Federated Learning with Unlimited StalenessTechnical
- Target Scanpath-Guided 360-Degree Image EnhancementTechnical
- Target Semantics Clustering via Text Representations for Robust Universal Domain AdaptationTechnical
- Task-Agnostic Language Model Watermarking via High Entropy Passthrough LayersTechnical
- Task-Specific Preconditioner for Cross-Domain Few-Shot LearningTechnical
- Task-level Distributionally Robust Optimization for Large Language Model-based Dense RetrievalTechnical
- TdAttenMix: Top-Down Attention Guided MixupTechnical
- Teacher-guided Edge Discriminator for Personalized Graph Masked AutoencoderTechnical
- Teaching Models to Improve on TapeTechnical
- TechSinger: Technique Controllable Multilingual Singing Voice Synthesis via Flow MatchingTechnical
- Template-Driven LLM-Paraphrased Framework for Tabular Math Word Problem GenerationTechnical
- Temporal Action Localization with Cross Layer Task Decoupling and RefinementTechnical
- Temporal Causal Reasoning with (Non-Recursive) Structural Equation ModelsTechnical
- Temporal Coherent Object Flow for Multi-Object TrackingTechnical
- Temporal Conjunctive Query Answering via RewritingTechnical
- Temporal Fair DivisionTechnical
- Temporal Numeric Planning with PatternsTechnical
- Temporal Specification Optimisation for the Event CalculusTechnical
- Temporal Streaming Batch Principal Component Analysis for Time Series Classification (Student Abstract)Technical
- Temporal Task and Motion Planning with Metric Time for Multiple Object NavigationTechnical
- Temporal Triadic Closure: Finding Dense Substructures in Social Networks That Evolve over TimeTechnical
- Temporal-Aware Evaluation and Learning for Temporal Graph Neural NetworksTechnical
- Tensor Decomposition Meets Knowledge Compilation: A Study Comparing Tensor Trains with OBDDsTechnical
- Tensorized Attention for Understanding Multi-Object RelationshipsTechnical
- Tensorized Label Learning Based Fast Fuzzy ClusteringTechnical
- Testing Causal Models with Hidden Variables in Polynomial Delay via Conditional IndependenciesTechnical
- Text and Image Are Mutually Beneficial: Enhancing Training-Free Few-Shot Classification with CLIPTechnical
- Text to Point Cloud Localization with Multi-Level Negative Contrastive LearningTechnical
- Text-Diffusion Red-Teaming of Large Language Models: Unveiling Harmful Behaviors with Proximity ConstraintsTechnical
- Text-Guided Fine-grained Counterfactual Inference for Short Video Fake News DetectionTechnical
- Text-Guided Nonverbal Enhancement Based on Modality-Invariant and -Specific Representations for Video Speaking Style RecognitionTechnical
- TextRefiner: Internal Visual Feature as Efficient Refiner for Vision-Language Models Prompt TuningTechnical
- Textualize Visual Prompt for Image Editing via Diffusion BridgeTechnical
- The (Exact) Price of Cardinality for Indivisible Goods: A Parametric PerspectiveTechnical
- The Adaptive Q-Network for Recommendation Tasks with Dynamic Item SpaceTechnical
- The Complexity of Extending Fair Allocations of Indivisible GoodsTechnical
- The Distributional Reward Critic Framework for Reinforcement Learning Under Perturbed RewardsTechnical
- The Dynamic Duo of Collaborative Masking and Target for Advanced Masked Autoencoder LearningTechnical
- The Essentials of AI for Life and Society: An AI Literacy Course for the University CommunityTechnical
- The Gradient of Algebraic Model CountingTechnical
- The Impact of Literal Sorting on Cardinality Constraint EncodingsTechnical
- The Indoor-Training Effect: Unexpected Gains from Distribution Shifts in the Transition FunctionTechnical
- The Mainstays of Trustworthy Machine LearningTechnical
- The Master Key Filters Hypothesis: Deep Filters Are GeneralTechnical
- The POWER of Ikigai: Optimizing Life Fulfillment with an Integrated User Simulator and Adaptive Hobby RecommenderTechnical
- The Parables of the Mustard Seed and the Yeast: Extremely Low-Budget, High-Performance Nighttime Semantic SegmentationTechnical
- The Pitfalls of “Security by Obscurity” and What They Mean for Transparent AITechnical
- The Surprising Effectiveness of Infinite-Width NTKs for Characterizing and Improving Model TrainingTechnical
- The VOROS: Lifting ROC Curves to 3D to Summarize Unbalanced Classifier PerformanceTechnical
- The Value of Recall in Extensive-Form GamesTechnical
- Thermal-Aware Low-Light Image Enhancement: A Real-World Benchmark and a New Light-Weight ModelTechnical
- Things Machine Learning Models Know That They Don’t KnowTechnical
- Thinking Racial Bias in Fair Forgery Detection: Models, Datasets and EvaluationsTechnical
- Thinking in Granularity: Dynamic Quantization for Image Super-Resolution by Intriguing Multi-Granularity CluesTechnical
- Threshold UCT: Cost-Constrained Monte Carlo Tree Search with Pareto CurvesTechnical
- Tilted Quantile Gradient Updates for Quantile-Constrained Reinforcement LearningTechnical
- TimeCHEAT: A Channel Harmony Strategy for Irregularly Sampled Multivariate Time Series AnalysisTechnical
- Times2D: Multi-Period Decomposition and Derivative Mapping for General Time Series ForecastingTechnical
- TinyFoA: Memory Efficient Forward-Only Algorithm for On-Device LearningTechnical
- TinySubNets: An Efficient and Low Capacity Continual Learning StrategyTechnical
- To Measure or Not: A Cost-Sensitive, Selective Measuring Environment for Agricultural Management Decisions with Reinforcement LearningTechnical
- To Predict or Not to Predict? Proportionally Masked Autoencoders for Tabular Data ImputationTechnical
- ToMATO: Verbalizing the Mental States of Role-Playing LLMs for Benchmarking Theory of MindTechnical
- TokenMatcher: Diverse Tokens Matching for Unsupervised Visible-Infrared Person Re-IdentificationTechnical
- Tokenization, Fusion, and Augmentation: Towards Fine-grained Multi-modal Entity RepresentationTechnical
- Tokenphormer: Structure-aware Multi-token Graph Transformer for Node ClassificationTechnical
- Told You That Will Not Work: Optimal Corrections to Planning Domains Using Counter-Example PlansTechnical
- Top-one Recommendation with Anonymous User Behaviors (Student Abstract)Technical
- Topo2Seq: Enhanced Topology Reasoning via Topology Sequence LearningTechnical
- Topology-Aware 3D Gaussian Splatting: Leveraging Persistent Homology for Optimized Structural IntegrityTechnical
- Toward Causal Generative Modeling: From Representation to GenerationTechnical
- Toward Efficient Data-Free UnlearningTechnical
- Toward Improving Robustness and Accuracy in Unsupervised Domain AdaptationTechnical
- Toward Verifiable Instruction-Following Alignment for Retrieval Augmented GenerationTechnical
- Towards Accurate Binary Spiking Neural Networks: Learning with Adaptive Gradient Modulation MechanismTechnical
- Towards Addressing Frontiers in Graph GenerationTechnical
- Towards Audio-Visual Navigation in Noisy Environments: A Large-Scale Benchmark Dataset and an Architecture Considering Multiple Sound-SourcesTechnical
- Towards Autonomous Network Management: AI-Driven Framework for Intelligent Log Analysis, Troubleshooting and DocumentationTechnical
- Towards Better Robot Learners: Leveraging Implicit and Explicit Human Feedback Together in Human Robot InteractionsTechnical
- Towards Better Robustness Against Natural Corruptions in Document Tampering LocalizationTechnical
- Towards Better Spherical Sliced-Wasserstein Distance Learning with Data-Adaptive Discriminative Projection DirectionTechnical
- Towards Building Human-like Smart Agents in Modern 3D Video Games (Student Abstract)Technical
- Towards Computational ForeseeabilityTechnical
- Towards Efficient Low-Order Hybrid Optimizer for Language Model Fine-TuningTechnical
- Towards Efficient Object Re-Identification with a Novel Cloud-Edge Collaborative FrameworkTechnical
- Towards Efficient and Intelligent Laser Weeding: Method and Dataset for Weed Stem DetectionTechnical
- Towards Enhancing Road Safety in South Carolina Using Insights from Traffic and Driver-Education Data (Student Abstract)Technical
- Towards Global-Topology Relation Graph for Inductive Knowledge Graph CompletionTechnical
- Towards Learnable Anchor for Deep Multi-View ClusteringTechnical
- Towards Loss-Resilient Image Coding for Unstable Satellite NetworksTechnical
- Towards Macro-AUC Oriented Imbalanced Multi-Label Continual LearningTechnical
- Towards More Discriminative Feature Learning in SNNs with Temporal-Self-Erasing SupervisionTechnical
- Towards Multimodal Sentiment Analysis via Hierarchical Correlation Modeling with Semantic Distribution ConstraintsTechnical
- Towards Practical Classical Planning Compilations of Numeric PlanningTechnical
- Towards Precise Prediction Uncertainty in GNNs: Refining GNNs with Topology-grouping StrategyTechnical
- Towards Projected and Incremental Pseudo-Boolean Model CountingTechnical
- Towards Real-Time Approximate CountingTechnical
- Towards Realistic Semi-supervised Medical Image ClassificationTechnical
- Towards Robust Visual Question Answering via Prompt-Driven Geometric HarmonizationTechnical
- Towards Robust, Efficient, and Practical Decision-Making: From Reward-Maximizing Deep Reinforcement Learning to Reward-Matching GFlowNetsTechnical
- Towards Runtime Analysis of Population-Based Co-evolutionary Algorithms on Sparse Binary Zero-Sum GameTechnical
- Towards Scalable and Deep Graph Neural Networks via Noise MaskingTechnical
- Towards Ship License Plate Recognition in the Wild: A Large Benchmark and Strong BaselineTechnical
- Towards S²-Challenges Underlying LLM-Based Augmentation for Personalized News RecommendationTechnical
- Towards Trustworthy Machine Learning Under Distribution ShiftsTechnical
- Towards Trustworthy, Efficient, and Scalable Machine LearningTechnical
- Towards Unbiased Information Extraction and Adaptation in Cross-Domain RecommendationTechnical
- Towards Universal Rainy Image Restoration: Benchmark and BaselineTechnical
- Towards Verifiable Text Generation with Generative AgentTechnical
- Towards a Multimodal Large Language Model with Pixel-Level Insight for BiomedicineTechnical
- Towards an AI Course Based on Neural NetworksTechnical
- Toy-GS: Assembling Local Gaussians for Precisely Rendering Large-Scale Free Camera TrajectoriesTechnical
- Tracking Everything Everywhere across Multiple CamerasTechnical
- Tracking and Identifying International Propaganda and Influence Networks OnlineTechnical
- Trade-Offs Between Information and Crowding in Sequential Decisions (Student Abstract)Technical
- Trading Off Quality and Uncertainty Through Multi-Objective Optimisation in Batch Bayesian OptimisationTechnical
- Tradutor: Building a Variety Specific Translation ModelTechnical
- Traffic Scenario Logic: A Spatial-Temporal Logic for Modeling and Reasoning of Urban Traffic ScenariosTechnical
- Training Consistent Mixture-of-Experts-Based Prompt Generator for Continual LearningTechnical
- Training Deep Neural Networks with Virtual Smoothing ClassesTechnical
- Training Matting Models Without Alpha LabelsTechnical
- Training Recurrent Neural Networks with Inherent Missing Data for Wearable Device Applications (Student Abstract)Technical
- Training Verification-Friendly Neural Networks via Neuron Behavior ConsistencyTechnical
- Training-and-Prompt-Free General Painterly Harmonization via Zero-Shot Disentenglement on Style and Content ReferencesTechnical
- Training-free Open-Vocabulary Semantic Segmentation via Diverse Prototype Construction and Sub-region MatchingTechnical
- Transfer Learning Meets Functional Linear Regression: No Negative Transfer Under Posterior DriftTechnical
- Transfer Learning in Financial Time Series with Gramian Angular Field (Student Abstract)Technical
- Transfer Learning of Real Image Features with Soft Contrastive Loss for Fake Image DetectionTechnical
- Transforming Healthcare Decision Making Using Artificial IntelligenceTechnical
- Treasures in Discarded Weights for LLM QuantizationTechnical
- TreeEval: Benchmark-Free Evaluation of Large Language Models through Tree PlanningTechnical
- TrojanDec: Data-free Detection of Trojan Inputs in Self-supervised LearningTechnical
- Truncated Gaussian Policy for Debiased Continuous ControlTechnical
- Trust-GRS: A Trustworthy Training Framework for Graph Neural Network Based Recommender Systems Against Shilling AttacksTechnical
- Trustworthy AI Meets Educational Assessment: Challenges and OpportunitiesTechnical
- Trustworthy and Practical AI for Healthcare: A Guided Deferral System with Large Language ModelsTechnical
- Truth Behind the Scene: Designing Evaluations Benchmarks to Assess LLMs’ Task-Specific Understanding over Test-Taking StrategiesTechnical
- Two Sides of the Same Coin: Learning the Backdoor to Remove the BackdoorTechnical
- Two-Timescale Critic-Actor for Average Reward MDPs with Function ApproximationTechnical
- Two-stream Beats One-stream: Asymmetric Siamese Network for Efficient Visual TrackingTechnical
- Twofold Debiasing Enhances Fine-Grained Learning with Coarse LabelsTechnical
- UACOF: A USV-AUV Collaboration Framework for Underwater Tasks Under Extreme Sea Conditions (Student Abstract)Technical
- UAWTrack: Universal 3D Single Object Tracking in Adverse WeatherTechnical
- UCF-Crime-DVS: A Novel Event-Based Dataset for Video Anomaly Detection with Spiking Neural NetworksTechnical
- UFID: A Unified Framework for Black-box Input-level Backdoor Detection on Diffusion ModelsTechnical
- UFO: Enhancing Diffusion-Based Video Generation with a Uniform Frame OrganizerTechnical
- UN-DETR: Promoting Objectness Learning via Joint Supervision for Unknown Object DetectionTechnical
- UP-Restorer: When Unrolling Meets Prompts for Unified Image RestorationTechnical
- Ultra-High-Definition Dynamic Multi-Exposure Image Fusion via Infinite Pixel LearningTechnical
- Unaligned Message-Passing and Contextualized-Pretraining for Robust Geo-Entity ResolutionTechnical
- Uncertainty-Aware Contrastive Learning with Hard Negative Sampling for Code Search TasksTechnical
- Uncertainty-Aware Global-View Reconstruction for Multi-View Multi-Label Feature SelectionTechnical
- Uncertainty-Aware Self-Training for CTC-Based Automatic Speech RecognitionTechnical
- Uncertainty-aware Knowledge TracingTechnical
- Uncommon Belief in RationalityTechnical
- Understanding AdvertisementsTechnical
- Understanding Annotator Perception: Modeling Psychological Inference from First- and Third-Person Annotations (Student Abstract)Technical
- Understanding EFX Allocations: Counting and VariantsTechnical
- Understanding GenAI for Teaching and Learning in Secondary ClassroomsTechnical
- Understanding Individual Agent Importance in Multi-Agent System via Counterfactual ReasoningTechnical
- Understanding K-12 Teachers’ Needs for AI Education: A Survey-Based StudyTechnical
- Understanding Microtargeting Pattern on Social MediaTechnical
- Understanding Unique Behavioral Patterns through Multimodal Analysis of Eye-Hand Coordination in Autistic Children (Student Abstract)Technical
- UniFORM: Towards Unified Framework for Anomaly Detection on GraphsTechnical
- UniPCGC: Towards Practical Point Cloud Geometry Compression via an Efficient Unified ApproachTechnical
- UniTR: A Unified Framework for Joint Representation Learning of Trajectories and Road NetworksTechnical
- Unified Coding for Both Human Perception and Generalized Machine Analytics with CLIP SupervisionTechnical
- Unified Graph Neural Networks Pre-training for Multi-domain GraphsTechnical
- Unified Knowledge Maintenance Pruning and Progressive Recovery with Weight Recalling for Large Vision-Language ModelsTechnical
- Union Is Strength! Unite the Power of LLMs and MLLMs for Chart Question AnsweringTechnical
- Universal Domain Adaptive Object Detection via Dual Probabilistic AlignmentTechnical
- Universal Features Guided Zero-Shot Category-Level Object Pose EstimationTechnical
- Universal Post-Processing Networks for Joint Optimization of Modules in Task-Oriented Dialogue SystemsTechnical
- Unleashing the Potential of Model Bias for Generalized Category DiscoveryTechnical
- Unleashing the Power of Visual Foundation Models for Generalizable Semantic SegmentationTechnical
- Unlocking Better Closed-Set Alignment Based on Neural Collapse for Open-Set RecognitionTechnical
- Unlocking the Game: Estimating Games in Möbius Representation for Explanation and High-Order Interaction DetectionTechnical
- Unlocking the Potential of Black-box Pre-trained GNNs for Graph Few-shot LearningTechnical
- Unlocking the Power of Patch: Patch-Based MLP for Long-Term Time Series ForecastingTechnical
- Unraveling the Influence of Training Data and Internal Structures in Large Language Models for Enhanced Explainability (Student Abstract)Technical
- Unravelling Causal Genetic Biomarkers of Alzheimer’s Disease via Neuron to Gene-token Backtracking in Neural Architecture: A Groundbreaking Reverse-Gene-Finder ApproachTechnical
- Unsupervised Anomaly Detection for Tabular Data Using Deep Noise EvaluationTechnical
- Unsupervised Degradation Representation Aware Transform for Real-World Blind Image Super-ResolutionTechnical
- Unsupervised Diffusion-Based Degradation Modeling for Real-World Super-ResolutionTechnical
- Unsupervised Domain Adaptive Person Search via Dual Self-CalibrationTechnical
- Unsupervised Kernel-based Multi-view Feature Selection with Robust Self-representation and Binary HashingTechnical
- Unsupervised Photometric-Consistent Depth Estimation from Endoscopic Monocular VideoTechnical
- Unsupervised Region-Based Image Editing of Denoising Diffusion ModelsTechnical
- Unsupervised Self-Prior Embedding Neural Representation for Iterative Sparse-View CT ReconstructionTechnical
- Unsupervised Translation of Emergent CommunicationTechnical
- Unveiling Multi-View Anomaly Detection: Intra-view Decoupling and Inter-view FusionTechnical
- Unveiling the Knowledge of CLIP for Training-Free Open-Vocabulary Semantic SegmentationTechnical
- UrbanWaste: In-the-Bin Dataset for Waste Disposal Inspection with Multi-Granularity Hierarchical LabelsTechnical
- User Preference Meets Pareto-Optimality in Multi-Objective Bayesian OptimizationTechnical
- Using Machine Learning to Improve Research in the Agriculture IndustryTechnical
- Using Next Sentence Prediction to Test ChatGPT’s Text Comprehension (Student Abstract)Technical
- Utilizing Vision-Language Models for Detection of Leaf-Based Diseases in TomatoesTechnical
- Utterance-level Emotion Recognition in Conversation with Conversation-level SupervisionTechnical
- V2C-CBM: Building Concept Bottlenecks with Vision-to-Concept TokenizerTechnical
- VA-AR: Learning Velocity-Aware Action Representations with Mixture of Window AttentionTechnical
- VCR: A “Cone of Experience” Driven Synthetic Data Generation Framework for Mathematical ReasoningTechnical
- VEGAS: Towards Visually Explainable and Grounded Artificial Social IntelligenceTechnical
- VERO: Verification and Zero-Shot Feedback Acquisition for Few-Shot Multimodal Aspect-Level Sentiment ClassificationTechnical
- VERSE: Verification-based Self-Play for Code InstructionsTechnical
- VFM-Adapter: Adapting Visual Foundation Models for Dense Prediction with Dynamic Hybrid Operation MappingTechnical
- VG-TVP: Multimodal Procedural Planning via Visually Grounded Text-Video PromptingTechnical
- VLScene: Vision-Language Guidance Distillation for Camera-Based 3D Semantic Scene CompletionTechnical
- VOILA: Complexity-Aware Universal Segmentation of CT Images by Voxel Interacting with LanguageTechnical
- VQLTI: Long-Term Tropical Cyclone Intensity Forecasting with Physical ConstraintsTechnical
- VQTalker: Towards Multilingual Talking Avatars Through Facial Motion TokenizationTechnical
- VRVVC: Variable-Rate NeRF-Based Volumetric Video CompressionTechnical
- VVRec: Reconstruction Attacks on DL-based Volumetric Video Upstreaming via Latent Diffusion Model with Gamma DistributionTechnical
- VarCMP: Adapting Cross-Modal Pre-Training Models for Video Anomaly RetrievalTechnical
- Verification of Neural Networks Against Convolutional Perturbations via Parameterised KernelsTechnical
- VersaFusion: A Versatile Diffusion-Based Framework for Fine-Grained Image Editing and EnhancementTechnical
- VersaGen: Unleashing Versatile Visual Control for Text-to-Image SynthesisTechnical
- ViFactCheck: A New Benchmark Dataset and Methods for Multi-Domain News Fact-Checking In VietnameseTechnical
- VidChain: Chain-of-Tasks with Metric-based Direct Preference Optimization for Dense Video CaptioningTechnical
- VidEvent: A Large Dataset for Understanding Dynamic Evolution of Events in VideosTechnical
- VidSole: A Multimodal Dataset for Joint Kinetics Quantification and Disease Detection with Deep LearningTechnical
- Video Summarization Using Denoising Diffusion Probabilistic ModelTechnical
- Vietnamese Words Are Not Constructed from Syllables: Rethinking the Role of Word Segmentation in Natural Language Processing for Vietnamese TextsTechnical
- View Transformation Robustness for Multi-View 3D Object Reconstruction with Reconstruction Error-Guided View SelectionTechnical
- Virtual Nodes Can Help: Tackling Distribution Shifts in Federated Graph LearningTechnical
- VisRec: A Semi-Supervised Approach to Visibility Data Reconstruction in Radio AstronomyTechnical
- Vision Transformers Beat WideResNets on Small Scale Datasets Adversarial RobustnessTechnical
- Vision-Based Generic Potential Function for Policy Alignment in Multi-Agent Reinforcement LearningTechnical
- Vision-aware Multimodal Prompt Tuning for Uploadable Multi-source Few-shot Domain AdaptationTechnical
- Vision-guided Text Mining for Unsupervised Cross-modal Hashing with Community Similarity QuantizationTechnical
- Visual Perturbation for Text-Based Person SearchTechnical
- Visual Question Answering for Peruvian Cuisine in Regional SpanishTechnical
- Voter Priming Campaigns: Strategies, Equilibria, and AlgorithmsTechnical
- Vox-UDA: Voxel-wise Unsupervised Domain Adaptation for Cryo-Electron Subtomogram Segmentation with Denoised Pseudo-LabelingTechnical
- WHALE-FL: Wireless and Heterogeneity Aware Latency Efficient Federated Learning over Mobile Devices via Adaptive Subnetwork SchedulingTechnical
- WST: Wavelet-Based Multi-scale Tuning for Visual Transfer LearningTechnical
- Walking the Web of Concept-Class Relationships in Incrementally Trained Interpretable ModelsTechnical
- Wasserstein Distance Constraint and Parameter Sparsification for Batched and Iterative Knowledge EditingTechnical
- WatE: A Wasserstein t-distributed Embedding Method for Information-enriched Graph VisualizationTechnical
- Watch Out for Your Guidance on Generation! Exploring Conditional Backdoor Attacks against Large Language ModelsTechnical
- Watch Video, Catch Keyword: Context-aware Keyword Attention for Moment Retrieval and Highlight DetectionTechnical
- WaterDiffusion: Learning a Prior-involved Unrolling Diffusion for Joint Underwater Saliency Detection and Visual RestorationTechnical
- WaveLoss: An Adaptive Dynamic Loss for Deep Gait RecognitionTechnical
- WaveMixSR-V2: Enhancing Super-resolution with Higher Efficiency (Student Abstract)Technical
- Wavelet-Assisted Multi-Frequency Attention Network for PansharpeningTechnical
- Wavelet-Driven Masked Image Modeling: A Path to Efficient Visual RepresentationTechnical
- WaveletMixer: A Multi-Resolution Wavelets Based MLP-Mixer for Multivariate Long-Term Time Series ForecastingTechnical
- We Are AI: Taking Control of TechnologyTechnical
- Weak Strategyproofness in Randomized Social ChoiceTechnical
- Weakly Supervised Gland Segmentation with Class Semantic Consistency and Purified Labels FiltrationTechnical
- Weapon Activity RecognitionTechnical
- Weighted Embeddings for Low-Dimensional Graph RepresentationTechnical
- Weighted Poisson-disk Resampling on Large-Scale Point CloudsTechnical
- Welfare-Optimal Serial Dictatorships Have Polynomial Query ComplexityTechnical
- What Can Youth Learn About Artificial Intelligence and Machine Learning in One Hour? Examining How Hour of Code Activities Address the Five Big Ideas of AITechnical
- What Do Machine Learning Researchers Mean by “Reproducible”?Technical
- What Is a Good Question? Assessing Question Quality via Meta-Fact CheckingTechnical
- What to Preserve and What to Transfer: Faithful, Identity-Preserving Diffusion-based Hairstyle TransferTechnical
- When Neutral Summaries Are Not That Neutral: Quantifying Political Neutrality in LLM-Generated News Summaries (Student Abstract)Technical
- When Open-Vocabulary Visual Question Answering Meets Causal Adapter: Benchmark and ApproachTechnical
- When Shadow Removal Meets Intrinsic Image Decomposition: A Joint Learning Framework Using Unpaired DataTechnical
- When to Learn and When to Stop: Quitting at the Optimal Time (Student Abstract)Technical
- Where Precision Meets Efficiency: Transformation Diffusion Model for Point Cloud RegistrationTechnical
- Who’s the (Multi-)Fairest of Them All: Rethinking Interpolation-Based Data Augmentation Through the Lens of MulticalibrationTechnical
- WildFake: A Large-Scale and Hierarchical Dataset for AI-Generated Images DetectionTechnical
- Wills Aligner: Multi-Subject Collaborative Brain Visual DecodingTechnical
- Word2Vec4Kids: Interactive Challenges to Introduce Middle School Students to Word EmbeddingsTechnical
- XCotton: Advancing AI-Enabled Hardware/Software Integrated System for Foreign Fiber CleaningTechnical
- XTSFormer: Cross-Temporal-Scale Transformer for Irregular-Time Event Prediction in Clinical ApplicationsTechnical
- You Should Learn to Stop Denoising on Point Clouds in AdvanceTechnical
- Zero-Shot Conditioning of Score-Based Diffusion Models by Neuro-Symbolic ConstraintsTechnical
- Zero-Shot Image Captioning with Multi-type Entity RepresentationsTechnical
- Zero-Shot Learning for Materials Science Texts: Leveraging Duck Typing PrinciplesTechnical
- Zero-Shot Learning in Industrial Scenarios: New Large-Scale Benchmark, Challenges and BaselineTechnical
- Zero-Shot Low-Light Image Enhancement via Latent Diffusion ModelsTechnical
- Zero-Shot Noise2Mean: Gap Minimization for Efficient Denoising from a Single Noisy ImageTechnical
- Zero-shot Depth Completion via Test-time Alignment with Affine-invariant Depth PriorTechnical
- ZeroHAR: Sensor Context Augments Zero-Shot Wearable Action RecognitionTechnical
- Zeroth-Order Methods for Nonconvex Stochastic Problems with Decision-Dependent DistributionsTechnical
- ZoRI: Towards Discriminative Zero-Shot Remote Sensing Instance SegmentationTechnical
- dyAb: Flow Matching for Flexible Antibody Design with AlphaFold-driven Pre-binding AntigenTechnical
- iLLuMinaTE: An LLM-XAI Framework Leveraging Social Science Explanation Theories Towards Actionable Student Performance FeedbackTechnical
- iMoT: Inertial Motion Transformer for Inertial NavigationTechnical
- k-HyperEdge Medoids for Clustering EnsembleTechnical
- mRNA2vec: mRNA Embedding with Language Model in the 5'UTR-CDS for mRNA DesignTechnical
- ml4xcube: Machine Learning Toolkits for Earth System Data CubesTechnical
- mmFAS: Multimodal Face Anti-Spoofing Using Multi-Level Alignment and Switch-Attention FusionTechnical
- nach0-pc: Multi-task Language Model with Molecular Point Cloud EncoderTechnical
- pFedES: Generalized Proxy Feature Extractor Sharing for Model Heterogeneous Personalized Federated LearningTechnical
- pFedGPA: Diffusion-based Generative Parameter Aggregation for Personalized Federated LearningTechnical
- scMBERT: A Pre-Trained Deep Learning Model for Single-Cell Multiomic Data Representation and Prediction (Student Abstract)Technical
- xPatch: Dual-Stream Time Series Forecasting with Exponential Seasonal-Trend DecompositionTechnical
- “AlphAI”: Teaching AI Algorithms to K12 by Training Learning Robots and Visualizing How It WorksTechnical
AAAI accepted papers in other years
Looking for submission deadlines instead? See the conference deadline calendar.