ECCV 2024 Accepted Papers
The full list of 2,385 papers accepted at ECCV 2024 (European Conference on Computer Vision). Click any title for details, similar papers, and links to the original source. You can also search these papers by meaning, not just keywords.
Poster: 2,185Oral: 200
- Surface-Centric Modeling for High-Fidelity Generalizable Neural Surface ReconstructionPoster1 citations
- Syn-to-Real Domain Adaptation for Point Cloud Completion via Part-based ApproachPoster1 citations
- Sync from the Sea: Retrieving Alignable Videos from Large-Scale DatasetsOral1 citations
- Synchronous Diffusion for Unsupervised Smooth Non-Rigid 3D Shape MatchingPoster1 citations
- T-CorresNet: Template Guided 3D Point Cloud Completion with Correspondence Pooling Query Generation StrategyPoster1 citations
- TAG: Text Prompt Augmentation for Zero-Shot Out-of-Distribution DetectionPoster1 citations
- TCC-Det: Temporarily consistent cues for weakly-supervised 3D detectionPoster1 citations
- Temporal As a Plugin: Unsupervised Video Denoising with Pre-Trained Image DenoisersPoster1 citations
- Tensorial template matching for fast cross-correlation with rotations and its application for tomographyPoster1 citations
- Textual Grounding for Open-vocabulary Visual Information Extraction in Layout-diversified DocumentsPoster1 citations
- The Gaussian Discriminant Variational Autoencoder (GdVAE): A Self-Explainable Model with Counterfactual ExplanationsPoster1 citations
- The Role of Masking for Efficient Supervised Knowledge Distillation of Vision TransformersPoster1 citations
- Think before Placement: Common Sense Enhanced Transformer for Object PlacementPoster1 citations
- Tight and Efficient Upper Bound on Spectral Norm of Convolutional LayersPoster1 citations
- TimeLens-XL: Real-time Event-based Video Frame Interpolation with Large MotionPoster1 citations
- Topology-Preserving Downsampling of Binary ImagesPoster1 citations
- Towards High-Quality 3D Motion Transfer with Realistic Apparel AnimationPoster1 citations
- Towards Model-Agnostic Dataset Condensation by Heterogeneous ModelsOral1 citations
- Towards Reliable Evaluation and Fast Training of Robust Semantic Segmentation ModelsPoster1 citations
- Towards a Density Preserving Objective Function for Learning on Point SetsPoster1 citations
- TrafficNight : An Aerial Multimodal Benchmark For Nighttime Vehicle SurveillancePoster1 citations
- Training A Secure Model against Data-Free Model ExtractionPoster1 citations
- Training-free Composite Scene Generation for Layout-to-Image SynthesisPoster1 citations
- TrajPrompt: Aligning Color Trajectory with Vision-Language RepresentationsPoster1 citations
- Transferable 3D Adversarial Shape Completion using Diffusion ModelsPoster1 citations
- Two-Stage Active Learning for Efficient Temporal Action SegmentationPoster1 citations
- UNIT: Backdoor Mitigation via Automated Neural Distribution TighteningPoster1 citations
- UPose3D: Uncertainty-Aware 3D Human Pose Estimation with Cross-View and Temporal CuesPoster1 citations
- Uncertainty Calibration with Energy Based Instance-wise Scaling in the Wild DatasetPoster1 citations
- Understanding Physical Dynamics with Counterfactual World ModelingPoster1 citations
- Uni3DL: A Unified Model for 3D Vision-Language UnderstandingPoster1 citations
- Unleashing the Potential of the Semantic Latent Space in Diffusion Models for Image DehazingPoster1 citations
- Unsupervised Multi-modal Medical Image Registration via Invertible TranslationPoster1 citations
- Unsupervised Variational Translator for Bridging Image Restoration and High-Level Vision TasksPoster1 citations
- Unveiling Advanced Frequency Disentanglement Paradigm for Low-Light Image EnhancementPoster1 citations
- Urban Waterlogging Detection: A Challenging Benchmark and Large-Small Model Co-AdapterPoster1 citations
- Using My Artistic Style? You Must Obtain My AuthorizationPoster1 citations
- VISAGE: Video Instance Segmentation with Appearance-Guided EnhancementPoster1 citations
- Versatile Incremental Learning: Towards Class and Domain-Agnostic Incremental LearningPoster1 citations
- VersatileGaussian: Real-time Neural Rendering for Versatile Tasks using Gaussian SplattingPoster1 citations
- ViG-Bias: Visually Grounded Bias Discovery and MitigationPoster1 citations
- ViPer: Visual Personalization of Generative Models via Individual Preference LearningPoster1 citations
- WAS: Dataset and Methods for Artistic Text SegmentationPoster1 citations
- Watching it in Dark: A Target-aware Representation Learning Framework for High-Level Vision Tasks in Low IlluminationPoster1 citations
- WeCromCL: Weakly Supervised Cross-Modality Contrastive Learning for Transcription-only Supervised Text SpottingPoster1 citations
- Weak-to-Strong Compositional Learning from Generative Models for Language-based Object DetectionPoster1 citations
- Weight Conditioning for Smooth Optimization of Neural NetworksPoster1 citations
- Zero-Shot Multi-Object Scene CompletionPoster1 citations
- cDP-MIL: Robust Multiple Instance Learning via Cascaded Dirichlet ProcessPoster1 citations
- "A Framework for Efficient Model Evaluation through Stratification, Sampling, and Estimation"Poster
- "Idling Neurons, Appropriately Lenient Workload During Fine-tuning Leads to Better Generalization"Poster
- "Refine, Discriminate and Align: Stealing Encoders via Sample-Wise Prototypes and Multi-Relational Extraction"Poster
- "Unsupervised, Online and On-The-Fly Anomaly Detection For Non-Stationary Image Distributions"Poster
- "Veil Privacy on Visual Data: Concealing Privacy for Humans, Unveiling for DNNs"Poster
- 3D Congealing: 3D-Aware Image Alignment in the WildPoster
- 3D Human Pose Estimation via Non-Causal Retentive NetworksPoster
- 3DFG-PIFu: 3D Feature Grids for Human Digitization from Sparse ViewsPoster
- 3DSA:Multi-View 3D Human Pose Estimation With 3D Space Attention MechanismsPoster
- A Probability-guided Sampler for Neural Implicit Surface RenderingPoster
- A Rotation-invariant Texture ViT for Fine-Grained Recognition of Esophageal Cancer Endoscopic Ultrasound ImagesPoster
- A Secure Image Watermarking Framework with Statistical Guarantees via Adversarial Attacks on Secret Key NetworksPoster
- A Unified Image Compression Method for Human Perception and Multiple Vision TasksPoster
- AID-AppEAL: Automatic Image Dataset and Algorithm for Content Appeal Enhancement and Assessment LabelingPoster
- APL: Anchor-based Prompt Learning for One-stage Weakly Supervised Referring Expression ComprehensionPoster
- AWOL: Analysis WithOut synthesis using LanguagePoster
- ActionSwitch: Class-agnostic Detection of Simultaneous Actions in Streaming VideosPoster
- Align before Collaborate: Mitigating Feature Misalignment for Robust Multi-Agent PerceptionOral
- Aligning Neuronal Coding of Dynamic Visual Scenes with Foundation Vision ModelsPoster
- All You Need is Your Voice: Emotional Face Representation with Audio Perspective for Emotional Talking Face GenerationPoster
- An Adaptive Screen-Space Meshing Approach for Normal IntegrationPoster
- An Explainable Vision Question Answer Model via Diffusion Chain-of-ThoughtPoster
- An Information Theoretical View for Out-Of-Distribution DetectionPoster
- An Optimal Control View of LoRA and Binary Controller Design for Vision TransformersPoster
- An Optimization Framework to Enforce Multi-View Consistency for Texturing 3D MeshesPoster
- Any2Point: Empowering Any-modality Transformers for Efficient 3D UnderstandingPoster
- AnyHome: Open-Vocabulary Large-Scale Indoor Scene Generation with First-Person View ExplorationPoster
- Asynchronous Bioplausible Neuron for Spiking Neural Networks for Event-Based VisionPoster
- BaSIC: BayesNet Structure Learning for Computational Scalable Neural Image CompressionPoster
- Bayesian Detector Combination for Object Detection with Crowdsourced AnnotationsPoster
- Bayesian Self-Training for Semi-Supervised 3D SegmentationPoster
- Beyond the Data Imbalance: Employing the Heterogeneous Datasets for Vehicle Maneuver PredictionPoster
- Binomial Self-compensation for Motion Error in Dynamic 3D ScanningPoster
- BlazeBVD: Make Scale-Time Equalization Great Again for Blind Video DeflickeringPoster
- Blind Image Deconvolution by Generative-based Kernel Prior and Initializer via Latent EncodingPoster
- Blind image deblurring with noise-robust kernel estimationPoster
- Bones Can't Be Triangles: Accurate and Efficient Vertebrae Keypoint Estimation through Collaborative Error RevisionPoster
- Bridging the Gap: Studio-like Avatar Creation from a Monocular Phone CaptureOral
- Bridging the Pathology Domain Gap: Efficiently Adapting CLIP for Pathology Image Analysis with Limited Labeled DataPoster
- CARB-Net: Camera-Assisted Radar-Based Network for Vulnerable Road User DetectionPoster
- CLEO: Continual Learning of Evolving OntologiesPoster
- CLR-GAN: Improving GANs Stability and Quality via Consistent Latent Representation and ReconstructionPoster
- COIN-Matting: Confounder Intervention for Image MattingPoster
- COM Kitchens: An Unedited Overhead-view Procedural Videos Dataset a Vision-Language BenchmarkPoster
- COSMU: Complete 3D human shape from monocular unconstrained imagesPoster
- CPT-VR: Improving Surface Rendering via Closest Point Transform with View-Reflection AppearancePoster
- CSOT: Cross-Scan Object Transfer for Semi-Supervised LiDAR Object DetectionPoster
- CanonicalFusion: Generating Drivable 3D Human Avatars from Multiple ImagesPoster
- Causal Subgraphs and Information Bottlenecks: Redefining OOD Robustness in Graph Neural NetworksPoster
- Characterizing Model Robustness via Natural Input GradientsPoster
- Classification Matters: Improving Video Action Detection with Class-Specific AttentionOral
- Co-Student: Collaborating Strong and Weak Students for Sparsely Annotated Object DetectionPoster
- Cocktail Universal Adversarial Attack on Deep Neural NetworksPoster
- Combining Generative and Geometry Priors for Wide-Angle Portrait CorrectionPoster
- Computing the Lipschitz constant needed for fast scene recovery from CASSI measurementsPoster
- Confidence-Based Iterative Generation for Real-World Image Super-ResolutionPoster
- Consistent 3D Line MappingPoster
- Constructing Concept-based Models to Mitigate Spurious Correlations with Minimal Human EffortPoster
- Contextual Correspondence Matters: Bidirectional Graph Matching for Video SummarizationPoster
- Continual Learning and Unknown Object Discovery in 3D Scenes via Self-DistillationPoster
- Continuous SO(3) Equivariant Convolution for 3D Point Cloud AnalysisPoster
- ControlNet++: Improving Conditional Controls with Efficient Consistency FeedbackPoster
- Convex Relaxations for Manifold-Valued Markov Random Fields with Approximation GuaranteesOral
- Correspondences of the Third Kind: Camera Pose Estimation from Object ReflectionOral
- CrossScore: A Multi-View Approach to Image Evaluation and ScoringPoster
- DMiT: Deformable Mipmapped Tri-Plane Representation for Dynamic ScenesPoster
- DSA: Discriminative Scatter Analysis for Early Smoke SegmentationPoster
- Data Overfitting for On-Device Super-Resolution with Dynamic Algorithm and Compiler Co-DesignPoster
- Debiasing surgeon: fantastic weights and how to find themPoster
- DecentNeRFs: Decentralized Neural Radiance Fields from Crowdsourced ImagesPoster
- Decomposition of Neural Discrete Representations for Large-Scale 3D MappingPoster
- Deep Companion Learning: Enhancing Generalization Through Historical ConsistencyPoster
- Deep Cost Ray Fusion for Sparse Depth Video CompletionPoster
- Deep Feature Surgery: Towards Accurate and Efficient Multi-Exit NetworksPoster
- DetailSemNet: Elevating Signature Verification through Detail-Semantic IntegrationPoster
- DiffPMAE: Diffusion Masked Autoencoders for Point Cloud ReconstructionPoster
- DiffSurf: A Transformer-based Diffusion Model for Generating and Reconstructing 3D Surfaces in PosePoster
- Differentiable Convex Polyhedra Optimization from Multi-view ImagesPoster
- Diffusion Bridges for 3D Point Cloud DenoisingPoster
- Diffusion-Generated Pseudo-Observations for High-Quality Sparse-View ReconstructionPoster
- Diffusion-Refined VQA Annotations for Semi-Supervised Gaze FollowingPoster
- Discovering Unwritten Visual Classifiers with Large Language ModelsPoster
- Distilling Knowledge from Large-Scale Image Models for Object DetectionPoster
- Distractor-Free Novel View Synthesis via Exploiting Memorization Effect in OptimizationPoster
- Distributed Active Client Selection With Noisy Clients Using Model Association ScoresPoster
- Distributed Semantic Segmentation with Efficient Joint Source and Task DecodingPoster
- Domain Reduction Strategy for Non-Line-of-Sight ImagingPoster
- Domesticating SAM for Breast Ultrasound Image Segmentation via Spatial-frequency Fusion and Uncertainty CorrectionPoster
- DreamReward: Aligning Human Preference in Text-to-3D GenerationPoster
- Dual-Decoupling Learning and Metric-Adaptive Thresholding for Semi-Supervised Multi-Label LearningPoster
- Dynamic Data Selection for Efficient SSL via Coarse-to-Fine RefinementPoster
- DεpS: Delayed ε-Shrinking for Faster Once-For-All TrainingPoster
- E3V-K5: An Authentic Benchmark for Redefining Video-Based Energy Expenditure EstimationPoster
- ELSE: Efficient Deep Neural Network Inference through Line-based Sparsity ExplorationPoster
- Echoes of the Past: Boosting Long-tail Recognition via Reflective LearningOral
- Efficient 3D-Aware Facial Image Editing via Attribute-Specific Prompt LearningPoster
- Efficient Active Domain Adaptation for Semantic Segmentation by Selecting Information-rich SuperpixelsOral
- Efficient Neural Video Representation with Temporally Coherent ModulationOral
- Efficient Pre-training for Localized Instruction Generation of Procedural VideosPoster
- Efficient Snapshot Spectral Imaging: Calibration-Free Parallel Structure with Aperture Diffraction FusionPoster
- Efficient Training of Spiking Neural Networks with Multi-Parallel Implicit Stream ArchitecturePoster
- Efficient Training with Denoised Neural WeightsPoster
- Efficient Unsupervised Visual Representation Learning with Explicit Cluster BalancingPoster
- Elysium: Exploring Object-level Perception in Videos through Semantic Integration Using MLLMsPoster
- Energy-Clibrated VAE with Test Time Free LunchPoster
- Energy-induced Explicit quantification for Multi-modality MRI fusionPoster
- Enhancing Optimization Robustness in 1-bit Neural Networks through Stochastic Sign DescentPoster
- EpipolarGAN: Omnidirectional Image Synthesis with Explicit Camera ControlPoster
- ExMatch: Self-guided Exploitation for Semi-Supervised Learning with Scarce Labeled SamplesPoster
- Exploiting Supervised Poison Vulnerability to Strengthen Self-Supervised DefensePoster
- Exploring Guided Sampling of Conditional GANsPoster
- Exploring the Feature Extraction and Relation Modeling For Light-Weight Transformer TrackingOral
- FAMOUS: High-Fidelity Monocular 3D Human Digitization Using View SynthesisPoster
- FMBoost: Boosting Latent Diffusion with Flow MatchingOral
- FRI-Net: Floorplan Reconstruction via Room-wise Implicit RepresentationPoster
- FTBC: Forward Temporal Bias Correction for Optimizing ANN-SNN ConversionPoster
- Face Reconstruction Transfer Attack as Out-of-Distribution GeneralizationPoster
- FairViT: Fair Vision Transformer via Adaptive MaskingPoster
- FedHARM: Harmonizing Model Architectural Diversity in Federated LearningPoster
- FedHide: Federated Learning by Hiding in the NeighborsPoster
- Finding NeMo: Negative-mined Mosaic Augmentation for Referring Image SegmentationPoster
- Fine-Grained Scene Graph Generation via Sample-Level Bias PredictionPoster
- Fine-grained Dynamic Network for Generic Event Boundary DetectionPoster
- Flatness-aware Sequential Learning Generates Resilient BackdoorsOral
- FlowCon: Out-of-Distribution Detection using Flow-based Contrastive LearningPoster
- Forbes: Face Obfuscation Rendering via Backpropagation Refinement SchemePoster
- Forecasting Future Videos from Novel Views via Disentangled 3D Scene RepresentationPoster
- Free-ATM: Harnessing Free Attention Masks for Representation Learning on Diffusion-Generated ImagesPoster
- Free-Viewpoint Video of Outdoor Sports Using a DronePoster
- FreeAugment: Data Augmentation Search Across All Degrees of FreedomPoster
- Freeview Sketching: View-Aware Fine-Grained Sketch-Based Image RetrievalPoster
- FuseTeacher: Modality-fused Encoders are Strong Vision SupervisorsPoster
- GAURA: Generalizable Approach for Unified Restoration and Rendering of Arbitrary ViewsPoster
- GENIXER: Empowering Multimodal Large Language Models as a Powerful Data GeneratorPoster
- GOEmbed: Gradient Origin Embeddings for Representation Agnostic 3D Feature LearningPoster
- GRAPE: Generalizable and Robust Multi-view Facial CapturePoster
- GTMS: A Gradient-driven Tree-guided Mask-free Referring Image Segmentation MethodPoster
- Gaze Target Detection Based on Head-Local-Global CoordinationPoster
- Generalizable Symbolic Optimizer LearningPoster
- Generalizing to Unseen Domains via Text-guided AugmentationPoster
- Get Your Embedding Space in Order: Domain-Adaptive Regression for Forest MonitoringPoster
- Global-to-Pixel Regression for Human Mesh RecoveryPoster
- GlobalPointer: Large-Scale Plane Adjustment with Bi-Convex RelaxationPoster
- Gradient-based Out-of-Distribution DetectionPoster
- Group Testing for Accurate and Efficient Range-Based Near Neighbor Search for Plagiarism DetectionPoster
- GroupDiff: Diffusion-based Group Portrait EditingPoster
- Harmonizing knowledge Transfer in Neural Network with Unified DistillationPoster
- Hierarchical Conditioning of Diffusion Models Using Tree-of-Life for Studying Species EvolutionPoster
- Hierarchical Separable Video Transformer for Snapshot Compressive ImagingPoster
- Hierarchically Structured Neural Bones for Reconstructing Animatable Objects from Casual VideosPoster
- High-Fidelity Modeling of Generalizable Wrinkle DeformationPoster
- High-Resolution and Few-shot View Synthesis from Asymmetric Dual-lens InputsPoster
- How Far Can a 1-Pixel Camera Go? Solving Vision Tasks using Photoreceptors and Computationally Designed Visual MorphologyPoster
- How Video Meetings Change Your ExpressionPoster
- Human Pose Recognition via Occlusion-Preserving Abstract ImagesPoster
- Human-in-the-Loop Visual Re-ID for Population Size EstimationPoster
- IAM-VFI : Interpolate Any Motion for Video Frame Interpolation with motion complexity mapPoster
- Implicit Steganography Beyond the Constraints of ModalityPoster
- Improving 3D Semi-supervised Learning by Effectively Utilizing All Unlabelled DataPoster
- Information Bottleneck Based Data Correction in Continual LearningPoster
- Instance-dependent Noisy-label Learning with Graphical Model Based Noise-rate EstimationPoster
- Inter-Class Topology Alignment for Efficient Black-Box Substitute AttacksPoster
- Interaction-centric Spatio-Temporal Context Reasoning for Multi-Person Video HOI RecognitionPoster
- Interactive 3D Object Detection with PromptsPoster
- Interpretability-Guided Test-Time Adversarial DefensePoster
- Investigating Style Similarity in Diffusion ModelsPoster
- Joint RGB-Spectral Decomposition Model Guided Image Enhancement in Mobile PhotographyPoster
- LASS3D: Language-Assisted Semi-Supervised 3D Semantic Segmentation with Progressive Unreliable Data ExploitationPoster
- LEROjD: Lidar Extended Radar-Only Object DetectionPoster
- LRSLAM: Low-rank Representation of Signed Distance Fields in Dense Visual SLAM SystemPoster
- Learn to Optimize Denoising Scores: A Unified and Improved Diffusion Prior for 3D GenerationPoster
- Learned Image Enhancement via Color NamingPoster
- Learning Dual-Level Deformable Implicit Representation for Real-World Scale Arbitrary Super-ResolutionPoster
- Learning Equilibrium Transformation for Gamut Expansion and Color RestorationPoster
- Learning Exhaustive Correlation for Spectral Super-Resolution: Where Spatial-Spectral Attention Meets Linear DependencePoster
- Learning Neural Deformation Representation for 4D Dynamic Shape GenerationPoster
- Learning Non-Linear Invariants for Unsupervised Out-of-Distribution DetectionPoster
- Learning Quantized Adaptive Conditions for Diffusion ModelsPoster
- Learning a Dynamic Privacy-preserving Camera Robust to Inversion AttacksOral
- Learning to Build by Building Your Own InstructionsPoster
- Learning to Robustly Reconstruct Dynamic Scenes from Low-light Spike StreamsPoster
- Learning with Counterfactual Explanations for Radiology Report GenerationPoster
- Learning-based Axial Video Motion MagnificationPoster
- Let the Avatar Talk using Texts without Paired Training DataPoster
- Leveraging Text Localization for Scene Text Removal via Text-aware Masked Image ModelingPoster
- Leveraging scale- and orientation-covariant features for planar motion estimationPoster
- LingoQA: Video Question Answering for Autonomous DrivingPoster
- Linking in Style: Understanding learned features in deep learning modelsPoster
- LoA-Trans: Enhancing Visual Grounding by Location-Aware TransformersPoster
- Loc3Diff: Local Diffusion for 3D Human Head Synthesis and EditingPoster
- M3DBench: Towards Omni 3D Assistant with Interleaved Multi-modal InstructionsPoster
- MC-PanDA: Mask Confidence for Panoptic Domain AdaptationPoster
- MOD-UV: Learning Mobile Object Detectors from Unlabeled VideosPoster
- MRSP: Learn Multi-Representations of Single Primitive for Compositional Zero-Shot LearningPoster
- MTaDCS: Moving Trace and Feature Density-based Confidence Sample Selection under Label NoisePoster
- MaRINeR: Enhancing Novel Views by Matching Rendered Images with Nearby ReferencesPoster
- McGrids: Monte Carlo-Driven Adaptive Grids for Iso-Surface ExtractionPoster
- MetaAT: Active Testing for Label-Efficient Evaluation of Dense Recognition TasksPoster
- MetaAug: Meta-Data Augmentation for Post-Training QuantizationPoster
- MetaWeather: Few-Shot Weather-Degraded Image RestorationPoster
- Mew: Multiplexed Immunofluorescence Image Analysis through an Efficient Multiplex NetworkPoster
- Modeling Label Correlations with Latent Context for Multi-Label RecognitionPoster
- Motion Keyframe Interpolation for Any Human Skeleton using Point Cloud-based Human Motion Data HomogenisationPoster
- Multi-Granularity Sparse Relationship Matrix Prediction Network for End-to-End Scene Graph GenerationPoster
- Multi-modal Relation Distillation for Unified 3D Representation LearningPoster
- Multi-scale Cross Distillation for Object Detection in Aerial ImagesPoster
- MultiGen: Zero-shot Image Generation from Multi-modal PromptsPoster
- Multimodal Label Relevance Ranking via Reinforcement LearningPoster
- Multiscale Graph Texture NetworkPoster
- NGP-RT: Fusing Multi-Level Hash Features with Lightweight Attention for Real-Time Novel View SynthesisPoster
- NeRF-XL: NeRF at Any Scale with Multi-GPUPoster
- NeRMo: Learning Implicit Neural Representations for 3D Human Motion PredictionOral
- Neural Poisson Solver: A Universal and Continuous Framework for Natural Signal BlendingPoster
- Neural graphics texture compression supporting random accessPoster
- Noise Calibration: Plug-and-play Content-Preserving Video Enhancement using Pre-trained Video Diffusion ModelsPoster
- Non-parametric Sensor Noise Modeling and SynthesisPoster
- Nymeria: A Massive Collection of Egocentric Multi-modal Human Motion in the WildPoster
- OLAF: A Plug-and-Play Framework for Enhanced Multi-object Multi-part Scene ParsingPoster
- Object-Aware NIR-to-Visible TranslationPoster
- Omni-Recon: Harnessing Image-based Rendering for General-Purpose Neural Radiance FieldsOral
- On Spectral Properties of Gradient-based Explanation MethodsPoster
- On the Evaluation Consistency of Attribution-based ExplanationsPoster
- On the Topology Awareness and Generalization Performance of Graph Neural NetworksOral
- On-the-fly Category Discovery for LiDAR Semantic SegmentationPoster
- Online Continuous Generalized Category DiscoveryPoster
- Online Temporal Action Localization with Memory-Augmented TransformerPoster
- Open-Vocabulary RGB-Thermal Semantic SegmentationPoster
- Open-set Domain Adaptation via Joint Error based Multi-class Positive and Unlabeled LearningPoster
- Optimal Transport of Diverse Unsupervised Tasks for Robust Learning from Noisy Few-Shot DataPoster
- Oulu Remote-photoplethysmography Physical Domain Attacks Database (ORPDAD)Poster
- PACE: Pose Annotations in Cluttered EnvironmentsPoster
- PAV: Personalized Head Avatar from Unstructured Video CollectionPoster
- POCA: Post-training Quantization with Temporal Alignment for Codec AvatarsPoster
- Panel-Specific Degradation Representation for Raw Under-Display Camera Image RestorationPoster
- Phase Concentration and Shortcut Suppression for Weakly Supervised Semantic SegmentationPoster
- Photon Inhibition for Energy-Efficient Single-Photon ImagingOral
- Physically Plausible Color Correction for Neural Radiance FieldsPoster
- Point-supervised Panoptic Segmentation via Estimating Pseudo Labels from Learnable DistancePoster
- PolyOculus: Simultaneous Multi-view Image-based Novel View SynthesisPoster
- PoseSOR: Human Pose Can Guide Our AttentionPoster
- Privacy-Preserving Adaptive Re-Identification without Image TransferOral
- Probabilistic Image-Driven Traffic Modeling via Remote SensingPoster
- Progressive Proxy Anchor Propagation for Unsupervised Semantic SegmentationPoster
- ProtoComp: Diverse Point Cloud Completion with Controllable PrototypePoster
- Pseudo-Embedding for Generalized Few-Shot Point Cloud SegmentationPoster
- Pseudo-Labelling Should Be Aware of Disguising Channel ActivationsPoster
- RANRAC: Robust Neural Scene Representations via Random Ray ConsensusPoster
- REFRAME: Reflective Surface Real-Time Rendering for Mobile DevicesPoster
- RING-NeRF : Rethinking Inductive Biases for Versatile and Efficient Neural FieldsPoster
- Region-Native Visual TokenizationPoster
- Region-aware Distribution Contrast: A Novel Approach to Multi-Task Partially Supervised LearningPoster
- Regularizing Dynamic Radiance Fields with Kinematic FieldsPoster
- Remove Projective LiDAR Depthmap Artifacts via Exploiting Epipolar GeometryPoster
- Removing Rows and Columns of Tokens in Vision Transformer enables Faster Dense Prediction without RetrainingPoster
- Representation Enhancement-Stabilization: Reducing Bias-Variance of Domain GeneralizationPoster
- Resolving Scale Ambiguity in Multi-view 3D Reconstruction using Dual-Pixel SensorsPoster
- Rethinking Data Bias: Dataset Copyright Protection via Embedding Class-wise Hidden BiasPoster
- Rethinking Directional Parameterization in Neural Implicit Surface ReconstructionPoster
- Rethinking Fast Adversarial Training: A Splitting Technique To Overcome Catastrophic OverfittingPoster
- Rethinking Features-Fused-Pyramid-Neck for Object DetectionPoster
- Rethinking Normalization Layers for Domain Generalizable Person Re-identificationPoster
- Retrieval Robust to Object Motion BlurPoster
- Revisit Self-supervision with Local Structure-from-MotionPoster
- Revisiting Feature Disentanglement Strategy in Diffusion Training and Breaking Conditional Independence Assumption in SamplingPoster
- Robust Fitting on a Gate Quantum ComputerOral
- Robustness Preserving Fine-tuning using Neuron ImportancePoster
- Rotated Orthographic Projection for Self-Supervised 3D Human Pose EstimationPoster
- SAH-SCI: Self-Supervised Adapter for Efficient Hyperspectral Snapshot Compressive ImagingPoster
- SCPNet: Unsupervised Cross-modal Homography Estimation via Intra-modal Self-supervised LearningPoster
- SDPT: Synchronous Dual Prompt Tuning for Fusion-based Visual-Language Pre-trained ModelsPoster
- SHINE: Saliency-aware HIerarchical NEgative Ranking for Compositional Temporal GroundingPoster
- SNP: Structured Neuron-level Pruning to Preserve Attention ScoresPoster
- ST-LDM: A Universal Framework for Text-Grounded Object Generation in Real ImagesPoster
- Scalar Function Topology Divergence: Comparing Topology of 3D ObjectsPoster
- Scaling Up Personalized Image Aesthetic Assessment via Task Vector CustomizationPoster
- SeA: Semantic Adversarial Augmentation for Last Layer Features from Unsupervised Representation LearningPoster
- Segmentation-guided Layer-wise Image Vectorization with Gradient FillsPoster
- SeiT++: Masked Token Modeling Improves Storage-efficient TrainingPoster
- Self-Supervised Video Copy Localization with Regional Token RepresentationPoster
- Self-supervised Shape Completion via Involution and Implicit CorrespondencesPoster
- Semantic-guided Robustness Tuning for Few-Shot Transfer Across Extreme Domain ShiftPoster
- Shapefusion: 3D localized human diffusion modelsPoster
- Single-Photon 3D Imaging with Equi-Depth Photon HistogramsPoster
- Sketch2Vox: Learning 3D Reconstruction from a Single Monocular Sketch ImagePoster
- Soft Shadow Diffusion (SSD): Physics-inspired Learning for 3D Computational PeriscopyPoster
- Solving the inverse problem of microscopy deconvolution with a residual Beylkin-Coifman-Rokhlin neural networkPoster
- SparseRadNet: Sparse Perception Neural Network on Subsampled Radar DataPoster
- SpatialFormer: Towards Generalizable Vision Transformers with Explicit Spatial UnderstandingPoster
- Spatio-Temporal Proximity-Aware Dual-Path Model for Panoramic Activity RecognitionPoster
- Spectral Subsurface Scattering for Material ClassificationPoster
- SpeedUpNet: A Plug-and-Play Adapter Network for Accelerating Text-to-Image Diffusion ModelsPoster
- Spline-based TransformersOral
- Stable Preference: Redefining training paradigm of human preference model for Text-to-Image SynthesisPoster
- Stepwise Multi-grained Boundary Detector for Point-supervised Temporal Action LocalizationPoster
- StereoGlue: Joint Feature Matching and Robust EstimationPoster
- Stripe Observation Guided Inference Cost-free Attention MechanismPoster
- SweepNet: Unsupervised Learning Shape Abstraction via Neural SweepersPoster
- Synthesizing Environment-Specific People in PhotographsPoster
- Synthesizing Time-varying BRDFs via Latent SpacePoster
- Textual-Visual Logic Challenge: Understanding and Reasoning in Text-to-Image GenerationPoster
- Time-Efficient and Identity-Consistent Virtual Try-On Using A Variant of Altered Diffusion ModelsPoster
- To Supervise or Not to Supervise: Understanding and Addressing the Key Challenges of Point Cloud Transfer LearningPoster
- Toward INT4 Fixed-Point Training via Exploring Quantization Error for GradientsPoster
- Towards Architecture-Agnostic Untrained Networks Priors for Image Reconstruction with Frequency RegularizationPoster
- Towards Certifiably Robust Face RecognitionPoster
- Towards Robust Full Low-bit Quantization of Super Resolution NetworksPoster
- Towards compact reversible image representations for neural style transferPoster
- Train Till You Drop: Towards Stable and Robust Source-free Unsupervised 3D Domain AdaptationPoster
- TreeSBA: Tree-Transformer for Self-Supervised Sequential Brick AssemblyPoster
- TurboEdit: Real-time text-based disentangled real image editingPoster
- UAV First-Person Viewers Are Radiance Field LearnersPoster
- URS-NeRF: Unordered Rolling Shutter Bundle Adjustment for Neural Radiance FieldsPoster
- Uncertainty-Driven Spectral Compressive Imaging with Spatial-Frequency TransformerPoster
- Understanding Multi-compositional learning in Vision and Language models via Category TheoryPoster
- UniVoxel: Fast Inverse Rendering by Unified Voxelization of Scene RepresentationPoster
- Unified Local-Cloud Decision-Making via Reinforcement LearningPoster
- Unleashing Text-to-Image Diffusion Prior for Zero-Shot Image CaptioningPoster
- Unrolled Decomposed Unpaired Learning for Controllable Low-Light Video EnhancementPoster
- Unsqueeze [CLS] Bottleneck to Learn Rich RepresentationsPoster
- Unsupervised Representation Learning by Balanced Self Attention MatchingPoster
- Unveiling Privacy Risks in Stochastic Neural Networks Training: Effective Image Reconstruction from GradientsPoster
- Upper-body Hierarchical Graph for Skeleton Based Emotion Recognition in Assistive DrivingPoster
- V-Trans4Style: Visual Transition Recommendation for Video Production Style AdaptationPoster
- VETRA: A Dataset for Vehicle Tracking in Aerial Imagery - New Challenges for Multi-Object TrackingOral
- VF-NeRF: Viewshed Fields for Rigid NeRF RegistrationPoster
- VP-SAM: Taming Segment Anything Model for Video Polyp Segmentation via Disentanglement and Spatio-temporal Side NetworkPoster
- VideoClusterNet: Self-Supervised and Adaptive Face Clustering for VideosPoster
- View-Consistent Hierarchical 3D Segmentation Using Ultrametric Feature FieldsPoster
- Visual Prompting via Partial Optimal TransportPoster
- Visual Relationship TransformationPoster
- Wavelength-Embedding-guided Filter-Array Transformer for Spectral DemosaicingPoster
- Weakly-Supervised 3D Hand Reconstruction with Knowledge Prior and Uncertainty GuidancePoster
- When and How do negative prompts take effect?Poster
- WindPoly: Polygonal Mesh Reconstruction via Winding NumbersPoster
- ZoLA: Zero-Shot Creative Long Animation Generation with Short Video ModelOral
- uCAP: An Unsupervised Prompting Method for Vision-Language ModelsOral
ECCV accepted papers in other years
Looking for submission deadlines instead? See the conference deadline calendar.