← All conferences

ECCV 2024 Accepted Papers

The full list of 2,385 papers accepted at ECCV 2024 (European Conference on Computer Vision). Click any title for details, similar papers, and links to the original source. You can also search these papers by meaning, not just keywords.

Poster: 2,185Oral: 200
  1. Surface-Centric Modeling for High-Fidelity Generalizable Neural Surface ReconstructionPoster1 citations
  2. Syn-to-Real Domain Adaptation for Point Cloud Completion via Part-based ApproachPoster1 citations
  3. Sync from the Sea: Retrieving Alignable Videos from Large-Scale DatasetsOral1 citations
  4. Synchronous Diffusion for Unsupervised Smooth Non-Rigid 3D Shape MatchingPoster1 citations
  5. T-CorresNet: Template Guided 3D Point Cloud Completion with Correspondence Pooling Query Generation StrategyPoster1 citations
  6. TAG: Text Prompt Augmentation for Zero-Shot Out-of-Distribution DetectionPoster1 citations
  7. TCC-Det: Temporarily consistent cues for weakly-supervised 3D detectionPoster1 citations
  8. Temporal As a Plugin: Unsupervised Video Denoising with Pre-Trained Image DenoisersPoster1 citations
  9. Tensorial template matching for fast cross-correlation with rotations and its application for tomographyPoster1 citations
  10. Textual Grounding for Open-vocabulary Visual Information Extraction in Layout-diversified DocumentsPoster1 citations
  11. The Gaussian Discriminant Variational Autoencoder (GdVAE): A Self-Explainable Model with Counterfactual ExplanationsPoster1 citations
  12. The Role of Masking for Efficient Supervised Knowledge Distillation of Vision TransformersPoster1 citations
  13. Think before Placement: Common Sense Enhanced Transformer for Object PlacementPoster1 citations
  14. Tight and Efficient Upper Bound on Spectral Norm of Convolutional LayersPoster1 citations
  15. TimeLens-XL: Real-time Event-based Video Frame Interpolation with Large MotionPoster1 citations
  16. Topology-Preserving Downsampling of Binary ImagesPoster1 citations
  17. Towards High-Quality 3D Motion Transfer with Realistic Apparel AnimationPoster1 citations
  18. Towards Model-Agnostic Dataset Condensation by Heterogeneous ModelsOral1 citations
  19. Towards Reliable Evaluation and Fast Training of Robust Semantic Segmentation ModelsPoster1 citations
  20. Towards a Density Preserving Objective Function for Learning on Point SetsPoster1 citations
  21. TrafficNight : An Aerial Multimodal Benchmark For Nighttime Vehicle SurveillancePoster1 citations
  22. Training A Secure Model against Data-Free Model ExtractionPoster1 citations
  23. Training-free Composite Scene Generation for Layout-to-Image SynthesisPoster1 citations
  24. TrajPrompt: Aligning Color Trajectory with Vision-Language RepresentationsPoster1 citations
  25. Transferable 3D Adversarial Shape Completion using Diffusion ModelsPoster1 citations
  26. Two-Stage Active Learning for Efficient Temporal Action SegmentationPoster1 citations
  27. UNIT: Backdoor Mitigation via Automated Neural Distribution TighteningPoster1 citations
  28. UPose3D: Uncertainty-Aware 3D Human Pose Estimation with Cross-View and Temporal CuesPoster1 citations
  29. Uncertainty Calibration with Energy Based Instance-wise Scaling in the Wild DatasetPoster1 citations
  30. Understanding Physical Dynamics with Counterfactual World ModelingPoster1 citations
  31. Uni3DL: A Unified Model for 3D Vision-Language UnderstandingPoster1 citations
  32. Unleashing the Potential of the Semantic Latent Space in Diffusion Models for Image DehazingPoster1 citations
  33. Unsupervised Multi-modal Medical Image Registration via Invertible TranslationPoster1 citations
  34. Unsupervised Variational Translator for Bridging Image Restoration and High-Level Vision TasksPoster1 citations
  35. Unveiling Advanced Frequency Disentanglement Paradigm for Low-Light Image EnhancementPoster1 citations
  36. Urban Waterlogging Detection: A Challenging Benchmark and Large-Small Model Co-AdapterPoster1 citations
  37. Using My Artistic Style? You Must Obtain My AuthorizationPoster1 citations
  38. VISAGE: Video Instance Segmentation with Appearance-Guided EnhancementPoster1 citations
  39. Versatile Incremental Learning: Towards Class and Domain-Agnostic Incremental LearningPoster1 citations
  40. VersatileGaussian: Real-time Neural Rendering for Versatile Tasks using Gaussian SplattingPoster1 citations
  41. ViG-Bias: Visually Grounded Bias Discovery and MitigationPoster1 citations
  42. ViPer: Visual Personalization of Generative Models via Individual Preference LearningPoster1 citations
  43. WAS: Dataset and Methods for Artistic Text SegmentationPoster1 citations
  44. Watching it in Dark: A Target-aware Representation Learning Framework for High-Level Vision Tasks in Low IlluminationPoster1 citations
  45. WeCromCL: Weakly Supervised Cross-Modality Contrastive Learning for Transcription-only Supervised Text SpottingPoster1 citations
  46. Weak-to-Strong Compositional Learning from Generative Models for Language-based Object DetectionPoster1 citations
  47. Weight Conditioning for Smooth Optimization of Neural NetworksPoster1 citations
  48. Zero-Shot Multi-Object Scene CompletionPoster1 citations
  49. cDP-MIL: Robust Multiple Instance Learning via Cascaded Dirichlet ProcessPoster1 citations
  50. "A Framework for Efficient Model Evaluation through Stratification, Sampling, and Estimation"Poster
  51. "Idling Neurons, Appropriately Lenient Workload During Fine-tuning Leads to Better Generalization"Poster
  52. "Refine, Discriminate and Align: Stealing Encoders via Sample-Wise Prototypes and Multi-Relational Extraction"Poster
  53. "Unsupervised, Online and On-The-Fly Anomaly Detection For Non-Stationary Image Distributions"Poster
  54. "Veil Privacy on Visual Data: Concealing Privacy for Humans, Unveiling for DNNs"Poster
  55. 3D Congealing: 3D-Aware Image Alignment in the WildPoster
  56. 3D Human Pose Estimation via Non-Causal Retentive NetworksPoster
  57. 3DFG-PIFu: 3D Feature Grids for Human Digitization from Sparse ViewsPoster
  58. 3DSA:Multi-View 3D Human Pose Estimation With 3D Space Attention MechanismsPoster
  59. A Probability-guided Sampler for Neural Implicit Surface RenderingPoster
  60. A Rotation-invariant Texture ViT for Fine-Grained Recognition of Esophageal Cancer Endoscopic Ultrasound ImagesPoster
  61. A Secure Image Watermarking Framework with Statistical Guarantees via Adversarial Attacks on Secret Key NetworksPoster
  62. A Unified Image Compression Method for Human Perception and Multiple Vision TasksPoster
  63. AID-AppEAL: Automatic Image Dataset and Algorithm for Content Appeal Enhancement and Assessment LabelingPoster
  64. APL: Anchor-based Prompt Learning for One-stage Weakly Supervised Referring Expression ComprehensionPoster
  65. AWOL: Analysis WithOut synthesis using LanguagePoster
  66. ActionSwitch: Class-agnostic Detection of Simultaneous Actions in Streaming VideosPoster
  67. Align before Collaborate: Mitigating Feature Misalignment for Robust Multi-Agent PerceptionOral
  68. Aligning Neuronal Coding of Dynamic Visual Scenes with Foundation Vision ModelsPoster
  69. All You Need is Your Voice: Emotional Face Representation with Audio Perspective for Emotional Talking Face GenerationPoster
  70. An Adaptive Screen-Space Meshing Approach for Normal IntegrationPoster
  71. An Explainable Vision Question Answer Model via Diffusion Chain-of-ThoughtPoster
  72. An Information Theoretical View for Out-Of-Distribution DetectionPoster
  73. An Optimal Control View of LoRA and Binary Controller Design for Vision TransformersPoster
  74. An Optimization Framework to Enforce Multi-View Consistency for Texturing 3D MeshesPoster
  75. Any2Point: Empowering Any-modality Transformers for Efficient 3D UnderstandingPoster
  76. AnyHome: Open-Vocabulary Large-Scale Indoor Scene Generation with First-Person View ExplorationPoster
  77. Asynchronous Bioplausible Neuron for Spiking Neural Networks for Event-Based VisionPoster
  78. BaSIC: BayesNet Structure Learning for Computational Scalable Neural Image CompressionPoster
  79. Bayesian Detector Combination for Object Detection with Crowdsourced AnnotationsPoster
  80. Bayesian Self-Training for Semi-Supervised 3D SegmentationPoster
  81. Beyond the Data Imbalance: Employing the Heterogeneous Datasets for Vehicle Maneuver PredictionPoster
  82. Binomial Self-compensation for Motion Error in Dynamic 3D ScanningPoster
  83. BlazeBVD: Make Scale-Time Equalization Great Again for Blind Video DeflickeringPoster
  84. Blind Image Deconvolution by Generative-based Kernel Prior and Initializer via Latent EncodingPoster
  85. Blind image deblurring with noise-robust kernel estimationPoster
  86. Bones Can't Be Triangles: Accurate and Efficient Vertebrae Keypoint Estimation through Collaborative Error RevisionPoster
  87. Bridging the Gap: Studio-like Avatar Creation from a Monocular Phone CaptureOral
  88. Bridging the Pathology Domain Gap: Efficiently Adapting CLIP for Pathology Image Analysis with Limited Labeled DataPoster
  89. CARB-Net: Camera-Assisted Radar-Based Network for Vulnerable Road User DetectionPoster
  90. CLEO: Continual Learning of Evolving OntologiesPoster
  91. CLR-GAN: Improving GANs Stability and Quality via Consistent Latent Representation and ReconstructionPoster
  92. COIN-Matting: Confounder Intervention for Image MattingPoster
  93. COM Kitchens: An Unedited Overhead-view Procedural Videos Dataset a Vision-Language BenchmarkPoster
  94. COSMU: Complete 3D human shape from monocular unconstrained imagesPoster
  95. CPT-VR: Improving Surface Rendering via Closest Point Transform with View-Reflection AppearancePoster
  96. CSOT: Cross-Scan Object Transfer for Semi-Supervised LiDAR Object DetectionPoster
  97. CanonicalFusion: Generating Drivable 3D Human Avatars from Multiple ImagesPoster
  98. Causal Subgraphs and Information Bottlenecks: Redefining OOD Robustness in Graph Neural NetworksPoster
  99. Characterizing Model Robustness via Natural Input GradientsPoster
  100. Classification Matters: Improving Video Action Detection with Class-Specific AttentionOral
  101. Co-Student: Collaborating Strong and Weak Students for Sparsely Annotated Object DetectionPoster
  102. Cocktail Universal Adversarial Attack on Deep Neural NetworksPoster
  103. Combining Generative and Geometry Priors for Wide-Angle Portrait CorrectionPoster
  104. Computing the Lipschitz constant needed for fast scene recovery from CASSI measurementsPoster
  105. Confidence-Based Iterative Generation for Real-World Image Super-ResolutionPoster
  106. Consistent 3D Line MappingPoster
  107. Constructing Concept-based Models to Mitigate Spurious Correlations with Minimal Human EffortPoster
  108. Contextual Correspondence Matters: Bidirectional Graph Matching for Video SummarizationPoster
  109. Continual Learning and Unknown Object Discovery in 3D Scenes via Self-DistillationPoster
  110. Continuous SO(3) Equivariant Convolution for 3D Point Cloud AnalysisPoster
  111. ControlNet++: Improving Conditional Controls with Efficient Consistency FeedbackPoster
  112. Convex Relaxations for Manifold-Valued Markov Random Fields with Approximation GuaranteesOral
  113. Correspondences of the Third Kind: Camera Pose Estimation from Object ReflectionOral
  114. CrossScore: A Multi-View Approach to Image Evaluation and ScoringPoster
  115. DMiT: Deformable Mipmapped Tri-Plane Representation for Dynamic ScenesPoster
  116. DSA: Discriminative Scatter Analysis for Early Smoke SegmentationPoster
  117. Data Overfitting for On-Device Super-Resolution with Dynamic Algorithm and Compiler Co-DesignPoster
  118. Debiasing surgeon: fantastic weights and how to find themPoster
  119. DecentNeRFs: Decentralized Neural Radiance Fields from Crowdsourced ImagesPoster
  120. Decomposition of Neural Discrete Representations for Large-Scale 3D MappingPoster
  121. Deep Companion Learning: Enhancing Generalization Through Historical ConsistencyPoster
  122. Deep Cost Ray Fusion for Sparse Depth Video CompletionPoster
  123. Deep Feature Surgery: Towards Accurate and Efficient Multi-Exit NetworksPoster
  124. DetailSemNet: Elevating Signature Verification through Detail-Semantic IntegrationPoster
  125. DiffPMAE: Diffusion Masked Autoencoders for Point Cloud ReconstructionPoster
  126. DiffSurf: A Transformer-based Diffusion Model for Generating and Reconstructing 3D Surfaces in PosePoster
  127. Differentiable Convex Polyhedra Optimization from Multi-view ImagesPoster
  128. Diffusion Bridges for 3D Point Cloud DenoisingPoster
  129. Diffusion-Generated Pseudo-Observations for High-Quality Sparse-View ReconstructionPoster
  130. Diffusion-Refined VQA Annotations for Semi-Supervised Gaze FollowingPoster
  131. Discovering Unwritten Visual Classifiers with Large Language ModelsPoster
  132. Distilling Knowledge from Large-Scale Image Models for Object DetectionPoster
  133. Distractor-Free Novel View Synthesis via Exploiting Memorization Effect in OptimizationPoster
  134. Distributed Active Client Selection With Noisy Clients Using Model Association ScoresPoster
  135. Distributed Semantic Segmentation with Efficient Joint Source and Task DecodingPoster
  136. Domain Reduction Strategy for Non-Line-of-Sight ImagingPoster
  137. Domesticating SAM for Breast Ultrasound Image Segmentation via Spatial-frequency Fusion and Uncertainty CorrectionPoster
  138. DreamReward: Aligning Human Preference in Text-to-3D GenerationPoster
  139. Dual-Decoupling Learning and Metric-Adaptive Thresholding for Semi-Supervised Multi-Label LearningPoster
  140. Dynamic Data Selection for Efficient SSL via Coarse-to-Fine RefinementPoster
  141. DεpS: Delayed ε-Shrinking for Faster Once-For-All TrainingPoster
  142. E3V-K5: An Authentic Benchmark for Redefining Video-Based Energy Expenditure EstimationPoster
  143. ELSE: Efficient Deep Neural Network Inference through Line-based Sparsity ExplorationPoster
  144. Echoes of the Past: Boosting Long-tail Recognition via Reflective LearningOral
  145. Efficient 3D-Aware Facial Image Editing via Attribute-Specific Prompt LearningPoster
  146. Efficient Active Domain Adaptation for Semantic Segmentation by Selecting Information-rich SuperpixelsOral
  147. Efficient Neural Video Representation with Temporally Coherent ModulationOral
  148. Efficient Pre-training for Localized Instruction Generation of Procedural VideosPoster
  149. Efficient Snapshot Spectral Imaging: Calibration-Free Parallel Structure with Aperture Diffraction FusionPoster
  150. Efficient Training of Spiking Neural Networks with Multi-Parallel Implicit Stream ArchitecturePoster
  151. Efficient Training with Denoised Neural WeightsPoster
  152. Efficient Unsupervised Visual Representation Learning with Explicit Cluster BalancingPoster
  153. Elysium: Exploring Object-level Perception in Videos through Semantic Integration Using MLLMsPoster
  154. Energy-Clibrated VAE with Test Time Free LunchPoster
  155. Energy-induced Explicit quantification for Multi-modality MRI fusionPoster
  156. Enhancing Optimization Robustness in 1-bit Neural Networks through Stochastic Sign DescentPoster
  157. EpipolarGAN: Omnidirectional Image Synthesis with Explicit Camera ControlPoster
  158. ExMatch: Self-guided Exploitation for Semi-Supervised Learning with Scarce Labeled SamplesPoster
  159. Exploiting Supervised Poison Vulnerability to Strengthen Self-Supervised DefensePoster
  160. Exploring Guided Sampling of Conditional GANsPoster
  161. Exploring the Feature Extraction and Relation Modeling For Light-Weight Transformer TrackingOral
  162. FAMOUS: High-Fidelity Monocular 3D Human Digitization Using View SynthesisPoster
  163. FMBoost: Boosting Latent Diffusion with Flow MatchingOral
  164. FRI-Net: Floorplan Reconstruction via Room-wise Implicit RepresentationPoster
  165. FTBC: Forward Temporal Bias Correction for Optimizing ANN-SNN ConversionPoster
  166. Face Reconstruction Transfer Attack as Out-of-Distribution GeneralizationPoster
  167. FairViT: Fair Vision Transformer via Adaptive MaskingPoster
  168. FedHARM: Harmonizing Model Architectural Diversity in Federated LearningPoster
  169. FedHide: Federated Learning by Hiding in the NeighborsPoster
  170. Finding NeMo: Negative-mined Mosaic Augmentation for Referring Image SegmentationPoster
  171. Fine-Grained Scene Graph Generation via Sample-Level Bias PredictionPoster
  172. Fine-grained Dynamic Network for Generic Event Boundary DetectionPoster
  173. Flatness-aware Sequential Learning Generates Resilient BackdoorsOral
  174. FlowCon: Out-of-Distribution Detection using Flow-based Contrastive LearningPoster
  175. Forbes: Face Obfuscation Rendering via Backpropagation Refinement SchemePoster
  176. Forecasting Future Videos from Novel Views via Disentangled 3D Scene RepresentationPoster
  177. Free-ATM: Harnessing Free Attention Masks for Representation Learning on Diffusion-Generated ImagesPoster
  178. Free-Viewpoint Video of Outdoor Sports Using a DronePoster
  179. FreeAugment: Data Augmentation Search Across All Degrees of FreedomPoster
  180. Freeview Sketching: View-Aware Fine-Grained Sketch-Based Image RetrievalPoster
  181. FuseTeacher: Modality-fused Encoders are Strong Vision SupervisorsPoster
  182. GAURA: Generalizable Approach for Unified Restoration and Rendering of Arbitrary ViewsPoster
  183. GENIXER: Empowering Multimodal Large Language Models as a Powerful Data GeneratorPoster
  184. GOEmbed: Gradient Origin Embeddings for Representation Agnostic 3D Feature LearningPoster
  185. GRAPE: Generalizable and Robust Multi-view Facial CapturePoster
  186. GTMS: A Gradient-driven Tree-guided Mask-free Referring Image Segmentation MethodPoster
  187. Gaze Target Detection Based on Head-Local-Global CoordinationPoster
  188. Generalizable Symbolic Optimizer LearningPoster
  189. Generalizing to Unseen Domains via Text-guided AugmentationPoster
  190. Get Your Embedding Space in Order: Domain-Adaptive Regression for Forest MonitoringPoster
  191. Global-to-Pixel Regression for Human Mesh RecoveryPoster
  192. GlobalPointer: Large-Scale Plane Adjustment with Bi-Convex RelaxationPoster
  193. Gradient-based Out-of-Distribution DetectionPoster
  194. Group Testing for Accurate and Efficient Range-Based Near Neighbor Search for Plagiarism DetectionPoster
  195. GroupDiff: Diffusion-based Group Portrait EditingPoster
  196. Harmonizing knowledge Transfer in Neural Network with Unified DistillationPoster
  197. Hierarchical Conditioning of Diffusion Models Using Tree-of-Life for Studying Species EvolutionPoster
  198. Hierarchical Separable Video Transformer for Snapshot Compressive ImagingPoster
  199. Hierarchically Structured Neural Bones for Reconstructing Animatable Objects from Casual VideosPoster
  200. High-Fidelity Modeling of Generalizable Wrinkle DeformationPoster
  201. High-Resolution and Few-shot View Synthesis from Asymmetric Dual-lens InputsPoster
  202. How Far Can a 1-Pixel Camera Go? Solving Vision Tasks using Photoreceptors and Computationally Designed Visual MorphologyPoster
  203. How Video Meetings Change Your ExpressionPoster
  204. Human Pose Recognition via Occlusion-Preserving Abstract ImagesPoster
  205. Human-in-the-Loop Visual Re-ID for Population Size EstimationPoster
  206. IAM-VFI : Interpolate Any Motion for Video Frame Interpolation with motion complexity mapPoster
  207. Implicit Steganography Beyond the Constraints of ModalityPoster
  208. Improving 3D Semi-supervised Learning by Effectively Utilizing All Unlabelled DataPoster
  209. Information Bottleneck Based Data Correction in Continual LearningPoster
  210. Instance-dependent Noisy-label Learning with Graphical Model Based Noise-rate EstimationPoster
  211. Inter-Class Topology Alignment for Efficient Black-Box Substitute AttacksPoster
  212. Interaction-centric Spatio-Temporal Context Reasoning for Multi-Person Video HOI RecognitionPoster
  213. Interactive 3D Object Detection with PromptsPoster
  214. Interpretability-Guided Test-Time Adversarial DefensePoster
  215. Investigating Style Similarity in Diffusion ModelsPoster
  216. Joint RGB-Spectral Decomposition Model Guided Image Enhancement in Mobile PhotographyPoster
  217. LASS3D: Language-Assisted Semi-Supervised 3D Semantic Segmentation with Progressive Unreliable Data ExploitationPoster
  218. LEROjD: Lidar Extended Radar-Only Object DetectionPoster
  219. LRSLAM: Low-rank Representation of Signed Distance Fields in Dense Visual SLAM SystemPoster
  220. Learn to Optimize Denoising Scores: A Unified and Improved Diffusion Prior for 3D GenerationPoster
  221. Learned Image Enhancement via Color NamingPoster
  222. Learning Dual-Level Deformable Implicit Representation for Real-World Scale Arbitrary Super-ResolutionPoster
  223. Learning Equilibrium Transformation for Gamut Expansion and Color RestorationPoster
  224. Learning Exhaustive Correlation for Spectral Super-Resolution: Where Spatial-Spectral Attention Meets Linear DependencePoster
  225. Learning Neural Deformation Representation for 4D Dynamic Shape GenerationPoster
  226. Learning Non-Linear Invariants for Unsupervised Out-of-Distribution DetectionPoster
  227. Learning Quantized Adaptive Conditions for Diffusion ModelsPoster
  228. Learning a Dynamic Privacy-preserving Camera Robust to Inversion AttacksOral
  229. Learning to Build by Building Your Own InstructionsPoster
  230. Learning to Robustly Reconstruct Dynamic Scenes from Low-light Spike StreamsPoster
  231. Learning with Counterfactual Explanations for Radiology Report GenerationPoster
  232. Learning-based Axial Video Motion MagnificationPoster
  233. Let the Avatar Talk using Texts without Paired Training DataPoster
  234. Leveraging Text Localization for Scene Text Removal via Text-aware Masked Image ModelingPoster
  235. Leveraging scale- and orientation-covariant features for planar motion estimationPoster
  236. LingoQA: Video Question Answering for Autonomous DrivingPoster
  237. Linking in Style: Understanding learned features in deep learning modelsPoster
  238. LoA-Trans: Enhancing Visual Grounding by Location-Aware TransformersPoster
  239. Loc3Diff: Local Diffusion for 3D Human Head Synthesis and EditingPoster
  240. M3DBench: Towards Omni 3D Assistant with Interleaved Multi-modal InstructionsPoster
  241. MC-PanDA: Mask Confidence for Panoptic Domain AdaptationPoster
  242. MOD-UV: Learning Mobile Object Detectors from Unlabeled VideosPoster
  243. MRSP: Learn Multi-Representations of Single Primitive for Compositional Zero-Shot LearningPoster
  244. MTaDCS: Moving Trace and Feature Density-based Confidence Sample Selection under Label NoisePoster
  245. MaRINeR: Enhancing Novel Views by Matching Rendered Images with Nearby ReferencesPoster
  246. McGrids: Monte Carlo-Driven Adaptive Grids for Iso-Surface ExtractionPoster
  247. MetaAT: Active Testing for Label-Efficient Evaluation of Dense Recognition TasksPoster
  248. MetaAug: Meta-Data Augmentation for Post-Training QuantizationPoster
  249. MetaWeather: Few-Shot Weather-Degraded Image RestorationPoster
  250. Mew: Multiplexed Immunofluorescence Image Analysis through an Efficient Multiplex NetworkPoster
  251. Modeling Label Correlations with Latent Context for Multi-Label RecognitionPoster
  252. Motion Keyframe Interpolation for Any Human Skeleton using Point Cloud-based Human Motion Data HomogenisationPoster
  253. Multi-Granularity Sparse Relationship Matrix Prediction Network for End-to-End Scene Graph GenerationPoster
  254. Multi-modal Relation Distillation for Unified 3D Representation LearningPoster
  255. Multi-scale Cross Distillation for Object Detection in Aerial ImagesPoster
  256. MultiGen: Zero-shot Image Generation from Multi-modal PromptsPoster
  257. Multimodal Label Relevance Ranking via Reinforcement LearningPoster
  258. Multiscale Graph Texture NetworkPoster
  259. NGP-RT: Fusing Multi-Level Hash Features with Lightweight Attention for Real-Time Novel View SynthesisPoster
  260. NeRF-XL: NeRF at Any Scale with Multi-GPUPoster
  261. NeRMo: Learning Implicit Neural Representations for 3D Human Motion PredictionOral
  262. Neural Poisson Solver: A Universal and Continuous Framework for Natural Signal BlendingPoster
  263. Neural graphics texture compression supporting random accessPoster
  264. Noise Calibration: Plug-and-play Content-Preserving Video Enhancement using Pre-trained Video Diffusion ModelsPoster
  265. Non-parametric Sensor Noise Modeling and SynthesisPoster
  266. Nymeria: A Massive Collection of Egocentric Multi-modal Human Motion in the WildPoster
  267. OLAF: A Plug-and-Play Framework for Enhanced Multi-object Multi-part Scene ParsingPoster
  268. Object-Aware NIR-to-Visible TranslationPoster
  269. Omni-Recon: Harnessing Image-based Rendering for General-Purpose Neural Radiance FieldsOral
  270. On Spectral Properties of Gradient-based Explanation MethodsPoster
  271. On the Evaluation Consistency of Attribution-based ExplanationsPoster
  272. On the Topology Awareness and Generalization Performance of Graph Neural NetworksOral
  273. On-the-fly Category Discovery for LiDAR Semantic SegmentationPoster
  274. Online Continuous Generalized Category DiscoveryPoster
  275. Online Temporal Action Localization with Memory-Augmented TransformerPoster
  276. Open-Vocabulary RGB-Thermal Semantic SegmentationPoster
  277. Open-set Domain Adaptation via Joint Error based Multi-class Positive and Unlabeled LearningPoster
  278. Optimal Transport of Diverse Unsupervised Tasks for Robust Learning from Noisy Few-Shot DataPoster
  279. Oulu Remote-photoplethysmography Physical Domain Attacks Database (ORPDAD)Poster
  280. PACE: Pose Annotations in Cluttered EnvironmentsPoster
  281. PAV: Personalized Head Avatar from Unstructured Video CollectionPoster
  282. POCA: Post-training Quantization with Temporal Alignment for Codec AvatarsPoster
  283. Panel-Specific Degradation Representation for Raw Under-Display Camera Image RestorationPoster
  284. Phase Concentration and Shortcut Suppression for Weakly Supervised Semantic SegmentationPoster
  285. Photon Inhibition for Energy-Efficient Single-Photon ImagingOral
  286. Physically Plausible Color Correction for Neural Radiance FieldsPoster
  287. Point-supervised Panoptic Segmentation via Estimating Pseudo Labels from Learnable DistancePoster
  288. PolyOculus: Simultaneous Multi-view Image-based Novel View SynthesisPoster
  289. PoseSOR: Human Pose Can Guide Our AttentionPoster
  290. Privacy-Preserving Adaptive Re-Identification without Image TransferOral
  291. Probabilistic Image-Driven Traffic Modeling via Remote SensingPoster
  292. Progressive Proxy Anchor Propagation for Unsupervised Semantic SegmentationPoster
  293. ProtoComp: Diverse Point Cloud Completion with Controllable PrototypePoster
  294. Pseudo-Embedding for Generalized Few-Shot Point Cloud SegmentationPoster
  295. Pseudo-Labelling Should Be Aware of Disguising Channel ActivationsPoster
  296. RANRAC: Robust Neural Scene Representations via Random Ray ConsensusPoster
  297. REFRAME: Reflective Surface Real-Time Rendering for Mobile DevicesPoster
  298. RING-NeRF : Rethinking Inductive Biases for Versatile and Efficient Neural FieldsPoster
  299. Region-Native Visual TokenizationPoster
  300. Region-aware Distribution Contrast: A Novel Approach to Multi-Task Partially Supervised LearningPoster
  301. Regularizing Dynamic Radiance Fields with Kinematic FieldsPoster
  302. Remove Projective LiDAR Depthmap Artifacts via Exploiting Epipolar GeometryPoster
  303. Removing Rows and Columns of Tokens in Vision Transformer enables Faster Dense Prediction without RetrainingPoster
  304. Representation Enhancement-Stabilization: Reducing Bias-Variance of Domain GeneralizationPoster
  305. Resolving Scale Ambiguity in Multi-view 3D Reconstruction using Dual-Pixel SensorsPoster
  306. Rethinking Data Bias: Dataset Copyright Protection via Embedding Class-wise Hidden BiasPoster
  307. Rethinking Directional Parameterization in Neural Implicit Surface ReconstructionPoster
  308. Rethinking Fast Adversarial Training: A Splitting Technique To Overcome Catastrophic OverfittingPoster
  309. Rethinking Features-Fused-Pyramid-Neck for Object DetectionPoster
  310. Rethinking Normalization Layers for Domain Generalizable Person Re-identificationPoster
  311. Retrieval Robust to Object Motion BlurPoster
  312. Revisit Self-supervision with Local Structure-from-MotionPoster
  313. Revisiting Feature Disentanglement Strategy in Diffusion Training and Breaking Conditional Independence Assumption in SamplingPoster
  314. Robust Fitting on a Gate Quantum ComputerOral
  315. Robustness Preserving Fine-tuning using Neuron ImportancePoster
  316. Rotated Orthographic Projection for Self-Supervised 3D Human Pose EstimationPoster
  317. SAH-SCI: Self-Supervised Adapter for Efficient Hyperspectral Snapshot Compressive ImagingPoster
  318. SCPNet: Unsupervised Cross-modal Homography Estimation via Intra-modal Self-supervised LearningPoster
  319. SDPT: Synchronous Dual Prompt Tuning for Fusion-based Visual-Language Pre-trained ModelsPoster
  320. SHINE: Saliency-aware HIerarchical NEgative Ranking for Compositional Temporal GroundingPoster
  321. SNP: Structured Neuron-level Pruning to Preserve Attention ScoresPoster
  322. ST-LDM: A Universal Framework for Text-Grounded Object Generation in Real ImagesPoster
  323. Scalar Function Topology Divergence: Comparing Topology of 3D ObjectsPoster
  324. Scaling Up Personalized Image Aesthetic Assessment via Task Vector CustomizationPoster
  325. SeA: Semantic Adversarial Augmentation for Last Layer Features from Unsupervised Representation LearningPoster
  326. Segmentation-guided Layer-wise Image Vectorization with Gradient FillsPoster
  327. SeiT++: Masked Token Modeling Improves Storage-efficient TrainingPoster
  328. Self-Supervised Video Copy Localization with Regional Token RepresentationPoster
  329. Self-supervised Shape Completion via Involution and Implicit CorrespondencesPoster
  330. Semantic-guided Robustness Tuning for Few-Shot Transfer Across Extreme Domain ShiftPoster
  331. Shapefusion: 3D localized human diffusion modelsPoster
  332. Single-Photon 3D Imaging with Equi-Depth Photon HistogramsPoster
  333. Sketch2Vox: Learning 3D Reconstruction from a Single Monocular Sketch ImagePoster
  334. Soft Shadow Diffusion (SSD): Physics-inspired Learning for 3D Computational PeriscopyPoster
  335. Solving the inverse problem of microscopy deconvolution with a residual Beylkin-Coifman-Rokhlin neural networkPoster
  336. SparseRadNet: Sparse Perception Neural Network on Subsampled Radar DataPoster
  337. SpatialFormer: Towards Generalizable Vision Transformers with Explicit Spatial UnderstandingPoster
  338. Spatio-Temporal Proximity-Aware Dual-Path Model for Panoramic Activity RecognitionPoster
  339. Spectral Subsurface Scattering for Material ClassificationPoster
  340. SpeedUpNet: A Plug-and-Play Adapter Network for Accelerating Text-to-Image Diffusion ModelsPoster
  341. Spline-based TransformersOral
  342. Stable Preference: Redefining training paradigm of human preference model for Text-to-Image SynthesisPoster
  343. Stepwise Multi-grained Boundary Detector for Point-supervised Temporal Action LocalizationPoster
  344. StereoGlue: Joint Feature Matching and Robust EstimationPoster
  345. Stripe Observation Guided Inference Cost-free Attention MechanismPoster
  346. SweepNet: Unsupervised Learning Shape Abstraction via Neural SweepersPoster
  347. Synthesizing Environment-Specific People in PhotographsPoster
  348. Synthesizing Time-varying BRDFs via Latent SpacePoster
  349. Textual-Visual Logic Challenge: Understanding and Reasoning in Text-to-Image GenerationPoster
  350. Time-Efficient and Identity-Consistent Virtual Try-On Using A Variant of Altered Diffusion ModelsPoster
  351. To Supervise or Not to Supervise: Understanding and Addressing the Key Challenges of Point Cloud Transfer LearningPoster
  352. Toward INT4 Fixed-Point Training via Exploring Quantization Error for GradientsPoster
  353. Towards Architecture-Agnostic Untrained Networks Priors for Image Reconstruction with Frequency RegularizationPoster
  354. Towards Certifiably Robust Face RecognitionPoster
  355. Towards Robust Full Low-bit Quantization of Super Resolution NetworksPoster
  356. Towards compact reversible image representations for neural style transferPoster
  357. Train Till You Drop: Towards Stable and Robust Source-free Unsupervised 3D Domain AdaptationPoster
  358. TreeSBA: Tree-Transformer for Self-Supervised Sequential Brick AssemblyPoster
  359. TurboEdit: Real-time text-based disentangled real image editingPoster
  360. UAV First-Person Viewers Are Radiance Field LearnersPoster
  361. URS-NeRF: Unordered Rolling Shutter Bundle Adjustment for Neural Radiance FieldsPoster
  362. Uncertainty-Driven Spectral Compressive Imaging with Spatial-Frequency TransformerPoster
  363. Understanding Multi-compositional learning in Vision and Language models via Category TheoryPoster
  364. UniVoxel: Fast Inverse Rendering by Unified Voxelization of Scene RepresentationPoster
  365. Unified Local-Cloud Decision-Making via Reinforcement LearningPoster
  366. Unleashing Text-to-Image Diffusion Prior for Zero-Shot Image CaptioningPoster
  367. Unrolled Decomposed Unpaired Learning for Controllable Low-Light Video EnhancementPoster
  368. Unsqueeze [CLS] Bottleneck to Learn Rich RepresentationsPoster
  369. Unsupervised Representation Learning by Balanced Self Attention MatchingPoster
  370. Unveiling Privacy Risks in Stochastic Neural Networks Training: Effective Image Reconstruction from GradientsPoster
  371. Upper-body Hierarchical Graph for Skeleton Based Emotion Recognition in Assistive DrivingPoster
  372. V-Trans4Style: Visual Transition Recommendation for Video Production Style AdaptationPoster
  373. VETRA: A Dataset for Vehicle Tracking in Aerial Imagery - New Challenges for Multi-Object TrackingOral
  374. VF-NeRF: Viewshed Fields for Rigid NeRF RegistrationPoster
  375. VP-SAM: Taming Segment Anything Model for Video Polyp Segmentation via Disentanglement and Spatio-temporal Side NetworkPoster
  376. VideoClusterNet: Self-Supervised and Adaptive Face Clustering for VideosPoster
  377. View-Consistent Hierarchical 3D Segmentation Using Ultrametric Feature FieldsPoster
  378. Visual Prompting via Partial Optimal TransportPoster
  379. Visual Relationship TransformationPoster
  380. Wavelength-Embedding-guided Filter-Array Transformer for Spectral DemosaicingPoster
  381. Weakly-Supervised 3D Hand Reconstruction with Knowledge Prior and Uncertainty GuidancePoster
  382. When and How do negative prompts take effect?Poster
  383. WindPoly: Polygonal Mesh Reconstruction via Winding NumbersPoster
  384. ZoLA: Zero-Shot Creative Long Animation Generation with Short Video ModelOral
  385. uCAP: An Unsupervised Prompting Method for Vision-Language ModelsOral

ECCV accepted papers in other years

Looking for submission deadlines instead? See the conference deadline calendar.