The multi-volume set of LNCS books with volume numbers 15059 up to 15147 constitutes the refereed proceedings of the 18th European Conference on Computer Vision, ECCV 2024, held in Milan, Italy, during September 29October 4, 2024.
The 2387 papers presented in these proceedings were carefully reviewed and selected from a total of 8585 submissions. They deal with topics such as computer vision; machine learning; deep neural networks; reinforcement learning; object recognition; image classification; image processing; object detection; semantic segmentation; human pose estimation; 3d reconstruction; stereo vision; computational photography; neural networks; image coding; image reconstruction; motion estimation.
Ex2Eg-MAE: A Framework for Adaptation of Exocentric Video Masked Autoencoders for Egocentric Social Role Understanding.- Self-Supervised Audio-Visual Soundscape Stylization.- SAVE: Protagonist Diversification with Structure Agnostic Video Editing.- VideoAgent: Long-form Video Understanding with Large Language Model as Agent.- Meta-optimized Angular Margin Contrastive Framework for Video-Language Representation Learning.- Source-Free Domain-Invariant Performance Prediction.- Improving Robustness to Model Inversion Attacks via Sparse Coding Architectures.- Constructing Concept-based Models to Mitigate Spurious Correlations with Minimal Human Effort.- Direct Distillation between Different Domains.- Contrastive ground-level image and remote sensing pre-training improves representation learning for natural world imagery.- V-Trans4Style: Visual Transition Recommendation for Video Production Style Adaptation.- GRiT: A Generative Region-to-text Transformer for Object Understanding.- LRSLAM: Low-rank Representation of Signed Distance Fields in Dense Visual SLAM System.- Learning Representation for Multitask Learning through Self-Supervised Auxiliary Learning.- Neural Poisson Solver: A Universal and Continuous Framewol³R