LEADER 05793nam 22008295 450 001 9910983329703321 005 20250626163527.0 010 $a3-031-72946-3 024 7 $a10.1007/978-3-031-72946-1 035 $a(CKB)36251569200041 035 $a(MiAaPQ)EBC31696316 035 $a(Au-PeEL)EBL31696316 035 $a(DE-He213)978-3-031-72946-1 035 $a(OCoLC)1458758568 035 $a(EXLCZ)9936251569200041 100 $a20241002d2025 u| 0 101 0 $aeng 135 $aur||||||||||| 181 $ctxt$2rdacontent 182 $cc$2rdamedia 183 $acr$2rdacarrier 200 10$aComputer Vision ? ECCV 2024 $e18th European Conference, Milan, Italy, September 29?October 4, 2024, Proceedings, Part XLII /$fedited by Ale? Leonardis, Elisa Ricci, Stefan Roth, Olga Russakovsky, Torsten Sattler, Gül Varol 205 $a1st ed. 2025. 210 1$aCham :$cSpringer Nature Switzerland :$cImprint: Springer,$d2025. 215 $a1 online resource (583 pages) 225 1 $aLecture Notes in Computer Science,$x1611-3349 ;$v15100 311 08$a3-031-72945-5 327 $aOpen-Set Recognition in the Age of Vision-Language Models -- Unsqueeze [CLS] Bottleneck to Learn Rich Representations -- Robust Multimodal Learning via Representation Decoupling -- Object-Conditioned Energy-Based Attention Map Alignment in Text-to-Image Diffusion Models -- WiMANS: A Benchmark Dataset for WiFi-based Multi-user Activity Sensing -- Embedding-Free Transformer with Inference Spatial Reduction for Efficient Semantic Segmentation -- VeCLIP: Improving CLIP Training via Visual-enriched Captions -- Three Things We Need to Know About Transferring Stable Diffusion to Visual Dense Prediciton Tasks -- Learning Representations from Foundation Models for Domain Generalized Stereo Matching -- Spike-Temporal Latent Representation for Energy-Efficient Event-to-Video Reconstruction -- Effective Lymph Nodes Detection in CT Scans Using Location Debiased Query Selection and Contrastive Query Representation in Transformer -- Chat-Edit-3D: Interactive 3D Scene Editing via Text Prompts -- Event-Adapted Video Super-Resolution -- Look Hear: Gaze Prediction for Speech-directed Human Attention -- Raising the Ceiling: Conflict-Free Local Feature Matching with Dynamic View Switching -- Q&A Prompts: Discovering Rich Visual Clues through Mining Question-Answer Prompts for VQA requiring Diverse World Knowledge -- Catastrophic Overfitting: A Potential Blessing in Disguise -- Long-range Turbulence Mitigation: A Large-scale Dataset and A Coarse-to-fine Framework -- SparseCtrl: Adding Sparse Controls to Text-to-Video Diffusion Models -- Visual Alignment Pre-training for Sign Language Translation -- Parrot Captions Teach CLIP to Spot Text -- Solving Motion Planning Tasks with a Scalable Generative Model -- Griffon: Spelling out All Object Locations at Any Granularity with Large Language Models -- Vision-Language Action Knowledge Learning for Semantic-Aware Action Quality Assessment -- Knowledge Transfer with Simulated Inter-Image Erasing for Weakly Supervised Semantic Segmentation -- BurstM: Deep Burst Multi-scale SR using Fourier Space with Optical Flow -- Diffusion Reward: Learning Rewards via Conditional Video Diffusion. 330 $aThe multi-volume set of LNCS books with volume numbers 15059 up to 15147 constitutes the refereed proceedings of the 18th European Conference on Computer Vision, ECCV 2024, held in Milan, Italy, during September 29?October 4, 2024. The 2387 papers presented in these proceedings were carefully reviewed and selected from a total of 8585 submissions. The papers deal with topics such as computer vision; machine learning; deep neural networks; reinforcement learning; object recognition; image classification; image processing; object detection; semantic segmentation; human pose estimation; 3d reconstruction; stereo vision; computational photography; neural networks; image coding; image reconstruction; motion estimation. . 410 0$aLecture Notes in Computer Science,$x1611-3349 ;$v15100 606 $aImage processing$xDigital techniques 606 $aComputer vision 606 $aImage processing 606 $aComputer networks 606 $aUser interfaces (Computer systems) 606 $aHuman-computer interaction 606 $aMachine learning 606 $aComputers, Special purpose 606 $aComputer Imaging, Vision, Pattern Recognition and Graphics 606 $aImage Processing 606 $aComputer Communication Networks 606 $aUser Interfaces and Human Computer Interaction 606 $aMachine Learning 606 $aSpecial Purpose and Application-Based Systems 615 0$aImage processing$xDigital techniques. 615 0$aComputer vision. 615 0$aImage processing. 615 0$aComputer networks. 615 0$aUser interfaces (Computer systems) 615 0$aHuman-computer interaction. 615 0$aMachine learning. 615 0$aComputers, Special purpose. 615 14$aComputer Imaging, Vision, Pattern Recognition and Graphics. 615 24$aImage Processing. 615 24$aComputer Communication Networks. 615 24$aUser Interfaces and Human Computer Interaction. 615 24$aMachine Learning. 615 24$aSpecial Purpose and Application-Based Systems. 676 $a006 700 $aLeonardis$b Ales?$01757791 701 $aRicci$b Elisa$0216674 701 $aRoth$b S?tefan$00 701 $aRussakovsky$b Olga$01767663 701 $aSattler$b Torsten$01767664 701 $aVarol$b Gül$01767665 801 0$bMiAaPQ 801 1$bMiAaPQ 801 2$bMiAaPQ 906 $aBOOK 912 $a9910983329703321 996 $aComputer Vision ? ECCV 2024$94213977 997 $aUNINA