LEADER 05842nam 22008295 450 001 9910983305203321 005 20250626164128.0 010 $a3-031-73242-1 024 7 $a10.1007/978-3-031-73242-3 035 $a(CKB)36431149400041 035 $a(MiAaPQ)EBC31743876 035 $a(Au-PeEL)EBL31743876 035 $a(DE-He213)978-3-031-73242-3 035 $a(OCoLC)1467876671 035 $a(EXLCZ)9936431149400041 100 $a20241028d2025 u| 0 101 0 $aeng 135 $aur||||||||||| 181 $ctxt$2rdacontent 182 $cc$2rdamedia 183 $acr$2rdacarrier 200 10$aComputer Vision ? ECCV 2024 $e18th European Conference, Milan, Italy, September 29?October 4, 2024, Proceedings, Part VIII /$fedited by Ale? Leonardis, Elisa Ricci, Stefan Roth, Olga Russakovsky, Torsten Sattler, Gül Varol 205 $a1st ed. 2025. 210 1$aCham :$cSpringer Nature Switzerland :$cImprint: Springer,$d2025. 215 $a1 online resource (584 pages) 225 1 $aLecture Notes in Computer Science,$x1611-3349 ;$v15066 311 08$a3-031-73241-3 327 $aWalker: Self-supervised Multiple Object Tracking by Walking on Temporal Object Appearance Graphs -- Spatio-Temporal Proximity-Aware Dual-Path Model for Panoramic Activity Recognition -- DiffiT: Diffusion Vision Transformers for Image Generation -- WebRPG: Automatic Web Rendering Parameters Generation for Visual Presentation -- GPSFormer: A Global Perception and Local Structure Fitting-based Transformer for Point Cloud Understanding -- FreeMotion: A Unified Framework for Number-free Text-to-Motion Synthesis -- FSD-BEV: Foreground Self-Distillation for Multi-view 3D Object Detection -- SceneGraphLoc: Cross-Modal Coarse Visual Localization on 3D Scene Graphs -- ScanReason: Empowering 3D Visual Grounding with Reasoning Capabilities -- MathVerse: Does Your Multi-modal LLM Truly See the Diagrams in Visual Math Problems? -- See and Think: Embodied Agent in Virtual Environment -- PISR: Polarimetric Neural Implicit Surface Reconstruction for Textureless and Specular Objects -- Bridging the Gap Between Human Motion and Action Semantics via Kinematics Phrases -- VisFocus: Prompt-Guided Vision Encoders for OCR-Free Dense Document Understanding -- Masked Angle-Aware Autoencoder for Remote Sensing Images -- Infinite-ID: Identity-preserved Personalization via ID-semantics Decoupling Paradigm -- MultiGen: Zero-shot Image Generation from Multi-modal Prompts -- GazeXplain: Learning to Predict Natural Language Explanations of Visual Scanpaths -- Learning Chain of Counterfactual Thought for Bias-Robust Vision-Language Reasoning -- SegGen: Supercharging Segmentation Models with Text2Mask and Mask2Img Synthesis -- Sync from the Sea: Retrieving Alignable Videos from Large-Scale Datasets -- FinePseudo: Improving Pseudo-Labelling through Temporal-Alignablity for Semi-Supervised Fine-Grained Action Recognition -- Elegantly Written: Disentangling Writer and Character Styles for Enhancing Online Chinese Handwriting -- UniCode : Learning a Unified Codebook for Multimodal Large Language Models -- When Do We Not Need Larger Vision Models? -- GVGEN: Text-to-3D Generation with Volumetric Representation -- Bidirectional Stereo Image Compression with Cross-Dimensional Entropy Model. 330 $aThe multi-volume set of LNCS books with volume numbers 15059 up to 15147 constitutes the refereed proceedings of the 18th European Conference on Computer Vision, ECCV 2024, held in Milan, Italy, during September 29?October 4, 2024. The 2387 papers presented in these proceedings were carefully reviewed and selected from a total of 8585 submissions. The papers deal with topics such as computer vision; machine learning; deep neural networks; reinforcement learning; object recognition; image classification; image processing; object detection; semantic segmentation; human pose estimation; 3d reconstruction; stereo vision; computational photography; neural networks; image coding; image reconstruction; motion estimation. 410 0$aLecture Notes in Computer Science,$x1611-3349 ;$v15066 606 $aImage processing$xDigital techniques 606 $aComputer vision 606 $aImage processing 606 $aComputer networks 606 $aUser interfaces (Computer systems) 606 $aHuman-computer interaction 606 $aMachine learning 606 $aComputers, Special purpose 606 $aComputer Imaging, Vision, Pattern Recognition and Graphics 606 $aImage Processing 606 $aComputer Communication Networks 606 $aUser Interfaces and Human Computer Interaction 606 $aMachine Learning 606 $aSpecial Purpose and Application-Based Systems 615 0$aImage processing$xDigital techniques. 615 0$aComputer vision. 615 0$aImage processing. 615 0$aComputer networks. 615 0$aUser interfaces (Computer systems) 615 0$aHuman-computer interaction. 615 0$aMachine learning. 615 0$aComputers, Special purpose. 615 14$aComputer Imaging, Vision, Pattern Recognition and Graphics. 615 24$aImage Processing. 615 24$aComputer Communication Networks. 615 24$aUser Interfaces and Human Computer Interaction. 615 24$aMachine Learning. 615 24$aSpecial Purpose and Application-Based Systems. 676 $a006.37 700 $aLeonardis$b Ales?$01757791 701 $aRicci$b Elisa$0216674 701 $aRoth$b S?tefan$00 701 $aRussakovsky$b Olga$01767663 701 $aSattler$b Torsten$01767664 701 $aVarol$b Gül$01767665 801 0$bMiAaPQ 801 1$bMiAaPQ 801 2$bMiAaPQ 906 $aBOOK 912 $a9910983305203321 996 $aComputer Vision ? ECCV 2024$94213977 997 $aUNINA