東華大學圖書館 |

Computer vision - ECCV 2024 = 18th European Conference, Milan, Italy, September 29-October 4, 2024 : proceedings.. Part VIII /

Record Type:	Electronic resources : Monograph/item
Title/Author:	Computer vision - ECCV 2024/ edited by Aleš Leonardis ... [et al.].
Reminder of title:	18th European Conference, Milan, Italy, September 29-October 4, 2024 : proceedings.
remainder title:	ECCV 2024
other author:	Leonardis, Aleš.
corporate name:	European Conference on Computer Vision
Published:	Cham :Springer Nature Switzerland : : 2025.,
Description:	lxxxv, 499 p. :ill. (chiefly color), digital ;24 cm.
[NT 15003449]:	Walker: Self-supervised Multiple Object Tracking by Walking on Temporal Object Appearance Graphs -- Spatio-Temporal Proximity-Aware Dual-Path Model for Panoramic Activity Recognition -- DiffiT: Diffusion Vision Transformers for Image Generation -- WebRPG: Automatic Web Rendering Parameters Generation for Visual Presentation -- GPSFormer: A Global Perception and Local Structure Fitting-based Transformer for Point Cloud Understanding -- FreeMotion: A Unified Framework for Number-free Text-to-Motion Synthesis -- FSD-BEV: Foreground Self-Distillation for Multi-view 3D Object Detection -- SceneGraphLoc: Cross-Modal Coarse Visual Localization on 3D Scene Graphs -- ScanReason: Empowering 3D Visual Grounding with Reasoning Capabilities -- MathVerse: Does Your Multi-modal LLM Truly See the Diagrams in Visual Math Problems? -- See and Think: Embodied Agent in Virtual Environment -- PISR: Polarimetric Neural Implicit Surface Reconstruction for Textureless and Specular Objects -- Bridging the Gap Between Human Motion and Action Semantics via Kinematics Phrases -- VisFocus: Prompt-Guided Vision Encoders for OCR-Free Dense Document Understanding -- Masked Angle-Aware Autoencoder for Remote Sensing Images -- Infinite-ID: Identity-preserved Personalization via ID-semantics Decoupling Paradigm -- MultiGen: Zero-shot Image Generation from Multi-modal Prompts -- GazeXplain: Learning to Predict Natural Language Explanations of Visual Scanpaths -- Learning Chain of Counterfactual Thought for Bias-Robust Vision-Language Reasoning -- SegGen: Supercharging Segmentation Models with Text2Mask and Mask2Img Synthesis -- Sync from the Sea: Retrieving Alignable Videos from Large-Scale Datasets -- FinePseudo: Improving Pseudo-Labelling through Temporal-Alignablity for Semi-Supervised Fine-Grained Action Recognition -- Elegantly Written: Disentangling Writer and Character Styles for Enhancing Online Chinese Handwriting -- UniCode : Learning a Unified Codebook for Multimodal Large Language Models -- When Do We Not Need Larger Vision Models? -- GVGEN: Text-to-3D Generation with Volumetric Representation -- Bidirectional Stereo Image Compression with Cross-Dimensional Entropy Model.
Contained By:	Springer Nature eBook
Subject:	Computer vision - Congresses. -
Online resource:	https://doi.org/10.1007/978-3-031-73242-3
ISBN:	9783031732423

Computer vision - ECCV 2024 = 18th European Conference, Milan, Italy, September 29-October 4, 2024 : proceedings.. Part VIII /
Computer vision - ECCV 202418th European Conference, Milan, Italy, September 29-October 4, 2024 : proceedings.Part VIII /[electronic resource] :ECCV 2024edited by Aleš Leonardis ... [et al.]. - Cham :Springer Nature Switzerland :2025. - lxxxv, 499 p. :ill. (chiefly color), digital ;24 cm. - Lecture notes in computer science,150661611-3349 ;. - Lecture notes in computer science ;15066..

Walker: Self-supervised Multiple Object Tracking by Walking on Temporal Object Appearance Graphs -- Spatio-Temporal Proximity-Aware Dual-Path Model for Panoramic Activity Recognition -- DiffiT: Diffusion Vision Transformers for Image Generation -- WebRPG: Automatic Web Rendering Parameters Generation for Visual Presentation -- GPSFormer: A Global Perception and Local Structure Fitting-based Transformer for Point Cloud Understanding -- FreeMotion: A Unified Framework for Number-free Text-to-Motion Synthesis -- FSD-BEV: Foreground Self-Distillation for Multi-view 3D Object Detection -- SceneGraphLoc: Cross-Modal Coarse Visual Localization on 3D Scene Graphs -- ScanReason: Empowering 3D Visual Grounding with Reasoning Capabilities -- MathVerse: Does Your Multi-modal LLM Truly See the Diagrams in Visual Math Problems? -- See and Think: Embodied Agent in Virtual Environment -- PISR: Polarimetric Neural Implicit Surface Reconstruction for Textureless and Specular Objects -- Bridging the Gap Between Human Motion and Action Semantics via Kinematics Phrases -- VisFocus: Prompt-Guided Vision Encoders for OCR-Free Dense Document Understanding -- Masked Angle-Aware Autoencoder for Remote Sensing Images -- Infinite-ID: Identity-preserved Personalization via ID-semantics Decoupling Paradigm -- MultiGen: Zero-shot Image Generation from Multi-modal Prompts -- GazeXplain: Learning to Predict Natural Language Explanations of Visual Scanpaths -- Learning Chain of Counterfactual Thought for Bias-Robust Vision-Language Reasoning -- SegGen: Supercharging Segmentation Models with Text2Mask and Mask2Img Synthesis -- Sync from the Sea: Retrieving Alignable Videos from Large-Scale Datasets -- FinePseudo: Improving Pseudo-Labelling through Temporal-Alignablity for Semi-Supervised Fine-Grained Action Recognition -- Elegantly Written: Disentangling Writer and Character Styles for Enhancing Online Chinese Handwriting -- UniCode : Learning a Unified Codebook for Multimodal Large Language Models -- When Do We Not Need Larger Vision Models? -- GVGEN: Text-to-3D Generation with Volumetric Representation -- Bidirectional Stereo Image Compression with Cross-Dimensional Entropy Model.

The multi-volume set of LNCS books with volume numbers 15059 up to 15147 constitutes the refereed proceedings of the 18th European Conference on Computer Vision, ECCV 2024, held in Milan, Italy, during September 29-October 4, 2024. The 2387 papers presented in these proceedings were carefully reviewed and selected from a total of 8585 submissions. The papers deal with topics such as computer vision; machine learning; deep neural networks; reinforcement learning; object recognition; image classification; image processing; object detection; semantic segmentation; human pose estimation; 3d reconstruction; stereo vision; computational photography; neural networks; image coding; image reconstruction; motion estimation.

ISBN: 9783031732423

Standard No.: 10.1007/978-3-031-73242-3doiSubjects--Topical Terms:

570734
Computer vision
--Congresses.

LC Class. No.: TA1634

Dewey Class. No.: 006.37

Computer vision - ECCV 2024 = 18th European Conference, Milan, Italy, September 29-October 4, 2024 : proceedings.. Part VIII /
LDR:04076nmm a2200349 a 4500 001 2407896
003 DE-He213
005 20241028115754.0
006 m d
007 cr nn 008maaau
008 260204s2025 sz s 0 eng d
020 $a 9783031732423 $q (electronic bk.)
020 $a 9783031732416 $q (paper)
024 7 $a 10.1007/978-3-031-73242-3 $2 doi
035 $a 978-3-031-73242-3
040 $a GP $c GP
041 0 $a eng
050 4 $a TA1634
072 7 $a UYT $2 bicssc
072 7 $a COM016000 $2 bisacsh
072 7 $a UYT $2 thema
082 0 4 $a 006.37 $2 23
090 $a TA1634 $b .E89 2024
111 2 $a European Conference on Computer Vision $n (18th : $d 2024 : $c Milan, Italy) $3 3733323
245 1 0 $a Computer vision - ECCV 2024 $h [electronic resource] : $b 18th European Conference, Milan, Italy, September 29-October 4, 2024 : proceedings. $n Part VIII / $c edited by Aleš Leonardis ... [et al.].
246 3 $a ECCV 2024
260 $a Cham : $b Springer Nature Switzerland : $b Imprint: Springer, $c 2025.
300 $a lxxxv, 499 p. : $b ill. (chiefly color), digital ; $c 24 cm.
490 1 $a Lecture notes in computer science, $x 1611-3349 ; $v 15066
505 0 $a Walker: Self-supervised Multiple Object Tracking by Walking on Temporal Object Appearance Graphs -- Spatio-Temporal Proximity-Aware Dual-Path Model for Panoramic Activity Recognition -- DiffiT: Diffusion Vision Transformers for Image Generation -- WebRPG: Automatic Web Rendering Parameters Generation for Visual Presentation -- GPSFormer: A Global Perception and Local Structure Fitting-based Transformer for Point Cloud Understanding -- FreeMotion: A Unified Framework for Number-free Text-to-Motion Synthesis -- FSD-BEV: Foreground Self-Distillation for Multi-view 3D Object Detection -- SceneGraphLoc: Cross-Modal Coarse Visual Localization on 3D Scene Graphs -- ScanReason: Empowering 3D Visual Grounding with Reasoning Capabilities -- MathVerse: Does Your Multi-modal LLM Truly See the Diagrams in Visual Math Problems? -- See and Think: Embodied Agent in Virtual Environment -- PISR: Polarimetric Neural Implicit Surface Reconstruction for Textureless and Specular Objects -- Bridging the Gap Between Human Motion and Action Semantics via Kinematics Phrases -- VisFocus: Prompt-Guided Vision Encoders for OCR-Free Dense Document Understanding -- Masked Angle-Aware Autoencoder for Remote Sensing Images -- Infinite-ID: Identity-preserved Personalization via ID-semantics Decoupling Paradigm -- MultiGen: Zero-shot Image Generation from Multi-modal Prompts -- GazeXplain: Learning to Predict Natural Language Explanations of Visual Scanpaths -- Learning Chain of Counterfactual Thought for Bias-Robust Vision-Language Reasoning -- SegGen: Supercharging Segmentation Models with Text2Mask and Mask2Img Synthesis -- Sync from the Sea: Retrieving Alignable Videos from Large-Scale Datasets -- FinePseudo: Improving Pseudo-Labelling through Temporal-Alignablity for Semi-Supervised Fine-Grained Action Recognition -- Elegantly Written: Disentangling Writer and Character Styles for Enhancing Online Chinese Handwriting -- UniCode : Learning a Unified Codebook for Multimodal Large Language Models -- When Do We Not Need Larger Vision Models? -- GVGEN: Text-to-3D Generation with Volumetric Representation -- Bidirectional Stereo Image Compression with Cross-Dimensional Entropy Model.
520 $a The multi-volume set of LNCS books with volume numbers 15059 up to 15147 constitutes the refereed proceedings of the 18th European Conference on Computer Vision, ECCV 2024, held in Milan, Italy, during September 29-October 4, 2024. The 2387 papers presented in these proceedings were carefully reviewed and selected from a total of 8585 submissions. The papers deal with topics such as computer vision; machine learning; deep neural networks; reinforcement learning; object recognition; image classification; image processing; object detection; semantic segmentation; human pose estimation; 3d reconstruction; stereo vision; computational photography; neural networks; image coding; image reconstruction; motion estimation.
650 0 $a Computer vision $x Congresses. $3 570734
650 0 $a Pattern recognition systems $v Congresses. $3 563039
650 1 4 $a Computer Imaging, Vision, Pattern Recognition and Graphics. $3 890871
650 2 4 $a Image Processing. $3 891209
650 2 4 $a Computer Communication Networks. $3 775497
650 2 4 $a User Interfaces and Human Computer Interaction. $3 892554
650 2 4 $a Machine Learning. $3 3382522
650 2 4 $a Special Purpose and Application-Based Systems. $3 892492
700 1 $a Leonardis, Aleš. $3 3733324
710 2 $a SpringerLink (Online service) $3 836513
773 0 $t Springer Nature eBook
830 0 $a Lecture notes in computer science ; $v 15066. $3 3780055
856 4 0 $u https://doi.org/10.1007/978-3-031-73242-3
950 $a Computer Science (SpringerNature-11645)