東華大學圖書館 |

Computer vision - ECCV 2024 = 18th European Conference, Milan, Italy, September 29-October 4, 2024 : proceedings.. Part XXIII /

Record Type:	Electronic resources : Monograph/item
Title/Author:	Computer vision - ECCV 2024/ edited by Aleš Leonardis ... [et al.].
Reminder of title:	18th European Conference, Milan, Italy, September 29-October 4, 2024 : proceedings.
remainder title:	ECCV 2024
other author:	Leonardis, Aleš.
corporate name:	European Conference on Computer Vision
Published:	Cham :Springer Nature Switzerland : : 2025.,
Description:	lxxxv, 496 p. :ill. (chiefly color), digital ;24 cm.
[NT 15003449]:	Weak-to-Strong Compositional Learning from Generative Models for Language-based Object Detection -- Domesticating SAM for Breast Ultrasound Image Segmentation via Spatial-frequency Fusion and Uncertainty Correction -- CanonicalFusion: Generating Drivable 3D Human Avatars from Multiple Images -- Camera Height Doesn't Change: Unsupervised Training for Metric Monocular Road-Scene Depth Estimation -- Uni3DL: A Unified Model for 3D Vision-Language Understanding -- Object-Aware NIR-to-Visible Translation -- PaPr: Training-Free One-Step Patch Pruning with Lightweight ConvNets for Faster Inference -- GENIXER: Empowering Multimodal Large Language Models as a Powerful Data Generator -- BLINK: Multimodal Large Language Models Can See but Not Perceive -- AFF-ttention! Affordances and Attention models for Short-Term Object Interaction Anticipation -- PreLAR: World Model Pre-training with Learnable Action Representation -- Multi-HMR: Multi-Person Whole-Body Human Mesh Recovery in a Single Shot -- De-confounded Gaze Estimation -- Diffusion Models for Monocular Depth Estimation: Overcoming Challenging Conditions -- FreestyleRet: Retrieving Images from Style-Diversified Queries -- ReGround: Improving Textual and Spatial Grounding at No Cost -- CardiacNet: Learning to Reconstruct Abnormalities for Cardiac Disease Assessment from Echocardiogram Videos -- LaMI-DETR: Open-Vocabulary Detection with Language Model Instruction -- Unrolled Decomposed Unpaired Learning for Controllable Low-Light Video Enhancement -- Efficient Image Pre-Training with Siamese Cropped Masked Autoencoders -- VP-SAM: Taming Segment Anything Model for Video Polyp Segmentation via Disentanglement and Spatio-temporal Side Network -- Dataset Enhancement with Instance-Level Augmentations -- FreeMotion: MoCap-Free Human Motion Synthesis with Multimodal Large Language Models -- Chameleon: A Data-Efficient Generalist for Dense Visual Prediction in the Wild -- Reliability in Semantic Segmentation: Can We Use Synthetic Data? -- SCPNet: Unsupervised Cross-modal Homography Estimation via Intra-modal Self-supervised Learning -- SCAPE: A Simple and Strong Category-Agnostic Pose Estimator.
Contained By:	Springer Nature eBook
Subject:	Computer vision - Congresses. -
Online resource:	https://doi.org/10.1007/978-3-031-73337-6
ISBN:	9783031733376

Computer vision - ECCV 2024 = 18th European Conference, Milan, Italy, September 29-October 4, 2024 : proceedings.. Part XXIII /
Computer vision - ECCV 202418th European Conference, Milan, Italy, September 29-October 4, 2024 : proceedings.Part XXIII /[electronic resource] :ECCV 2024edited by Aleš Leonardis ... [et al.]. - Cham :Springer Nature Switzerland :2025. - lxxxv, 496 p. :ill. (chiefly color), digital ;24 cm. - Lecture notes in computer science,150811611-3349 ;. - Lecture notes in computer science ;15081..

Weak-to-Strong Compositional Learning from Generative Models for Language-based Object Detection -- Domesticating SAM for Breast Ultrasound Image Segmentation via Spatial-frequency Fusion and Uncertainty Correction -- CanonicalFusion: Generating Drivable 3D Human Avatars from Multiple Images -- Camera Height Doesn't Change: Unsupervised Training for Metric Monocular Road-Scene Depth Estimation -- Uni3DL: A Unified Model for 3D Vision-Language Understanding -- Object-Aware NIR-to-Visible Translation -- PaPr: Training-Free One-Step Patch Pruning with Lightweight ConvNets for Faster Inference -- GENIXER: Empowering Multimodal Large Language Models as a Powerful Data Generator -- BLINK: Multimodal Large Language Models Can See but Not Perceive -- AFF-ttention! Affordances and Attention models for Short-Term Object Interaction Anticipation -- PreLAR: World Model Pre-training with Learnable Action Representation -- Multi-HMR: Multi-Person Whole-Body Human Mesh Recovery in a Single Shot -- De-confounded Gaze Estimation -- Diffusion Models for Monocular Depth Estimation: Overcoming Challenging Conditions -- FreestyleRet: Retrieving Images from Style-Diversified Queries -- ReGround: Improving Textual and Spatial Grounding at No Cost -- CardiacNet: Learning to Reconstruct Abnormalities for Cardiac Disease Assessment from Echocardiogram Videos -- LaMI-DETR: Open-Vocabulary Detection with Language Model Instruction -- Unrolled Decomposed Unpaired Learning for Controllable Low-Light Video Enhancement -- Efficient Image Pre-Training with Siamese Cropped Masked Autoencoders -- VP-SAM: Taming Segment Anything Model for Video Polyp Segmentation via Disentanglement and Spatio-temporal Side Network -- Dataset Enhancement with Instance-Level Augmentations -- FreeMotion: MoCap-Free Human Motion Synthesis with Multimodal Large Language Models -- Chameleon: A Data-Efficient Generalist for Dense Visual Prediction in the Wild -- Reliability in Semantic Segmentation: Can We Use Synthetic Data? -- SCPNet: Unsupervised Cross-modal Homography Estimation via Intra-modal Self-supervised Learning -- SCAPE: A Simple and Strong Category-Agnostic Pose Estimator.

The multi-volume set of LNCS books with volume numbers 15059 up to 15147 constitutes the refereed proceedings of the 18th European Conference on Computer Vision, ECCV 2024, held in Milan, Italy, during September 29-October 4, 2024. The 2387 papers presented in these proceedings were carefully reviewed and selected from a total of 8585 submissions. They deal with topics such as computer vision; machine learning; deep neural networks; reinforcement learning; object recognition; image classification; image processing; object detection; semantic segmentation; human pose estimation; 3d reconstruction; stereo vision; computational photography; neural networks; image coding; image reconstruction; motion estimation.

ISBN: 9783031733376

Standard No.: 10.1007/978-3-031-73337-6doiSubjects--Topical Terms:

570734
Computer vision
--Congresses.

LC Class. No.: TA1634

Dewey Class. No.: 006.37

Computer vision - ECCV 2024 = 18th European Conference, Milan, Italy, September 29-October 4, 2024 : proceedings.. Part XXIII /
LDR:04050nmm a2200349 a 4500 001 2407910
003 DE-He213
005 20241031115742.0
006 m d
007 cr nn 008maaau
008 260204s2025 sz s 0 eng d
020 $a 9783031733376 $q (electronic bk.)
020 $a 9783031733369 $q (paper)
024 7 $a 10.1007/978-3-031-73337-6 $2 doi
035 $a 978-3-031-73337-6
040 $a GP $c GP
041 0 $a eng
050 4 $a TA1634
072 7 $a UYT $2 bicssc
072 7 $a COM016000 $2 bisacsh
072 7 $a UYT $2 thema
082 0 4 $a 006.37 $2 23
090 $a TA1634 $b .E89 2024
111 2 $a European Conference on Computer Vision $n (18th : $d 2024 : $c Milan, Italy) $3 3733323
245 1 0 $a Computer vision - ECCV 2024 $h [electronic resource] : $b 18th European Conference, Milan, Italy, September 29-October 4, 2024 : proceedings. $n Part XXIII / $c edited by Aleš Leonardis ... [et al.].
246 3 $a ECCV 2024
260 $a Cham : $b Springer Nature Switzerland : $b Imprint: Springer, $c 2025.
300 $a lxxxv, 496 p. : $b ill. (chiefly color), digital ; $c 24 cm.
490 1 $a Lecture notes in computer science, $x 1611-3349 ; $v 15081
505 0 $a Weak-to-Strong Compositional Learning from Generative Models for Language-based Object Detection -- Domesticating SAM for Breast Ultrasound Image Segmentation via Spatial-frequency Fusion and Uncertainty Correction -- CanonicalFusion: Generating Drivable 3D Human Avatars from Multiple Images -- Camera Height Doesn't Change: Unsupervised Training for Metric Monocular Road-Scene Depth Estimation -- Uni3DL: A Unified Model for 3D Vision-Language Understanding -- Object-Aware NIR-to-Visible Translation -- PaPr: Training-Free One-Step Patch Pruning with Lightweight ConvNets for Faster Inference -- GENIXER: Empowering Multimodal Large Language Models as a Powerful Data Generator -- BLINK: Multimodal Large Language Models Can See but Not Perceive -- AFF-ttention! Affordances and Attention models for Short-Term Object Interaction Anticipation -- PreLAR: World Model Pre-training with Learnable Action Representation -- Multi-HMR: Multi-Person Whole-Body Human Mesh Recovery in a Single Shot -- De-confounded Gaze Estimation -- Diffusion Models for Monocular Depth Estimation: Overcoming Challenging Conditions -- FreestyleRet: Retrieving Images from Style-Diversified Queries -- ReGround: Improving Textual and Spatial Grounding at No Cost -- CardiacNet: Learning to Reconstruct Abnormalities for Cardiac Disease Assessment from Echocardiogram Videos -- LaMI-DETR: Open-Vocabulary Detection with Language Model Instruction -- Unrolled Decomposed Unpaired Learning for Controllable Low-Light Video Enhancement -- Efficient Image Pre-Training with Siamese Cropped Masked Autoencoders -- VP-SAM: Taming Segment Anything Model for Video Polyp Segmentation via Disentanglement and Spatio-temporal Side Network -- Dataset Enhancement with Instance-Level Augmentations -- FreeMotion: MoCap-Free Human Motion Synthesis with Multimodal Large Language Models -- Chameleon: A Data-Efficient Generalist for Dense Visual Prediction in the Wild -- Reliability in Semantic Segmentation: Can We Use Synthetic Data? -- SCPNet: Unsupervised Cross-modal Homography Estimation via Intra-modal Self-supervised Learning -- SCAPE: A Simple and Strong Category-Agnostic Pose Estimator.
520 $a The multi-volume set of LNCS books with volume numbers 15059 up to 15147 constitutes the refereed proceedings of the 18th European Conference on Computer Vision, ECCV 2024, held in Milan, Italy, during September 29-October 4, 2024. The 2387 papers presented in these proceedings were carefully reviewed and selected from a total of 8585 submissions. They deal with topics such as computer vision; machine learning; deep neural networks; reinforcement learning; object recognition; image classification; image processing; object detection; semantic segmentation; human pose estimation; 3d reconstruction; stereo vision; computational photography; neural networks; image coding; image reconstruction; motion estimation.
650 0 $a Computer vision $x Congresses. $3 570734
650 0 $a Pattern recognition systems $v Congresses. $3 563039
650 1 4 $a Computer Imaging, Vision, Pattern Recognition and Graphics. $3 890871
650 2 4 $a Image Processing. $3 891209
650 2 4 $a Computer Communication Networks. $3 775497
650 2 4 $a User Interfaces and Human Computer Interaction. $3 892554
650 2 4 $a Machine Learning. $3 3382522
650 2 4 $a Special Purpose and Application-Based Systems. $3 892492
700 1 $a Leonardis, Aleš. $3 3733324
710 2 $a SpringerLink (Online service) $3 836513
773 0 $t Springer Nature eBook
830 0 $a Lecture notes in computer science ; $v 15081. $3 3780071
856 4 0 $u https://doi.org/10.1007/978-3-031-73337-6
950 $a Computer Science (SpringerNature-11645)