東華大學圖書館 |

Computer vision - ECCV 2024 = 18th European Conference, Milan, Italy, September 29-October 4, 2024 : proceedings.. Part XLII /

紀錄類型:	書目-電子資源 : Monograph/item
正題名/作者:	Computer vision - ECCV 2024/ edited by Aleš Leonardis ... [et al.].
其他題名:	18th European Conference, Milan, Italy, September 29-October 4, 2024 : proceedings.
其他題名:	ECCV 2024
其他作者:	Leonardis, Aleš.
團體作者:	European Conference on Computer Vision
出版者:	Cham :Springer Nature Switzerland : : 2025.,
面頁冊數:	lxxxv, 499 p. :ill., digital ;24 cm.
內容註:	Open-Set Recognition in the Age of Vision-Language Models -- Unsqueeze [CLS] Bottleneck to Learn Rich Representations -- Robust Multimodal Learning via Representation Decoupling -- Object-Conditioned Energy-Based Attention Map Alignment in Text-to-Image Diffusion Models -- WiMANS: A Benchmark Dataset for WiFi-based Multi-user Activity Sensing -- Embedding-Free Transformer with Inference Spatial Reduction for Efficient Semantic Segmentation -- VeCLIP: Improving CLIP Training via Visual-enriched Captions -- Three Things We Need to Know About Transferring Stable Diffusion to Visual Dense Prediciton Tasks -- Learning Representations from Foundation Models for Domain Generalized Stereo Matching -- Spike-Temporal Latent Representation for Energy-Efficient Event-to-Video Reconstruction -- Effective Lymph Nodes Detection in CT Scans Using Location Debiased Query Selection and Contrastive Query Representation in Transformer -- Chat-Edit-3D: Interactive 3D Scene Editing via Text Prompts -- Event-Adapted Video Super-Resolution -- Look Hear: Gaze Prediction for Speech-directed Human Attention -- Raising the Ceiling: Conflict-Free Local Feature Matching with Dynamic View Switching -- Q&A Prompts: Discovering Rich Visual Clues through Mining Question-Answer Prompts for VQA requiring Diverse World Knowledge -- Catastrophic Overfitting: A Potential Blessing in Disguise -- Long-range Turbulence Mitigation: A Large-scale Dataset and A Coarse-to-fine Framework -- SparseCtrl: Adding Sparse Controls to Text-to-Video Diffusion Models -- Visual Alignment Pre-training for Sign Language Translation -- Parrot Captions Teach CLIP to Spot Text -- Solving Motion Planning Tasks with a Scalable Generative Model -- Griffon: Spelling out All Object Locations at Any Granularity with Large Language Models -- Vision-Language Action Knowledge Learning for Semantic-Aware Action Quality Assessment -- Knowledge Transfer with Simulated Inter-Image Erasing for Weakly Supervised Semantic Segmentation -- BurstM: Deep Burst Multi-scale SR using Fourier Space with Optical Flow -- Diffusion Reward: Learning Rewards via Conditional Video Diffusion.
Contained By:	Springer Nature eBook
標題:	Computer vision - Congresses. -
電子資源:	https://doi.org/10.1007/978-3-031-72946-1
ISBN:	9783031729461

Computer vision - ECCV 2024 = 18th European Conference, Milan, Italy, September 29-October 4, 2024 : proceedings.. Part XLII /
Computer vision - ECCV 202418th European Conference, Milan, Italy, September 29-October 4, 2024 : proceedings.Part XLII /[electronic resource] :ECCV 2024edited by Aleš Leonardis ... [et al.]. - Cham :Springer Nature Switzerland :2025. - lxxxv, 499 p. :ill., digital ;24 cm. - Lecture notes in computer science,151001611-3349 ;. - Lecture notes in computer science ;15100..

Open-Set Recognition in the Age of Vision-Language Models -- Unsqueeze [CLS] Bottleneck to Learn Rich Representations -- Robust Multimodal Learning via Representation Decoupling -- Object-Conditioned Energy-Based Attention Map Alignment in Text-to-Image Diffusion Models -- WiMANS: A Benchmark Dataset for WiFi-based Multi-user Activity Sensing -- Embedding-Free Transformer with Inference Spatial Reduction for Efficient Semantic Segmentation -- VeCLIP: Improving CLIP Training via Visual-enriched Captions -- Three Things We Need to Know About Transferring Stable Diffusion to Visual Dense Prediciton Tasks -- Learning Representations from Foundation Models for Domain Generalized Stereo Matching -- Spike-Temporal Latent Representation for Energy-Efficient Event-to-Video Reconstruction -- Effective Lymph Nodes Detection in CT Scans Using Location Debiased Query Selection and Contrastive Query Representation in Transformer -- Chat-Edit-3D: Interactive 3D Scene Editing via Text Prompts -- Event-Adapted Video Super-Resolution -- Look Hear: Gaze Prediction for Speech-directed Human Attention -- Raising the Ceiling: Conflict-Free Local Feature Matching with Dynamic View Switching -- Q&A Prompts: Discovering Rich Visual Clues through Mining Question-Answer Prompts for VQA requiring Diverse World Knowledge -- Catastrophic Overfitting: A Potential Blessing in Disguise -- Long-range Turbulence Mitigation: A Large-scale Dataset and A Coarse-to-fine Framework -- SparseCtrl: Adding Sparse Controls to Text-to-Video Diffusion Models -- Visual Alignment Pre-training for Sign Language Translation -- Parrot Captions Teach CLIP to Spot Text -- Solving Motion Planning Tasks with a Scalable Generative Model -- Griffon: Spelling out All Object Locations at Any Granularity with Large Language Models -- Vision-Language Action Knowledge Learning for Semantic-Aware Action Quality Assessment -- Knowledge Transfer with Simulated Inter-Image Erasing for Weakly Supervised Semantic Segmentation -- BurstM: Deep Burst Multi-scale SR using Fourier Space with Optical Flow -- Diffusion Reward: Learning Rewards via Conditional Video Diffusion.

The multi-volume set of LNCS books with volume numbers 15059 up to 15147 constitutes the refereed proceedings of the 18th European Conference on Computer Vision, ECCV 2024, held in Milan, Italy, during September 29-October 4, 2024. The 2387 papers presented in these proceedings were carefully reviewed and selected from a total of 8585 submissions. The papers deal with topics such as computer vision; machine learning; deep neural networks; reinforcement learning; object recognition; image classification; image processing; object detection; semantic segmentation; human pose estimation; 3d reconstruction; stereo vision; computational photography; neural networks; image coding; image reconstruction; motion estimation.

ISBN: 9783031729461

Standard No.: 10.1007/978-3-031-72946-1doiSubjects--Topical Terms:

570734
Computer vision
--Congresses.

LC Class. No.: TA1634

Dewey Class. No.: 006.37

Computer vision - ECCV 2024 = 18th European Conference, Milan, Italy, September 29-October 4, 2024 : proceedings.. Part XLII /
LDR:04011nmm a2200349 a 4500 001 2407824
003 DE-He213
005 20241002130237.0
006 m d
007 cr nn 008maaau
008 260204s2025 sz s 0 eng d
020 $a 9783031729461 $q (electronic bk.)
020 $a 9783031729454 $q (paper)
024 7 $a 10.1007/978-3-031-72946-1 $2 doi
035 $a 978-3-031-72946-1
040 $a GP $c GP
041 0 $a eng
050 4 $a TA1634
072 7 $a UYT $2 bicssc
072 7 $a COM016000 $2 bisacsh
072 7 $a UYT $2 thema
082 0 4 $a 006.37 $2 23
090 $a TA1634 $b .E89 2024
111 2 $a European Conference on Computer Vision $n (18th : $d 2024 : $c Milan, Italy) $3 3733323
245 1 0 $a Computer vision - ECCV 2024 $h [electronic resource] : $b 18th European Conference, Milan, Italy, September 29-October 4, 2024 : proceedings. $n Part XLII / $c edited by Aleš Leonardis ... [et al.].
246 3 $a ECCV 2024
260 $a Cham : $b Springer Nature Switzerland : $b Imprint: Springer, $c 2025.
300 $a lxxxv, 499 p. : $b ill., digital ; $c 24 cm.
490 1 $a Lecture notes in computer science, $x 1611-3349 ; $v 15100
505 0 $a Open-Set Recognition in the Age of Vision-Language Models -- Unsqueeze [CLS] Bottleneck to Learn Rich Representations -- Robust Multimodal Learning via Representation Decoupling -- Object-Conditioned Energy-Based Attention Map Alignment in Text-to-Image Diffusion Models -- WiMANS: A Benchmark Dataset for WiFi-based Multi-user Activity Sensing -- Embedding-Free Transformer with Inference Spatial Reduction for Efficient Semantic Segmentation -- VeCLIP: Improving CLIP Training via Visual-enriched Captions -- Three Things We Need to Know About Transferring Stable Diffusion to Visual Dense Prediciton Tasks -- Learning Representations from Foundation Models for Domain Generalized Stereo Matching -- Spike-Temporal Latent Representation for Energy-Efficient Event-to-Video Reconstruction -- Effective Lymph Nodes Detection in CT Scans Using Location Debiased Query Selection and Contrastive Query Representation in Transformer -- Chat-Edit-3D: Interactive 3D Scene Editing via Text Prompts -- Event-Adapted Video Super-Resolution -- Look Hear: Gaze Prediction for Speech-directed Human Attention -- Raising the Ceiling: Conflict-Free Local Feature Matching with Dynamic View Switching -- Q&A Prompts: Discovering Rich Visual Clues through Mining Question-Answer Prompts for VQA requiring Diverse World Knowledge -- Catastrophic Overfitting: A Potential Blessing in Disguise -- Long-range Turbulence Mitigation: A Large-scale Dataset and A Coarse-to-fine Framework -- SparseCtrl: Adding Sparse Controls to Text-to-Video Diffusion Models -- Visual Alignment Pre-training for Sign Language Translation -- Parrot Captions Teach CLIP to Spot Text -- Solving Motion Planning Tasks with a Scalable Generative Model -- Griffon: Spelling out All Object Locations at Any Granularity with Large Language Models -- Vision-Language Action Knowledge Learning for Semantic-Aware Action Quality Assessment -- Knowledge Transfer with Simulated Inter-Image Erasing for Weakly Supervised Semantic Segmentation -- BurstM: Deep Burst Multi-scale SR using Fourier Space with Optical Flow -- Diffusion Reward: Learning Rewards via Conditional Video Diffusion.
520 $a The multi-volume set of LNCS books with volume numbers 15059 up to 15147 constitutes the refereed proceedings of the 18th European Conference on Computer Vision, ECCV 2024, held in Milan, Italy, during September 29-October 4, 2024. The 2387 papers presented in these proceedings were carefully reviewed and selected from a total of 8585 submissions. The papers deal with topics such as computer vision; machine learning; deep neural networks; reinforcement learning; object recognition; image classification; image processing; object detection; semantic segmentation; human pose estimation; 3d reconstruction; stereo vision; computational photography; neural networks; image coding; image reconstruction; motion estimation.
650 0 $a Computer vision $x Congresses. $3 570734
650 0 $a Pattern recognition systems $v Congresses. $3 563039
650 1 4 $a Computer Imaging, Vision, Pattern Recognition and Graphics. $3 890871
650 2 4 $a Image Processing. $3 891209
650 2 4 $a Computer Communication Networks. $3 775497
650 2 4 $a User Interfaces and Human Computer Interaction. $3 892554
650 2 4 $a Machine Learning. $3 3382522
650 2 4 $a Special Purpose and Application-Based Systems. $3 892492
700 1 $a Leonardis, Aleš. $3 3779889
710 2 $a SpringerLink (Online service) $3 836513
773 0 $t Springer Nature eBook
830 0 $a Lecture notes in computer science ; $v 15100. $3 3779903
856 4 0 $u https://doi.org/10.1007/978-3-031-72946-1
950 $a Computer Science (SpringerNature-11645)