My Account Log in

1 option

Computer Vision – ECCV 2024 : 18th European Conference, Milan, Italy, September 29–October 4, 2024, Proceedings, Part XV / edited by Aleš Leonardis, Elisa Ricci, Stefan Roth, Olga Russakovsky, Torsten Sattler, Gül Varol.

Springer Nature - Springer Computer Science (R0) eBooks 2025 English International Available online

View online
Format:
Book
Author/Creator:
Leonardis, Aleš.
Contributor:
Ricci, Elisa.
Roth, Ștefan.
Russakovsky, Olga.
Sattler, Torsten.
Varol, Gül.
Series:
Lecture Notes in Computer Science, 1611-3349 ; 15073
Language:
English
Subjects (All):
Image processing--Digital techniques.
Image processing.
Computer vision.
Computer networks.
User interfaces (Computer systems).
Human-computer interaction.
Machine learning.
Computers, Special purpose.
Computer Imaging, Vision, Pattern Recognition and Graphics.
Image Processing.
Computer Communication Networks.
User Interfaces and Human Computer Interaction.
Machine Learning.
Special Purpose and Application-Based Systems.
Local Subjects:
Computer Imaging, Vision, Pattern Recognition and Graphics.
Image Processing.
Computer Communication Networks.
User Interfaces and Human Computer Interaction.
Machine Learning.
Special Purpose and Application-Based Systems.
Physical Description:
1 online resource (0 pages)
Edition:
1st ed. 2025.
Place of Publication:
Cham : Springer Nature Switzerland : Imprint: Springer, 2025.
Summary:
The multi-volume set of LNCS books with volume numbers 15059 up to 15147 constitutes the refereed proceedings of the 18th European Conference on Computer Vision, ECCV 2024, held in Milan, Italy, during September 29–October 4, 2024. The 2387 papers presented in these proceedings were carefully reviewed and selected from a total of 8585 submissions. The papers deal with topics such as computer vision; machine learning; deep neural networks; reinforcement learning; object recognition; image classification; image processing; object detection; semantic segmentation; human pose estimation; 3d reconstruction; stereo vision; computational photography; neural networks; image coding; image reconstruction; motion estimation.
Contents:
GRM: Large Gaussian Reconstruction Model for Efficient 3D Reconstruction and Generation
IRGen: Generative Modeling for Image Retrieval
Learning Trimodal Relation for Audio-Visual Question Answering with Missing Modality
FastCAD: Real-Time CAD Retrieval and Alignment from Scans and Videos
A Simple Latent Diffusion Approach for Panoptic Segmentation and Mask Inpainting
VISA: Reasoning Video Object Segmentation via Large Language Model
Lego: Learning to Disentangle and Invert Personalized Concepts Beyond Object Appearance in Text-to-Image Diffusion Models
IDOL: Unified Dual-Modal Latent Diffusion for Human-Centric Joint Video-Depth Generation
Scaling Backwards: Minimal Synthetic Pre-training?
BAMM: Bidirectional Autoregressive Motion Model
Event-based Head Pose Estimation: Benchmark and Method
Avatar Fingerprinting for Authorized Use of Synthetic Talking-Head Videos
Towards Multi-modal Transformers in Federated Learning
Fisher Calibration for Backdoor-Robust Heterogeneous Federated Learning
QueryCDR: Query-based Controllable Distortion Rectification Network for Fisheye Images
Latent-INR: A Flexible Framework for Implicit Representations of Videos with Discriminative Semantics
DCDM: Diffusion-Conditioned-Diffusion Model for Scene Text Image Super-Resolution
Do not move together: per-Gaussian Deformation for 4DGS
DreamMover: Leveraging the Prior of Diffusion Models for Image Interpolation with Large Motion
CoLA: Conditional Dropout and Language-driven Robust Dual-modal Salient Object Detection
Image-Feature Weak-to-Strong Consistency: An Enhanced Paradigm for Semi-Supervised Learning
RPBG: Towards Robust Neural Point-based Graphics in the Wild
GaussReg: Fast 3D Registration with Gaussian Splatting
Efficient Diffusion Transformer with Step-wise Dynamic Attention Mediators
Open Vocabulary 3D Scene Understanding via Geometry Guided Self-Distillation
IAM-VFI : Interpolate Any Motion for Video Frame Interpolation with motion complexity map
TIP: Tabular-Image Pre-training for Multimodal Classification with Incomplete Data.
Notes:
Description based on publisher supplied metadata and other sources.
ISBN:
9783031726330
3031726332
OCLC:
1474244087

The Penn Libraries is committed to describing library materials using current, accurate, and responsible language. If you discover outdated or inaccurate language, please fill out this feedback form to report it and suggest alternative language.

Find

Home Release notes

My Account

Shelf Request an item Bookmarks Fines and fees Settings

Guides

Using the Find catalog Using Articles+ Using your account