Buch, Englisch, 701 Seiten, Format (B × H): 155 mm x 235 mm
19th European Conference, Malmö, Sweden, September 8–12, 2026, Proceedings, Part LXXI
Buch, Englisch, 701 Seiten, Format (B × H): 155 mm x 235 mm
Reihe: Lecture Notes in Computer Science
ISBN: 978-3-032-37015-0
Verlag: Springer
The multi-volume set of LNCS books with volume numbers 17001 up to 17083 constitutes the refereed proceedings of the 19th European Conference on Computer Vision, ECCV 2026, held in Malmö, Sweden, during September 8–12, 2026.
The 2866 papers presented in these proceedings were carefully reviewed and selected from a total of 10,473 submissions. They deal with topics such as computer vision; machine learning; deep neural networks; reinforcement learning; object recognition; image classification; image processing; object detection; semantic segmentation; human pose estimation; 3d reconstruction; stereo vision; computational photography; neural networks; image coding; image reconstruction; motion estimation.
Zielgruppe
Research
Autoren/Hrsg.
Fachgebiete
- Technische Wissenschaften Elektronik | Nachrichtentechnik Elektronik
- Mathematik | Informatik EDV | Informatik Informatik Mensch-Maschine-Interaktion
- Mathematik | Informatik EDV | Informatik Informatik Künstliche Intelligenz Maschinelles Lernen
- Mathematik | Informatik EDV | Informatik Informatik Bildsignalverarbeitung
- Technische Wissenschaften Elektronik | Nachrichtentechnik Nachrichten- und Kommunikationstechnik Signalverarbeitung
Weitere Infos & Material
TreeSRNF: Square-Root Normal Fields for Generative Modelling of the Geometric and Structural Variability in Tree-like 3D Objects.- What CLIP Knows but Cannot Say: Recovering Negation from Frozen Intermediate Features.- Scaling Verification Can Be More Effective than Scaling Policy Learning for Vision-Language-Action Alignment.- Be Tangential to Manifold: Discovering Riemannian Metric for Diffusion Models.- GUI-AIMA: Aligning Intrinsic Multimodal Attention with a Context Anchor for GUI Grounding.- HippoCamp: Benchmarking Contextual Agents on Personal Computers.- TMI: Text-to-Image Meets Image-to-Image for Complementary Data Synthesis to Boost Long-Tailed Instance Segmentation.- DA-F2F: Domain-Adaptive Object Detection with Feature-to-Feature Modulation and Alignment.- Cast and Attached Shadow Detection via Iterative Light and Geometry Reasoning.- RoadBench: Benchmarking MLLMs on Fine-Grained Spatial Understanding and Reasoning under Urban Road Scenarios.- LIIFusion: Coarse-to-fine Framework for Generative MEF via Implicit Neural Representation.- DLGStream: Dynamic Language-embedded Guassian Splatting for Open-vocabulary Enabled Free-viewpoint Video Streaming.- A scalar per patch from pre-trained ViTs enables fast moving navigation in the real world.- PACO: Stabilizing Vision Embeddings along Local Paths for Robust Vision-Language Models.- DocLayout-VL: A Foundational Model for Hierarchical, Open-set, and Promptable Document Layout Segmentation.- MetricAnything: Scaling Metric Depth Pretraining with Noisy Heterogeneous Sources.- ColorFM: An Optimization-to-Learning Framework for Color Transfer via Flow Matching.- RePer-360: Releasing Perspective Priors for 360° Depth Estimation via Self-Modulation.- Steerable Vision Transformers.- Single-Query Person-Centric Bimanual Hand-Object Interaction Detection.- SkyLume: A Large-Scale Multi-Illumination Aerial Benchmark for Urban Scene Reconstruction and Beyond.- Gaze-to-text Generation: Beyond Categorical Decoding of Human Attention.- CMDS-AD: Cross-Modal Dual-Stream Decoupling for Few-Shot Anomaly Detection.- DreamEdit3D: Personalization of Multi-View Diffusion Models for 3D Editing.- Composing Driving Worlds through Disentangled Control for Adversarial Scenario Generation.- ESC: Emotional Self-Correction for Reliable Vision-Language Models.- VLOD-TTA: Test-Time Adaptation of Vision-Language Object Detectors.- CoLT: Teaching Multi-Modal Models to Think with Chain of Latent Thoughts.- PixGS: Pixel-Space Diffusion for Direct 3D Gaussian Splat Generation.- CMuon: Accelerating and Stabilizing Diffusion Transformer Training via Chunked Momentum Orthogonalization.- BiCE-HG: A Bi-Conditional Egocentric Hand Gesture Dataset for Intelligent Reality Systems.- Ceptor: Vision-Language Model-Infused Diverse Guidance for Detecting Anything.- PointSplat: Compact Gaussian Splatting via Human-Centric Prediction.- SFKD: Spatial–Frequency Joint-Aware Heterogeneous Knowledge Distillation via Multi-Level Wavelet Spectral Interaction.- Extending a Large View Synthesis Model for Multi-view Panoptic Segmentation.- PriorEye: Geospatial Visual Priors for End-to-End Autonomous Driving.




