19th European Conference, Malmö, Sweden, September 8–12, 2026, Proceedings, Part LIII
Buch, Englisch, Format (B × H): 155 mm x 235 mm
Reihe: Lecture Notes in Computer Science
ISBN: 978-3-032-37254-3
Verlag: Springer
The multi-volume set of LNCS books with volume numbers 17001 up to 17083 constitutes the refereed proceedings of the 19th European Conference on Computer Vision, ECCV 2026, held in Malmö, Sweden, during September 8–12, 2026.
The 2866 papers presented in these proceedings were carefully reviewed and selected from a total of 10,473 submissions. They deal with topics such as computer vision; machine learning; deep neural networks; reinforcement learning; object recognition; image classification; image processing; object detection; semantic segmentation; human pose estimation; 3d reconstruction; stereo vision; computational photography; neural networks; image coding; image reconstruction; motion estimation.
Zielgruppe
Research
Autoren/Hrsg.
Fachgebiete
- Mathematik | Informatik EDV | Informatik Informatik Bildsignalverarbeitung
- Technische Wissenschaften Elektronik | Nachrichtentechnik Elektronik
- Mathematik | Informatik EDV | Informatik Informatik Künstliche Intelligenz Maschinelles Lernen
- Mathematik | Informatik EDV | Informatik Informatik Mensch-Maschine-Interaktion
- Technische Wissenschaften Elektronik | Nachrichtentechnik Nachrichten- und Kommunikationstechnik Signalverarbeitung
Weitere Infos & Material
SGP2: Coarse-to-Fine Controllable Multimodal Remote Sensing Image Generation.- Tiled Prompts: Overcoming Prompt Misguidance in Image and Video Super-Resolution.- GameWorlds: Towards Standardized and Verifiable Evaluation of Multimodal Game Agents.- InSpace: Structure-Aware 3D Indoor Scene Generation from a Single 360° Image.- Video Streaming Thinking: VideoLLMs Can Watch and Think Simultaneously.- DIVER: Disentangling Camera–Object and Active–Passive Motion for Video Generation.- Text-Conditioned Background Generation for Editable Multi-Layer Documents.- VoCa: Unified Autoregressive Modeling for Talking Audio-Video Generation.- Decompose, Compare, and Decide: Multimodal LLMs are Implicit Few-Shot Learners.- NarrativeTrack: Evaluating Entity-Centric Reasoning for Narrative Understanding.- Hi-DiT: Hybrid Latent-Pixel Diffusion Transformer for Image Generation.- Cube-Splat: High-Fidelity 360° Gaussian Splatting SLAM via Cubemap Factorization and Adjoint-Consistent Optimization.- Motion-aware Sparse Pipeline for Lightweight Object Tracking.- VideoTIR: Accurate Understanding for Long Videos with Efficient Tool-Integrated Reasoning.- Face Anything: 4D Face Reconstruction from Any Image Sequence.- ReDesign: Recovering Editable Design Structures from Raster Images via Agentic Decomposition.- Distill on a Diet: Efficient Knowledge Distillation via Learnable Data Pruning.- SPAR: Semantic-Pixel Self-Alignment and Adaptive Routing for Unified Multimodal Models.- VoxAnchor: Explicit Voxel-Semantic Grounding for Spatial Understanding in Videos.- LISA: Locality-Informed Speculative Decoding for Accelerating Autoregressive Image Generation.- Layering Virtual Try-On.- clean2green2clean: Synthesising Bad Composites to Learn Actor-Background Video Harmonisation.- EatVid-Bench: A Multimodal Fine-Grained Eating Behavior Video Dataset.- EAGS: Error-Aware Gaussian Splatting with Dual-Confidence-Guided Modeling for Uncalibrated Driving Scenes.- What Images Cannot Say: Language-Guided Olfactory Representation Learning.- PAI-Studio: Cinematic Video Background Replacement with Camera-Aware Motion.- Multi-modal Knowledge Preserving Adapter for Embedding Backward Compatibility.- Sparse-View Surface Reconstruction using Gaussian Splatting through High-Confidence Depth Propagation with Normal Priors.- Reasoning Path and Latent State Analysis for Multi-view Visual Spatial Reasoning: A Cognitive Science Perspective.- Calibrate Before Adapt: Training-Free Pseudo-Label Calibration for Semi-Supervised Cross-Domain Few-Shot Detection.- Atlas is Your Perfect Context: One-Shot Customization for Generalizable Foundational Medical Image Segmentation.- ICLAgent: Integrated Circuit Footprint Geometry Labeling via LMM-empowered Multi-Agent Framework.- Frozen CLIP Priors for Robust Self-Supervised Poisson Inverse Problems.- CurveStream: Boosting Streaming Video Understanding in MLLMs via Curvature-Aware Hierarchical Visual Memory Management.- 3D Field of Junctions: A Noise-Robust, Training-Free Structural Prior for Volumetric Inverse Problems.- Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation.- Benchmarking MLLMs on Mistake Recognition and Explanation in Single-Step Components of Cooking.




