Buch, Englisch, 701 Seiten, Format (B × H): 155 mm x 235 mm
19th European Conference, Malmö, Sweden, September 8 –12, 2026, Proceedings, Part LXX
Buch, Englisch, 701 Seiten, Format (B × H): 155 mm x 235 mm
Reihe: Lecture Notes in Computer Science
ISBN: 978-3-032-37012-9
Verlag: Springer
The multi-volume set of LNCS books with volume numbers 17001 up to 17083 constitutes the refereed proceedings of the 19th European Conference on Computer Vision, ECCV 2026, held in Malmö, Sweden, during September 8–12, 2026.
The 2866 papers presented in these proceedings were carefully reviewed and selected from a total of 10,473 submissions. They deal with topics such as computer vision; machine learning; deep neural networks; reinforcement learning; object recognition; image classification; image processing; object detection; semantic segmentation; human pose estimation; 3d reconstruction; stereo vision; computational photography; neural networks; image coding; image reconstruction; motion estimation.
Zielgruppe
Research
Autoren/Hrsg.
Fachgebiete
- Mathematik | Informatik EDV | Informatik Informatik Bildsignalverarbeitung
- Technische Wissenschaften Elektronik | Nachrichtentechnik Nachrichten- und Kommunikationstechnik Signalverarbeitung
- Mathematik | Informatik EDV | Informatik Informatik Mensch-Maschine-Interaktion
- Technische Wissenschaften Elektronik | Nachrichtentechnik Elektronik
- Mathematik | Informatik EDV | Informatik Informatik Künstliche Intelligenz Maschinelles Lernen
Weitere Infos & Material
LinStereo: Linear-Complexity Global Attention for Multi-Scale Iterative Stereo Matching.- EMOTE: Expressive Motion and Shape Disentanglement for Human Animation.- SCoT: Similarity-guided Conflict-aware Task Consolidation for Continual VQA.- FixAnything: 3D-Consistent Rendering Refinement via Video Generative Priors.- Solving Diffusion Inverse Problems with Restart Posterior Sampling.- JointHOI: Jointly Generating Contact Maps Enhances Hand Object Interaction Generation.- Thermo-JEPA: Learning a Geometry-Grounded Thermal World Model via Cross-Modal Privileged Masking.- Enhancing Embodied Reasoning and Grounding by Novel View Synthesis.- Robust 3DGS-based SLAM via Adaptive Kernel Smoothing.- AVQ-Attention: Adaptive Vector-Quantized Attention.- Aligning Human Sense: Calibrated Distributional Reward Learning for Video Generation.- ReasonCLIP-58M: Visually Grounded Commonsense Reasoning Supervision for CLIP.- LaMP: Learning Vision-Language-Action Policies with 3D Scene Flow as Latent Motion Prior.- VOID: Video Object and Interaction Deletion.- Natural Language Camera Movement Understanding.- MuCHeR: Multi-Person Camera-Centric Human Detection, Mesh Recovery and Tracking.- SA-ResGS: Self-Augmented Residual 3D Gaussian Splatting for Next Best View Selection.- Policy-Based Tuning of Autoregressive Image Models with Instance- and Distribution-Level Rewards.- Unified Prediction and Planning via Conflict-Aware Disjoint Parameter Training.- Fast-dVLA: Accelerating Discrete Diffusion VLA to Real-Time Performance.- CL-Anomaly: Layer-Adaptive Mixture-of-Experts with Multimodal Large Language Model for Continual Learning in Anomaly Detection.- GridVQA-X: A Diagnostic Framework for Evaluating Multimodal Explainability Methods.- WildProp: Visual Estimation of Wildlife Body Proportions at Scale.- DriveWeaver: Point-Conditioned Video Inpainting for Controllable Vehicle Insertion in Autonomous Driving Simulation.- OpenGround: Planning-based Online Perception for Open-World 3D Visual Grounding.- WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing.- TDSR-VLA: Transition-aware Denoising Sequence Representations for Vision-Language-Action.- RelAfford6D: Relational 6D Affordance Graphs for Constraint-Driven Robotic Manipulation.- From Script to Shot: A Benchmark for Grounding Screenplays in Movies.- GeoNVS: Geometry Grounded Video Diffusion for Novel View Synthesis.- TETO: Tracking Events with Teacher Observation for Motion Estimation and Frame Interpolation.- Weight Feedback Computes the Exact Jacobian Transpose in Modern Deep Networks.- CogniCred: A Dataset and Benchmark for Cognitive Credential Forgery Detection.- JacobianAvatar: Temporally Consistent Semi-rigid Avatar Reconstruction from a Monocular Video.- LEGO: Leveled Language Gaussian Splatting.- MobileOcc: A Human-Aware Semantic Occupancy Dataset for Mobile Robots.




