19th European Conference, Malmö, Sweden, September 8–12, 2026, Proceedings, Part LIV
Buch, Englisch, Format (B × H): 155 mm x 235 mm
Reihe: Lecture Notes in Computer Science
ISBN: 978-3-032-37257-4
Verlag: Springer
The multi-volume set of LNCS books with volume numbers 17001 up to 17083 constitutes the refereed proceedings of the 19th European Conference on Computer Vision, ECCV 2026, held in Malmö, Sweden, during September 8–12, 2026.
The 2866 papers presented in these proceedings were carefully reviewed and selected from a total of 10,473 submissions. They deal with topics such as computer vision; machine learning; deep neural networks; reinforcement learning; object recognition; image classification; image processing; object detection; semantic segmentation; human pose estimation; 3d reconstruction; stereo vision; computational photography; neural networks; image coding; image reconstruction; motion estimation.
Zielgruppe
Research
Autoren/Hrsg.
Fachgebiete
- Mathematik | Informatik EDV | Informatik Informatik Bildsignalverarbeitung
- Mathematik | Informatik EDV | Informatik Informatik Mensch-Maschine-Interaktion
- Technische Wissenschaften Elektronik | Nachrichtentechnik Elektronik
- Mathematik | Informatik EDV | Informatik Informatik Künstliche Intelligenz Maschinelles Lernen
- Technische Wissenschaften Elektronik | Nachrichtentechnik Nachrichten- und Kommunikationstechnik Signalverarbeitung
Weitere Infos & Material
NegAS: Negative Label Guided Attention and Scoring for Out-of-Distribution Object Detection with Vision-Language Models.- Scientific Image Synthesis: Benchmarking, Methodologies, and Downstream Utility.- Boxer: Robust Lifting of Open-World 2D Bounding Boxes to 3D.- FlowerDance: MeanFlow for Efficient and Refined 3D Dance Generation.- Rethinking Pseudo-Labels: Multi-Granularity Supervision for Domain Adaptive Object Detection.- STANCE: Controllable Video Generation for Structured Dynamics via Sparse-To-dense ANChored Encoding.- Path-JEPA: Path Signature Based Predictive Learning for Skeleton Action Recognition.- Out of Sight, Out of Mind? Evaluating State Evolution in Video World Models.- Learning Generatable Mutual Distance for Scene-Aware Human Motion Generation.- TaxoMIL: Taxonomy-Constrained Learning for Hierarchical Whole Slide Image Analysis.- Learning to Deny: Action Denial in Multimodal Large Language Models.- SUM: Unified Geometric Surgery on Spatio-Temporal Adaptation Vectors for Federated Class Incremental Learning.- Self-Evolving Just-In-Time Memory for Proactive Embodied Safety.- Pay Attention to Attention Distribution: A New Local Lipschitz Bound for Transformers.- Filterless Snapshot Hyperspectral Imaging using Guided Patch Diffusion.- Learning Consistency in Reward Modeling for Multi-Modal Reasoning.- Forge4D: Feed-Forward 4D Human Reconstruction and Interpolation from Uncalibrated Sparse-View Videos.- World Reconstruction From Inconsistent Views.- HybridSim: A Physics–Learning Hybrid Digital Twin for mmWave Human Sensing.- WiFi-JEPA: Self-supervised Learning for WiFi-CSI 3D Human Pose Estimation.- CoSyncDiT: Cognitive Synchronous Diffusion Transformer for Movie Dubbing.- MagnetGS-Mesh: High-Quality Multi-Object Mesh Reconstruction via Adaptive Surface Optimization.- SFDATrack: Generalized Source-Free Domain Adaptive Tracking Under Adverse Weather Conditions.- BrepCoder: A Unified Multimodal Large Language Model for Multi-task B-rep Reasoning.- Decoupled Illumination Priors for Spatially Controllable Multi-View Indoor Scene Relighting.- The Map Is Not the Territory: Embedding-Coverage Blacklists for Safe Diffusion Steering.- GTR: Guide-Then-Refine Token Compression for Training-Free Acceleration of Video-LLMs.- Keep It Simple: Multi-Key Episodic Memory Retrieval for Ultra-Long Video Understanding.- KineBench: Benchmarking Embodied World Models via IDM-Free Kinematic Grounding.- M4-SAR: A Multi-Resolution, Multi-Polarization, Multi-Scene, Multi-Source Dataset and Benchmark for optical-SAR Object Detection.- Rotate Your Character: Revisiting Video Diffusion Models for High-Quality 3D Character Generation.- Anatomy of a Lie: A Multi-Stage Diagnostic Framework for Tracing Hallucinations in Vision-Language Models.- Sim, Yet Same: Physics-Aligned Simulator as Zero-Shot Data Scaler in Deformable Worlds.- Defending from GeoLocalization through Adversarial Road Trips.- BRepFacetGen: Reverse Engineering B-Reps By Generative Face Segmentation.- PanoSAM2: Lightweight Distortion- and Memory-aware Adaptions of SAM2 for 360 Video Object Segmentation.- Dual-Generalization-aware Minimization for Continual Fine-Tuning of Vision-Language Models.- InSeg: Interactive Refinement via Intent Propagation for Point Cloud Semantic Segmentation.




