Buch, Englisch, 699 Seiten, Format (B × H): 155 mm x 235 mm
19th European Conference, Malmö, Sweden, September 8–12, 2026, Proceedings, Part XLVIII
Buch, Englisch, 699 Seiten, Format (B × H): 155 mm x 235 mm
Reihe: Lecture Notes in Computer Science
ISBN: 978-3-032-37371-7
Verlag: Springer
The multi-volume set of LNCS books with volume numbers 17001 up to 17083 constitutes the refereed proceedings of the 19th European Conference on Computer Vision, ECCV 2026, held in Malmö, Sweden, during September 8–12, 2026.
The 2866 papers presented in these proceedings were carefully reviewed and selected from a total of 10,473 submissions. They deal with topics such as computer vision; machine learning; deep neural networks; reinforcement learning; object recognition; image classification; image processing; object detection; semantic segmentation; human pose estimation; 3d reconstruction; stereo vision; computational photography; neural networks; image coding; image reconstruction; motion estimation.
Zielgruppe
Research
Autoren/Hrsg.
Fachgebiete
- Technische Wissenschaften Elektronik | Nachrichtentechnik Elektronik
- Mathematik | Informatik EDV | Informatik Informatik Bildsignalverarbeitung
- Technische Wissenschaften Elektronik | Nachrichtentechnik Nachrichten- und Kommunikationstechnik Signalverarbeitung
- Mathematik | Informatik EDV | Informatik Informatik Mensch-Maschine-Interaktion
- Mathematik | Informatik EDV | Informatik Informatik Künstliche Intelligenz Maschinelles Lernen
Weitere Infos & Material
Latent Visual Diffusion Reasoning with Monte Carlo Tree Search.- SPICE: Simple Polysemantic feature Interpretation via Clustering-based Explanations.- Large-Scale High-Quality 3D Gaussian Head Reconstruction from Multi-View Captures.- Estimating Individual Tree Height and Species from UAV Imagery.- Advancing WordArt-Oriented Scene Text Recognition: Datasets and Methods.- OneWorld: Taming Scene Generation with 3D Unified Representation Autoencoder.- AFFMAE: Scalable Vision Pre-Training for High-Resolution Microscopy Segmentation on Desktop Hardware.- Unsafe by Reciprocity: How Generation–Understanding Coupling Undermines Safety in Unified Multimodal Models.- QST-SAM: Leveraging Cross-modal Instructions for Few-shot Referring Video Object Segmentation.- ORBIT: Overcoming Hallucination Risks via Bi-manifold Interaction and Traction.- SymbOmni: Evolving Agentic Omni Models via Symbolic Concept Learning.- NUN: Nested Unfolding Network for Real-World Concealed Object Segmentation.- AutoV: Loss-Oriented Ranking for Visual Prompt Retrieval in LVLMs.- OSVE: One Step Video Editing with One Step Diffusion Models.- OneHSI: A Unified Hyperspectral Foundation Model with Physical Consistency.- DualDiff3D: Dual Structure-Appearance Diffusion Priors for Reliability-Enhanced 3D Gaussian Splatting.- Learning to Tessellate: Point Cloud Generation via Recursive Spectral Partitioning.- Spanning the Visual Analogy Space with a Weight Basis of LoRAs.- Think While Watching: Online Streaming Segment-Level Memory for Multi-Turn Video Reasoning in Multimodal Large Language Models.- 2D Features Are All You Need for 3D Shape Understanding.- NavWM: A Unified Navigation World Model for Foresight-Driven Planning.- MA-VLA: Multi-Arm Vision-Language-Action Model for Collaboration and Compositional Generalization.- FUSE: Filter-Free Unified Spatiotemporal Estimation of SpO2 via Wave-Transport Modeling.- P-MTP: Efficient Document Parsing via Multi-Token Prediction with Progressive Depth Scaling.- MV-Forcing: Long Multi-View Video Generation via 4D-Grounded Spatio-Temporal Self-Forcing.- Wake up for Touch! Mask-isolated Tactile Alignment Learning in MLLMs.- Diffusion Models are Open-World Affordance Learners: Leveraging Generative Priors for 3D Affordance Learning.- Making Partial-Label Datasets Easier: A Simple Yet Highly Effective Data Augmentation for Deep Partial-Label Learning.- On the Reliability of Cue Conflict and Beyond.- StructSplat: Generalizable 3D Gaussian Splatting from Uncalibrated Sparse Views.- Read or Ignore? A Unified Benchmark for Typographic-Attack Robustness and Text Recognition in Vision-Language Models.- WALL-EVE: World Alignment with Rule Learning in Visual Environments.- SkillSpotter: Pose-Aware Multi-View Skilled Action Detection and Grading in Ego-Exo Videos.- IACD: Iterative Adversarial Collaborative Detection via Dual-Perspective Blind Spot Discovery.- MV-STRIDE: Enabling MLLMs to Master Multi-View Spatial Reasoning via Hierarchical Capability Modeling.- AlphaRad: Grounded Zero-Shot Classification in Chest Radiology via a-Corrected Binary Cross Entropy and Factorized Latent Supervision.




