Buch, Englisch, 706 Seiten, Format (B × H): 155 mm x 235 mm
19th European Conference, Malmö, Sweden, September 8–12, 2026, Proceedings, Part LII
Buch, Englisch, 706 Seiten, Format (B × H): 155 mm x 235 mm
Reihe: Lecture Notes in Computer Science
ISBN: 978-3-032-37234-5
Verlag: Springer
The multi-volume set of LNCS books with volume numbers 17001 up to 17083 constitutes the refereed proceedings of the 19th European Conference on Computer Vision, ECCV 2026, held in Malmö, Sweden, during September 8–12, 2026.
The 2866 papers presented in these proceedings were carefully reviewed and selected from a total of 10,473 submissions. They deal with topics such as computer vision; machine learning; deep neural networks; reinforcement learning; object recognition; image classification; image processing; object detection; semantic segmentation; human pose estimation; 3d reconstruction; stereo vision; computational photography; neural networks; image coding; image reconstruction; motion estimation.
Zielgruppe
Research
Autoren/Hrsg.
Fachgebiete
- Technische Wissenschaften Elektronik | Nachrichtentechnik Elektronik
- Technische Wissenschaften Elektronik | Nachrichtentechnik Nachrichten- und Kommunikationstechnik Signalverarbeitung
- Mathematik | Informatik EDV | Informatik Informatik Bildsignalverarbeitung
- Mathematik | Informatik EDV | Informatik Informatik Künstliche Intelligenz Maschinelles Lernen
- Mathematik | Informatik EDV | Informatik Informatik Mensch-Maschine-Interaktion
Weitere Infos & Material
VD-LoRA: Adaptive Reuse of Low-Rank Directions for Continual Learning.- Stylized Video Generation via Decoupled Data Synthesis and Gated Style Token Injection.- RawGen: Learning Camera Raw Image Generation.- Color Pass-Through via Camera-Display Coupling.- Video Generation Models Are Inherent Lighting Estimators.- ZMIS-SAM: Segment Anything Model Enhanced With Wavelet Transform For Zooplankton Microscopy Image Instance Segmentation.- VIGA: View-Conditioned and Identity-Guided Adaptation for Aerial-Ground Person Re-Identification.- ProMSA:Progressive Multimodal Search Agents for Knowledge-Based Visual Question Answering.- EgoEverything: A Benchmark for Human Behavior–Inspired Long-Context Egocentric Video Understanding in AR Environment.- AnyFlow: Any-Step Video Diffusion Model with On-Policy Flow Map Distillation.- DIGS: Differentiable, Incremental, Global, Scalable Pruning for Language Models.- PhysConvex: Physics-Informed Dynamic Convex Fields for Reconstruction and Simulation.- Objects as Audio-Visual Modal Sound Fields.- Reweighting Framewise Attention in Video Transformers for Facial Emotion Recognition.- Generation Models Know Space: Unleashing Implicit 3D Priors for Scene Understanding.- RADIANCE: Relative Adaptive Denoising with IP-Adapter for Novel Concept Enhancement.- Unlocking Complex Image Editing via Natively Interleaved Visual Textual CoT with Deep Confidence Reasoning.- Towards Sparsely Annotated Open World Object Detection.- OmniX: Any-view and Any-time 4D reconstruction via Feed-forward Trajectory Fields.- MSVS-VAE: Multi-Scale Anchored VecSet for High-Fidelity 3D Reconstruction.- Aligning Anything: Hierarchical Motion Estimation for Video Frame Interpolation.- TanGO: Training-Free 3D Editing via Tangent-Space Guidance and Optimization.- ReQuest: Rethinking-based Question-Aware Frame Selection for Long-Form Video QA.- Active View Selection with Perturbed Gaussian Ensemble for Tomographic Reconstruction.- SpaMEM: Benchmarking Dynamic Spatial Reasoning via Perception–Memory Integration in Embodied Environments.- InFlux++: Real and Synthetic Data for Estimating Dynamic Camera Intrinsics.- Capturing Spectral and Spatial Patterns for Federated Remote Sensing Segmentation.- Boosting 3D Foundation Models with Featureless Pose Optimization.- Doe-2: 3D Representation World Model for Unified Driving Scene Forecasting.- ProactiveBench: Benchmarking Proactiveness in Multimodal Large Language Models.- CASA: Cross-Attention over Self-Attention for Efficient Vision-Language Fusion.- Continuous Speculative Decoding for Autoregressive Image Generation.- H-Adapter: Pose-Robust Hairstyle Transfer via Attention-Derived, Source-Aligned Hair Masks.- PLOT: Pseudo-Labeling via Object Tracking for Monocular 3D Object Detection.- Domain Generalization via Text-Anchored Information Bottleneck.- There and Back Again: A Flexible-Frame Transformer for Multi-Exposure Fusion.- Difficulty-Conditioned Attribute-Specific Restoration for Low-Light Image Enhancement.- Proximity-CLIP: Text-Guided Semantic Proximity Learning for Zero-Shot Anomaly Detection.




