{"product_id":"computer-vision-eccv-2026-paolo-favaro-ebook-25","title":"Computer Vision - ECCV 2026","description":"\u003cp\u003eSGP2: Coarse-to-Fine Controllable Multimodal Remote Sensing Image Generation.- Tiled Prompts: Overcoming Prompt Misguidance in Image and Video Super-Resolution.- GameWorlds: Towards Standardized and Verifiable Evaluation of Multimodal Game Agents.- InSpace: Structure-Aware 3D Indoor Scene Generation from a Single 360° Image.- Video Streaming Thinking: VideoLLMs Can Watch and Think Simultaneously.- DIVER: Disentangling Camera–Object and Active–Passive Motion for Video Generation.- Text-Conditioned Background Generation for Editable Multi-Layer Documents.- VoCa: Unified Autoregressive Modeling for Talking Audio-Video Generation.- Decompose, Compare, and Decide: Multimodal LLMs are Implicit Few-Shot Learners.- NarrativeTrack: Evaluating Entity-Centric Reasoning for Narrative Understanding.- Hi-DiT: Hybrid Latent-Pixel Diffusion Transformer for Image Generation.- Cube-Splat: High-Fidelity 360° Gaussian Splatting SLAM via Cubemap Factorization and Adjoint-Consistent Optimization.- Motion-aware Sparse Pipeline for Lightweight Object Tracking.- VideoTIR: Accurate Understanding for Long Videos with Efficient Tool-Integrated Reasoning.- Face Anything: 4D Face Reconstruction from Any Image Sequence.- ReDesign: Recovering Editable Design Structures from Raster Images via Agentic Decomposition.- Distill on a Diet: Efficient Knowledge Distillation via Learnable Data Pruning.- SPAR: Semantic-Pixel Self-Alignment and Adaptive Routing for Unified Multimodal Models.- VoxAnchor: Explicit Voxel-Semantic Grounding for Spatial Understanding in Videos.- LISA: Locality-Informed Speculative Decoding for Accelerating Autoregressive Image Generation.- Layering Virtual Try-On.- clean2green2clean: Synthesising Bad Composites to Learn Actor-Background Video Harmonisation.- EatVid-Bench: A Multimodal Fine-Grained Eating Behavior Video Dataset.- EAGS: Error-Aware Gaussian Splatting with Dual-Confidence-Guided Modeling for Uncalibrated Driving Scenes.- What Images Cannot Say: Language-Guided Olfactory Representation Learning.- PAI-Studio: Cinematic Video Background Replacement with Camera-Aware Motion.- Multi-modal Knowledge Preserving Adapter for Embedding Backward Compatibility.- Sparse-View Surface Reconstruction using Gaussian Splatting through High-Confidence Depth Propagation with Normal Priors.- Reasoning Path and Latent State Analysis for Multi-view Visual Spatial Reasoning: A Cognitive Science Perspective.- Calibrate Before Adapt: Training-Free Pseudo-Label Calibration for Semi-Supervised Cross-Domain Few-Shot Detection.- Atlas is Your Perfect Context: One-Shot Customization for Generalizable Foundational Medical Image Segmentation.- ICLAgent:  Integrated Circuit Footprint Geometry Labeling via LMM-empowered Multi-Agent Framework.- Frozen CLIP Priors for Robust Self-Supervised Poisson Inverse Problems.- CurveStream: Boosting Streaming Video Understanding in MLLMs via Curvature-Aware Hierarchical Visual Memory Management.- 3D Field of Junctions: A Noise-Robust, Training-Free Structural Prior for Volumetric Inverse Problems.- Token-level Response-visual Attention Guidance for Multimodal LLMs Knowledge Distillation.- Benchmarking MLLMs on Mistake Recognition and Explanation in Single-Step Components of Cooking.\u003c\/p\u003e","brand":"Paolo Favaro","offers":[{"title":"Default Title","offer_id":55097174688071,"sku":"9783032372550","price":96.29,"currency_code":"EUR","in_stock":true}],"thumbnail_url":"\/\/cdn.shopify.com\/s\/files\/1\/0920\/5455\/2903\/files\/computer-vision-eccv-2026-ebook-cover-paolo-favaro-springer-2026-v19.webp?v=1789885561","url":"https:\/\/www.cinebuch.de\/products\/computer-vision-eccv-2026-paolo-favaro-ebook-25","provider":"CineBuch","version":"1.0","type":"link"}