Tagged articles

ECCV 2026

8 articles · Page 1 of 1
Machine Heart
Machine Heart
Aug 5, 2026 · Artificial Intelligence

Atomic Dance: Explainable Music-to-Dance Generation Using Atomic Movements

The paper introduces Atomic Dance, a two‑stage framework that first discovers repeatable, semantically labeled atomic movements and then plans and completes dance sequences, achieving more coherent, rhythm‑aligned and editable music‑driven choreography, as demonstrated on the AIST++ benchmark.

Atomic MovementsChoreographyECCV 2026
0 likes · 9 min read
Atomic Dance: Explainable Music-to-Dance Generation Using Atomic Movements
Machine Learning Algorithms & Natural Language Processing
Machine Learning Algorithms & Natural Language Processing
Jul 20, 2026 · Artificial Intelligence

How VGGRPO Uses 4D Latent Rewards for World‑Consistent Video Generation

VGGRPO introduces a latent‑space geometry model and two 4D rewards—camera motion smoothness and geometry reprojection consistency—to improve geometric consistency in video diffusion models without sacrificing pre‑training generalization, achieving smoother camera paths and coherent scene structures even in dynamic scenarios.

4D rewardECCV 2026VGGRPO
0 likes · 9 min read
How VGGRPO Uses 4D Latent Rewards for World‑Consistent Video Generation
Machine Heart
Machine Heart
Jul 17, 2026 · Artificial Intelligence

VGGRPO: 4D Latent Rewards for World‑Consistent Video Generation (ECCV 2026)

VGGRPO introduces a latent‑space geometry model and two 4D rewards—camera‑motion smoothness and geometry‑reprojection consistency—to eliminate drift and improve structural coherence in video diffusion models without altering their pretrained architecture, achieving state‑of‑the‑art results on static and dynamic benchmarks.

4D rewardECCV 2026Video Generation
0 likes · 7 min read
VGGRPO: 4D Latent Rewards for World‑Consistent Video Generation (ECCV 2026)
Machine Learning Algorithms & Natural Language Processing
Machine Learning Algorithms & Natural Language Processing
Jul 15, 2026 · Artificial Intelligence

Is Video Generation the ‘Next Token Prediction’ for Vision? Insights from the GenCeption Paper

The ECCV 2026 paper by He Kaiming, Andrew Zisserman and collaborators proposes GenCeption, a text‑to‑video model repurposed as a single‑forward DiT‑based vision learner that unifies depth, segmentation, pose and 3D tasks, and evaluates its multi‑task performance against specialized baselines.

DiTECCV 2026GenCeption
0 likes · 10 min read
Is Video Generation the ‘Next Token Prediction’ for Vision? Insights from the GenCeption Paper
Machine Heart
Machine Heart
Jul 11, 2026 · Artificial Intelligence

Real-Time Multi-Shot Long Video Generation: Introducing ShotStream (ECCV 2026)

ShotStream tackles the high latency and zero‑interaction problems of multi‑shot long video generation by proposing a streaming architecture with a dual‑cache memory, discontinuous RoPE, and a two‑stage self‑forcing distillation, achieving over 25× speedup to 16 FPS on a single H200 GPU and outperforming existing bidirectional and autoregressive models.

ECCV 2026Real-time StreamingShotStream
0 likes · 8 min read
Real-Time Multi-Shot Long Video Generation: Introducing ShotStream (ECCV 2026)
Machine Heart
Machine Heart
Jul 7, 2026 · Artificial Intelligence

Unlocking Free‑View Video Virtual Try‑On with TryOnCrafter’s 4D Try‑On Proxy

TryOnCrafter introduces a camera‑controllable video virtual try‑on framework that builds a renderable 4D proxy to enable free‑view, 360° and bullet‑time effects while preserving structural stability, texture consistency, and realistic motion across arbitrary camera trajectories.

3D Gaussian Splatting4D try-on proxyECCV 2026
0 likes · 16 min read
Unlocking Free‑View Video Virtual Try‑On with TryOnCrafter’s 4D Try‑On Proxy
Machine Heart
Machine Heart
Jul 4, 2026 · Artificial Intelligence

LinStereo Bridges the Last Mile of Stereo Matching (ECCV 2026)

LinStereo replaces ConvGRU with a position‑aware linear attention module, adds a multi‑scale cost volume and monocular depth initialization, cutting Middlebury occlusion error by 37%, outperforming larger models, and achieving strong zero‑shot underwater performance while remaining parameter‑efficient.

ECCV 2026Linear AttentionStereo Matching
0 likes · 10 min read
LinStereo Bridges the Last Mile of Stereo Matching (ECCV 2026)
Amap Tech
Amap Tech
Jun 30, 2026 · Artificial Intelligence

Six ECCV 2026 Papers – Vision, Video Generation, Visual‑Language Navigation

ECCV 2026 received 10,473 submissions and accepted 2,883 (27.5%); Gaode contributed six papers spanning computer vision, generative video, and visual‑language navigation, each presenting novel reinforcement‑learning or multimodal frameworks, new datasets, and benchmark results that outperform prior state‑of‑the‑art methods.

ECCV 2026Video Generationcomputer vision
0 likes · 13 min read
Six ECCV 2026 Papers – Vision, Video Generation, Visual‑Language Navigation