Tagged articles

Streaming video

4 articles · Page 1 of 1
Machine Heart
Machine Heart
Aug 2, 2026 · Artificial Intelligence

Stream3D Enables Streaming 3D Reconstruction and Generation Without Retraining

Stream3D, a collaborative effort from Harvard, MIT and HKUST, introduces a training‑free, memory‑efficient mechanism that continuously extracts reliable multi‑view evidence from video streams to drive frozen view‑conditioned 3D generators, achieving consistent, complete 3D reconstructions superior to both pure reconstruction and generation baselines.

3D ReconstructionEvidence memoryMulti-view learning
0 likes · 14 min read
Stream3D Enables Streaming 3D Reconstruction and Generation Without Retraining
Data Party THU
Data Party THU
Jul 26, 2026 · Artificial Intelligence

VideoChat3: Open-Source Full-Stack Video Understanding Model Linking Perception, Understanding, and Interaction

VideoChat3 is a 4‑billion‑parameter multimodal large model that unifies short‑video, long‑video, and streaming video understanding through a native spatiotemporal encoder, adaptive resolution budgeting, and four‑stage training, achieving competitive accuracy while dramatically reducing visual token count and inference cost.

Adaptive ResolutionI3D-ViTLong Video
0 likes · 15 min read
VideoChat3: Open-Source Full-Stack Video Understanding Model Linking Perception, Understanding, and Interaction
Machine Heart
Machine Heart
Jul 22, 2026 · Artificial Intelligence

Open VideoChat3: A Full‑Stack 4B Video Understanding Model for Short, Long, and Streaming Videos

VideoChat3 is a 4‑billion‑parameter multimodal model that introduces an Inflated 3D Vision Transformer and an adaptive frame‑resolution mechanism to efficiently handle short clips, hour‑long videos, and live streams, achieving competitive benchmarks while being fully open‑source.

Adaptive Frame ResolutionI3D-ViTLong Video
0 likes · 13 min read
Open VideoChat3: A Full‑Stack 4B Video Understanding Model for Short, Long, and Streaming Videos
Machine Learning Algorithms & Natural Language Processing
Machine Learning Algorithms & Natural Language Processing
Jun 28, 2026 · Artificial Intelligence

Om AI Unveils Three Edge AI Models for Continuous Perception to Action

Om AI announced a three‑model VLX suite—VLX‑Flow, VLX‑Seek and VLX‑Go—designed to keep video streams continuously feeding a device‑side brain, using incremental visual memory and linear attention to meet the low‑latency, resource‑constrained demands of real‑world cameras, drones and robots.

Linear AttentionOm AIStreaming video
0 likes · 12 min read
Om AI Unveils Three Edge AI Models for Continuous Perception to Action