Tagged articles

multitask vision

2 articles · Page 1 of 1
SuanNi
SuanNi
Jul 15, 2026 · Artificial Intelligence

DeepMind’s Video Generation Model Becomes a General Visual Intelligence – He Kaiming’s Involvement

GenCeption repurposes a 140‑billion‑parameter text‑to‑video diffusion model into a single‑step feed‑forward visual system that handles depth, segmentation, pose and other tasks via text prompts, achieves state‑of‑the‑art results with far fewer training frames, and demonstrates strong out‑of‑domain generalisation using synthetic data.

GenCeptionmultitask visionsynthetic data
0 likes · 10 min read
DeepMind’s Video Generation Model Becomes a General Visual Intelligence – He Kaiming’s Involvement
AIWalker
AIWalker
Jun 4, 2026 · Artificial Intelligence

How YOLO26 Redefines Real‑Time Detection: NMS‑Free Dual‑Head Architecture Beats YOLO11

YOLO26 eliminates NMS and DFL, adopts a dual‑head design, MuSGD optimizer, progressive loss weighting, and STAL small‑object assignment, achieving 57.5 mAP with 1.7 ms latency on COCO while unifying detection, segmentation, pose, OBB and open‑set tasks, as shown by extensive ablations.

MuSGD optimizerSTAL small-object assignmentYOLO26
0 likes · 14 min read
How YOLO26 Redefines Real‑Time Detection: NMS‑Free Dual‑Head Architecture Beats YOLO11