Tagged articles

Multimodal RL

2 articles · Page 1 of 1
Machine Heart
Machine Heart
Sep 3, 2026 · Artificial Intelligence

EASE: Teaching Multimodal RL Where to Look, Not Just What to Answer

EASE introduces evidence-anchored spatial attention supervision to multimodal reinforcement learning with verifiable rewards, using annotated evidence bounding boxes to guide model attention toward relevant image regions during training, improving accuracy on visual reasoning benchmarks without inference overhead.

Attention MechanismEMNLP 2026Evidence Grounding
0 likes · 13 min read
EASE: Teaching Multimodal RL Where to Look, Not Just What to Answer
Machine Learning Algorithms & Natural Language Processing
Machine Learning Algorithms & Natural Language Processing
Jun 18, 2026 · Artificial Intelligence

UniRL: Tencent Hunyuan’s Open‑Source Framework Unifying Multimodal RL Training

UniRL is an open‑source, distributed reinforcement‑learning post‑training framework that consolidates fragmented pipelines for image, video, and language‑vision models, offering a unified rollout‑reward‑advantage‑train‑sync contract, extensive model support, built‑in algorithms, and multi‑modal reward components to lower engineering barriers in AIGC research.

Distributed TrainingLLMMultimodal RL
0 likes · 10 min read
UniRL: Tencent Hunyuan’s Open‑Source Framework Unifying Multimodal RL Training