Tagged articles

Meituan LongCat

3 articles · Page 1 of 1
Meituan Technology Team
Meituan Technology Team
Jun 25, 2026 · Artificial Intelligence

Meituan LongCat’s VitaBench 2.0: A New Benchmark for Long‑Term Dynamic Agents

VitaBench 2.0, an open‑source benchmark from Meituan LongCat, evaluates large language models on long‑term, dynamic user interactions using 56 realistic users, 819 tasks, over 2 000 evolving preferences across up to 1 580 days, and reveals that even top models struggle with memory, personalization and proactive behavior.

LLM evaluationMeituan LongCatVitaBench 2.0
0 likes · 17 min read
Meituan LongCat’s VitaBench 2.0: A New Benchmark for Long‑Term Dynamic Agents
Design Hub
Design Hub
May 25, 2026 · Artificial Intelligence

How Meituan’s Open‑Source Avatar Redefines Digital Human Voice‑Over Costs (Beyond HeyGen)

LongCat‑Video‑Avatar‑1.5, Meituan’s open‑source audio‑driven video generation model, upgrades its encoder, stability, multi‑character support and 8‑step distillation, provides a detailed workflow, benchmark evaluation, and examines its impact on designers, operators, e‑commerce and marketing while highlighting deployment and compliance challenges.

AI videoAudio-driven Video GenerationContent Automation
0 likes · 22 min read
How Meituan’s Open‑Source Avatar Redefines Digital Human Voice‑Over Costs (Beyond HeyGen)
Meituan Technology Team
Meituan Technology Team
May 14, 2026 · Artificial Intelligence

General 365: Meituan LongCat’s Open‑Source Benchmark Redefines LLM Reasoning Evaluation

The General 365 benchmark, built from 365 original seed questions and 1,095 variants across eight reasoning challenges, reveals that most mainstream large language models struggle with everyday logical tasks, achieving at most 62.8% accuracy and requiring far more tokens than on traditional subject‑specific tests.

AI reasoningGeneral 365LLM evaluation
0 likes · 9 min read
General 365: Meituan LongCat’s Open‑Source Benchmark Redefines LLM Reasoning Evaluation