Tagged articles

long-term planning

7 articles · Page 1 of 1
AI Engineering
AI Engineering
Sep 1, 2026 · Artificial Intelligence

Why Long-Term LLM Agents Fail: The Same Design Choice Behind Two Deaths

Long‑term LLM agents suffer from ever‑slowing execution and context poisoning because they continuously append every observation, action, and reasoning step to the prompt, but the SKILL.state approach replaces this growing history with a compact mutable state, dramatically cutting token usage while boosting accuracy and robustness across diverse benchmarks.

GeminiGemmaLLM Agents
0 likes · 11 min read
Why Long-Term LLM Agents Fail: The Same Design Choice Behind Two Deaths
Xiaohongshu Tech REDtech
Xiaohongshu Tech REDtech
Aug 14, 2026 · Artificial Intelligence

dots3-note Preview: A First Step Toward Long‑Term Agents for Real‑World Service

The open‑source dots3-note Preview model, a 280B‑parameter multimodal agent with 512K context, introduces the TEMPO training scheme to improve long‑term reinforcement learning, achieves benchmark gains of up to 31.5% over baselines, and is evaluated on new VibeSearchBench and VibeLifeBench suites while acknowledging current limitations.

Agentic AIReinforcement Learningbenchmark
0 likes · 27 min read
dots3-note Preview: A First Step Toward Long‑Term Agents for Real‑World Service
Machine Heart
Machine Heart
Jul 17, 2026 · Artificial Intelligence

Qianxun’s Spirit v1.6 Beats “Goldfish Memory” to Clean a Whole Living Room

At WAIC, Qianxun showcased its Moz1 and Moz2 robots performing a full living‑room tidy‑up, demonstrating long‑range task execution, dynamic replanning, VLA‑world‑model integration, abstract action tokens, and a data‑driven pipeline that bridges industrial deployment and future service‑robot applications.

embodied intelligenceindustrial automationlong-term planning
0 likes · 15 min read
Qianxun’s Spirit v1.6 Beats “Goldfish Memory” to Clean a Whole Living Room
Machine Learning Algorithms & Natural Language Processing
Machine Learning Algorithms & Natural Language Processing
Jun 10, 2026 · Artificial Intelligence

Why Code Is the Core of Agent Harness: Deep Insights from UIUC, Meta, and Stanford

The article explains how code serves as the executable, inspectable, and stateful medium that links reasoning, action, feedback, verification, and collaboration in long‑term AI agents, detailing the harness interface, planning‑execute‑verify loop, multi‑agent coordination, and open research challenges.

AI AgentAgent HarnessCode as Interface
0 likes · 14 min read
Why Code Is the Core of Agent Harness: Deep Insights from UIUC, Meta, and Stanford
Architecture and Beyond
Architecture and Beyond
Jul 19, 2025 · Product Management

Mastering Endgame Thinking: Align Tactics with Long‑Term Strategy

The article explores the concept of “endgame thinking,” illustrating how starting from the desired future outcome and working backwards can improve product decisions, team management, technical architecture, and long‑term strategy, while warning against short‑sighted tactics, over‑design, and neglecting real constraints.

Product Strategydecision makingendgame thinking
0 likes · 16 min read
Mastering Endgame Thinking: Align Tactics with Long‑Term Strategy
ITPUB
ITPUB
Jul 31, 2023 · Databases

How to Choose the Right Database: Key Steps for Successful Selection

This guide walks you through the essential stages of database selection—from assessing project requirements and comparing candidate systems to performance testing, long‑term impact analysis, and making the final decision—ensuring you pick a solution that fits both current and future needs.

BenchmarkingNoSQLSQL
0 likes · 10 min read
How to Choose the Right Database: Key Steps for Successful Selection