Tagged articles

DeepSeek V4 Pro

6 articles · Page 1 of 1
PaperAgent
PaperAgent
Sep 13, 2026 · Artificial Intelligence

Trace as State: Simple Inference-Time Scaling Lifts DeepSeek V4 Pro 29%→82%

Z.ai and Tsinghua's Tang Jie propose Trace as State, a training-free inference-time scaling method that places reasoning traces before long context, achieving 26 out of 27 wins across models and benchmarks, boosting DeepSeek V4 Pro from 29.2% to 81.8% and GLM-5.2 to 100% on GraphWalks Parents.

DeepSeek V4 ProGLM-5.2GraphWalks
0 likes · 12 min read
Trace as State: Simple Inference-Time Scaling Lifts DeepSeek V4 Pro 29%→82%
Smart Era Software Development
Smart Era Software Development
Aug 15, 2026 · Artificial Intelligence

From Model Params to Full‑System 'Model+Harness': DeepSeek V4 Pro Agent Engineering Deep Dive

The report reveals how Agent competition has shifted from pure model‑parameter races to a full‑system "model+Harness" battle, detailing DeepSeek V4 Pro's technical breakthroughs, massive cost advantage, four‑stage development roadmap, benchmark improvements, industry trends, expert insights, and commercial pathways for AI Agents.

AI agentsAgent EngineeringAgent commercialization
0 likes · 42 min read
From Model Params to Full‑System 'Model+Harness': DeepSeek V4 Pro Agent Engineering Deep Dive
Design Hub
Design Hub
Aug 14, 2026 · Artificial Intelligence

Four AI Releases in One Day: What’s Shaping the Emerging AI Delivery Stack?

On a single day, Google, DeepSeek, and MiniMax unveiled Gemini 3.7 Flash, V4‑Pro, the Harness runtime, and Music 3, each illustrating how AI is shifting from headline‑grabbing benchmarks toward cost‑effective agents, plug‑in runtimes, and controllable content generation for real‑world workflows.

AI agentsDeepSeek HarnessDeepSeek V4 Pro
0 likes · 13 min read
Four AI Releases in One Day: What’s Shaping the Emerging AI Delivery Stack?
DataFunTalk
DataFunTalk
Aug 13, 2026 · Artificial Intelligence

Overnight DeepSeek V4 Pro Test Reveals Disappointing Performance – Not Fit for Codex

After integrating the newly released DeepSeek V4 Pro into Codex and running 41 million tokens, the author finds the model’s silent operation, excessive context copying, weak frontend and writing abilities, and higher latency make it unsuitable as a primary model despite solid throughput and low cost.

DeepSeek V4 ProLLM evaluationcontext reuse
0 likes · 7 min read
Overnight DeepSeek V4 Pro Test Reveals Disappointing Performance – Not Fit for Codex
Top Architecture Tech Stack
Top Architecture Tech Stack
Apr 27, 2026 · Artificial Intelligence

DeepSeek V4 Pro vs GPT‑5.3 Codex High: Direct Code‑Generation Test Reveals the Gap

A two‑stage evaluation compares DeepSeek V4 Pro and GPT‑5.3 Codex High on a TypeScript LRU‑Cache task and a markdown‑inspection CLI project, showing DeepSeek leads on basic code correctness while GPT‑5.3 delivers a more complete engineering solution, with detailed scores and analysis.

Agent EngineeringDeepSeek V4 ProGPT-5.3 Codex High
0 likes · 13 min read
DeepSeek V4 Pro vs GPT‑5.3 Codex High: Direct Code‑Generation Test Reveals the Gap