Old Zhang's AI Learning
Author

Old Zhang's AI Learning

AI practitioner specializing in large-model evaluation and on-premise deployment, agents, AI programming, Vibe Coding, general AI, and broader tech trends, with daily original technical articles.

307
Articles
2
Likes
3.9k
Views
0
Comments
Recent Articles

Latest from Old Zhang's AI Learning

100 recent articles max
Old Zhang's AI Learning
Old Zhang's AI Learning
Sep 10, 2026 · Artificial Intelligence

DeepSeek-V4.1-Flash: 8B Activated Model Beats 1.6T V4-Pro, Local Deployment Tested

DeepSeek-V4.1-Flash open-sourced with CED architecture and 8B/16B activated parameters outperforms its 1.6T predecessor V4-Pro on coding benchmarks, approaches GPT-6 Astra on DeepSWE, but requires 510GB FP8 weights needing 8×H200 for full-context local deployment; author tests across six harnesses finding Claude Code/Codex integration near peak performance.

AI model evaluationCED architectureDeepSWE
0 likes · 8 min read
DeepSeek-V4.1-Flash: 8B Activated Model Beats 1.6T V4-Pro, Local Deployment Tested
Old Zhang's AI Learning
Old Zhang's AI Learning
Sep 7, 2026 · Artificial Intelligence

Doubao Work Deep Dive: 5 Advanced AI Workflows for PPT, Data & App Building

After two weeks of intensive use and watching official livestreams, the author reveals five advanced Doubao Work workflows: data-first PPT creation, browser automation for tedious data collection, serverless shareable app building, custom skill development, and candid product feedback for this AI productivity agent.

AI productivity agentDoubao WorkPPT automation
0 likes · 10 min read
Doubao Work Deep Dive: 5 Advanced AI Workflows for PPT, Data & App Building
Old Zhang's AI Learning
Old Zhang's AI Learning
Sep 3, 2026 · Artificial Intelligence

Gemini 3.8 Flash Review: Benchmark Leader but Real-World Letdown?

The author evaluates Google's Gemini 3.8 Flash model through hands-on coding, creative, and reasoning tests, finding it excels on benchmarks and cost-efficiency but falls short in agent capabilities, context retention, and practical tasks compared to rivals like Doubao and DeepSeek, while criticizing Google's product management and subscription value.

AI model benchmarkingGemini 3.8 FlashGoogle AI products
0 likes · 6 min read
Gemini 3.8 Flash Review: Benchmark Leader but Real-World Letdown?
Old Zhang's AI Learning
Old Zhang's AI Learning
Sep 1, 2026 · Artificial Intelligence

Doubao-Seed-Evolving Tested: Coding, Multimodal, and Long-Horizon Agent Skills

The author evaluates Doubao-Seed-Evolving's latest upgrades across coding, agent, multimodal, and long-horizon skill execution using real-world tasks like PPT-to-website conversion, personal project management, SVG generation, 3D Rubik's cube animation, and a complex 30-minute web scraping skill, finding strong autonomous task decomposition, testing, and error recovery.

AI model testingDoubao-Seed-EvolvingLLM evaluation
0 likes · 10 min read
Doubao-Seed-Evolving Tested: Coding, Multimodal, and Long-Horizon Agent Skills
Old Zhang's AI Learning
Old Zhang's AI Learning
Aug 29, 2026 · Artificial Intelligence

ChatGPT's Computer History: AI That Watches Your Workflow and Writes Automations

ChatGPT's new Computer History feature on macOS captures clicks, keystrokes, and app switches via Accessibility API, builds a searchable timeline, summarizes activity into local markdown memories, and suggests automations for repetitive tasks—available only to Pro, Business, and Enterprise users with granular privacy controls and notable prompt-injection risks.

AI AgentsAccessibility APIAutomation
0 likes · 10 min read
ChatGPT's Computer History: AI That Watches Your Workflow and Writes Automations