Old Zhang's AI Learning
Author

Old Zhang's AI Learning

AI practitioner specializing in large-model evaluation and on-premise deployment, agents, AI programming, Vibe Coding, general AI, and broader tech trends, with daily original technical articles.

307
Articles
2
Likes
3.9k
Views
0
Comments
Recent Articles

Latest from Old Zhang's AI Learning

100 recent articles max
Old Zhang's AI Learning
Old Zhang's AI Learning
Jun 26, 2026 · Artificial Intelligence

Claude‑style 9B Model with 1M‑Token Context Runs Locally

Qwythos‑9B, a Qwen3.5‑9B model fine‑tuned with over 500 M Claude‑style tokens, offers a 1 M‑token YaRN context, native function calling and tool‑augmented self‑correction, outperforms its base on MMLU and gsm8k benchmarks, and provides GGUF quantizations for consumer‑grade GPU deployment.

1M tokenClaudeFunction Calling
0 likes · 15 min read
Claude‑style 9B Model with 1M‑Token Context Runs Locally
Old Zhang's AI Learning
Old Zhang's AI Learning
Jun 25, 2026 · Industry Insights

Beyond WorkBuddy: Tencent’s Hidden AI Agent Play with AiPy

The article analyzes Tencent’s dual‑track AI Agent strategy, detailing how the consumer‑focused WorkBuddy leverages the company’s ecosystem while the newly acquired AiPy from security firm Zhidao Chuangyu targets enterprise and government markets with on‑premise, code‑as‑agent technology, and evaluates the competitive landscape and future prospects.

AI AgentAiPyCode is Agent
0 likes · 14 min read
Beyond WorkBuddy: Tencent’s Hidden AI Agent Play with AiPy
Old Zhang's AI Learning
Old Zhang's AI Learning
Jun 24, 2026 · Industry Insights

OpenAI Unveils Its Own AI Inference Chip: What It Means for the Industry

OpenAI has partnered with Broadcom to launch Jalapeño, a purpose‑built AI inference ASIC designed in nine months, promising superior performance‑per‑watt, integrated networking, and a full‑stack AI hardware‑software optimization cycle that could lower inference costs and reshape future data‑center deployments.

AI hardwareAI inference chipASIC
0 likes · 6 min read
OpenAI Unveils Its Own AI Inference Chip: What It Means for the Industry
Old Zhang's AI Learning
Old Zhang's AI Learning
Jun 24, 2026 · Artificial Intelligence

Universal Video Download Skill Evolves into Full‑Video Summarization (z‑video‑study‑webpage‑qwen)

The author open‑sources a universal video‑download Skill and then introduces a companion Skill that automatically extracts audio, frames, and visual insights from a local MP4, runs Whisper and qwen3.7‑plus to generate a structured summary webpage with player, key points, timeline and actionable items.

Whispermultimodal AIopen source
0 likes · 3 min read
Universal Video Download Skill Evolves into Full‑Video Summarization (z‑video‑study‑webpage‑qwen)
Old Zhang's AI Learning
Old Zhang's AI Learning
Jun 22, 2026 · Artificial Intelligence

How Codex Became My Ultimate Computer Assistant

The author demonstrates how OpenAI Codex can serve as a full‑featured computer manager on macOS, automating cache cleaning, software uninstall, startup service control, large‑file detection, browser data cleanup, material organization, and daily inspections through tailored prompts and screenshots.

AI-powered PC managementOpenAI CodexSystem Automation
0 likes · 10 min read
How Codex Became My Ultimate Computer Assistant
Old Zhang's AI Learning
Old Zhang's AI Learning
Jun 21, 2026 · Artificial Intelligence

Finding the ‘Father’ of Any Concept: My Father’s Day AI Skill

On Father’s Day the author built an AI Agent skill called z‑father‑concept that, given any term, traces its lineage through concrete ancestors, functional roles, societal issues and finally a philosophical theme, illustrating the process with examples from fans to loneliness.

AI AgentSkill Designconcept hierarchy
0 likes · 12 min read
Finding the ‘Father’ of Any Concept: My Father’s Day AI Skill
Old Zhang's AI Learning
Old Zhang's AI Learning
Jun 19, 2026 · Artificial Intelligence

Gemma‑4‑12B‑v2 (Fable 5 Clone) Achieves 3.5× Telecom Benchmark Boost

The author reproduces Anthropic’s Fable 5 using Gemma‑4‑12B‑v2, showing a 3.5× improvement on the telecom tau2‑bench versus the base model, details the agentic, coding, and general training data, compares quantization sizes, provides llama.cpp launch commands, and notes speed gains from speculative MTP decoding and current limitations.

Agentic AIFable 5Gemma-4-12B
0 likes · 9 min read
Gemma‑4‑12B‑v2 (Fable 5 Clone) Achieves 3.5× Telecom Benchmark Boost