Tagged articles

MoE models

3 articles · Page 1 of 1
Lao Guo's Learning Space
Lao Guo's Learning Space
Sep 23, 2026 · Artificial Intelligence

Mac mini 16GB to Studio 256GB: Which Mac Runs Your LLMs Best?

This article analyzes Mac mini and Mac Studio configurations for local LLM inference, showing how unified memory capacity determines which models fit and memory bandwidth dictates token generation speed, with real-world benchmarks for 8B to 235B models across M6, M5 Pro, M5 Max, and M5 Ultra chips.

Apple SiliconLLM inferenceMac Studio
0 likes · 13 min read
Mac mini 16GB to Studio 256GB: Which Mac Runs Your LLMs Best?
Architects' Tech Alliance
Architects' Tech Alliance
Apr 30, 2026 · Artificial Intelligence

Token Era Unpacked: The ‘One Chip, Two Models, Three Clouds’ Blueprint for AI Agents

The article analyzes how the rise of AI agents transforms the industry from dialogue‑centric models to 24/7 digital employees, driving a shift toward CPU‑centric compute, domestic MoE models with strong coding abilities, and cloud platforms that become the core deployment and billing ecosystem, all fueled by massive token inflation.

AI AgentsAI hardwareCloud AI
0 likes · 13 min read
Token Era Unpacked: The ‘One Chip, Two Models, Three Clouds’ Blueprint for AI Agents