Tagged articles

System Design

799 articles · Page 1 of 8
Java Captain
Java Captain
Sep 30, 2026 · Interview Experience

WXG First-Round Interview: 60 Questions from Algorithms to Agent Design

A shared WXG first-round interview experience lists 60 questions covering self-introduction, algorithms (IP-to-uint64, linked-list folding, uniform sampling), system design, AI tools, large-model hallucinations, RAG, vector search, collaborative filtering, 12306 ticketing architecture, Redis internals, MySQL B+ trees, concurrency locks, and design patterns.

AlgorithmsB+ TreeDesign Patterns
0 likes · 7 min read
WXG First-Round Interview: 60 Questions from Algorithms to Agent Design
dbaplus Community
dbaplus Community
Sep 27, 2026 · Fundamentals

Why the Creators of Linux, Python, Redis, Nginx & SQLite Are True Geniuses

The article argues that creators of foundational software like Linux, Python, Redis, Nginx, and SQLite are true geniuses due to their exceptional abstraction skills and system-level boundary control, illustrated by real-world cases where their minimal, convergent designs outperform modern complex solutions, and warns that over-reliance on AI tools erodes deep system understanding essential for solving hard production issues.

AI-assisted codingLinuxNginx
0 likes · 14 min read
Why the Creators of Linux, Python, Redis, Nginx & SQLite Are True Geniuses
Ops Development & AI Practice
Ops Development & AI Practice
Sep 25, 2026 · Interview Experience

Why Production Heroes Fail Interviews: Converting Systemic Intuition into Architectural Proof

This article explains why experienced engineers who excel at real-world troubleshooting often struggle in technical interviews due to a structural mismatch between systemic debugging intuition and rote memorization tests, and provides a three-layer drill-down model plus a five-step diagnostic derivation framework to translate practical expertise into compelling architectural narratives that interviewers value.

System Designarchitectural communicationcareer development
0 likes · 15 min read
Why Production Heroes Fail Interviews: Converting Systemic Intuition into Architectural Proof
Java Captain
Java Captain
Sep 25, 2026 · Backend Development

Tencent Architect's Decade of Insights: Simple, Evolvable Architecture Principles

A Tencent architect with nearly ten years experience shares principles for simple, evolvable architecture: modularization, layering, focusing on core business needs like QQ's social chat, avoiding premature optimization, service design with monitoring and canary releases, and code design with clear structure and single responsibility.

QQSystem DesignTencent
0 likes · 4 min read
Tencent Architect's Decade of Insights: Simple, Evolvable Architecture Principles
Ops Development & AI Practice
Ops Development & AI Practice
Sep 25, 2026 · Interview Experience

Forgetting Details ≠ Incompetence: Senior Engineers' Cognitive Compression Playbook

This article explains why senior engineers naturally forget micro-details through the brain's lossy compression, and provides a three-part framework — pre-interview 'deep-water nails' (real failure cases, core parameters, personal artifacts) and three on-the-spot defense frameworks (deduction, methodology anchoring, pivoting) — to demonstrate unforgeable engineering depth in high-stakes technical interviews.

GoSystem Designarchitecture review
0 likes · 18 min read
Forgetting Details ≠ Incompetence: Senior Engineers' Cognitive Compression Playbook
Cloud Architecture
Cloud Architecture
Sep 24, 2026 · Backend Development

Scaling Spring Boot Sign-In to 100M Users: Distributed Architecture Evolution & Production Hardening

This article details the evolution of a Spring Boot daily sign-in system from 10K to 100M users, covering Redis bitmap sharding, Lua atomic operations, command outbox pattern, Kafka async reward processing, idempotency guarantees, fault tolerance strategies, and production-grade observability with real incident postmortems.

BitmapCommand OutboxKafka
0 likes · 36 min read
Scaling Spring Boot Sign-In to 100M Users: Distributed Architecture Evolution & Production Hardening
Data Bricklaying Diary
Data Bricklaying Diary
Sep 21, 2026 · Backend Development

Scaling Isn't Just Adding Instances: End-to-End Capacity, Traffic & Cost Design

This article argues that true system scaling requires end-to-end capacity planning across ingress, task processing, dependencies, and recovery—not merely adding instances—and shows how to translate business growth into workload models, manage backpressure, set per-link capacity budgets, validate with realistic load tests, and balance cost against performance.

Load TestingSystem Designauto-scaling
0 likes · 24 min read
Scaling Isn't Just Adding Instances: End-to-End Capacity, Traffic & Cost Design
Weekly Large Model Application
Weekly Large Model Application
Sep 21, 2026 · Artificial Intelligence

Full-Duplex Voice Agents: When to Commit Parameters? Eager, Conservative, Two-Phase Compared

This article analyzes three commit strategies—Eager, Conservative, and Two-phase—for tool calling in full-duplex voice agents, explaining how each handles user self-correction, side-effect safety, and latency trade-offs, and provides selection guidelines based on operation reversibility.

System DesignTool Callingconservative commit
0 likes · 8 min read
Full-Duplex Voice Agents: When to Commit Parameters? Eager, Conservative, Two-Phase Compared
James' Growth Diary
James' Growth Diary
Sep 18, 2026 · Backend Development

Container Routing: Why the Same Agent Lands in Different Containers

This article details a four-step container routing mechanism (authentication, permission, routing, resolution) for an Agent platform, explaining how requests reach different container forms based on identity and instance health, and why distinguishing hard vs soft failures is critical for debugging.

AuthenticationAuthorizationSystem Design
0 likes · 27 min read
Container Routing: Why the Same Agent Lands in Different Containers
IT Services Circle
IT Services Circle
Sep 15, 2026 · Fundamentals

Why Linus Won't Put a National Anti-Fraud Center in the Linux Kernel

The article explains why integrating a national anti-fraud center into the Linux kernel violates core OS principles: kernels must remain neutral, minimal, and secure; adding application-specific logic with kernel-level privileges breaks least privilege, introduces instability, and undermines the open-source consensus that keeps Linux universal.

Linux kernelSystem Designkernel architecture
0 likes · 8 min read
Why Linus Won't Put a National Anti-Fraud Center in the Linux Kernel
dbaplus Community
dbaplus Community
Sep 13, 2026 · R&D Management

Uber EM Reveals 14 Principles for System Design, Team Leadership & Career Transition

Uber engineering manager Sendil Nellaiyapen shares 14 hard-won principles covering system design foundations, scaling from thousands to millions, MVP trade-offs, clarity over seniority, avoiding experience traps, hypothesis-driven culture, latency misconceptions, guardrails for autonomy, and the mindset shift from engineer to manager.

Career TransitionMVPSystem Design
0 likes · 24 min read
Uber EM Reveals 14 Principles for System Design, Team Leadership & Career Transition
Data Party THU
Data Party THU
Sep 13, 2026 · Artificial Intelligence

Agentic AI Systems: Reasoning Loops, Tools & Guardrails Explained

This article contrasts traditional RAG pipelines with agentic AI systems, detailing the four core components—orchestrator, tool calling, memory, and guardrails—and demonstrates how reasoning loops enable multi-step problem solving for system design interviews.

Agentic AIMemory ManagementOrchestrator
0 likes · 16 min read
Agentic AI Systems: Reasoning Loops, Tools & Guardrails Explained
TonyBai
TonyBai
Sep 13, 2026 · Backend Development

2 Engineers, AI, and a Rust Rewrite: Scaling OpenAI's Storage to 1B Users

OpenAI's Habitat storage system evolved from a Python library to a distributed platform handling 70M requests/second for 1B users, with engineers detailing scaling challenges, asyncio tuning, connection pool fixes, and a 2-engineer Rust rewrite using Codex and GPT-5.5 that boosted CPU efficiency 6x and memory efficiency 15x.

AI-assisted codingCodexGPT-5.5
0 likes · 26 min read
2 Engineers, AI, and a Rust Rewrite: Scaling OpenAI's Storage to 1B Users
Architect
Architect
Sep 12, 2026 · Artificial Intelligence

Google's Multi-Agent Research: Task Structure, Not Agent Count, Determines Architecture Value

Google's research on 260 multi-agent configurations across six benchmarks shows centralized architectures improve parallel tasks by 81% but hurt sequential planning by 39-70%. Teamwork framework adds critique-synthesis loops that retain failed branches. The key insight: agent count isn't an architecture metric—task decomposability, verifiable sub-results, and coordination costs should drive design.

AI agentsAgent ArchitectureGoogle Research
0 likes · 18 min read
Google's Multi-Agent Research: Task Structure, Not Agent Count, Determines Architecture Value
TechVision Expert Circle
TechVision Expert Circle
Sep 9, 2026 · Industry Insights

AI Isn't Eliminating Tech Jobs—It's Quietly Compressing Headcount

In 2026, AI tools like Claude Code and GitHub Copilot Workspace are silently compressing junior tech roles by automating coding, testing, and ops tasks, reducing headcount needs while increasing workload for remaining engineers, making system design, incident judgment, and domain expertise the new irreplaceable skills.

AI-assisted developmentModel Context ProtocolSystem Design
0 likes · 13 min read
AI Isn't Eliminating Tech Jobs—It's Quietly Compressing Headcount
Data Bricklaying Diary
Data Bricklaying Diary
Sep 9, 2026 · Artificial Intelligence

Three Graphs, Three Jobs: Loop, Task Graph & Plugin Runtime in Agent Systems

This article argues that complex Agent systems must separate three distinct structures: Loop handles node-level convergence, Task Graph manages work dependencies and coordination, and Plugin Runtime binds capabilities like models and tools; mixing them into a monolithic Agent leads to invisible boundaries and unrecoverable failures.

Agent ArchitectureComponent GraphContext
0 likes · 17 min read
Three Graphs, Three Jobs: Loop, Task Graph & Plugin Runtime in Agent Systems
Code Farming
Code Farming
Sep 8, 2026 · Backend Development

QR Code Payments Decoded: 4 Systems, 6 Institutions, and the Async Architecture Behind Every Scan

This article breaks down the complete technical chain behind a QR code payment, revealing how authentication tokens, cross-bank clearing with central bank settlement, foreign exchange conversion, and asynchronous messaging systems work together to move money across institutions and currencies without synchronous responses.

QR code paymentSystem Designasynchronous systems
0 likes · 9 min read
QR Code Payments Decoded: 4 Systems, 6 Institutions, and the Async Architecture Behind Every Scan
Code Farming
Code Farming
Sep 7, 2026 · Backend Development

How Encyclopedia Systems Survive Data Center Fires: Architecture Deep Dive

This article breaks down the four-step architecture design of a high-concurrency encyclopedia system, covering latency-driven multi-data-center deployment, a five-layer request chain with caching at each level, a three-tier optimization pyramid, and master-slave synchronization with automatic failover for disaster recovery.

CDN cachingCanalGeoDNS
0 likes · 8 min read
How Encyclopedia Systems Survive Data Center Fires: Architecture Deep Dive
ITPUB
ITPUB
Sep 6, 2026 · Backend Development

How WeChat Resets 1 Billion Step Counts at Midnight Without Crashing

WeChat avoids server crashes during midnight step-count resets for 1 billion users by using logical time-based versioning instead of physical updates, a custom PaxosStore for atomic increments, delayed double-write buffers for clock skew, Redis ZSet sharding for rankings, and asynchronous cold-data archival during low-traffic hours.

PaxosStoreRedisSystem Design
0 likes · 18 min read
How WeChat Resets 1 Billion Step Counts at Midnight Without Crashing
James' Growth Diary
James' Growth Diary
Sep 5, 2026 · Artificial Intelligence

Why Build Your Own Agent: From Chat to Reliable Execution

This article argues that chat APIs alone are insufficient for AI agents; true agent systems require execution environments, tool loops, multi-tenant isolation, and layered architecture to move from demo to production-grade reliability across multiple entry points.

AI agentsAgent ArchitectureExecution Environment
0 likes · 22 min read
Why Build Your Own Agent: From Chat to Reliable Execution
Design Hub
Design Hub
Sep 1, 2026 · Artificial Intelligence

How to Build an AI Agent That Won’t Fall Apart with Harness Engineering

The article explains that AI agents often fail because they lack a reliable runtime environment—called a Harness—and outlines a systematic Harness Engineering approach, including seven core responsibilities, a practical checklist, and concrete examples to turn failures into reusable infrastructure.

AI agentsAgent ReliabilityHarness Engineering
0 likes · 19 min read
How to Build an AI Agent That Won’t Fall Apart with Harness Engineering
DeepNoMind
DeepNoMind
Aug 30, 2026 · Interview Experience

System Design Interview Prep: 7 Resources & a 7-Week Roadmap to Offer

This article outlines a structured 7-week preparation plan for system design interviews, recommending seven key resources — from foundational primers to real-world case studies — and emphasizes deliberate practice over passive video consumption to master trade-off reasoning under pressure.

7-week planAlex XuByteByteGo
0 likes · 12 min read
System Design Interview Prep: 7 Resources & a 7-Week Roadmap to Offer
dbaplus Community
dbaplus Community
Aug 23, 2026 · R&D Management

Why a “Big Mud Ball” Can Be a Good Architecture – Insights from a Former Google/AWS Architect

In this interview, former Google and AWS senior architect Gregor Hohpe explains that architects should amplify their teams rather than be the smartest, emphasizes risk reduction as the core value, advocates simplicity, visual thinking with paper and pen, warns against over‑reliance on AI, and shows how even a “big mud ball” can be a pragmatic architectural choice.

AI pitfallsSystem Designleadership
0 likes · 25 min read
Why a “Big Mud Ball” Can Be a Good Architecture – Insights from a Former Google/AWS Architect
samdeepthink
samdeepthink
Aug 23, 2026 · Databases

Is Splitting Data Across Machines Distributed? Understand Sharding vs Distributed Systems

The article explains that merely placing data on multiple machines constitutes sharding—a way to split data for capacity—but true distributed systems require coordinated nodes that communicate, replicate, and handle failures, illustrated with an e‑commerce warehouse analogy and guidance on choosing between sharding and distributed databases.

System Designdatabase architecturedistributed systems
0 likes · 8 min read
Is Splitting Data Across Machines Distributed? Understand Sharding vs Distributed Systems
samdeepthink
samdeepthink
Aug 21, 2026 · R&D Management

Why Architects Who Haven’t Coded in Years Miss Critical System Realities

The article explains how architects who stop writing or reading code lose essential insight into the system, leading them to suggest impractical solutions like a non‑existent message center, and offers concrete practices to stay connected with the codebase.

System Designarchitecturecode review
0 likes · 8 min read
Why Architects Who Haven’t Coded in Years Miss Critical System Realities
Data Bricklaying Diary
Data Bricklaying Diary
Aug 11, 2026 · Artificial Intelligence

Three-Layer Ontology Intelligence: Separating Semantics, Decisions, and Actions

This article explains why ontology intelligence systems require a three-layer architecture—semantic layer for defining business meaning, decision layer for forming explainable action plans, and action layer for controlled state changes—with explicit handoff contracts to avoid mixing rules and side effects into prompts.

AI architectureDecision LayerSemantic Layer
0 likes · 20 min read
Three-Layer Ontology Intelligence: Separating Semantics, Decisions, and Actions
Code Farming
Code Farming
Aug 9, 2026 · Backend Development

How a System Handles 30 Million Simultaneous Video Views

The article breaks down how QuickTok supports 30 million concurrent video streams by calculating QPS, storage and bandwidth needs, then applying HDFS with HBase indexing and aggressive CDN pre‑warming to shrink traffic from 88 Tbps to under 4 Tbps.

CDNHDFSSystem Design
0 likes · 5 min read
How a System Handles 30 Million Simultaneous Video Views
samdeepthink
samdeepthink
Aug 7, 2026 · Frontend Development

Why Backend Engineers Mistake Frontend for Simple—and What That Reveals

The article explains that backend engineers often view frontend work as easy because it changes rapidly with user preferences, while backend deals with stable, long‑standing challenges like data consistency, making each side complex in fundamentally different ways.

BackendComplexitySystem Design
0 likes · 4 min read
Why Backend Engineers Mistake Frontend for Simple—and What That Reveals
Code Farming
Code Farming
Aug 5, 2026 · Backend Development

How to Generate Billions of Conflict‑Free Short URLs

The article breaks down a real‑world architecture for a short‑URL service that must handle 12 billion entries and 40 k QPS, showing how to calculate capacity, compare generation algorithms, use Bloom filters for offline de‑duplication, and employ a three‑layer cache‑plus‑storage design to meet performance goals.

Bloom FilterHBaseRedis
0 likes · 7 min read
How to Generate Billions of Conflict‑Free Short URLs
AI Engineer Programming
AI Engineer Programming
Aug 4, 2026 · Artificial Intelligence

Why Agents Call Unneeded Tools and How to Tackle It as a System‑Engineering Problem

The article defines tool hallucination in LLM agents, analyses training bias, context pollution, loop feedback and dialogue inertia as root causes, and proposes multi‑layer defenses—including visibility control, intent verification, runtime gating, architectural isolation, and feedback loops—framed as a system‑engineering challenge rather than mere prompt tweaking.

AgentLLMRuntime Guard
0 likes · 17 min read
Why Agents Call Unneeded Tools and How to Tackle It as a System‑Engineering Problem
Code Farming
Code Farming
Jul 30, 2026 · Backend Development

Can Your System Survive a Sudden 500K Live Viewers? 4 Proven Traffic‑Splitting Techniques

When a live broadcast suddenly attracts 500,000 viewers, the system faces 100,000 QPS likes and rapid reward transactions; this article breaks down a battle‑tested traffic‑splitting architecture—room‑based sharding, queue buffering, multi‑layer write caching, and consistent‑hash sharding—showing how each component controls load, ensures consistency, and enables seamless scaling.

System Designconsistent hashingdistributed cache
0 likes · 6 min read
Can Your System Survive a Sudden 500K Live Viewers? 4 Proven Traffic‑Splitting Techniques
CTO Full-Stack Academy
CTO Full-Stack Academy
Jul 28, 2026 · Backend Development

Mastering Batch Processing: Core Principles, Pitfalls, and Real‑World Cases

This article explains the fundamentals of batch processing, outlines its non‑real‑time nature, lists typical scenarios and pros‑cons, provides practical implementation guidelines, and presents three detailed industry case studies illustrating how large‑scale offline jobs are built and maintained.

BackendBatch ProcessingSystem Design
0 likes · 11 min read
Mastering Batch Processing: Core Principles, Pitfalls, and Real‑World Cases
samdeepthink
samdeepthink
Jul 25, 2026 · Backend Development

Do You Need Strong Coding Skills to Be an Effective Architect?

The author argues that, in most cases, a software architect must possess solid coding abilities because architecture decisions are validated through implementation, and without hands‑on coding the design often fails to meet real‑world constraints.

Backend DevelopmentMicroservicesSystem Design
0 likes · 4 min read
Do You Need Strong Coding Skills to Be an Effective Architect?
samdeepthink
samdeepthink
Jul 24, 2026 · Industry Insights

How Should Companies Hire Programmers in the AI Era?

The article argues that as AI takes over most business code, hiring will shift from pure coding tests to evaluating system‑design skills and the ability to collaborate with AI, emphasizing decision‑making over rote implementation.

AIAI collaborationSystem Design
0 likes · 5 min read
How Should Companies Hire Programmers in the AI Era?
Liangxu Linux
Liangxu Linux
Jul 22, 2026 · Interview Experience

Top Embedded Engineer Interview Questions Every Interviewer Should Ask

This article outlines the essential embedded‑engineer interview questions—covering pointers, memory management, interrupts, peripheral protocols, debugging strategies, and system design—while explaining the reasoning behind each question and what answers reveal about a candidate's depth of knowledge.

C languageMemory ManagementSystem Design
0 likes · 8 min read
Top Embedded Engineer Interview Questions Every Interviewer Should Ask
samdeepthink
samdeepthink
Jul 22, 2026 · Backend Development

Why You Should Minimize Local Cache Usage

The article argues that local caches add significant consistency and management complexity, so they should be avoided unless a genuine performance bottleneck exists, illustrating the point with real‑world promotion spikes, GC concerns, and careful off‑heap testing.

CachingGCPerformance
0 likes · 4 min read
Why You Should Minimize Local Cache Usage
Wu Shixiong's Large Model Academy
Wu Shixiong's Large Model Academy
Jul 21, 2026 · Artificial Intelligence

How to Decompose a Production‑Ready RAG System for Interview Success

The article outlines a production‑ready RAG architecture by separating offline ingestion and online query pipelines, detailing nine ingestion steps, online request flow, data storage responsibilities, failure‑handling, monitoring, and acceptance criteria, all illustrated with concrete examples and traceable state machines.

RAGSystem Designfailure handling
0 likes · 29 min read
How to Decompose a Production‑Ready RAG System for Interview Success
Code Farming
Code Farming
Jul 20, 2026 · Backend Development

How a Message Queue Keeps Flash‑Sale Systems Stable Under 10k Orders per Second

The article explains how using a message queue as a buffer, asynchronous processor, and decoupling layer enables flash‑sale systems to handle tens of thousands of orders per second, reducing database overload, cutting response time from 500 ms to 50 ms, and preventing cascade failures.

System Designasynchronous-processingdecoupling
0 likes · 5 min read
How a Message Queue Keeps Flash‑Sale Systems Stable Under 10k Orders per Second
Nightwalker Tech
Nightwalker Tech
Jul 20, 2026 · Artificial Intelligence

Designing Reliable AI Agents: From a Single Prompt to Stable Delivery

The article explains why treating complex AI agents as a single long prompt leads to instability, and proposes a reusable closed‑loop architecture—scheduler, planner, executor, evaluator, repairer, and finalizer—that makes agents explainable, recoverable, and safely deliverable in production.

AI AgentClosed-loop ArchitectureSystem Design
0 likes · 21 min read
Designing Reliable AI Agents: From a Single Prompt to Stable Delivery
Ray's Galactic Tech
Ray's Galactic Tech
Jul 19, 2026 · Artificial Intelligence

Why the Real Production Bottleneck for AI Agents Is the Harness, Not the Model

The article explains that when AI agents move from prototype to production, failures usually stem from the execution harness—issues like multi‑step orchestration, tool integration, context overflow, and lack of observability—rather than the underlying language model itself, and it provides a concrete seven‑layer framework (ETCLOVG) to diagnose and engineer a reliable harness.

AI agentsETCLOVGHarness Engineering
0 likes · 32 min read
Why the Real Production Bottleneck for AI Agents Is the Harness, Not the Model
Data Bricklaying Diary
Data Bricklaying Diary
Jul 19, 2026 · R&D Management

AI Programming Era: Design Judgment Is the Real Scarcity, Not Code

The article argues that AI lowers code generation cost but amplifies the need for clear requirements, architecture, and verification; it advocates Spec-Driven Development (SDD) with controlled agents and test feedback loops, and shares a practical workflow separating design, implementation, and testing roles to ensure systems are correctly designed, implemented, and validated.

AI programmingHarness EngineeringLoop Engineering
0 likes · 20 min read
AI Programming Era: Design Judgment Is the Real Scarcity, Not Code
Random Bulletin
Random Bulletin
Jul 17, 2026 · Backend Development

Why Moving from Real‑Time to Batch Is Essential for Scaling to Tens of Millions QPS

Scaling a service from millions to tens of millions of queries per second fails not because of data size but due to per‑request fixed costs, and the article shows how batching aggregates these costs, dramatically boosts throughput, reduces latency, and introduces new challenges such as memory pressure and partial failures.

Batch ProcessingSystem Designhigh QPS
0 likes · 17 min read
Why Moving from Real‑Time to Batch Is Essential for Scaling to Tens of Millions QPS
Infinite Tech Management
Infinite Tech Management
Jul 15, 2026 · R&D Management

Practical Guide to Drawing Architecture Diagrams: Using an AI Customer Service System as an Example

This article walks through a step‑by‑step method for creating a technical architecture diagram of an AI‑powered customer service system, covering how to identify the core business flow, map states and data ownership, decide sync/async boundaries, place modules, and annotate risks and support capabilities.

AI chatbotArchitecture DiagramSystem Design
0 likes · 18 min read
Practical Guide to Drawing Architecture Diagrams: Using an AI Customer Service System as an Example
YiSu Grain
YiSu Grain
Jul 15, 2026 · R&D Management

From Tech Terms to Real Architecture Design: Day22‑Day42 Learning Roadmap

After completing the first 21 days of foundational topics, the second stage (Day22‑Day42) shifts focus to core architecture design, covering stakeholders, requirements, trade‑offs, evaluation methods, and a detailed daily syllabus that guides learners from theory to practical system design.

Learning RoadmapSystem Designarchitecture evaluation
0 likes · 12 min read
From Tech Terms to Real Architecture Design: Day22‑Day42 Learning Roadmap
Code Farming
Code Farming
Jul 13, 2026 · Backend Development

Mastering Traffic Control: Keeping Systems Stable Under a Million Concurrent Requests

The article explains how to prevent system crashes during massive traffic spikes by applying three core techniques—rate limiting, circuit breaking, and graceful degradation—detailing algorithm choices, state machines, and practical implementation steps for high‑concurrency back‑end services.

BackendCircuit BreakerSystem Design
0 likes · 6 min read
Mastering Traffic Control: Keeping Systems Stable Under a Million Concurrent Requests
CTO Full-Stack Academy
CTO Full-Stack Academy
Jul 11, 2026 · Operations

Designing a Universal OMS: End‑to‑End Order Management Architecture

This article presents a comprehensive, industry‑agnostic design for a universal Order Management System (OMS), detailing its four core goals, complete order lifecycle flow, layered architecture with four‑level functional modules, implementation steps, and practical solutions to common deployment challenges.

OMSSystem Designe-commerce
0 likes · 22 min read
Designing a Universal OMS: End‑to‑End Order Management Architecture
samdeepthink
samdeepthink
Jul 10, 2026 · R&D Management

Why I Proactively Gave a Young Engineer a Raise and Promotion

Over two years I promoted a junior developer to mid‑level, raised his salary, and increased his bonus after he consistently delivered clean DDD‑styled code, handled simple tasks flawlessly, proactively addressed third‑party integration challenges, and demonstrated the ability to take on greater responsibility, illustrating the true criteria leaders use for advancement.

DDDSystem Designcareer growth
0 likes · 8 min read
Why I Proactively Gave a Young Engineer a Raise and Promotion
IT Learning Made Simple
IT Learning Made Simple
Jul 9, 2026 · Fundamentals

Master System Permissions and Boundary Architecture by Analyzing a Popular Family Drama

The article uses the plot of the viral family series ‘Researcher’s Daughter‑in‑Law Fixes the Toxic Relatives’ to illustrate core IT concepts such as three‑tier B/S architecture, the principle of least privilege, permission‑fuse mechanisms, and network boundary isolation, helping beginners grasp these topics without code.

IT educationPermission ManagementSystem Design
0 likes · 7 min read
Master System Permissions and Boundary Architecture by Analyzing a Popular Family Drama
AI Illustrated Series
AI Illustrated Series
Jul 8, 2026 · Interview Experience

How to Design a Multi‑Agent Collaborative Office Assistant System

The article outlines a practical interview‑style design for a multi‑agent office assistant, detailing role specialization, task allocation flow, three communication patterns, conflict‑resolution strategies, and concrete usage scenarios such as meeting scheduling, weekly reporting, and sales data analysis.

Multi-agentSystem Designcommunication patterns
0 likes · 5 min read
How to Design a Multi‑Agent Collaborative Office Assistant System
Linyb Geek Road
Linyb Geek Road
Jul 3, 2026 · Artificial Intelligence

Production-Ready AI Agent Harness: Architecture and Design Principles

The article explains why the stability of AI agents depends on the harness rather than the model, outlines a five‑layer production‑grade harness architecture (Environment, Tool, Control, Memory, Evaluation), and presents five engineering principles to build a reliable, observable, and maintainable agent runtime system.

AI AgentHarness EngineeringMemory Management
0 likes · 18 min read
Production-Ready AI Agent Harness: Architecture and Design Principles
Subtle Storm
Subtle Storm
Jul 2, 2026 · R&D Management

How to Gracefully Transition from Programmer to Architect

The article explains that moving from programmer to architect requires shifting focus from isolated code elegance to system‑wide thinking, mastering trade‑offs, learning from production incidents, broadening technical breadth, and communicating designs in business terms, offering a step‑by‑step roadmap for the transition.

System Designcareer developmentcommunication
0 likes · 6 min read
How to Gracefully Transition from Programmer to Architect
IT Learning Made Simple
IT Learning Made Simple
Jul 2, 2026 · Operations

Process View: The Heartbeat of System Runtime

The article explains the process view, which reveals how a system operates at runtime, covering processes, threads, inter‑process communication, concurrency models, synchronization mechanisms, performance indicators, and design principles, illustrated with diagrams and a concrete e‑commerce case study.

IPCPerformanceSystem Design
0 likes · 9 min read
Process View: The Heartbeat of System Runtime
Thought Artisan
Thought Artisan
Jun 23, 2026 · Fundamentals

First Principles Design: How Top Engineers Rethink System Architecture

This article compiles first-principles design philosophies from ten software engineering leaders, including LLVM creator Chris Lattner and Perfetto expert Lalit Maganti, showing how to break problems to fundamental constraints rather than copying existing solutions.

Chris LattnerDomain-Driven DesignLLVM
0 likes · 15 min read
First Principles Design: How Top Engineers Rethink System Architecture
Software Engineering 3.0 Era
Software Engineering 3.0 Era
Jun 21, 2026 · Fundamentals

What Is the First Principle of Software Engineering and Why It Matters

The article explains that software engineering’s recurring problems stem from three inherent contradictions—state‑space explosion versus human cognition, inevitable iteration versus entropy, and collective production versus information loss—and presents a four‑layer failure model and a concrete first‑principle framework to guide sustainable system design, even in the AI era.

AI toolsSystem Designcomplexity management
0 likes · 14 min read
What Is the First Principle of Software Engineering and Why It Matters
IT Learning Made Simple
IT Learning Made Simple
Jun 21, 2026 · Fundamentals

Step-by-Step Guide to Creating Your First Architecture Diagram

This tutorial walks you through why beginners struggle with blank canvases, how to define the diagram’s purpose, gather system details, use common shapes and connectors, and build a complete e‑commerce architecture diagram in Draw.io while avoiding common pitfalls.

Architecture DiagramSystem Designbest practices
0 likes · 7 min read
Step-by-Step Guide to Creating Your First Architecture Diagram
Hacker Afternoon Tea
Hacker Afternoon Tea
Jun 20, 2026 · Artificial Intelligence

How Multica Turns 2 People and 10 Agents into the Output of a 20‑Person Team

Multica treats AI agents as colleagues rather than tools, running them on users' machines while the server only stores data, queues tasks, and broadcasts events, and uses a four‑layer frontend, dual data streams, and a set of core modules to let a small team manage dozens of agents with real‑time visibility and measurable productivity.

AI agentsAgent OrchestrationGo backend
0 likes · 12 min read
How Multica Turns 2 People and 10 Agents into the Output of a 20‑Person Team
JavaGuide
JavaGuide
Jun 18, 2026 · Artificial Intelligence

From AI Coding to Full‑Stack AI Apps: Master Claude, Codex, Agents, and Skills

AIGuide is a free, open‑source handbook that walks Java, Go, frontend, testing, and architecture professionals through the entire AI application development lifecycle—from LLM fundamentals and RAG to agents, system design, and practical AI‑assisted coding—providing real‑world scenarios, key parameters, pitfalls, and interview preparation.

AI Application DevelopmentAI agentsLLM
0 likes · 14 min read
From AI Coding to Full‑Stack AI Apps: Master Claude, Codex, Agents, and Skills
Subtle Storm
Subtle Storm
Jun 15, 2026 · Backend Development

Caching, Rate Limiting, Smoothing, and Idempotency: Solving Concurrency Problems

The article breaks down how caching reduces repeated slow‑resource access, rate limiting protects systems from overload, smoothing (peak shaving) buffers burst traffic with queues, and idempotency prevents duplicate operations, using a milk‑tea shop analogy to illustrate each technique’s role in high‑concurrency environments.

BackendCachingPeak Shaving
0 likes · 7 min read
Caching, Rate Limiting, Smoothing, and Idempotency: Solving Concurrency Problems
Chen Tian Universe
Chen Tian Universe
Jun 15, 2026 · Operations

Ready‑to‑Use Settlement System Design Cases: 7 Real‑World Examples

This article presents a reusable 10‑step framework for designing settlement systems and showcases seven detailed real‑world case studies—including enterprise welfare, government services, promotion platforms, bank acquiring, points e‑commerce, highway ETC, and consumer finance—illustrating data sources, settlement flows, rule configuration, document design, and integration points.

System Designcase studyfintech
0 likes · 35 min read
Ready‑to‑Use Settlement System Design Cases: 7 Real‑World Examples
IT Learning Made Simple
IT Learning Made Simple
Jun 8, 2026 · R&D Management

The Essential Gear to Become a Software Architect

This guide maps the complete skill tree for aspiring software architects, detailing foundational knowledge, core competencies such as system design and performance tuning, extended expertise in cloud‑native and big‑data technologies, and a staged learning roadmap to help newcomers acquire the necessary gear.

Performance OptimizationSystem Designbig data
0 likes · 9 min read
The Essential Gear to Become a Software Architect
AI Engineer Programming
AI Engineer Programming
Jun 8, 2026 · Artificial Intelligence

When to Use Small Models: A System Design Perspective

Small models are chosen based on deployment constraints rather than absolute parameter counts; the article outlines how resource limits, latency, cost, privacy, and task characteristics define their suitability, compares their strengths and weaknesses to large models, and offers system‑level design patterns for effective use.

Inference OptimizationLLM deploymentRAG
0 likes · 20 min read
When to Use Small Models: A System Design Perspective
IT Learning Made Simple
IT Learning Made Simple
Jun 7, 2026 · Industry Insights

A Day in the Life of an Architect: Meetings, Diagrams, and Taking the Blame

The article walks through a typical software architect’s day—from early‑morning stand‑ups and requirement reviews, through writing design documents, ad‑hoc incident triage, architecture review meetings, code reviews, and technical‑debt cleanup—highlighting time allocation, required skills, pain points, and practical advice for aspiring architects.

System Designarchitecture-designcareer advice
0 likes · 9 min read
A Day in the Life of an Architect: Meetings, Diagrams, and Taking the Blame
Architectural Methodology
Architectural Methodology
Jun 5, 2026 · Interview Experience

How to Ace Architecture Interview Questions with Structured Thinking

The article explains that unlike developer interviews, architecture interviews focus on structured thinking, problem decomposition, and communication, and provides a step‑by‑step framework—break the problem, layer the design, weigh trade‑offs, and use a STAR‑plus‑technical format—along with concrete examples and common pitfalls to avoid.

DDDSTAR methodSystem Design
0 likes · 7 min read
How to Ace Architecture Interview Questions with Structured Thinking
IT Learning Made Simple
IT Learning Made Simple
Jun 2, 2026 · R&D Management

What Exactly Does a Software Architect Do? The Role That Involves More PPTs Than Coding

The article demystifies the software architect role by outlining core duties such as system design, technology selection, solving technical challenges, cross‑team coordination, a typical daily schedule, and how it differs from senior developers, while emphasizing that architects are not omnipotent but facilitators.

System Designrole comparisonsoftware architecture
0 likes · 8 min read
What Exactly Does a Software Architect Do? The Role That Involves More PPTs Than Coding
Subtle Storm
Subtle Storm
May 31, 2026 · Artificial Intelligence

Essential AI Knowledge Every Top Architect Must Master

The article outlines the AI topics that modern architects need to master—including fundamentals, weak and narrow AI, generative models, large language models, Transformers, prompt engineering, multimodal concepts, intelligent agents, end‑to‑end system design, MLOps, distributed high‑performance computing, and technology‑cost trade‑offs—highlighting why AI expertise is now a core requirement for architectural roles.

AIDistributed ComputingMLOps
0 likes · 5 min read
Essential AI Knowledge Every Top Architect Must Master
Linyb Geek Road
Linyb Geek Road
May 31, 2026 · Artificial Intelligence

From Prompt to Harness: The Three Evolutions of AI Engineering

The article traces AI engineering's three-stage evolution—from single‑turn Prompt Engineering, through multi‑turn Context Engineering, to system‑level Harness Engineering—explaining the problems each stage solves, the techniques introduced, concrete examples, and why the shift matters for scalable, reliable AI agents.

AI EngineeringAgentContext Engineering
0 likes · 11 min read
From Prompt to Harness: The Three Evolutions of AI Engineering
Linyb Geek Road
Linyb Geek Road
May 29, 2026 · Artificial Intelligence

A Panoramic Look at Harness Engineering: The Engineering Paradigm for Production‑Grade AI Agents

The article explains why Harness Engineering is needed, defines its core concepts, details a five‑layer architecture with concrete mechanisms, outlines design principles and practical steps for building stable, observable AI agents, and discusses future opportunities and limitations.

AI EngineeringAI agentsHarness Engineering
0 likes · 13 min read
A Panoramic Look at Harness Engineering: The Engineering Paradigm for Production‑Grade AI Agents
Subtle Storm
Subtle Storm
May 26, 2026 · Cloud Native

Structuring a High-Concurrency System Design Paper for the 2026 Soft Exam

The article outlines a step‑by‑step framework for writing a high‑concurrency system design paper, covering project background, performance challenges, six concrete technical solutions—including multi‑level caching, async processing, rate limiting, database optimization, microservice decomposition, and elastic scaling—and how to quantify their impact with real data.

CachingKubernetesMicroservices
0 likes · 6 min read
Structuring a High-Concurrency System Design Paper for the 2026 Soft Exam
Su San Talks Tech
Su San Talks Tech
May 19, 2026 · Interview Experience

Designing a Hundred‑Billion‑Scale Message Queue: A ByteDance Interview Walkthrough

This article walks through the interview question of designing a message queue that handles billions of messages daily and peaks at millions of QPS, covering traffic calculations, core roles, storage and throughput techniques, scalability, high availability, observability, framework comparisons, a real‑world case study, and key follow‑up interview topics.

KafkaPulsarRocketMQ
0 likes · 12 min read
Designing a Hundred‑Billion‑Scale Message Queue: A ByteDance Interview Walkthrough
Infinite Tech Management
Infinite Tech Management
May 17, 2026 · Fundamentals

Tech Lead’s Guide: How to Create Effective System Architecture Diagrams

This article explains why a proper technical architecture diagram is essential, defines its purpose, outlines three hierarchical view levels, presents a six‑step method for drawing clear, layered diagrams, highlights common pitfalls, and shows how to keep diagrams up‑to‑date as living documentation for teams and stakeholders.

Architecture DiagramSystem Designlayered architecture
0 likes · 13 min read
Tech Lead’s Guide: How to Create Effective System Architecture Diagrams
AgentGuide
AgentGuide
May 9, 2026 · Artificial Intelligence

Interview Question: What Is Harness Engineering and How to Answer It

The article defines Harness Engineering—also called "驾驭工程"—as a set of engineering methods that create a structured environment for AI agents, addressing issues like missing context, tool access, feedback loops, and security, and contrasts it with prompt engineering while providing concrete implementation steps.

AI AgentAgent EnvironmentHarness Engineering
0 likes · 8 min read
Interview Question: What Is Harness Engineering and How to Answer It
Tinker Programmer
Tinker Programmer
May 3, 2026 · Artificial Intelligence

Why 99% of AI Agents Fail and How to Avoid Common Pitfalls

Most developers mistake model capability for system capability, leading to unstable agents; this article breaks down six essential modules—four‑layer architecture, execution model, memory system, framework choice, multi‑agent design, and observability—to guide engineers toward production‑ready AI agents.

AI agentsObservabilitySystem Design
0 likes · 6 min read
Why 99% of AI Agents Fail and How to Avoid Common Pitfalls
Smart Workplace Lab
Smart Workplace Lab
May 2, 2026 · Industry Insights

Prompt Engineer Layoffs: How to Re‑Engineer Your Career Path

As large language models mature, prompt‑writing roles are disappearing, prompting engineers to shift from crafting prompts to designing end‑to‑end AI workflows; this article outlines a three‑step system‑reconstruction protocol, common pitfalls, and practical guidelines for transitioning into workflow architecture.

AI workflowCareer TransitionLLM
0 likes · 6 min read
Prompt Engineer Layoffs: How to Re‑Engineer Your Career Path
Data Party THU
Data Party THU
Apr 29, 2026 · Artificial Intelligence

Claude Opus 4.7 System Prompt Leak: Decoding Its 10 Core Design Decisions

The article dissects the leaked Claude Opus 4.7 system prompt, revealing ten intertwined design decisions—from treating psychological reconstruction as a danger signal to dynamic safety‑policy upgrades—that together shape the model’s self‑restraint, tool‑use, memory handling, and risk‑aware behavior.

AI safetyClaudeSystem Design
0 likes · 8 min read
Claude Opus 4.7 System Prompt Leak: Decoding Its 10 Core Design Decisions
Infinite Tech Management
Infinite Tech Management
Apr 28, 2026 · R&D Management

From Business Boxes to Microservices: A Technical Leader’s Guide to Application Architecture Diagrams

The article explains how to translate business architecture boxes into concrete micro‑service designs, outlines a four‑step mapping method, defines service granularity, contracts, cross‑cutting concerns, diagram notation, and presents two diagram styles with real‑world pitfalls and best‑practice recommendations.

Architecture DiagramMicroservicesService design
0 likes · 18 min read
From Business Boxes to Microservices: A Technical Leader’s Guide to Application Architecture Diagrams
ITPUB
ITPUB
Apr 25, 2026 · Interview Experience

How to Design a Billion‑Scale URL Shortening System for an Interview

This article walks through the complete interview‑style design of a billion‑scale URL shortener, covering requirements, capacity estimation, API definitions, database schema, short‑code generation algorithms, sharding, caching, load balancing, rate limiting, and expiration handling, while illustrating each step with concrete examples and calculations.

API DesignCachingSystem Design
0 likes · 24 min read
How to Design a Billion‑Scale URL Shortening System for an Interview
Infinite Tech Management
Infinite Tech Management
Apr 24, 2026 · R&D Management

From Pig Pens to Skyscrapers: How Programmers Become Architects

The article explains why coding skill alone isn’t enough, outlines the four essential architectural mindsets, shows how neglecting business, people, technology, and operational complexity turns systems into spaghetti, and offers concrete career advice for engineers at every level.

System Designarchitectural thinkingcareer development
0 likes · 10 min read
From Pig Pens to Skyscrapers: How Programmers Become Architects
ZhiKe AI
ZhiKe AI
Apr 22, 2026 · Artificial Intelligence

Why Harness Engineering Is the Hottest AI Engineering Paradigm in 2026

The article explains how the emerging "Harness Engineering" paradigm—highlighted by OpenAI, Stripe and Anthropic—shifts AI development from prompt tweaking to building full control systems, promising ten‑fold efficiency gains, new architectural components, and both opportunities and risks for developers.

AI EngineeringHarness EngineeringSystem Design
0 likes · 9 min read
Why Harness Engineering Is the Hottest AI Engineering Paradigm in 2026
FunTester
FunTester
Apr 20, 2026 · Artificial Intelligence

Why Self‑Evaluating Agents Fail and How to Build Reliable Multi‑Agent Systems

The article analyzes why letting the same AI Agent generate and self‑evaluate results in over‑confident but flawed outputs, especially for subjective tasks, and proposes a three‑stage multi‑agent architecture with independent evaluation, concrete standards, and prompt‑based calibration to improve reliability as models evolve.

AIMulti-agentSystem Design
0 likes · 9 min read
Why Self‑Evaluating Agents Fail and How to Build Reliable Multi‑Agent Systems
DeepHub IMBA
DeepHub IMBA
Apr 20, 2026 · Artificial Intelligence

What 10 Core Design Decisions the Claude Opus 4.7 Prompt Leak Reveals

The leaked Claude Opus 4.7 system prompt exposes ten intertwined design choices—ranging from treating psychological reconstruction as a danger signal to prohibiting over‑politeness, treating tool calls as cost‑free, using natural language as memory cues, and dynamically upgrading safety—illustrating a pattern of self‑regulation rather than pure capability enhancement.

AI safetyBehavioral ConstraintsClaude
0 likes · 8 min read
What 10 Core Design Decisions the Claude Opus 4.7 Prompt Leak Reveals
AI Waka
AI Waka
Apr 20, 2026 · Artificial Intelligence

Why the Hidden ‘Agent Harness’ Beats Bigger Models in AI Performance

The article explains how the often‑overlooked Agent Harness—an orchestration layer surrounding large language models—determines AI agent success, detailing its five core components, real‑world case studies, and why system design now outweighs raw model size.

AI agentsAgent ArchitectureHarness Engineering
0 likes · 17 min read
Why the Hidden ‘Agent Harness’ Beats Bigger Models in AI Performance
Code Mala Tang
Code Mala Tang
Apr 19, 2026 · Artificial Intelligence

Why Real‑World Constraints Define the Success of Claude Code Agents

The analysis of the arXiv paper “Dive into Claude Code” reveals that beyond model loops, the decisive factors for coding agents are practical system design issues such as permission control, context compression, safety, user intervention, and reliable execution in real environments.

AI architectureClaude CodeContext Management
0 likes · 5 min read
Why Real‑World Constraints Define the Success of Claude Code Agents
Architecture Breakthrough
Architecture Breakthrough
Apr 16, 2026 · Backend Development

Mastering Asynchronous Processing: Design Principles, Patterns, and Risks

This comprehensive guide explains the purpose, core concepts, suitable scenarios, common patterns, benefits, and potential pitfalls of asynchronous processing, offering detailed design, development, review, and operational principles to help teams build reliable, high‑throughput systems.

PerformanceSystem Designasynchronous-processing
0 likes · 22 min read
Mastering Asynchronous Processing: Design Principles, Patterns, and Risks
Architect Practice
Architect Practice
Apr 7, 2026 · Backend Development

The 8 Characters in a Text Message Reveal Hidden Challenges in Short‑Link System Design

This article walks through the end‑to‑end design of a lightweight short‑link service, covering why short links are needed, the choice of 302 redirects, key‑generation strategies (MurmurHash vs ID‑based base62), caching layers, security threats, scaling techniques, and practical pitfalls, all illustrated with concrete code and benchmark numbers.

CachingSystem Designbase62 encoding
0 likes · 25 min read
The 8 Characters in a Text Message Reveal Hidden Challenges in Short‑Link System Design
DeepNoMind
DeepNoMind
Apr 5, 2026 · Fundamentals

Stop Buying Courses: 5 GitHub Repos That Outperform a $5,000 Class

The article argues that instead of spending money on video courses, developers can break free from "tutorial hell" by tackling five carefully chosen GitHub repositories that force active coding, system‑level thinking, and deep JavaScript mastery, ultimately building real‑world engineering skills.

GitHubJavaScriptSystem Design
0 likes · 7 min read
Stop Buying Courses: 5 GitHub Repos That Outperform a $5,000 Class
JavaEdge
JavaEdge
Apr 3, 2026 · Artificial Intelligence

Why Harness Engineering Is the Next Frontier for AI Agents

This article analyzes the rise of Harness Engineering for AI agents, contrasting it with Prompt and Context Engineering, detailing how leading companies like Anthropic, OpenAI, Google DeepMind, Windsurf, and Stripe design comprehensive runtime systems, and offering practical steps for teams to build robust agent harnesses.

AI agentsAgent ArchitectureContext Engineering
0 likes · 12 min read
Why Harness Engineering Is the Next Frontier for AI Agents
o-ai.tech
o-ai.tech
Apr 1, 2026 · Artificial Intelligence

How CE Turns Engineering Experience into a Compound, Reusable System

CE proposes that instead of storing experience only in chat logs, an agent system should convert it into consumable, maintainable, refreshable, and discoverable assets, organized into three durable artifact layers—brainstorms, plans, and solutions—so that future tasks become easier, faster, and less error‑prone.

AI agentsCompound EngineeringKnowledge Refresh
0 likes · 19 min read
How CE Turns Engineering Experience into a Compound, Reusable System
o-ai.tech
o-ai.tech
Mar 31, 2026 · Artificial Intelligence

CE System Design: From Workflow to AI Agent Engineering

The article examines the Compound‑Engineering (CE) system, showing how it structures complex engineering tasks into layered workflows, specialist agents, and reusable documentation, contrasting its systematic approach with the Superpowers framework and offering concrete insights for building robust AI agent pipelines.

AI agentsAgent OrchestrationCompound Engineering
0 likes · 12 min read
CE System Design: From Workflow to AI Agent Engineering
LuTiao Programming
LuTiao Programming
Mar 25, 2026 · Backend Development

From Confusion to Mastery: A Structured Path for System Design Skills

The article explains why many developers get stuck when moving from writing business code to system design, outlines a step‑by‑step engineering learning path that covers core components, hands‑on examples, trade‑off analysis, interview preparation, and communication techniques to build a holistic system‑design mindset.

MicroservicesSystem Designbackend architecture
0 likes · 7 min read
From Confusion to Mastery: A Structured Path for System Design Skills