Tagged articles

distributed systems

2274 articles · Page 21 of 23
21CTO
21CTO
Sep 19, 2017 · Fundamentals

Why Git Dominates Version Control: From Linus’s DIY System to Modern Workflows

This article traces the history of Linux's source‑code management, explains why Linus Torvalds created Git as a distributed version‑control system, compares it with centralized systems like CVS and SVN, and highlights Git’s key advantages and ecosystem.

Gitdistributed systemsopen source
0 likes · 10 min read
Why Git Dominates Version Control: From Linus’s DIY System to Modern Workflows
Java High-Performance Architecture
Java High-Performance Architecture
Sep 18, 2017 · Big Data

Can Kafka Safely Serve as Long‑Term Storage? Answers and Real‑World Scenarios

This article explains why Kafka can be used for permanent data retention, outlines practical use cases such as event logging, cache rebuilding, stream recomputation, and database change capture, and clarifies why Kafka is a stream‑processing platform rather than a traditional message queue or database.

Long‑term Storagedata retentiondistributed systems
0 likes · 6 min read
Can Kafka Safely Serve as Long‑Term Storage? Answers and Real‑World Scenarios
WeChat Backend Team
WeChat Backend Team
Sep 12, 2017 · Backend Development

How PhxQueue Achieves High‑Throughput, High‑Reliability Distributed Queuing with Paxos

PhxQueue, an open‑source, Paxos‑based distributed queue from WeChat, delivers at‑least‑once delivery, synchronous disk flushing, strict ordering, multi‑subscription, and high availability, outperforming Kafka in reliability and latency while maintaining comparable throughput, as demonstrated through detailed design, performance, and failover analyses.

KafkaPaxosPerformance Comparison
0 likes · 26 min read
How PhxQueue Achieves High‑Throughput, High‑Reliability Distributed Queuing with Paxos
Architecture Digest
Architecture Digest
Aug 30, 2017 · Backend Development

Evolution of System Architecture and Key Distributed Service Technologies

This article outlines the progressive stages of system architecture—from a single‑server LAMP setup through application‑data separation, caching, clustering, read/write splitting, CDN, distributed storage, NoSQL, business decomposition, and finally distributed services—while detailing essential technologies such as message queues, service frameworks, service buses, communication patterns, and governance mechanisms like Dubbo and OSB.

Message Queuebackend developmentdistributed systems
0 likes · 14 min read
Evolution of System Architecture and Key Distributed Service Technologies
ITFLY8 Architecture Home
ITFLY8 Architecture Home
Aug 27, 2017 · Backend Development

Scaling Web Systems to 100M Daily Visits: Load Balancing and Caching

This article explains how a web system can evolve from handling 100,000 daily visits to over 100 million by progressively implementing multi‑level caching, various load‑balancing techniques—including HTTP redirects, reverse‑proxy, IP‑level, DNS, and GSLB—and optimizing MySQL through indexing, connection pooling, sharding, replication, and integrating memory caches such as Redis to ensure high performance and reliability.

MySQLRedisWeb Scaling
0 likes · 21 min read
Scaling Web Systems to 100M Daily Visits: Load Balancing and Caching
ITFLY8 Architecture Home
ITFLY8 Architecture Home
Aug 22, 2017 · Backend Development

Mastering Flash Sale Systems: High-Concurrency Strategies and Real-World Solutions

This article explores the core concepts, design patterns, and practical techniques—including caching, distributed coordination, front‑end scaling, and back‑end queuing—to build robust high‑concurrency flash‑sale systems while preventing oversell and performance bottlenecks.

High Concurrencybackend developmentdistributed systems
0 likes · 8 min read
Mastering Flash Sale Systems: High-Concurrency Strategies and Real-World Solutions
Architecture Digest
Architecture Digest
Aug 22, 2017 · Backend Development

Strengthening Backend Fundamentals: Distributed Service Architecture and Personal Growth

In this talk, chief architect Li Yanpeng shares his career background, outlines the goals and design principles of distributed service architecture—including high availability, performance, scalability, extensibility, security, and consistency—and offers a methodology for cultivating both technical and personal inner skills for engineers.

BackendEngineering Methodologydistributed systems
0 likes · 22 min read
Strengthening Backend Fundamentals: Distributed Service Architecture and Personal Growth
ITFLY8 Architecture Home
ITFLY8 Architecture Home
Aug 21, 2017 · Backend Development

Designing a Scalable Insurance O2O Platform: Architecture, Services, and Storage

This article presents a comprehensive design for a high‑traffic insurance O2O platform, detailing functional requirements, system analysis, storage and caching strategies, logical and service architectures, distributed transaction handling, and the chosen development stack, emphasizing simplicity, scalability, and high concurrency.

O2Odistributed systemsinsurance
0 likes · 20 min read
Designing a Scalable Insurance O2O Platform: Architecture, Services, and Storage
ITFLY8 Architecture Home
ITFLY8 Architecture Home
Aug 17, 2017 · Backend Development

Mastering High-Concurrency Flash Sale Systems: Architecture, Challenges, and Solutions

This article dissects the technical challenges of building a high‑concurrency flash‑sale (seckill) system—covering business analysis, traffic isolation, static page caching, CDN bandwidth, dynamic order URLs, request throttling, database sharding, optimistic locking, and anti‑cheat mechanisms—while presenting concrete architectural principles and code examples.

High ConcurrencyOptimistic Lockdistributed systems
0 likes · 36 min read
Mastering High-Concurrency Flash Sale Systems: Architecture, Challenges, and Solutions
Qunar Tech Salon
Qunar Tech Salon
Aug 17, 2017 · Databases

Evolution of Meituan-Dianping MySQL High‑Availability Architecture: From MMM to MHA+Zebra and Beyond

This article reviews the evolution of Meituan‑Dianping's MySQL high‑availability architecture over recent years, detailing the transition from the MMM replication manager to MHA, the integration of Zebra and Proxy middleware, and future design considerations such as distributed agents, semi‑synchronous replication, and MySQL Group Replication.

Database ReplicationMHAMySQL
0 likes · 11 min read
Evolution of Meituan-Dianping MySQL High‑Availability Architecture: From MMM to MHA+Zebra and Beyond
ITFLY8 Architecture Home
ITFLY8 Architecture Home
Aug 6, 2017 · Backend Development

How Meizu Scales Real‑Time Push to 600 M Messages/min: Architecture, Pitfalls & Solutions

The article details Meizu's massive real‑time push system handling 25 million online users and 600 million messages per minute, explains its four‑layer architecture, and shares how the team tackled phone power consumption, mobile network instability, massive connections, monitoring, and gray‑release deployment.

High Concurrencydistributed systemsgray release
0 likes · 13 min read
How Meizu Scales Real‑Time Push to 600 M Messages/min: Architecture, Pitfalls & Solutions
Architecture Digest
Architecture Digest
Aug 4, 2017 · Backend Development

Common Architectural Patterns for Large-Scale Websites

The article outlines essential website architecture patterns—layered design, separation, distribution, clustering, caching, asynchronous processing, redundancy, automation, and security—explaining how each contributes to high concurrency, scalability, reliability, and maintainability of large web applications.

Cachingdistributed systemslayered architecture
0 likes · 7 min read
Common Architectural Patterns for Large-Scale Websites
WeChat Backend Team
WeChat Backend Team
Aug 3, 2017 · Databases

How PhxSQL Achieves Strong Consistency and High Availability for MySQL

This article explains the design and implementation of PhxSQL, a MySQL‑compatible high‑availability solution that uses a reliable log storage based on Paxos, Proxy request forwarding, automatic master election, and other mechanisms to overcome native MySQL replication flaws and provide strong data consistency and fault‑tolerant performance.

Database ReplicationMySQLPaxos
0 likes · 17 min read
How PhxSQL Achieves Strong Consistency and High Availability for MySQL
Architecture Digest
Architecture Digest
Aug 3, 2017 · Backend Development

Evolution of Large-Scale Website Architecture: From Single Server to Distributed Services

The article outlines the progressive architectural stages of large‑scale websites—starting with a single‑server setup and advancing through service separation, caching, load balancing, database read/write splitting, CDN/reverse proxy, distributed storage, NoSQL, business splitting, and distributed services—to illustrate how high concurrency, massive traffic, high availability, and massive data are handled.

Cachingdistributed systemsload balancing
0 likes · 6 min read
Evolution of Large-Scale Website Architecture: From Single Server to Distributed Services
Tongcheng Travel Technology Center
Tongcheng Travel Technology Center
Aug 1, 2017 · Blockchain

Blockchain Overview and Merkle Tree Algorithm for Data Verification

This article introduces the origin and key alliances of blockchain, explains its core characteristics such as decentralization, consensus mechanisms and transaction transparency, and provides a detailed description of the Merkle Tree algorithm, its implementation, verification steps, and applications in distributed file synchronization.

ConsensusMerkle treeblockchain
0 likes · 9 min read
Blockchain Overview and Merkle Tree Algorithm for Data Verification
dbaplus Community
dbaplus Community
Jul 26, 2017 · Databases

How Ele.me Achieved Sub‑Second MySQL Multi‑Active Replication with DRC

This article details Ele.me's design and implementation of a MySQL bidirectional replication component (DRC) that enables sub‑second, high‑throughput data synchronization across Beijing and Shanghai data centers, addressing latency, consistency, and failover challenges in a multi‑active environment.

Multi-ActiveMySQLdata replication
0 likes · 18 min read
How Ele.me Achieved Sub‑Second MySQL Multi‑Active Replication with DRC
ITFLY8 Architecture Home
ITFLY8 Architecture Home
Jul 26, 2017 · Big Data

Inside Taobao’s Massive Data Architecture: From Hadoop “Cloud Ladder” to Real‑Time “Galaxy”

This article details Taobao’s multi‑layer massive data platform, covering its five‑tier architecture, the 1500‑node Hadoop “Cloud Ladder” for batch processing, the low‑latency “Galaxy” stream engine, MySQL‑based MyFOX, HBase‑based Prom storage, the glider middle‑layer, and sophisticated caching strategies that together support petabytes of data and millions of daily queries.

CachingHBaseHadoop
0 likes · 16 min read
Inside Taobao’s Massive Data Architecture: From Hadoop “Cloud Ladder” to Real‑Time “Galaxy”
Architecture Digest
Architecture Digest
Jul 25, 2017 · Fundamentals

Understanding Distributed System Consistency: CAP Theorem, ACID, Distributed Transactions, and 2PC/3PC Protocols

This article explains the core concepts of distributed system consistency—including the CAP theorem, ACID properties, various distributed transaction techniques, and the two‑phase and three‑phase commit protocols—while illustrating practical implementations with message queues and local message tables.

2PC3PCCAP theorem
0 likes · 13 min read
Understanding Distributed System Consistency: CAP Theorem, ACID, Distributed Transactions, and 2PC/3PC Protocols
Architecture Digest
Architecture Digest
Jul 22, 2017 · Big Data

Popular Big Data Tools and Their Descriptions

This article provides an extensive overview of more than ninety open‑source and commercial big‑data tools—including ETL platforms, resource managers, storage systems, messaging queues, processing engines, and visualization libraries—detailing their core functions, typical use cases, and notable adopters.

AnalyticsETLbig data
0 likes · 26 min read
Popular Big Data Tools and Their Descriptions
Architecture Digest
Architecture Digest
Jul 18, 2017 · Backend Development

Design and Implementation of Ctrip Real‑Time User Data Collection System

This article describes the design, technology selection, and performance evaluation of Ctrip's real‑time user behavior data collection platform, covering Netty‑based network handling, Kafka/Hermes messaging, encryption, compression, Avro backup, and related analytics products, with detailed feasibility analysis and benchmark results.

KafkaNettybackend architecture
0 likes · 17 min read
Design and Implementation of Ctrip Real‑Time User Data Collection System
ITFLY8 Architecture Home
ITFLY8 Architecture Home
Jul 17, 2017 · Big Data

Mastering Data Sync, Real‑Time Processing, and Scalable Storage for Modern Systems

This article explores practical techniques for synchronizing heterogeneous data sources, performing batch and incremental analytics with Hadoop and Spark, designing low‑latency real‑time computation pipelines, implementing push notifications, and choosing appropriate storage solutions—from in‑memory caches to distributed databases—while addressing performance, reliability, and scalability challenges.

big datadata synchronizationdatabases
0 likes · 25 min read
Mastering Data Sync, Real‑Time Processing, and Scalable Storage for Modern Systems
Architecture Digest
Architecture Digest
Jul 16, 2017 · Operations

Fault Governance in Distributed Systems: Dependency Failures, Strong/Weak Dependency, and Fault‑Injection Practices

This article presents a comprehensive overview of fault governance in large‑scale distributed systems, covering classic dependency failures, the concept of strong and weak dependencies, experimental observations, the evolution of fault‑injection techniques, and best practices for building reliable fault‑drill platforms.

chaos engineeringdependency managementdistributed systems
0 likes · 20 min read
Fault Governance in Distributed Systems: Dependency Failures, Strong/Weak Dependency, and Fault‑Injection Practices
ITFLY8 Architecture Home
ITFLY8 Architecture Home
Jul 16, 2017 · Operations

How Top E‑Commerce Platforms Engineer Scalable, High‑Performance Architecture

This article consolidates e‑commerce platform architecture practices, covering design principles, multi‑level caching, indexing strategies, parallel and distributed computing, high availability, scaling techniques, resource optimization, static blueprint, component analysis, and supporting middleware such as load balancers, routers, HA, messaging, caching, buffering, search, and log collection.

Cachingarchitecturedistributed systems
0 likes · 24 min read
How Top E‑Commerce Platforms Engineer Scalable, High‑Performance Architecture
21CTO
21CTO
Jul 12, 2017 · Backend Development

Designing Scalable Game Server Architecture: From Login to Distributed Systems

This article outlines the principles and components of a scalable game server architecture, covering login management, region selection, gateway handling, communication protocols, publish‑subscribe, RPC, server merging, and the evolution from single‑threaded to cloud‑native designs.

architecturebackend developmentdistributed systems
0 likes · 24 min read
Designing Scalable Game Server Architecture: From Login to Distributed Systems
Architecture Digest
Architecture Digest
Jul 6, 2017 · Fundamentals

PacificA: Microsoft’s General Replication Framework for Large‑Scale Distributed Storage Systems

PacificA is Microsoft’s generic replication framework for large‑scale distributed storage systems that provides strong consistency, separates configuration management from data replication, and uses a primary‑secondary model with lease‑based fault detection to ensure availability, correctness, and efficient operation across heterogeneous nodes.

ConsistencyPacificAReplication
0 likes · 14 min read
PacificA: Microsoft’s General Replication Framework for Large‑Scale Distributed Storage Systems
Meituan Technology Team
Meituan Technology Team
Jun 30, 2017 · Operations

How Meituan‑Dianping Evolved MySQL HA from MMM to MHA‑Zebra and Beyond

This article traces Meituan‑Dianping's MySQL high‑availability journey from the early MMM replication manager to the modern MHA‑Zebra and MHA‑Proxy solutions, compares each architecture, highlights their shortcomings, and outlines future directions such as distributed agents, semi‑sync replication, and Paxos‑based MySQL Group Replication.

MHAMMMMySQL
0 likes · 12 min read
How Meituan‑Dianping Evolved MySQL HA from MMM to MHA‑Zebra and Beyond
Architecture Digest
Architecture Digest
Jun 24, 2017 · Backend Development

Unit Architecture Practice: Design, Benefits, and Implementation at Weibo

This article explains why architecture practice is essential, introduces the concept of unit (cell) architecture, discusses its performance and cost advantages, and details how Weibo applied it to its fan service platform, including handling of partitioning and job management.

architecturedistributed systemsservice design
0 likes · 11 min read
Unit Architecture Practice: Design, Benefits, and Implementation at Weibo
Meituan Technology Team
Meituan Technology Team
Jun 23, 2017 · Backend Development

Fault Drill: Traffic Replication and Fault Injection Platform for Hotel Backend

The Fault‑Drill platform for hotel back‑end services combines real‑time traffic replication to shadow clusters with UI‑driven fault injection via a java‑agent, enabling developers to validate incident‑response plans, measure latency impacts, and reduce MTTR by testing normal and abnormal conditions on live traffic.

Backend Engineeringdistributed systemsfault injection
0 likes · 13 min read
Fault Drill: Traffic Replication and Fault Injection Platform for Hotel Backend
Architecture Digest
Architecture Digest
Jun 23, 2017 · Backend Development

Evolution of a Startup's Backend Architecture: From Single Server to Distributed System

After leaving Baidu in 2015, the author recounts nearly two years of evolving a startup’s backend architecture—from a single‑server Tomcat setup to multi‑node clusters with Nginx reverse proxy, database sharding, caching, session sharing, and service‑oriented design—highlighting challenges, optimizations, and lessons learned.

BackendNginxTomcat
0 likes · 5 min read
Evolution of a Startup's Backend Architecture: From Single Server to Distributed System
Architecture Digest
Architecture Digest
Jun 8, 2017 · Fundamentals

Fundamental Principles of System Design, Complexity, and Performance

The article discusses how expert programmers must continuously learn and balance humility with confidence, explores the end‑to‑end design principle, examines complexity, layering, and componentization in large systems, and highlights performance considerations, distributed‑system realities, and management lessons for building robust software.

Complexitydistributed systemsmanagement
0 likes · 20 min read
Fundamental Principles of System Design, Complexity, and Performance
ITFLY8 Architecture Home
ITFLY8 Architecture Home
Jun 4, 2017 · Operations

How eBay Scales to Billions: 7 Proven Practices for Massive Web Systems

This article outlines eBay's seven scalability best practices—including functional partitioning, horizontal sharding, avoiding distributed transactions, asynchronous decoupling, streaming, virtualization, and smart caching—to help large‑scale web services achieve reliable, cost‑effective growth.

Asynchronous ArchitectureCachingdistributed systems
0 likes · 14 min read
How eBay Scales to Billions: 7 Proven Practices for Massive Web Systems
ITFLY8 Architecture Home
ITFLY8 Architecture Home
May 30, 2017 · Fundamentals

CAP Theory, Shared‑Nothing, Load Balancing & High Availability Explained

This article explores core distributed system design principles, detailing the CAP theorem and its implications, the BASE extension, shared‑nothing architecture, various load‑balancing algorithms and deployment modes, as well as high‑availability strategies such as active‑standby, active‑active, and clustering to eliminate single points of failure.

CAP theoremdistributed systemshigh availability
0 likes · 18 min read
CAP Theory, Shared‑Nothing, Load Balancing & High Availability Explained
ITPUB
ITPUB
May 27, 2017 · Databases

Building a High‑Performance OAuth Token Service with Tarantool, Raft, and Sharding

This article explains how we designed and implemented a scalable OAuth token storage and refresh system using Tarantool’s in‑memory database, Raft leader election, sharding across multiple data centers, and a custom lightweight queue to handle high‑throughput token updates while maintaining consistency and fault tolerance.

OAuthRaftTarantool
0 likes · 20 min read
Building a High‑Performance OAuth Token Service with Tarantool, Raft, and Sharding
Alibaba Cloud Developer
Alibaba Cloud Developer
May 24, 2017 · Databases

Rethinking Future Database Architecture: Insights from Alibaba’s Lead Engineer

In this comprehensive talk, Alibaba’s database chief Zhang Rui shares the challenges of high‑availability, cost, and elasticity in massive transaction systems, outlines innovations like AliSQL X‑Cluster, Paxos‑based consistency, X‑KV, dual‑engine storage, and discusses the evolving role of DBAs toward automation and intelligent optimization.

AliSQLDBA automationPaxos
0 likes · 23 min read
Rethinking Future Database Architecture: Insights from Alibaba’s Lead Engineer
21CTO
21CTO
May 23, 2017 · Backend Development

How to Build a High‑Concurrency, High‑Availability E‑Commerce Platform

This article outlines the design principles and architectural strategies for constructing a high‑concurrency, high‑availability e‑commerce platform, covering space‑time tradeoffs, caching layers, indexing techniques, parallel and distributed computing, load balancing, stateless services, resource optimization, fault tolerance, data storage options, and real‑time processing components.

CachingHigh Concurrencydatabase design
0 likes · 45 min read
How to Build a High‑Concurrency, High‑Availability E‑Commerce Platform
ITPUB
ITPUB
May 19, 2017 · Backend Development

How Message Queues Turn Slow Email Sends into Fast, Reliable Services

This article tells the story of a developer who first used a blocking email call, then switched to multithreading, and finally adopted a message queue to achieve asynchronous processing, decoupling, scalability, and reliability for user registration emails.

Java MultithreadingMessage Queueasynchronous-processing
0 likes · 6 min read
How Message Queues Turn Slow Email Sends into Fast, Reliable Services
MaGe Linux Operations
MaGe Linux Operations
May 16, 2017 · Operations

How Distributed Clusters Achieve Load Balancing: Principles and Practices

This article explains the concepts of distributed clusters and load balancing, contrasting clusters and distributed systems with real‑world analogies, describing various load‑balancing techniques such as DNS, LVS, and reverse proxies, and offers practical guidance on designing simple, reliable, and efficient load‑balancing solutions for distributed back‑ends.

clustersdistributed systemsload balancing
0 likes · 11 min read
How Distributed Clusters Achieve Load Balancing: Principles and Practices
Meituan Technology Team
Meituan Technology Team
May 12, 2017 · Cloud Native

Design and Implementation of the HULK Container Platform Scheduling System

The HULK Container Platform scheduling system, built for Meituan‑Dianping, combines a hybrid, actor‑based scheduler with filter‑and‑rank logic, configurable trade‑offs, and dynamic over‑commit to balance resource utilization, high availability, and massive concurrent placement decisions for thousands of containerized services.

Cloud-nativeDockerHULK
0 likes · 17 min read
Design and Implementation of the HULK Container Platform Scheduling System
Efficient Ops
Efficient Ops
May 7, 2017 · Fundamentals

How Paxos Powers Zookeeper: A Simple Island Analogy Explained

This article uses a vivid island metaphor to break down the Paxos consensus algorithm, maps its concepts to Zookeeper components such as servers, leaders, and Zxid, and illustrates conflict resolution and leader election through clear examples and diagrams.

ConsensusLeader ElectionPaxos
0 likes · 7 min read
How Paxos Powers Zookeeper: A Simple Island Analogy Explained
MaGe Linux Operations
MaGe Linux Operations
May 5, 2017 · Backend Development

How Meituan Scaled Its Food‑Delivery Order System to Millions of Daily Orders

This article chronicles the evolution of Meituan's food‑delivery order system from a simple modular prototype to a distributed, high‑performance, highly available architecture, detailing the business characteristics, architectural milestones, performance optimizations, consistency safeguards, scalability techniques, and intelligent operations that enable handling millions of orders per day.

distributed systemshigh availabilityorder-system
0 likes · 20 min read
How Meituan Scaled Its Food‑Delivery Order System to Millions of Daily Orders
Qunar Tech Salon
Qunar Tech Salon
May 5, 2017 · Backend Development

WeChat MQ 2.0: Enhanced Asynchronous Queue Design and Optimizations

The article introduces WeChat's self‑developed MQ 2.0 asynchronous queue, detailing its architecture, cross‑machine consumption model, improved task scheduling, efficient processing frameworks—including a MapReduce‑style engine and streaming tasks—and robust overload protection mechanisms that together boost reliability and performance for large‑scale backend services.

MapReduceMessage QueueStreaming Tasks
0 likes · 12 min read
WeChat MQ 2.0: Enhanced Asynchronous Queue Design and Optimizations
21CTO
21CTO
Apr 28, 2017 · Backend Development

How Small Websites Grow into Scalable Giants: A Step‑by‑Step Architecture Guide

This article walks through the evolution of a website from a single‑server setup to a distributed, high‑performance architecture, covering service separation, caching strategies, server clustering, load balancing, database replication, CDN acceleration, distributed storage, NoSQL adoption, and modular business decomposition.

BackendCachingDatabase Replication
0 likes · 8 min read
How Small Websites Grow into Scalable Giants: A Step‑by‑Step Architecture Guide
Architects' Tech Alliance
Architects' Tech Alliance
Apr 27, 2017 · Big Data

Curated List of Big Data Learning Resources from w3cschool

This article presents a comprehensive, Chinese‑language collection of big‑data resources—including relational databases, distributed file systems, key‑value stores, distributed programming tools, file data models, and key‑map frameworks—compiled by w3cschool to help programmers deepen their understanding of big data technologies.

LearningResourcesbig data
0 likes · 6 min read
Curated List of Big Data Learning Resources from w3cschool
Meituan Technology Team
Meituan Technology Team
Apr 21, 2017 · Backend Development

Design and Implementation of Meituan's Distributed ID Generation System Leaf

Meituan’s Leaf system merges segment‑based caching with Snowflake‑style bit fields to deliver globally unique, trend‑increasing 64‑bit IDs at ultra‑low latency, using double‑buffered DB segments, master‑slave MySQL replication, Zookeeper‑assigned worker IDs, and clock‑rollback safeguards, achieving ~50 k QPS and 1 ms 99.9th‑percentile response across billions of daily IDs.

BackendLeafSnowflake
0 likes · 18 min read
Design and Implementation of Meituan's Distributed ID Generation System Leaf
21CTO
21CTO
Apr 16, 2017 · Operations

Which Load‑Balancing Strategy Guarantees the Highest Reliability?

This article explains common load‑balancing strategies—round‑robin, random, minimum response time, minimum concurrency, and hash—detailing their principles, advantages, drawbacks, and mathematical reliability analysis, including probability formulas and visual illustrations to help choose the most fault‑tolerant approach for distributed systems.

distributed systemsload balancingminimum concurrency
0 likes · 9 min read
Which Load‑Balancing Strategy Guarantees the Highest Reliability?
Architecture Digest
Architecture Digest
Apr 16, 2017 · Operations

Common Load‑Balancing Strategies and Their Reliability Analysis in Distributed Systems

The article reviews hardware and software load‑balancing, explains classic strategies such as round‑robin, random, minimum‑response‑time, least‑connections and hash, and quantitatively evaluates their fault‑tolerance using probability formulas and example scenarios in distributed systems.

Least Connectionsdistributed systemsfault tolerance
0 likes · 10 min read
Common Load‑Balancing Strategies and Their Reliability Analysis in Distributed Systems
21CTO
21CTO
Apr 11, 2017 · Backend Development

Simulating 10 Billion Red‑Packet Requests on One Server: Achieving 60k QPS with Go

This article details how to design, implement, and benchmark a single‑machine backend capable of handling up to 1 million concurrent connections and 60 000 queries per second while simulating the shake‑and‑send red‑packet workflow of a large‑scale messaging app, including capacity calculations, architecture choices, Go‑based implementation, and multi‑stage performance testing.

BackendHigh ConcurrencyQPS
0 likes · 18 min read
Simulating 10 Billion Red‑Packet Requests on One Server: Achieving 60k QPS with Go
Architecture Digest
Architecture Digest
Apr 11, 2017 · Backend Development

Design and Practice of a High‑Throughput Spring Festival Red‑Packet System Supporting One Million Connections

This article describes how to design, implement, and evaluate a backend system that simulates the Spring Festival red‑packet service, achieving up to 6 × 10⁴ QPS on a single server while handling one million concurrent connections, and discusses the hardware, software, architecture, and performance results.

BackendQPSdistributed systems
0 likes · 18 min read
Design and Practice of a High‑Throughput Spring Festival Red‑Packet System Supporting One Million Connections
ITFLY8 Architecture Home
ITFLY8 Architecture Home
Apr 11, 2017 · Databases

Why High Availability Triggers a Consistency‑Performance Trade‑off in Distributed Databases

The article explains how achieving high availability through data redundancy introduces consistency challenges that in turn affect performance, and it reviews partitioning, mirroring, consistency models, replication architectures, and two/three‑phase commit protocols in distributed systems.

ReplicationTransaction Processingdata consistency
0 likes · 18 min read
Why High Availability Triggers a Consistency‑Performance Trade‑off in Distributed Databases
ITFLY8 Architecture Home
ITFLY8 Architecture Home
Apr 9, 2017 · Backend Development

Mastering Distributed Locks and Idempotency for High‑Concurrency Systems

This article explores the challenges of mutual exclusion and idempotency in distributed environments, explains the underlying principles of locks in multi‑threaded and multi‑process contexts, and presents practical implementations using Zookeeper, Redis, Tair, and the Cerberus and GTIS frameworks to ensure reliable, scalable operations.

IdempotencyRedisZooKeeper
0 likes · 35 min read
Mastering Distributed Locks and Idempotency for High‑Concurrency Systems
Architecture Digest
Architecture Digest
Apr 6, 2017 · Fundamentals

Distributed Service System Consistency: Best Practices and Patterns

This article examines the challenges of achieving consistency in large‑scale distributed service systems, outlines common inconsistency scenarios such as split‑brain and lost updates, and presents practical patterns—including ACID/BASE trade‑offs, two‑phase and three‑phase commit, TCC, query, compensation, and reliable messaging—to guide engineers in designing robust, eventually consistent architectures.

ACIDBASEConsistency
0 likes · 35 min read
Distributed Service System Consistency: Best Practices and Patterns
Architecture Digest
Architecture Digest
Apr 1, 2017 · Backend Development

Distributed Consistency and Transactional Messaging Solutions

This article explains the challenges of achieving consistency in distributed systems and presents practical solutions such as two‑phase commit, asynchronous assurance, compensating transactions, message retry mechanisms, idempotent designs, and a custom Redis‑based delayed queue (DelayQ) with a transactional proxy (TMQProxy) to provide reliable transactional messaging.

IdempotencyMessage RetryRedis delay queue
0 likes · 19 min read
Distributed Consistency and Transactional Messaging Solutions
21CTO
21CTO
Mar 31, 2017 · Backend Development

How to Build Highly Available and Scalable Distributed Systems

This article explains the key challenges of high availability and scalability in distributed architectures and provides practical solutions for each layer—entry, business, cache, and database—using techniques such as heartbeat IPs, stateless services, consistent hashing, and sharding.

backend architecturedistributed systemsload balancing
0 likes · 18 min read
How to Build Highly Available and Scalable Distributed Systems
Efficient Ops
Efficient Ops
Mar 26, 2017 · Operations

How Google Scales App Engine: Lessons in Cloud Scalability and SRE

The article shares Google SRE veteran Minghua Ye’s insights on App Engine’s evolution, emphasizing the critical role of automatic scalability, distributed locks, service discovery, load balancing, and open‑source tools like gRPC, Protobuf, gflags, glog, and Googletest in building reliable, high‑traffic cloud services.

Google App EngineProtobufSRE
0 likes · 12 min read
How Google Scales App Engine: Lessons in Cloud Scalability and SRE
ITPUB
ITPUB
Mar 22, 2017 · Backend Development

What Makes Taobao’s Massive Scale Demand Hundreds of Elite Engineers?

The article explains how a high‑traffic e‑commerce platform like Taobao relies on distributed storage, search engines, massive caching, load‑balancing, CDN, sophisticated advertising and analytics systems, all of which require large teams of top engineers to design, implement, and operate.

BackendCachingdistributed systems
0 likes · 12 min read
What Makes Taobao’s Massive Scale Demand Hundreds of Elite Engineers?
Alibaba Cloud Infrastructure
Alibaba Cloud Infrastructure
Mar 15, 2017 · Operations

Alibaba IDC and Network Monitoring System Architecture and Practices

The article details Alibaba's globally distributed IDC and network monitoring systems, describing their fully distributed data collection, centralized computation, storage strategies, alarm mechanisms, and frontend visualization that together enable real‑time infrastructure and network health management for large‑scale operations.

IDCInfrastructuredistributed systems
0 likes · 13 min read
Alibaba IDC and Network Monitoring System Architecture and Practices
Architecture Digest
Architecture Digest
Mar 9, 2017 · Backend Development

Common Design Issues and Best Practices for Distributed System Interfaces

The article outlines key challenges in distributed API design—including date formatting, decimal precision, response structures, idempotency, security, and naming consistency—and provides practical recommendations to improve usability, scalability, and maintainability across backend services.

API DesignBackendIdempotency
0 likes · 12 min read
Common Design Issues and Best Practices for Distributed System Interfaces
21CTO
21CTO
Feb 21, 2017 · Backend Development

How WeChat and Alibaba Handle Billions of Red Packets: High‑Concurrency Architecture Secrets

This article examines the high‑availability architectures behind massive online transaction systems such as Alibaba's Double 11 sales, Alipay and WeChat red packets, detailing the challenges of billions of requests and the engineering solutions that ensure performance, reliability, and security.

High ConcurrencyWeChatdistributed systems
0 likes · 20 min read
How WeChat and Alibaba Handle Billions of Red Packets: High‑Concurrency Architecture Secrets
ITFLY8 Architecture Home
ITFLY8 Architecture Home
Feb 16, 2017 · Backend Development

How VIPshop Evolved from Monolithic LAMP to Distributed Service Architecture

This article examines VIPshop's transformation from a single‑application LAMP system to a vertically split and finally distributed service‑oriented architecture, detailing the business model, key design requirements, platform governance, and the technical services that enable a scalable e‑commerce operation.

Cloud Computingdistributed systemse-commerce
0 likes · 13 min read
How VIPshop Evolved from Monolithic LAMP to Distributed Service Architecture
ITPUB
ITPUB
Feb 10, 2017 · Backend Development

How to Generate Globally Unique IDs in Distributed Systems: Snowflake and Its Variants

This article explains the challenges of generating globally unique IDs across distributed shards, outlines the requirements for such IDs, and details Twitter's Snowflake algorithm—including its structure, generation process, and clock handling—before exploring three notable Snowflake variants and their trade‑offs.

BackendSnowflakeUnique ID
0 likes · 10 min read
How to Generate Globally Unique IDs in Distributed Systems: Snowflake and Its Variants
Architects' Tech Alliance
Architects' Tech Alliance
Feb 10, 2017 · Industry Insights

Inside Scality Ring: How Its Scale‑Out Architecture Powers Object, File, and Block Storage

The article provides a detailed technical overview of Scality Ring 6.0, explaining its three‑layer software stack, X86‑based scale‑out hardware, diverse connectors, storage node design, management tools, routing protocol, data durability, multi‑site deployment models, and consistency guarantees.

Data durabilityMulti-site DeploymentObject Storage
0 likes · 13 min read
Inside Scality Ring: How Its Scale‑Out Architecture Powers Object, File, and Block Storage
ITFLY8 Architecture Home
ITFLY8 Architecture Home
Feb 7, 2017 · Operations

Master System Architecture: CAP Theory, Shared‑Nothing, Load Balancing & HA

This article explores core system architecture concepts—including the CAP theorem and its BASE extension, the shared‑nothing design, various load‑balancing algorithms and deployment modes, and high‑availability patterns such as active‑standby, active‑active and clustering—providing practical guidance for building scalable, reliable distributed applications.

CAP theoremdistributed systemshigh availability
0 likes · 22 min read
Master System Architecture: CAP Theory, Shared‑Nothing, Load Balancing & HA
Alibaba Cloud Developer
Alibaba Cloud Developer
Feb 4, 2017 · Cloud Computing

Inside Alibaba’s Feitian Middleware: Powering Massive E‑Commerce and Cloud Innovation

Alibaba’s transformation from e‑commerce leader to tech powerhouse is highlighted by its Feitian middleware platform, a cloud‑based, highly available solution that supports diverse industries, enables massive transaction volumes, and exemplifies the evolution of large‑scale distributed architectures pioneered since Alibaba’s early IOE days.

Alibaba Clouddistributed systemsenterprise architecture
0 likes · 7 min read
Inside Alibaba’s Feitian Middleware: Powering Massive E‑Commerce and Cloud Innovation
dbaplus Community
dbaplus Community
Jan 15, 2017 · Databases

How JD’s JIMDB Achieves Zero‑Downtime Scaling and Automatic Failover for Massive Caches

JIMDB is JD’s in‑house distributed cache platform that combines automatic fault detection, seamless online scaling, multi‑language support, and containerized deployment to replace traditional Memcached/Redis solutions, offering features such as one‑click cluster creation, elastic expansion, lossless scaling, and comprehensive monitoring for high‑traffic e‑commerce services.

Elastic Scalingcachedistributed systems
0 likes · 23 min read
How JD’s JIMDB Achieves Zero‑Downtime Scaling and Automatic Failover for Massive Caches
Architecture Digest
Architecture Digest
Jan 12, 2017 · Backend Development

Evolution of Internet Technical Architecture: From Single‑Server to Distributed Microservices

This article traces the evolution of internet‑scale technical architecture across three eras—single‑machine, cluster, and distributed—detailing the motivations, core patterns, advantages, and drawbacks of monolithic, layered, data‑separated, cached, load‑balanced, CDN‑accelerated, redundant, service‑oriented, sharded, and microservice designs.

Backenddistributed systemsmicroservices
0 likes · 12 min read
Evolution of Internet Technical Architecture: From Single‑Server to Distributed Microservices
21CTO
21CTO
Jan 8, 2017 · Backend Development

Unlocking High‑Availability: A Sneak Peek at the New Internet Architecture Series

The author announces a forthcoming series on Internet high‑availability architecture, outlining topics such as CAP theory, distributed caching, SOA, message queues, search systems, and real‑world case studies, and invites readers to suggest additional content while promising detailed, valuable guidance for developers and architects.

Cachingdistributed systemsmessage queues
0 likes · 3 min read
Unlocking High‑Availability: A Sneak Peek at the New Internet Architecture Series
360 Zhihui Cloud Developer
360 Zhihui Cloud Developer
Jan 6, 2017 · Operations

How Qcmd Revolutionizes Automated Operations for 7,000+ Servers

Qcmd, the command execution system behind 360’s private HULK cloud platform, replaces SaltStack with an asynchronous, Golang‑based architecture that ensures high‑availability, encrypted messaging, and reliable mass‑host command execution across thousands of servers, dramatically reducing task timeouts and operational overhead.

Cloud OperationsCommand Executiondistributed systems
0 likes · 10 min read
How Qcmd Revolutionizes Automated Operations for 7,000+ Servers
21CTO
21CTO
Jan 4, 2017 · Operations

How to Build Truly High‑Availability Systems: Principles and Practices

This article explains what high availability means for distributed systems, outlines common availability tiers, and describes how redundancy, load balancing, and automatic failover across a typical Internet architecture can achieve reliable, scalable services.

distributed systemsoperationsreliability
0 likes · 6 min read
How to Build Truly High‑Availability Systems: Principles and Practices
Architecture Digest
Architecture Digest
Dec 30, 2016 · Operations

Zero‑Point Battle: Evolution of Alibaba's Double 11 High‑Availability Architecture

The talk details how Alibaba tackled the massive technical challenges of Double 11 over eight years by evolving a highly available, scalable architecture through capacity planning, distributed middleware, hybrid‑cloud deployment, online stress testing, and fine‑grained traffic control to balance cost, performance, and user experience.

AlibabaDouble 11capacity planning
0 likes · 22 min read
Zero‑Point Battle: Evolution of Alibaba's Double 11 High‑Availability Architecture
Qunar Tech Salon
Qunar Tech Salon
Dec 30, 2016 · Operations

Mesos Architecture and Its Practical Use at Qunar: Framework Unification and Operational Insights

This article explains the Mesos distributed system kernel, its resource‑allocation workflow, and how Qunar engineers applied and evolved Mesos, Marathon, and custom frameworks to achieve fine‑grained scheduling, high availability, service discovery, and multi‑tenant management in a large‑scale production environment.

MarathonMesosResource Allocation
0 likes · 14 min read
Mesos Architecture and Its Practical Use at Qunar: Framework Unification and Operational Insights
Architects' Tech Alliance
Architects' Tech Alliance
Dec 23, 2016 · Fundamentals

Advanced Distributed Systems Theory: Paxos, Raft, and Zab

This article provides an in‑depth exploration of distributed consensus protocols, detailing the basics of Paxos, extending to Multi‑Paxos, and comparing it with Raft and Zab while discussing leader election, quorum, lease mechanisms, and practical considerations for implementing these algorithms in real‑world systems.

ConsensusLeader ElectionPaxos
0 likes · 21 min read
Advanced Distributed Systems Theory: Paxos, Raft, and Zab
Architects' Tech Alliance
Architects' Tech Alliance
Dec 22, 2016 · Fundamentals

Fundamentals of Distributed Systems: Consensus, 2PC/3PC, CAP Theorem, and Logical Clocks

This article introduces core distributed‑system concepts—including the definition of consensus, the two‑phase and three‑phase commit protocols, the CAP theorem and its engineering implications, and logical‑clock mechanisms such as Lamport timestamps, vector clocks, and version vectors—explaining their models, challenges, and practical trade‑offs.

CAP theoremConsensusLamport timestamps
0 likes · 21 min read
Fundamentals of Distributed Systems: Consensus, 2PC/3PC, CAP Theorem, and Logical Clocks
ITFLY8 Architecture Home
ITFLY8 Architecture Home
Dec 20, 2016 · Backend Development

Designing Scalable Web Architectures: Key Principles and Practices

This article explains the essential design principles, trade‑offs, and core components—such as availability, performance, reliability, scalability, manageability, and cost—required to build large‑scale, high‑availability web systems and illustrates them with an image‑hosting example.

Cachingdistributed systemsweb architecture
0 likes · 37 min read
Designing Scalable Web Architectures: Key Principles and Practices
ITFLY8 Architecture Home
ITFLY8 Architecture Home
Dec 13, 2016 · Backend Development

How to Build a High‑Performance Flash Sale System: Strategies & Pitfalls

This article outlines the key technical challenges of flash‑sale (秒杀) systems—high concurrency, cache usage, distributed locking, database pressure, and overselling—and presents practical front‑end and back‑end design patterns, including atomic counters, memcached decrements, queueing, Redis off‑loading, and two‑phase commit solutions.

CachingHigh Concurrencydistributed systems
0 likes · 8 min read
How to Build a High‑Performance Flash Sale System: Strategies & Pitfalls
Architecture Digest
Architecture Digest
Dec 9, 2016 · Cloud Native

Deep Dive into Etcd Architecture, Consistency, Storage, Watch Mechanisms, and Comparison with Zookeeper and Consul

This article analyzes Etcd's distributed architecture, Raft‑based consistency, storage implementation, watch and lease mechanisms, differences between v2 and v3, and compares it with Zookeeper and Consul, providing practical usage tips and surrounding tooling for developers of distributed systems.

Consuldistributed systemsetcd
0 likes · 18 min read
Deep Dive into Etcd Architecture, Consistency, Storage, Watch Mechanisms, and Comparison with Zookeeper and Consul
Alibaba Cloud Developer
Alibaba Cloud Developer
Dec 7, 2016 · Big Data

How Alibaba Handled Real‑Time Billions of Events During Double 11

This article outlines Alibaba Cloud's big‑data platform challenges and solutions during the 2016 Double 11 event, covering sub‑second real‑time processing, multi‑million‑records‑per‑second throughput, full‑day high availability, and massive offline workloads exceeding hundreds of petabytes.

AlibabaMaxComputedistributed systems
0 likes · 3 min read
How Alibaba Handled Real‑Time Billions of Events During Double 11
Art of Distributed System Architecture Design
Art of Distributed System Architecture Design
Dec 2, 2016 · Backend Development

Mastering Distributed Transaction Consistency: From CAP to Message‑Based Compensation

This article examines the fundamental challenges of achieving consistency in distributed systems, explains the CAP theorem, compares two‑phase and three‑phase commit protocols, explores XA transactions, and presents practical compensation patterns such as local message tables, non‑transactional and transactional MQ designs, highlighting their trade‑offs and applicability.

CAP theoremMessage Queuedistributed systems
0 likes · 15 min read
Mastering Distributed Transaction Consistency: From CAP to Message‑Based Compensation
Architecture Digest
Architecture Digest
Dec 2, 2016 · Fundamentals

Fundamentals of Distributed Version Control with Git

This article explains the core concepts of distributed version control, compares it with centralized systems, describes repository structures, outlines the advantages of Git, and provides step‑by‑step command examples for initializing, committing, branching, merging, cloning, pulling, and pushing changes in a collaborative development workflow.

collaborationdistributed systemssoftware development
0 likes · 21 min read
Fundamentals of Distributed Version Control with Git
dbaplus Community
dbaplus Community
Nov 29, 2016 · Fundamentals

Essential Distributed System Components: ZooKeeper, Queues, Docker & Logs

Distributed systems rely on coordinated services such as ZooKeeper for state management, message queues like ActiveMQ for inter‑process communication, robust transaction handling, automated deployment tools like Docker, and comprehensive logging solutions, each playing a critical role in achieving high availability, scalability, and operational visibility.

distributed systemslogging
0 likes · 16 min read
Essential Distributed System Components: ZooKeeper, Queues, Docker & Logs
Weidian Tech Team
Weidian Tech Team
Nov 28, 2016 · Big Data

How We Built the Mars Big Data Platform to Boost Development Efficiency

The article explains why Weidian needed a new big data development platform, outlines the functional features of the Mars system, describes its architecture, scheduling mechanisms, task execution flow, and discusses remaining challenges and future enhancements.

HadoopPlatform Architecturedistributed systems
0 likes · 11 min read
How We Built the Mars Big Data Platform to Boost Development Efficiency
Architects' Tech Alliance
Architects' Tech Alliance
Nov 25, 2016 · Databases

Why NoSQL Matters: From ACID to CAP and Beyond

An in‑depth overview of NoSQL databases explains the limitations of traditional relational systems, details ACID properties, introduces the CAP theorem and BASE model, compares RDBMS with NoSQL, outlines advantages, disadvantages, history, classifications, and real‑world usage examples.

ACIDBASECAP theorem
0 likes · 12 min read
Why NoSQL Matters: From ACID to CAP and Beyond
Architecture Digest
Architecture Digest
Nov 23, 2016 · Backend Development

Evolution of .NET Web Architecture: From Single Server to Distributed Cloud Services

The article outlines the step‑by‑step evolution of a .NET‑based web system, describing how a single‑server setup grows into a multi‑tier, load‑balanced, clustered, stateless, micro‑service architecture that leverages caching, NoSQL, search engines, cloud services, Docker and CDN to handle large‑scale traffic and data processing.

BackendCachingcloud
0 likes · 10 min read
Evolution of .NET Web Architecture: From Single Server to Distributed Cloud Services
Ctrip Technology
Ctrip Technology
Nov 22, 2016 · Backend Development

Evolution and Service Decomposition of Qunar's Payment System (1.0 → 2.0)

The article outlines the five‑year evolution of Qunar's payment platform from a tightly coupled monolith (1.0) to a highly available, service‑oriented distributed architecture (2.0), detailing component breakdown, challenges, and the resulting core transaction, payment, cashier, and API layers.

Financial Technologybackend architecturedistributed systems
0 likes · 9 min read
Evolution and Service Decomposition of Qunar's Payment System (1.0 → 2.0)