Tagged articles

low latency

146 articles · Page 2 of 2
Didi Tech
Didi Tech
Dec 21, 2020 · Big Data

HBase Availability and Latency Optimizations: Replication‑Based Multi‑Read and ZGC Adoption

To overcome HBase’s weak availability and GC‑induced latency spikes, the DiDi team introduced a replication‑based client multi‑read (hedged‑read) mechanism and migrated to the Z Garbage Collector, which together dramatically cut maximum and 99.9th‑percentile latencies while keeping services online during region disruptions.

Big DataHBaseMulti-Read
0 likes · 12 min read
HBase Availability and Latency Optimizations: Replication‑Based Multi‑Read and ZGC Adoption
High Availability Architecture
High Availability Architecture
Nov 5, 2020 · Backend Development

Why We Chose Java for Our High‑Frequency Trading Application

The article explains how a high‑frequency trading firm evaluated Java versus C++ for ultra‑low‑latency trading, discusses the challenges of JVM JIT compilation and garbage‑collection pauses, and shows how Azul Zing’s C4 collector delivers near‑C++ latency while preserving Java’s development productivity.

Azul ZingGarbage CollectionJVM
0 likes · 11 min read
Why We Chose Java for Our High‑Frequency Trading Application
Amap Tech
Amap Tech
Oct 30, 2020 · Mobile Development

Video Streaming Solution for the ARC Car Cloud Control Platform

The ARC Car Cloud Control platform now streams the vehicle’s screen using Android’s Virtual Display and a C++‑based H.264 hardware encoder, sending raw video over a TCP socket to a server that adaptively adjusts bitrate and frame rate, while the web client decodes the fragmented MP4 via MSE, dramatically lowering CPU usage and latency on low‑end head‑units.

AndroidH.264Media Source Extensions
0 likes · 8 min read
Video Streaming Solution for the ARC Car Cloud Control Platform
Youku Technology
Youku Technology
Aug 18, 2020 · Backend Development

How Youku Engineered a High‑Performance, Low‑Latency Marketing Platform

This article details Youku's membership marketing system architecture, covering complex marketing scenarios, high‑availability and low‑latency requirements, rule‑based QL engine, unified marketing framework, multi‑cache storage, multithreaded matching, asynchronous reward distribution, and distributed transaction mechanisms.

High AvailabilitySystem Designbackend
0 likes · 12 min read
How Youku Engineered a High‑Performance, Low‑Latency Marketing Platform
High Availability Architecture
High Availability Architecture
Aug 11, 2020 · Operations

Understanding and Optimizing ZGC (Z Garbage Collector) for Low‑Latency Java Services

This article examines the Z Garbage Collector (ZGC) introduced in JDK 11, detailing its low‑pause design goals, underlying concurrent marking‑copy algorithm, colored pointer and read‑barrier techniques, practical tuning parameters, real‑world case studies, and the performance impact of upgrading from CMS/G1 to ZGC in high‑throughput, low‑latency services.

Garbage CollectionJVMZGC
0 likes · 28 min read
Understanding and Optimizing ZGC (Z Garbage Collector) for Low‑Latency Java Services
Meituan Technology Team
Meituan Technology Team
Aug 6, 2020 · Backend Development

ZGC: Principles, Tuning Practices, and Production Upgrade Experience

The article explains how Meituan’s risk‑control platform eliminated frequent 40 ms CMS pauses by adopting JDK 11’s ZGC—detailing its concurrent mark‑copy design, practical tuning parameters, real‑world case fixes, and measured latency reductions of up to 74 % while noting trade‑offs.

Garbage CollectionJDK11ZGC
0 likes · 27 min read
ZGC: Principles, Tuning Practices, and Production Upgrade Experience
Tencent Tech
Tencent Tech
Jun 18, 2020 · Backend Development

Scaling Live‑Ecommerce Platforms: Architecture Behind Billions of Users

This article examines the rapid rise of live‑ecommerce during the 618 shopping festival, explains why the “live + ecommerce” model demands robust backend, streaming and CDN infrastructure, and details Tencent Cloud’s architectural solutions—including media processing, low‑latency protocols, bandwidth optimization and anti‑attack measures—to support massive concurrent traffic.

architecturee‑commercelive streaming
0 likes · 10 min read
Scaling Live‑Ecommerce Platforms: Architecture Behind Billions of Users
Tencent Tech
Tencent Tech
Jun 2, 2020 · Cloud Computing

How SRT Enables Low‑Latency, Reliable Live Streaming for Global eSports Events

The article explains how the Secure Reliable Transport (SRT) protocol, combined with Tencent Video Cloud’s optimized infrastructure, overcame latency and packet‑loss challenges to deliver stable, high‑quality live streams for the 2020 LPL Mid‑Season Cup and other large‑scale events, and describes its broader applications through MediaConnect.

MediaConnectNetwork TransmissionSRT
0 likes · 10 min read
How SRT Enables Low‑Latency, Reliable Live Streaming for Global eSports Events
Programmer DD
Programmer DD
May 22, 2020 · Backend Development

Can ZGC Deliver Sub‑10ms Pauses for Massive Java Heaps?

This article explains the design goals, architecture, key features, tuning options, and version history of Java's Z Garbage Collector (ZGC), highlighting its sub‑10 ms pause times for terabyte‑scale heaps, its use of colored pointers and load barriers, and the trade‑offs in throughput and configuration.

Garbage CollectionJVMZGC
0 likes · 16 min read
Can ZGC Deliver Sub‑10ms Pauses for Massive Java Heaps?
Tencent Cloud Developer
Tencent Cloud Developer
May 21, 2020 · Game Development

How Tencent’s Game Server Engine Tackles Low Latency and Cost in Multiplayer Games

This article analyzes the challenges of low‑latency, stable, and cost‑effective online multiplayer games and explains how Tencent's Game Server Engine (GSE) provides elastic scaling, near‑by scheduling, stateful shrinkage, multi‑region disaster recovery, and zero‑downtime updates to meet those demands.

Elastic ScalingTencent GSEcloud gaming
0 likes · 11 min read
How Tencent’s Game Server Engine Tackles Low Latency and Cost in Multiplayer Games
Big Data Technology Architecture
Big Data Technology Architecture
May 10, 2020 · Big Data

Understanding Apache Hudi: Incremental Processing and Low‑Latency Data Management on Hadoop

This article explains how Apache Hudi provides an incremental processing framework that enables efficient, low‑latency data ingestion, storage, and query capabilities on Hadoop, detailing its architecture, storage layout, compaction, write and read paths, and support for real‑time and batch analytics.

Data IngestionHadoopHudi
0 likes · 15 min read
Understanding Apache Hudi: Incremental Processing and Low‑Latency Data Management on Hadoop
iQIYI Technical Product Team
iQIYI Technical Product Team
Apr 24, 2020 · Cloud Computing

Technical Insights into Cloud Gaming Advertising Trials: Low‑Latency RTCDN Solutions by iQIYI Live Cloud

In an interview, iQIYI Live Cloud’s Chen Kunzhong explains how their WebRTC‑based RTCDN reduces encoding and transmission delay to achieve roughly 100 ms end‑to‑end latency for cloud‑gaming ads, supporting cross‑device play, instant click‑to‑play sessions, and future 5G‑enhanced high‑resolution experiences.

5GAdvertisingRTCDN
0 likes · 9 min read
Technical Insights into Cloud Gaming Advertising Trials: Low‑Latency RTCDN Solutions by iQIYI Live Cloud
Top Architect
Top Architect
Apr 9, 2020 · Backend Development

Low‑Latency and High‑Availability Design of RocketMQ: Evolution, Optimizations, and Capacity Planning

This article reviews the evolution of Alibaba's Aliware message engine, analyzes the low‑latency and high‑availability challenges faced during Double 11, and details the architectural, JVM, memory, rate‑limiting, and multi‑replica solutions that enabled RocketMQ to achieve sub‑millisecond write latency and five‑nine availability.

RocketMQcapacity planningdistributed systems
0 likes · 29 min read
Low‑Latency and High‑Availability Design of RocketMQ: Evolution, Optimizations, and Capacity Planning
dbaplus Community
dbaplus Community
Apr 7, 2020 · Databases

How Pharos Accelerates HBase Multi‑Condition Queries with Low‑Latency Indexing

This article examines Pharos, Everbright Bank's home‑grown HBase indexing middleware, detailing why existing secondary‑index solutions fall short, the design goals of low latency, simple architecture and non‑intrusiveness, and the concrete storage, pagination, and transaction‑consistency techniques that enable fast complex queries on massive data.

HBasePharosdistributed database
0 likes · 14 min read
How Pharos Accelerates HBase Multi‑Condition Queries with Low‑Latency Indexing
Architects' Tech Alliance
Architects' Tech Alliance
Feb 23, 2020 · Cloud Computing

Edge Computing and Its Relationship with 5G: Concepts, Value, Applications, and Future Outlook

This article explains edge computing, its distributed architecture, key advantages such as higher security, lower latency and reduced bandwidth costs, explores major application scenarios like smart manufacturing and autonomous driving, and analyzes how 5G both drives and benefits from edge computing development.

5GDistributed ComputingEdge computing
0 likes · 11 min read
Edge Computing and Its Relationship with 5G: Concepts, Value, Applications, and Future Outlook
Youku Technology
Youku Technology
Feb 19, 2020 · Cloud Native

Low‑Latency Live Streaming System: Challenges, Architecture, and Solutions

Youku’s low‑latency live‑streaming system replaces the traditional RTMP‑CDN pipeline with a controllable media‑transport architecture that combines private real‑time protocols, WebRTC, and edge nodes, cutting anchor‑to‑anchor latency below 300 ms and anchor‑to‑viewer latency to under 415 ms while preserving smooth playback.

CDNReal‑time communicationWebRTC
0 likes · 10 min read
Low‑Latency Live Streaming System: Challenges, Architecture, and Solutions
Alibaba Cloud Developer
Alibaba Cloud Developer
Feb 13, 2020 · Backend Development

How Alibaba Cut Live Stream Latency Below 300ms with a New Architecture

Facing pandemic-driven remote teaching, Alibaba’s live streaming team redesigned their media pipeline, combining CDN, custom real-time protocols, WebRTC, and cloud-native techniques to control transmission and playback buffers, achieving sub-300 ms host-to-host latency and under-600 ms host-to-viewer latency while maintaining smooth playback.

CDNWebRTCbackend-development
0 likes · 10 min read
How Alibaba Cut Live Stream Latency Below 300ms with a New Architecture
Architects' Tech Alliance
Architects' Tech Alliance
Feb 8, 2020 · Cloud Computing

Demystifying FPGA: Architecture, Performance, and Microsoft's Data Center Deployment

FPGA, a reconfigurable hardware architecture, offers low latency and high efficiency compared to CPUs, GPUs, and ASICs, making it ideal for both compute‑intensive and communication‑intensive tasks, and Microsoft’s multi‑stage data‑center deployments illustrate its scalability, flexibility, and impact on cloud services.

Data CenterFPGAHardware Acceleration
0 likes · 21 min read
Demystifying FPGA: Architecture, Performance, and Microsoft's Data Center Deployment
Big Data Technology & Architecture
Big Data Technology & Architecture
Jan 23, 2020 · Big Data

Understanding Apache Hudi: Incremental Processing and Low‑Latency Data Management on Hadoop

This article explains how Apache Hudi enables efficient, low‑latency incremental data ingestion and processing on Hadoop by providing a unified service layer, describing its motivation, architecture, storage components, write and read paths, compaction, fault recovery, and incremental query capabilities.

Apache HudiData IngestionHadoop
0 likes · 17 min read
Understanding Apache Hudi: Incremental Processing and Low‑Latency Data Management on Hadoop
Tencent Cloud Developer
Tencent Cloud Developer
Jul 11, 2019 · Industry Insights

How Real-Time Audio/Video Meets Traditional PSTN: Architecture and Low‑Latency Solutions

This article provides an in‑depth technical analysis of integrating real‑time audio/video (RTC) with legacy PSTN, covering latency sources, protocol and codec differences, adaptation layers, system architecture, and optimization techniques such as jitter buffering, ARQ/FEC, and automatic failover.

PSTN integrationRTCReal‑time communication
0 likes · 17 min read
How Real-Time Audio/Video Meets Traditional PSTN: Architecture and Low‑Latency Solutions
Tencent Cloud Developer
Tencent Cloud Developer
Jul 3, 2019 · Cloud Computing

Technical Overview of Low‑Latency Interactive Live Streaming and MLVBLiveRoom Solutions

The talk detailed Tencent Cloud’s low‑latency interactive live‑streaming techniques—using UDP‑accelerated RTMP, built‑in echo cancellation, cloud‑side video mixing, and scalable room management—to overcome latency, echo, and mixing challenges in multi‑host link‑mic scenarios, illustrated by the MLVBLiveRoom and TRTC large‑room solutions.

RTMP over UDPSDKTRTC
0 likes · 18 min read
Technical Overview of Low‑Latency Interactive Live Streaming and MLVBLiveRoom Solutions
iQIYI Technical Product Team
iQIYI Technical Product Team
May 24, 2019 · Industry Insights

iQIYI’s 8K VR Live Streaming: Cutting Bitrate 75% and Eliminating Motion Latency

The article examines iQIYI’s 8K VR live‑streaming pipeline, detailing how 5G connectivity, tiled encoding, ROI‑focused transmission, and hardware‑accelerated processing reduce bitrate by 75 % and bring motion‑to‑photon latency down to zero, while addressing resolution, bandwidth, and latency challenges of immersive VR broadcasts.

5G8K streamingTiling
0 likes · 9 min read
iQIYI’s 8K VR Live Streaming: Cutting Bitrate 75% and Eliminating Motion Latency
Architects' Tech Alliance
Architects' Tech Alliance
Apr 8, 2019 · Fundamentals

Understanding RDMA: Principles, Advantages, and Implementation Details

This article explains how RDMA (Remote Direct Memory Access) technology, originating from InfiniBand and extended to Ethernet (RoCE) and TCP/IP (iWARP), provides ultra‑low latency, high throughput, and minimal CPU usage for high‑performance computing and big‑data applications by bypassing traditional OS and protocol stack processing.

High-performance networkingRDMARoCE
0 likes · 8 min read
Understanding RDMA: Principles, Advantages, and Implementation Details
ITPUB
ITPUB
Mar 28, 2019 · Big Data

Why Pravega Matters: Native Stream Storage for Low‑Latency, Exactly‑Once Data Pipelines

Pravega, Dell’s native stream storage project, addresses the challenges of modern low‑latency, exactly‑once stream processing by combining tiered storage, Apache BookKeeper, and seamless Flink integration, offering a unified solution that reduces development, storage, and operational costs compared to traditional message systems like Kafka.

Apache FlinkKafka ComparisonPravega
0 likes · 10 min read
Why Pravega Matters: Native Stream Storage for Low‑Latency, Exactly‑Once Data Pipelines
Programmer DD
Programmer DD
Feb 12, 2019 · Fundamentals

How ZGC Achieves Sub‑10 ms Pauses: A Deep Dive into Java’s Low‑Latency GC

ZGC is a scalable, low‑latency Java garbage collector designed to keep pause times under 10 ms regardless of heap size, supporting up to 4 TB, and leveraging concurrent, region‑based, compacting, NUMA‑aware techniques, colored pointers, and load barriers, with detailed compilation and tuning guidance.

Garbage CollectionJDKNUMA
0 likes · 8 min read
How ZGC Achieves Sub‑10 ms Pauses: A Deep Dive into Java’s Low‑Latency GC
Alibaba Cloud Infrastructure
Alibaba Cloud Infrastructure
Nov 21, 2018 · Cloud Computing

Alibaba Data Center Network Architecture HAIL 5.1: High Availability, De‑stacking, and Low‑Latency RDMA Design

The article describes Alibaba's HAIL 5.1 data‑center network architecture introduced for the 2018 Double‑11 event, detailing its high‑availability de‑stacking design, low‑latency RDMA deployment, and future HAIL 2.0 evolution to support larger‑scale, intelligent, and high‑performance cloud networking.

Data CenterHigh AvailabilityRDMA
0 likes · 9 min read
Alibaba Data Center Network Architecture HAIL 5.1: High Availability, De‑stacking, and Low‑Latency RDMA Design
Architecture Digest
Architecture Digest
Sep 10, 2018 · Backend Development

Low‑Latency and High‑Availability Design of RocketMQ for Double‑11 Peak Traffic

This article reviews the evolution of Alibaba's Aliware message engine, analyzes the latency and availability challenges faced during Double‑11, and describes the low‑latency optimizations, capacity‑guarantee strategies, and multi‑replica high‑availability architecture implemented in RocketMQ to sustain trillion‑level message flows.

Message QueueRocketMQcapacity planning
0 likes · 22 min read
Low‑Latency and High‑Availability Design of RocketMQ for Double‑11 Peak Traffic
Alibaba Cloud Infrastructure
Alibaba Cloud Infrastructure
Aug 29, 2018 · Artificial Intelligence

Alibaba's FPGA-Based Ultra‑Low Latency, High‑Throughput Machine Learning Processor

Alibaba unveiled an FPGA‑designed machine‑learning accelerator that achieves sub‑millisecond inference latency and thousands of frames‑per‑second throughput, demonstrating how integrated hardware‑software optimizations can deliver real‑time AI performance surpassing conventional GPU and ASIC solutions.

AI acceleratorFPGAHigh Throughput
0 likes · 5 min read
Alibaba's FPGA-Based Ultra‑Low Latency, High‑Throughput Machine Learning Processor
Tencent Cloud Developer
Tencent Cloud Developer
Jul 30, 2018 · Game Development

How Real-Time Voice Changing Boosts Social Interaction in Games with Tencent GME

The article explains how Tencent Cloud’s Gaming Multimedia Engine (GME) introduces a built‑in real‑time voice‑changing feature for games, detailing the underlying pitch‑and‑timbre manipulation, latency‑reduction techniques that keep delay under 40 ms, and how developers can integrate the SDK to enrich player social interaction without external hardware.

Game DevelopmentSDKTencent Cloud
0 likes · 5 min read
How Real-Time Voice Changing Boosts Social Interaction in Games with Tencent GME
Tencent Cloud Developer
Tencent Cloud Developer
May 2, 2018 · Cloud Computing

Tencent Video Cloud Mini-program Audio/Video Solution: From Concept to Implementation

The article chronicles how Tencent Video Cloud built a low‑latency audio/video SDK for WeChat mini‑programs—using live‑pusher and live‑player components to capture, process, encode, and transmit streams via TCP/UDP, adding echo cancellation, QoS, and room‑based signaling to enable real‑time chat and multi‑party conferencing within a 500 ms end‑to‑end delay.

Mini-programRTCSDK
0 likes · 15 min read
Tencent Video Cloud Mini-program Audio/Video Solution: From Concept to Implementation
21CTO
21CTO
May 10, 2017 · Backend Development

How We Built a Scalable, Low‑Latency Ranking System for Millions of Users

This article describes the challenges and solutions behind designing a high‑availability, low‑latency ranking service that supports tens of thousands of leaderboards, optimizes storage engine choices, automates scheduling, and isolates resources using ZooKeeper, Redis, and container‑based deployments across multiple data centers.

RedisZookeepercontainerization
0 likes · 17 min read
How We Built a Scalable, Low‑Latency Ranking System for Millions of Users
Architecture Digest
Architecture Digest
Feb 9, 2017 · Backend Development

Low‑Latency and High‑Availability Design of RocketMQ: Evolution, Optimization, and Capacity Assurance

This article reviews the evolution of Alibaba's middleware message engine, analyzes the low‑latency and high‑availability challenges faced during Double‑11, and details the optimization techniques, capacity‑guarantee strategies, and HA architecture that enable RocketMQ to handle massive traffic spikes with millisecond‑level latency.

RocketMQlow latency
0 likes · 24 min read
Low‑Latency and High‑Availability Design of RocketMQ: Evolution, Optimization, and Capacity Assurance
Huawei Cloud Developer Alliance
Huawei Cloud Developer Alliance
Nov 15, 2016 · Cloud Computing

How Personal Live Streaming Works: Key Technologies and Performance Tips

This article examines the rapid growth of personal live streaming, outlines major market players, compares cloud provider offerings, and dives into essential technologies such as RTMP variants, source‑side processing, low‑power encoding, fast startup, and methods to reduce stutter for a seamless viewer experience.

CDNVideo Encodinglive streaming
0 likes · 7 min read
How Personal Live Streaming Works: Key Technologies and Performance Tips
Meituan Technology Team
Meituan Technology Team
Nov 4, 2016 · Big Data

Design and Implementation of a Low-Latency App Exception Monitoring Platform Using Spark Streaming, Kafka, and Elasticsearch

The paper presents a production‑grade, low‑cost mobile‑app exception monitoring platform built on Spark Streaming, Kafka, and Elasticsearch that achieves high availability through exactly‑once processing and checkpointing, minute‑level latency by decoupling raw and symbolized logs, high throughput via reservoir sampling, and dynamic scalability without code changes.

Big DataElasticsearchException Monitoring
0 likes · 11 min read
Design and Implementation of a Low-Latency App Exception Monitoring Platform Using Spark Streaming, Kafka, and Elasticsearch
High Availability Architecture
High Availability Architecture
Aug 19, 2016 · Fundamentals

Design and Implementation of a Sub‑500 ms Ultra‑HD Real‑Time Video Transmission System

This article details the architecture, encoding choices, network‑level optimizations, transmission model, measurement methods, and practical pitfalls involved in building a 1080p real‑time video streaming solution that consistently keeps end‑to‑end latency below 500 ms for interactive online education.

H.264Streaming ArchitectureUDP
0 likes · 27 min read
Design and Implementation of a Sub‑500 ms Ultra‑HD Real‑Time Video Transmission System
Architecture Digest
Architecture Digest
Aug 17, 2016 · Backend Development

Design and Optimization of Bilibili Live Chat (GOIM) System

The article presents a detailed overview of Bilibili's GOIM live chat architecture, covering its high‑stability, high‑availability, low‑latency design, component breakdown, memory and module optimizations, network improvements, and performance testing results to achieve scalable real‑time messaging.

Backend ArchitectureGoHigh Availability
0 likes · 13 min read
Design and Optimization of Bilibili Live Chat (GOIM) System
WeChat Client Technology Team
WeChat Client Technology Team
May 10, 2016 · Information Security

How We Built mmtls: A High‑Performance, Low‑Latency Secure Protocol for WeChat

mmtls is a custom, lightweight secure communication protocol designed for WeChat that encrypts all client‑to‑server traffic, offering confidentiality, integrity, low latency, scalability, and forward secrecy by adapting TLS 1.3 concepts with optimized handshake, key‑exchange, record, and replay‑protection mechanisms.

TLSWeChatauthentication
0 likes · 32 min read
How We Built mmtls: A High‑Performance, Low‑Latency Secure Protocol for WeChat
Qunar Tech Salon
Qunar Tech Salon
Jan 7, 2015 · Fundamentals

Understanding the C4 Garbage Collector: A Concurrent Continuously Compacting Collector for Low‑Latency Java Applications

This article explains the design, phases, and practical implications of the C4 concurrent continuously compacting garbage collector, comparing it with G1 and IBM's Balanced GC, and provides guidance on when to choose C4 for enterprise Java workloads requiring low pause times and high scalability.

C4Garbage CollectionJVM
0 likes · 21 min read
Understanding the C4 Garbage Collector: A Concurrent Continuously Compacting Collector for Low‑Latency Java Applications