Tagged articles

Cloud Native

3402 articles · Page 1 of 35
TonyBai
TonyBai
Oct 8, 2026 · Cloud Native

Cloud Native Not Dead: Kubernetes, CNCF, Cilium Redesign Foundations for AI Agents

Tony Bai analyzes how Kubernetes, CNCF, Cilium, and OpenTelemetry are fundamentally redesigning cloud-native primitives—scheduling, resource allocation, sandbox runtimes, gateways, and observability—to support stateful, bursty AI agents, proving cloud-native isn't obsolete but evolving its foundational interfaces.

AI agentsCNCFCilium
0 likes · 27 min read
Cloud Native Not Dead: Kubernetes, CNCF, Cilium Redesign Foundations for AI Agents
Linyb Geek Road
Linyb Geek Road
Oct 6, 2026 · Operations

AI Writes Kubernetes YAML in Seconds: The Real Value of Ops Engineers

The article tests AI tools like DeepSeek for generating Kubernetes YAML, finding they handle standard templates well but fail on cluster-specific configs, security, probes, resource quotas, and complex multi-CRD scenarios. It argues ops engineers' value lies in troubleshooting, architecture decisions, incident handling, setting standards, and building platforms—not writing YAML—and advises embracing AI for drafts while deepening core expertise.

AICloud NativeDevOps
0 likes · 13 min read
AI Writes Kubernetes YAML in Seconds: The Real Value of Ops Engineers
dbaplus Community
dbaplus Community
Sep 29, 2026 · Industry Insights

From Monoliths to AI Agents: Architecture Evolution's Two Patterns and New Challenges

This article traces software architecture evolution from monoliths through primitive distributed systems, SOA, microservices, and cloud-native, highlighting two recurring patterns — finer decoupling and stronger fault isolation — and a shift from zero-failure goals to designing for resilience, while outlining four unprecedented challenges AI agents introduce: semantic hallucinations, stateful context, dynamic orchestration, and observability gaps.

AI agentsCloud Nativedistributed systems
0 likes · 23 min read
From Monoliths to AI Agents: Architecture Evolution's Two Patterns and New Challenges
Ops Development & AI Practice
Ops Development & AI Practice
Sep 25, 2026 · Cloud Native

Apache APISIX Evolution: From Disrupting Kong to Premier AI Gateway

This article traces Apache APISIX's four-phase evolution: its etcd-based architecture achieving millisecond config updates versus Kong's seconds, its 9-month Apache graduation, multi-language plugin runner and Wasm support, Kubernetes-native ingress controller, and transformation into an AI Gateway with token-based rate limiting, protocol normalization, and SSE streaming optimization.

AI GatewayAPI GatewayApache APISIX
0 likes · 17 min read
Apache APISIX Evolution: From Disrupting Kong to Premier AI Gateway
ITPUB
ITPUB
Sep 23, 2026 · R&D Management

Architecture Under Constraints: CTO Weng Yifei on Tech-Business-AI Trade-offs

In this interview, CTO Weng Yifei shares insights on making architecture decisions under resource constraints, contrasting big-tech and startup environments, managing technical debt, integrating AI responsibly, and building governance mechanisms for sustainable technology adoption. He emphasizes controlling complexity, establishing lightweight AI governance, evaluating true ROI beyond local efficiency, and designing auditability for probabilistic systems.

AI GovernanceCloud NativeROI evaluation
0 likes · 27 min read
Architecture Under Constraints: CTO Weng Yifei on Tech-Business-AI Trade-offs
java1234
java1234
Sep 20, 2026 · Artificial Intelligence

Quarkus Embraces AI: Declarative Services, Tools, RAG, and MCP in Java

This article explores Quarkus's AI integration via the LangChain4j extension, demonstrating declarative AI services, function calling with tools, retrieval-augmented generation (RAG), and Model Context Protocol (MCP) support, all while retaining Quarkus's cloud-native benefits like fast startup and native compilation.

AICloud NativeLangChain4j
0 likes · 13 min read
Quarkus Embraces AI: Declarative Services, Tools, RAG, and MCP in Java
Ops Development & AI Practice
Ops Development & AI Practice
Sep 17, 2026 · Cloud Native

Why OpenTelemetry Helm Splits into 3 Releases: Agent, Cluster, Gateway Architecture Explained

This article explains why OpenTelemetry Helm charts now recommend deploying Collector as three separate releases—otel-agent (DaemonSet for node metrics), otel-cluster (singleton Deployment for cluster metrics), and otel-gateway (scalable Deployment for trace ingestion)—detailing Presets simplification, lifecycle isolation, failure domains, and when to consolidate to two releases.

Cloud NativeCollectorDaemonSet
0 likes · 24 min read
Why OpenTelemetry Helm Splits into 3 Releases: Agent, Cluster, Gateway Architecture Explained
Golang Shines
Golang Shines
Sep 17, 2026 · Operations

500 Essential Ops Terms: Kubernetes, Docker & SRE Glossary

This glossary defines 500 fundamental terms for operations engineers, covering Kubernetes core concepts, components, networking, and Docker container terminology with concise explanations for each term.

Cloud NativeContainer OrchestrationDevOps
0 likes · 5 min read
500 Essential Ops Terms: Kubernetes, Docker & SRE Glossary
BanTech Think Tank
BanTech Think Tank
Sep 16, 2026 · Cloud Native

ARM Cloud-Native Performance: Stress Testing Guide for Domestic Microservice Containerization

This article presents a practical guide for performance stress testing in domestic microservice containerization, addressing three core challenges—lack of unified computing power conversion, non-linear multi-core scalability, and frequent base software iterations—through standardized baselines, quantitative modeling, multi-core optimization, and continuous governance frameworks.

ARM architectureCloud NativeNUMA
0 likes · 17 min read
ARM Cloud-Native Performance: Stress Testing Guide for Domestic Microservice Containerization
Alibaba Cloud Observability
Alibaba Cloud Observability
Sep 14, 2026 · Cloud Native

How Lemon Retail Achieved 70% Alert Convergence and Minute-Level MTTR with AI-Driven Cloud-Native Observability

Lemon, a food retail SaaS provider serving 20,000+ stores, unified logs, metrics, and traces into a full-chain digital twin using Alibaba Cloud CloudMonitor 2.0 and STAROps, deploying four intelligent operations layers that cut alert noise by 70% and reduced incident response and MTTR to minutes.

AIOpsAlert GovernanceCloud Native
0 likes · 16 min read
How Lemon Retail Achieved 70% Alert Convergence and Minute-Level MTTR with AI-Driven Cloud-Native Observability
Java Tech Enthusiast
Java Tech Enthusiast
Sep 11, 2026 · Cloud Native

Java Cloud-Native Deployment 2026: 5 Paradigm Shifts Beyond JARs

This article analyzes how GraalVM Native Image 23.0 and Spring Boot 3.5 enable millisecond cold starts, 80% smaller containers, and production-ready serverless Java across AWS, Alibaba Cloud, Tencent Cloud, and Azure — with benchmarks, migration steps, and pitfall solutions.

AOT CompilationCloud NativeDocker
0 likes · 19 min read
Java Cloud-Native Deployment 2026: 5 Paradigm Shifts Beyond JARs
Full-Stack DevOps & Kubernetes
Full-Stack DevOps & Kubernetes
Sep 9, 2026 · Cloud Native

AI Low-Code Cuts K8s Platform Dev Time by 90%: A Practical Guide

The article demonstrates how AI low-code tools like Cursor accelerate building a Kubernetes management platform, reducing development from 7-10 days to 1-2 days, with code examples for Go/FastAPI backend, Vue3 frontend, plus cases for Prometheus alert forwarding and CRUD backends, while stressing human oversight for security and logic.

AI low-codeCloud NativeCursor
0 likes · 18 min read
AI Low-Code Cuts K8s Platform Dev Time by 90%: A Practical Guide
Alibaba Cloud Native
Alibaba Cloud Native
Sep 8, 2026 · Cloud Native

Lemon's Intelligent Ops: 70% Alert Convergence, Minute-Level MTTR for 20K+ Retail Stores

Food retail digitalizer Lemon unified logs, metrics, and traces into a digital twin using Alibaba Cloud CloudMonitor 2.0 and STAROps with UModel, deploying unified alert governance, natural language observability, automated inspections, and AI-driven root cause analysis to achieve 70% alert convergence and minute-level MTTR across 20,000+ stores.

AIOpsAlert GovernanceAutomated Inspection
0 likes · 16 min read
Lemon's Intelligent Ops: 70% Alert Convergence, Minute-Level MTTR for 20K+ Retail Stores
Golang Shines
Golang Shines
Sep 3, 2026 · Operations

Beyond Linux: 10 Skills That Define Senior Operations Engineers

An experienced operations engineer shares the key skills that differentiate senior professionals, including troubleshooting methodology, automation, cloud-native technologies, monitoring systems, security practices, business alignment, SRE principles, AI-assisted operations, and continuous learning, emphasizing that Linux is merely the foundation.

AI operationsAutomationCloud Native
0 likes · 12 min read
Beyond Linux: 10 Skills That Define Senior Operations Engineers
Woodpecker Software Testing
Woodpecker Software Testing
Sep 3, 2026 · Operations

Performance Testing Tools Compared: Real-World Lessons from 27 High-Compliance Projects

Based on 27 real projects across finance, healthcare, and government, this article compares JMeter, k6, Gatling, Locust, and Artillery across programmability, observability, scalability, and engineering integration, revealing why k6 and Locust excel in cloud-native DevOps while JMeter remains viable for Java-heavy teams.

Cloud NativeDevOpsGatling
0 likes · 7 min read
Performance Testing Tools Compared: Real-World Lessons from 27 High-Compliance Projects
dbaplus Community
dbaplus Community
Aug 31, 2026 · Interview Experience

Why Kubernetes Leader Kelsey Hightower Refused Microsoft and Retired at 43

From dropping out of college to becoming Google’s distinguished L9 engineer, Kelsey Hightower’s unconventional journey—spanning DSL modem work, A+ certification, data‑center roles, Puppet, CoreOS, and Kubernetes—reveals his belief that technology serves people, his refusal of Microsoft, and his early retirement at 43.

AICareerCloud Native
0 likes · 50 min read
Why Kubernetes Leader Kelsey Hightower Refused Microsoft and Retired at 43
Didi Tech
Didi Tech
Aug 31, 2026 · Cloud Native

How HUATUO Builds Kernel Panoramic Observability for Agent Sandboxes

DiDi's HUATUO project presents a kernel observability framework for agent sandboxes using eBPF and dynamic tracing to capture system-wide metrics, exceptions, auto-tracing, continuous profiling, and hardware faults, correlating kernel events with container identities for root-cause analysis in cloud-native and AI infrastructure.

AutoTracingCloud Nativeagent sandbox
0 likes · 15 min read
How HUATUO Builds Kernel Panoramic Observability for Agent Sandboxes
Alibaba Cloud Native
Alibaba Cloud Native
Aug 31, 2026 · Cloud Native

Kickstarting the Data Flywheel: Four Ways to Connect Agents to AgentLoop

This article explains how AgentLoop uses the OpenTelemetry protocol and probes to ingest high‑quality runtime data, offering four integration methods—one‑click generic agents, SDK framework integration, annotation‑based high‑code, and eBPF—demonstrated with a Claude Code customer‑service agent and end‑to‑end verification on the observation page.

AgentLoopCloud NativeData Ingestion
0 likes · 9 min read
Kickstarting the Data Flywheel: Four Ways to Connect Agents to AgentLoop
Tencent Technical Engineering
Tencent Technical Engineering
Aug 31, 2026 · Cloud Native

CubeSandbox v0.7.0: Cross-Machine Sandbox Migration & Multi-Version Runtime Support

CubeSandbox v0.7.0 introduces cross-machine pause/resume via S3 backend storage, allows component multi-version coexistence so upgrades don't break existing templates, merges NetworkAgent into Cubelet to cut RPC calls and speed cold starts, and separates control-plane scheduling from node operations via new CubeOps service.

Cloud NativeCubeSandboxKubernetes
0 likes · 10 min read
CubeSandbox v0.7.0: Cross-Machine Sandbox Migration & Multi-Version Runtime Support
TonyBai
TonyBai
Aug 30, 2026 · Cloud Native

Stop Struggling with Kubernetes Docs: An Amazon‑Warehouse Analogy That Reveals the Whole System

Using an Amazon‑warehouse analogy, the article explains Kubernetes’s core philosophy of declarative desired state and maps each control‑plane and node component—apiserver, etcd, controller‑manager, scheduler, kubelet, kube‑proxy—to intuitive warehouse roles, while detailing service discovery, EndpointSlices, and external traffic flow.

Cloud NativeControl PlaneDesired State
0 likes · 17 min read
Stop Struggling with Kubernetes Docs: An Amazon‑Warehouse Analogy That Reveals the Whole System
Alibaba Cloud Native
Alibaba Cloud Native
Aug 29, 2026 · Cloud Native

Ingress NGINX Retired Amid New Critical Vulnerabilities – Migrate to Alibaba Cloud API Gateway in 10 Minutes

Ingress NGINX has been retired and is plagued by multiple CVSS 8.1 high‑severity vulnerabilities that lack patches, prompting urgent migration to Alibaba Cloud's Cloud Native API Gateway, which now offers expanded CLB/NLB reuse, annotation compatibility analysis, and integrated traffic‑shifting and rollback workflows.

ACKAPI GatewayCloud Native
0 likes · 14 min read
Ingress NGINX Retired Amid New Critical Vulnerabilities – Migrate to Alibaba Cloud API Gateway in 10 Minutes
Alibaba Cloud Native
Alibaba Cloud Native
Aug 28, 2026 · Industry Insights

When Model Power Is Plenty, Real‑Time Data Pipelines Become the AI Production Bottleneck

As AI models become sufficiently capable, the primary obstacle to production shifts from model selection to the real‑time data link, requiring new load handling, sub‑second context availability, and tighter trust boundaries, prompting a redesign of streaming platforms, consumption models, and security mechanisms.

AICloud NativeEvent-driven AI
0 likes · 16 min read
When Model Power Is Plenty, Real‑Time Data Pipelines Become the AI Production Bottleneck
AI Engineering
AI Engineering
Aug 28, 2026 · Cloud Native

Why Round‑Robin Fails for LLM Inference and How llm‑d Fixes It

Round‑robin routing in Kubernetes wipes out KV‑cache benefits for LLM inference, but llm‑d introduces cache‑aware routing, hierarchical eviction, and prefill/decode separation, delivering up to three‑fold throughput gains and halving first‑token latency, as shown in Tesla's production rollout.

Cloud NativeKV CacheKubernetes
0 likes · 7 min read
Why Round‑Robin Fails for LLM Inference and How llm‑d Fixes It
AI Architecture Path
AI Architecture Path
Aug 28, 2026 · Industry Insights

Why Apache Superset’s 74.5K Stars Make It the Free, Open‑Source Choice for Enterprise Data Dashboards

The article explains how Apache Superset, a free open‑source BI platform with 74.5K GitHub stars, solves the high cost and lock‑in issues of commercial tools by offering extensive data‑source compatibility, dual no‑code and SQL‑Lab modes, cloud‑native architecture, fine‑grained security, and step‑by‑step deployment guidance for enterprise data dashboards.

Apache SupersetCloud NativeData Visualization
0 likes · 11 min read
Why Apache Superset’s 74.5K Stars Make It the Free, Open‑Source Choice for Enterprise Data Dashboards
Woodpecker Software Testing
Woodpecker Software Testing
Aug 27, 2026 · Cloud Native

Adversarial Performance Testing: A Hands‑On Guide to Boost System Resilience

In today’s high‑concurrency, microservice‑driven cloud‑native world, traditional load testing often misses real‑world failure modes, so this guide introduces adversarial performance testing—injecting faults, latency, and malicious traffic—to expose hidden bottlenecks and build resilient systems.

Cloud NativePerformance OptimizationResilience
0 likes · 8 min read
Adversarial Performance Testing: A Hands‑On Guide to Boost System Resilience
Cloud Architecture
Cloud Architecture
Aug 26, 2026 · Backend Development

From Zero to Production: High‑Concurrency Netty TCP Server for Cloud‑Native

This article walks through building a production‑grade Netty TCP server, covering protocol design, reactor threading, back‑pressure handling, session management, authentication, heartbeats, scaling to hundreds of thousands of connections, cloud‑native deployment, graceful shutdown, observability, reliability, and security considerations.

Cloud NativeHigh ConcurrencyNetty
0 likes · 45 min read
From Zero to Production: High‑Concurrency Netty TCP Server for Cloud‑Native
mikechen
mikechen
Aug 26, 2026 · Cloud Native

Kubernetes Architecture Deep Dive: Master-Node Components & Workflow

This article explains Kubernetes architecture, covering the master-node model, core components like API server, etcd, scheduler, controller manager, kubelet, and kube-proxy, and how they coordinate to deploy and manage containerized applications.

API ServerCloud NativeContainer Orchestration
0 likes · 6 min read
Kubernetes Architecture Deep Dive: Master-Node Components & Workflow
Tencent Architect
Tencent Architect
Aug 26, 2026 · Cloud Native

Zombie Memcg Exhausting Node Memory? TencentOS's Cross-Kernel Fix

TencentOS analyzes zombie memcg root causes — page cache, shmem, and swap entries — reviews existing solutions' limitations, and proposes a dual-track approach: upstream patches for new kernels (obj_cgroup binding, swap entry fix) and a kernel module for old kernels that asynchronously reparents dying memcg pages and swap entries to online parents, preserving cache while enabling zero-restart deployment.

Cloud NativeLinux kernelTencentOS
0 likes · 31 min read
Zombie Memcg Exhausting Node Memory? TencentOS's Cross-Kernel Fix
Woodpecker Software Testing
Woodpecker Software Testing
Aug 26, 2026 · Operations

2026 Open‑Source Performance Testing Tools: From Load to Diagnosis

The article evaluates the evolution and practical capabilities of leading 2026 open‑source performance testing tools across twelve real‑world scenarios—ranging from financial API stress tests to IoT clusters and LLM latency—using a five‑dimensional model that assesses protocol coverage, native cloud‑native integration, intelligent diagnosis, generative collaboration, and compliance readiness.

Cloud NativeLoad Testingbenchmark
0 likes · 8 min read
2026 Open‑Source Performance Testing Tools: From Load to Diagnosis
Alibaba Cloud Big Data AI Platform
Alibaba Cloud Big Data AI Platform
Aug 26, 2026 · Artificial Intelligence

How Alibaba Cloud Elasticsearch’s Cloud‑Native Vector Engine Tops VectorDBBench

Alibaba Cloud Elasticsearch on ES 9.4, using the FalconSeek HNSW engine, achieves 82,520 QPS at 0.98 recall with a 1.8 ms P99 latency in VectorDBBench, and the article explains the end‑to‑end architectural redesign—including quantized candidate recall, batch distance computation, hot‑data layout, on‑demand re‑ranking, and segment lifecycle integration—that makes these results possible.

AI SearchCloud NativeElasticsearch
0 likes · 17 min read
How Alibaba Cloud Elasticsearch’s Cloud‑Native Vector Engine Tops VectorDBBench
dbaplus Community
dbaplus Community
Aug 24, 2026 · Cloud Native

Why Docker Dominates Cloud Computing by Reusing Decades‑Old Technologies

Docker’s success stems from stitching together decades‑old OS primitives—chroot, Linux namespaces, layered tar filesystems, SLIRP, and QEMU—so developers can keep their existing workflows unchanged, allowing containers to become the de‑facto operating system of the cloud era.

Cloud NativeDockercontainers
0 likes · 15 min read
Why Docker Dominates Cloud Computing by Reusing Decades‑Old Technologies
CTO Full-Stack Academy
CTO Full-Stack Academy
Aug 24, 2026 · Cloud Native

Docker vs Traditional Physical Machine Deployment: 8 Key Differences Explained

The article compares Docker container deployment with traditional physical‑machine deployment across eight dimensions—environment consistency, speed, resource utilization, isolation, migration, operational overhead, failure recovery, and version rollback—highlighting how containers improve efficiency, cost, and reliability.

Cloud NativeDevOpsDocker
0 likes · 10 min read
Docker vs Traditional Physical Machine Deployment: 8 Key Differences Explained
Alibaba Cloud Native
Alibaba Cloud Native
Aug 23, 2026 · Artificial Intelligence

Why Is GPU Utilization Low? Try This Zero‑Intrusion AI Profiling Tool

The article introduces SysOM AI Profiling, a zero‑intrusive, cloud‑native performance observation and diagnosis solution for AI workloads that spans training to inference, single‑GPU to multi‑GPU clusters, and Python to GPU kernel layers, helping users pinpoint low GPU utilization, memory leaks, and communication bottlenecks.

AI profilingCloud NativeGPU utilization
0 likes · 14 min read
Why Is GPU Utilization Low? Try This Zero‑Intrusion AI Profiling Tool
Three Knives
Three Knives
Aug 22, 2026 · Cloud Native

Eliminating Startup Decisions: How Feat Moves Runtime Choices to Compile Time for Faster Cold Starts

The article explains how the Feat framework reduces Java cold‑start latency by shifting configuration and environment decisions from runtime to compile time, using a three‑layer config model, externalized placeholders, and compile‑time code generation to eliminate on‑the‑fly processing.

Cloud NativeCompile-time ConfigurationFeat framework
0 likes · 7 min read
Eliminating Startup Decisions: How Feat Moves Runtime Choices to Compile Time for Faster Cold Starts
ITPUB
ITPUB
Aug 20, 2026 · Industry Insights

2026 China Database Technology Conference Launches: Data Fusion and AI Leadership

The 17th China Database Technology Conference (DTCC 2026) ran from August 20‑22 in Beijing, gathering top experts to discuss database kernel innovations, cloud‑native and distributed practices, AI‑driven data, vector databases, real‑time warehouses, and the emerging Agent era, while showcasing cutting‑edge solutions from Dameng, Tencent Cloud, Alibaba Cloud, OceanBase and GoldenDB.

AIAgentCloud Native
0 likes · 15 min read
2026 China Database Technology Conference Launches: Data Fusion and AI Leadership
Airbnb Technology Team
Airbnb Technology Team
Aug 20, 2026 · Cloud Native

How Airbnb Built a Scalable, Reliable Kubernetes Sidecar for Dynamic Configuration

The article explains Airbnb's Sitar‑agent sidecar architecture, detailing the end‑to‑end configuration distribution lifecycle, key design choices such as sidecar versus in‑process deployment, pull‑model optimizations, and the migration from Sparkey to SQLite for robust, multi‑language support at massive scale.

Cloud NativeKubernetesRocksDB
0 likes · 15 min read
How Airbnb Built a Scalable, Reliable Kubernetes Sidecar for Dynamic Configuration
Tencent Cloud Middleware
Tencent Cloud Middleware
Aug 19, 2026 · Operations

How AI Gateway Makes Large-Model Calls Visible, Traceable, and Auditable

Enterprises deploying large-model APIs often struggle to see token usage, latency, and errors; the AI Gateway embeds metrics, structured logs, and distributed tracing at the gateway layer, providing token-level insights, request-level latency breakdowns, and full-chain auditability without code changes, as demonstrated in a real-world incident.

AI GatewayCloud NativeLLM
0 likes · 17 min read
How AI Gateway Makes Large-Model Calls Visible, Traceable, and Auditable
java1234
java1234
Aug 18, 2026 · Backend Development

Why Is Jakarta EE Gaining Traction in Enterprise Java?

Jakarta EE, the rebranded Java EE now governed by the Eclipse Foundation, offers open standards, built‑in enterprise capabilities, cloud‑native improvements, a smooth upgrade path for legacy systems, and familiar APIs, making it an attractive alternative to Spring for long‑term, maintainable applications.

CDICloud NativeEnterprise Java
0 likes · 9 min read
Why Is Jakarta EE Gaining Traction in Enterprise Java?
Smart Sea Tide
Smart Sea Tide
Aug 18, 2026 · Big Data

How to Choose and Architect a Data Lake Platform for Enterprise Digital Transformation

The article outlines the strategic need for a unified data lake in a digital‑focused enterprise, details functional and non‑functional requirements such as linear scalability, real‑time and batch processing, multi‑tenant support, security and governance, and presents a comprehensive architecture design that integrates storage, compute, and management components.

Cloud NativeData Lakebig data
0 likes · 19 min read
How to Choose and Architect a Data Lake Platform for Enterprise Digital Transformation
21CTO
21CTO
Aug 16, 2026 · Artificial Intelligence

Microsoft Joins Google in Adding Go Support for AI Agent Development

The article explains how Microsoft’s new Agent Framework for Go extends native AI agent capabilities—such as large‑model access, tool calls, and multi‑agent coordination—to the Go ecosystem, reflecting broader industry moves by Google and the growing demand for Go‑centric cloud‑native AI development.

AI agentsAgent FrameworkAzure OpenAI
0 likes · 6 min read
Microsoft Joins Google in Adding Go Support for AI Agent Development
Cloud Architecture
Cloud Architecture
Aug 14, 2026 · Cloud Native

Complete Guide to Go Microservice Logging and Tracing with OpenTelemetry (Industrial‑Grade Solution)

When an alarm rang at 2:17 AM, a Go order service’s P99 latency surged from 220 ms to 4.6 s and its error rate climbed to 1.8 %; the article explains why many teams still see limited value after adopting OpenTelemetry, identifies three missing pieces—stable trace IDs, end‑to‑end context propagation, and production‑ready pipelines—and delivers a step‑by‑step, code‑first blueprint for building an industrial‑grade observability stack that scales in Kubernetes.

Cloud NativeGoOpenTelemetry
0 likes · 40 min read
Complete Guide to Go Microservice Logging and Tracing with OpenTelemetry (Industrial‑Grade Solution)
TechVision Expert Circle
TechVision Expert Circle
Aug 13, 2026 · Cloud Native

Does a Three‑Hour Commute Actually Boost Anyone’s Efficiency?

A 2026 OPM survey of over 600,000 federal workers shows that mandatory office returns cut satisfaction and raise turnover, while modern cloud‑native collaboration tools and AI‑driven workflows enable remote teams to match or exceed on‑site productivity, revealing the hidden technical costs of forced commuting.

AI collaborationCRDTCloud Native
0 likes · 13 min read
Does a Three‑Hour Commute Actually Boost Anyone’s Efficiency?
Alibaba Cloud Native
Alibaba Cloud Native
Aug 13, 2026 · Artificial Intelligence

Which Model Should Handle Your Request? Alibaba Cloud AI Gateway Intelligent Routing Goes Live

As the number of LLMs grows, developers face the dilemma of selecting the right model for each request; Alibaba Cloud’s AI Gateway introduces intelligent routing that evaluates model capabilities, cost, speed, and task suitability, making a single, optimal decision per request while preserving context in multi‑turn and Agent workflows.

AI GatewayAgent WorkflowAlibaba Cloud
0 likes · 11 min read
Which Model Should Handle Your Request? Alibaba Cloud AI Gateway Intelligent Routing Goes Live
Woodpecker Software Testing
Woodpecker Software Testing
Aug 13, 2026 · Industry Insights

2026 Stress‑Testing ROI: When Is the Investment Worth It?

The article analyzes how AI‑assisted scenario generation, chaos‑as‑a‑service, and observability reshape stress‑testing costs in 2026, presenting ROI models, industry benchmarks, and critical thresholds that turn testing from a risk hedge into a growth lever.

AI-generated trafficCloud NativeROI
0 likes · 8 min read
2026 Stress‑Testing ROI: When Is the Investment Worth It?
Alibaba Middleware
Alibaba Middleware
Aug 12, 2026 · Operations

How STAROps Detects Unknown Anomalies with Intelligent Log Inspection

STAROps transforms raw logs into actionable insights by clustering log patterns, drilling down across dimensions with AI operators, and using an Agent that dynamically plans investigations, integrates UModel cross‑source mapping, and continuously refines findings to catch unknown anomalies before they become incidents.

AI operatorsCloud NativeUModel
0 likes · 17 min read
How STAROps Detects Unknown Anomalies with Intelligent Log Inspection
Alibaba Cloud Native
Alibaba Cloud Native
Aug 12, 2026 · Cloud Native

Alibaba Cloud and Datadog Release OpenTelemetry Go Compile‑Time Instrumentation v1 for Zero‑Code Observability

The OpenTelemetry Go Compile‑Time Instrumentation project, jointly launched by Alibaba Cloud and Datadog, fills the last observability gap for Go by injecting tracing and metrics code at build time, offering zero‑code instrumentation, no runtime overhead, and seamless CI/CD integration while comparing it with manual and eBPF approaches.

Cloud NativeCompile-Time InstrumentationGo
0 likes · 10 min read
Alibaba Cloud and Datadog Release OpenTelemetry Go Compile‑Time Instrumentation v1 for Zero‑Code Observability
Woodpecker Software Testing
Woodpecker Software Testing
Aug 12, 2026 · Cloud Native

Distributed vs Monolithic: Uncovering the Real Performance Trade‑offs

The article analyzes distributed and traditional monolithic architectures across response latency, throughput, scalability, and fault‑tolerance cost, revealing that distributed systems introduce network latency, serialization overhead, higher resource consumption, and complex failure modes that often offset their scalability benefits, and provides concrete case studies from Netflix, Alibaba, and LinkedIn.

Cloud NativeThroughputdistributed systems
0 likes · 7 min read
Distributed vs Monolithic: Uncovering the Real Performance Trade‑offs
samdeepthink
samdeepthink
Aug 12, 2026 · Operations

Why Observability Is More Than Monitoring: Finding the Root Cause Quickly

The article explains that observability goes beyond simple monitoring by combining metrics, logs, and traces to pinpoint where and why a system issue occurs, especially in microservice and cloud‑native environments, and stresses the importance of correlating data rather than merely collecting more.

AIOpsCloud NativeLogs
0 likes · 3 min read
Why Observability Is More Than Monitoring: Finding the Root Cause Quickly
Java Architecture Diary
Java Architecture Diary
Aug 12, 2026 · Cloud Native

Why Upgrading Your MCP Server to 2.0 Solves Stateless Session Issues

The article explains how MCP 1.x's stateful handshake caused node‑crash failures, sticky sessions, and serverless incompatibility, and how the 2.0 release removes the handshake, makes each request self‑describing via _meta and HTTP headers, introduces MRTR for multi‑round interactions, and provides a Java/TypeScript code walkthrough demonstrating the new stateless behavior.

Cloud NativeJavaKubernetes
0 likes · 8 min read
Why Upgrading Your MCP Server to 2.0 Solves Stateless Session Issues
Ray's Galactic Tech
Ray's Galactic Tech
Aug 10, 2026 · Cloud Native

Destruction and Rebirth: Deep Dive into ETCD Backup and Restore for Kubernetes Clusters

This article walks through a real‑world ETCD failure, explains why ETCD is the control‑plane brain, details the three‑layer ETCD architecture, exposes common backup pitfalls, and provides a production‑grade backup‑restore workflow—including snapshot API usage, Go implementation, verification steps, and post‑restore validation—for reliable Kubernetes disaster recovery.

Cloud NativeGoKubernetes
0 likes · 33 min read
Destruction and Rebirth: Deep Dive into ETCD Backup and Restore for Kubernetes Clusters
Random Bulletin
Random Bulletin
Aug 10, 2026 · Cloud Native

From Embedded SDKs to Sidecars: How Service Mesh Evolves Governance at Scale

The article examines how embedding service‑governance logic in application processes creates multi‑language, upgrade, coupling, and resource "taxes," and how moving that logic to sidecar proxies and a centralized Service Mesh shifts those costs while delivering language‑agnostic, upgrade‑decoupled, zero‑trust, and declarative traffic management, albeit with new performance, resource, and operational overheads.

Cloud NativeeBPFgovernance
0 likes · 21 min read
From Embedded SDKs to Sidecars: How Service Mesh Evolves Governance at Scale
Java Architect Handbook
Java Architect Handbook
Aug 10, 2026 · Cloud Native

Tired of XXL‑Job? Try This Elegant Nacos‑Based Scheduling Solution

The article analyses why XXL‑Job’s separate registration, configuration, and weak sharding cause state inconsistency, observability gaps, and duplicate processing, then proposes JobFlow – a lightweight scheduler that removes redundant components, adds full‑traceId tracing, true sharding with distributed locks, exponential retry, and cloud‑native configuration managed by Nacos, all illustrated with concrete code snippets and deployment diagrams.

Cloud NativeJavaNacos
0 likes · 21 min read
Tired of XXL‑Job? Try This Elegant Nacos‑Based Scheduling Solution
Alibaba Cloud Infrastructure
Alibaba Cloud Infrastructure
Aug 9, 2026 · Cloud Native

How ACK One Fleet Transforms Agent Sandbox from Single-Cluster to Multi-Cluster

The article explains how ACK One Fleet upgrades the AI Agent Sandbox from a single‑cluster Kubernetes setup to a multi‑cluster architecture, addressing capacity limits, fault‑domain risks, and scheduling inefficiencies while providing global capacity control, water‑level balancing, fault‑tolerant failover, and faster sandbox startup through E2B and CRD integrations.

ACK OneCloud NativeE2B
0 likes · 11 min read
How ACK One Fleet Transforms Agent Sandbox from Single-Cluster to Multi-Cluster
Machine Heart
Machine Heart
Aug 9, 2026 · Industry Insights

Why Codex-Style Harnesses Will Peak in Just Two Months—and Laptops Won’t Keep Up

OpenAI’s product chief warns that Codex‑based Harness agents will become a primitive tool within two to three months as notebook‑bound workflows hit compute, memory, uptime, and context‑concurrency limits, prompting a shift toward cloud‑native micro‑sandbox infrastructures for scalable AI agents.

AI agentsAnthropicCloud Native
0 likes · 6 min read
Why Codex-Style Harnesses Will Peak in Just Two Months—and Laptops Won’t Keep Up
Woodpecker Software Testing
Woodpecker Software Testing
Aug 8, 2026 · Operations

Shift‑Left Performance Testing vs Traditional: A Must‑Read for Test Experts

The article examines how shift‑left performance testing moves responsibility earlier in the development lifecycle, contrasting it with traditional post‑release load testing, and demonstrates through real‑world cases how this approach improves efficiency, cost, quality, and collaboration in cloud‑native, high‑availability systems.

CI/CDCloud NativeDevOps
0 likes · 8 min read
Shift‑Left Performance Testing vs Traditional: A Must‑Read for Test Experts
Alibaba Cloud Native
Alibaba Cloud Native
Aug 8, 2026 · Backend Development

Designing AI‑Friendly Backend Architecture for 24/7 Unattended Development

The article outlines a comprehensive roadmap for transforming traditional backend systems into AI‑friendly architectures, introducing concepts such as Architecture Maps, Service Cards, SKILL packages, multi‑layered testing, permission tiers, and a Harness framework to enable reliable, 24/7 unattended AI‑driven development and operations.

AICloud NativeDevOps
0 likes · 39 min read
Designing AI‑Friendly Backend Architecture for 24/7 Unattended Development
Xiaolin Talks Programming
Xiaolin Talks Programming
Aug 8, 2026 · Backend Development

Spring Boot & MinIO: Building an Enterprise File Platform with Multipart Upload & Multi-tenancy

This article details building an enterprise file service platform using Spring Boot and MinIO, covering cluster architecture, multi-tenant isolation, SDK tuning, multipart upload with resumable frontend direct upload, lifecycle management, security controls, production optimization, and async image processing pipelines.

Cloud NativeMultipart UploadObject Storage
0 likes · 23 min read
Spring Boot & MinIO: Building an Enterprise File Platform with Multipart Upload & Multi-tenancy
Cloud Native Technology Community
Cloud Native Technology Community
Aug 6, 2026 · Cloud Native

5 Production Challenges for Running AI Workloads on Kubernetes: From GPU Scheduling to Observability

Running AI workloads on Kubernetes introduces five production‑grade challenges—complex GPU and accelerator management, workload‑aware scheduling, inference autoscaling beyond CPU metrics, multi‑layer observability, and Day 2 governance—requiring platform teams to extend their capabilities beyond traditional container operations.

AI workloadsCloud NativeDay 2 operations
0 likes · 10 min read
5 Production Challenges for Running AI Workloads on Kubernetes: From GPU Scheduling to Observability
Alibaba Cloud Native
Alibaba Cloud Native
Aug 6, 2026 · Artificial Intelligence

OpenAgentPack: Managing and Migrating Cloud AI Agents Like Code

OpenAgentPack is an open‑source tool that lets you describe a cloud AI Agent’s entire workflow—including model, environment, skills, MCP, knowledge files and credentials—in a single agents.yaml file, version it in Git, preview changes with validate and plan commands, and redeploy the agent across providers while preserving reproducibility, collaboration and auditability.

AI agentsCLICloud Native
0 likes · 7 min read
OpenAgentPack: Managing and Migrating Cloud AI Agents Like Code
Ray's Galactic Tech
Ray's Galactic Tech
Aug 4, 2026 · Cloud Native

Ditch Glue Code: Skill‑Based Distributed Engine for Multi‑API Orchestration and Real‑Time DB Checks

The article analyzes why traditional script‑based API tests fail in microservice environments and proposes a Skill‑oriented, DAG‑driven distributed automation engine that unifies HTTP calls, database verification, message validation, retries, and observability into a scalable, governable platform.

AutomationCloud NativeDAG
0 likes · 37 min read
Ditch Glue Code: Skill‑Based Distributed Engine for Multi‑API Orchestration and Real‑Time DB Checks
Alibaba Cloud Native
Alibaba Cloud Native
Aug 4, 2026 · Artificial Intelligence

AI Innovation Forum Shanghai: Key Takeaways, Multi‑Agent Architecture, and PPT Resources

The AI Innovation Practice Forum in Shanghai gathered over 70 tech professionals to present deep dives on multi‑agent governance, the Agent Native Cloud three‑layer model, AgentTeams collaboration platform, AgentLoop lifecycle flywheel, a cloud‑native network foundation, and next‑gen AIOps, with PPTs available for download.

AI agentsAIOpsAgent Native Cloud
0 likes · 6 min read
AI Innovation Forum Shanghai: Key Takeaways, Multi‑Agent Architecture, and PPT Resources
Alibaba Cloud Native
Alibaba Cloud Native
Aug 4, 2026 · Operations

From Building Wheels to Embedding OpenAPI: Jingchen’s Choice of an Intelligent Ops Foundation

Facing exploding system complexity, unclear global topology, fragmented observability data, and noisy alerts, Jingchen migrated its full‑stack to the cloud and adopted Alibaba Cloud STAROps, a unified CMS 2.0 data base, UModel digital‑twin topology, and OpenAPI‑driven AI diagnostics to turn heavy‑lifting ops work into an automated, business‑focused capability.

Cloud NativeOpenAPISRE
0 likes · 10 min read
From Building Wheels to Embedding OpenAPI: Jingchen’s Choice of an Intelligent Ops Foundation
Alibaba Cloud Native
Alibaba Cloud Native
Aug 3, 2026 · Artificial Intelligence

Building a Financial‑Grade AI Agent Platform with AgentScope: A Practical Whitepaper

FinXScope, a financial‑grade AI‑native agent base built on AgentScope Java, serves as the core engine of the Agent Harness system, offering multi‑agent orchestration, dual‑mode execution, six‑layer architecture, high‑availability, security, observability and low‑code to high‑code pathways, and has already been adopted by dozens of leading financial institutions.

AI agentsAgentScopeCloud Native
0 likes · 33 min read
Building a Financial‑Grade AI Agent Platform with AgentScope: A Practical Whitepaper
ITPUB
ITPUB
Aug 3, 2026 · Industry Insights

Why Docker’s Patchwork of Legacy Tech Still Dominates Cloud Computing

Despite being built on decades‑old components such as chroot, Linux namespaces, SLIRP and QEMU, Docker has become the de‑facto operating system of the cloud because it demands almost no changes to developers' existing workflows, offering a pragmatic, “good enough” solution that scales globally.

Cloud NativeDockerLegacy Technology
0 likes · 16 min read
Why Docker’s Patchwork of Legacy Tech Still Dominates Cloud Computing
Golang Shines
Golang Shines
Aug 3, 2026 · Cloud Native

How I Built a Production‑Ready HA Kubernetes Cluster in Minutes

When my manager suddenly demanded a production‑grade, highly available Kubernetes cluster integrated with a private Harbor registry, I followed a comprehensive step‑by‑step guide to finish the entire setup within a few hours, and now share the 83‑page manual for anyone to replicate.

Cloud NativeCluster DeploymentHarbor
0 likes · 3 min read
How I Built a Production‑Ready HA Kubernetes Cluster in Minutes
Architect Chen
Architect Chen
Aug 2, 2026 · Cloud Native

All Essential kubectl Commands for 2026: A Complete Guide

This article provides a concise, step‑by‑step reference of the most frequently used kubectl commands—including get, describe, logs, exec, apply, port‑forward, rollout, scale, and delete—showing exact syntax and typical use cases for managing Kubernetes resources.

Cloud NativeContainer ManagementDevOps
0 likes · 4 min read
All Essential kubectl Commands for 2026: A Complete Guide
Alibaba Cloud Native
Alibaba Cloud Native
Aug 2, 2026 · Cloud Native

From Visibility to Self‑Healing: ChangjieTong’s Observability and AI‑Powered Ops Journey

ChangjieTong transformed its SaaS‑based, multi‑tenant finance cloud platform by building a five‑layer observability stack on Alibaba Cloud CloudMonitor 2.0, integrating a UModel digital‑twin, and deploying AI‑driven inspection, self‑healing and capacity‑prediction loops, which lifted SLA from 99.9% to 99.995% and cut average fault‑resolution time from over 10 minutes to under 30 seconds.

AIOpsCapacity PredictionCloud Native
0 likes · 15 min read
From Visibility to Self‑Healing: ChangjieTong’s Observability and AI‑Powered Ops Journey
Random Bulletin
Random Bulletin
Aug 1, 2026 · Cloud Native

Mixed‑Tenant Architecture: From Exclusive to Shared, Boosting Utilization to 45%

The article explains how moving from exclusive, peak‑sized clusters to a mixed‑tenant architecture—leveraging time‑shifted workloads, cgroup and hardware isolation (LLC, memory bandwidth), QoS tiers, elastic throttling and dynamic over‑commit—can raise CPU utilization from under 20% to over 45% and cut costs by about 30%, while introducing significant operational complexity and stability risks.

CPU utilizationCloud NativeQoS
0 likes · 18 min read
Mixed‑Tenant Architecture: From Exclusive to Shared, Boosting Utilization to 45%
Random Bulletin
Random Bulletin
Jul 31, 2026 · Cloud Native

Scaling at Ten‑Million QPS: From Manual to Automatic Autoscaling

The article analyzes why manual capacity adjustments break down at ten‑million‑QPS scale, then walks through metric‑driven autoscaling, anti‑flapping algorithms, headroom planning, predictive scaling, stateful service challenges, and multi‑dimensional strategies to achieve a cost‑stable dynamic balance.

AutoscalingCloud NativeKubernetes
0 likes · 20 min read
Scaling at Ten‑Million QPS: From Manual to Automatic Autoscaling
MaGe Linux Operations
MaGe Linux Operations
Jul 31, 2026 · Cloud Native

Choosing an Ingress Controller: Production Comparison of NGINX, Traefik, and APISIX

This article presents a production‑grade comparison of three Kubernetes Ingress controllers—NGINX, Traefik, and APISIX—by defining a four‑layer evaluation framework, detailing pre‑deployment checks, configuration examples, testing scripts, performance metrics, and rollout/rollback procedures to help teams select the most suitable solution.

APISIXCloud NativeKubernetes
0 likes · 24 min read
Choosing an Ingress Controller: Production Comparison of NGINX, Traefik, and APISIX
Alibaba Cloud Native
Alibaba Cloud Native
Jul 30, 2026 · Cloud Native

Replication‑Free Failover for RocketMQ: Achieving Second‑Level Takeover Without Data Copy

The ACM FSE‑2026 industry paper introduces a replication‑free failover mechanism for cloud‑native stateful services like Apache RocketMQ, using protocol‑level write isolation and multi‑attach storage to achieve second‑level recovery without extra data copies, while maintaining low cost and near‑native throughput.

Cloud NativeProtocol FencingReplication-Free Failover
0 likes · 8 min read
Replication‑Free Failover for RocketMQ: Achieving Second‑Level Takeover Without Data Copy
DevOps Operations Practice
DevOps Operations Practice
Jul 30, 2026 · Operations

Essential Velero Guide for Kubernetes Disaster Recovery

This article walks through using Velero to back up, restore, and migrate Kubernetes clusters, covering MinIO installation, Velero client and server setup, storage volume creation, backup location configuration, and execution of backup, restore, and scheduled backup commands.

Cloud NativeKubernetesVelero
0 likes · 11 min read
Essential Velero Guide for Kubernetes Disaster Recovery
Alibaba Cloud Native
Alibaba Cloud Native
Jul 29, 2026 · Artificial Intelligence

Launching a Multi‑Agent AI Platform in One Week with Alibaba Cloud AgentTeams and AI Gateway

In just one week, XinYongZhongHe partnered with Alibaba Cloud to build an enterprise‑grade multi‑agent AI platform using AgentTeams and the AI Gateway, enabling coordinated agents, unified model governance, secure access, cost control, and a range of internal services that move AI from simple Q&A to task execution.

AI GatewayAI GovernanceAgentTeams
0 likes · 10 min read
Launching a Multi‑Agent AI Platform in One Week with Alibaba Cloud AgentTeams and AI Gateway
YiSu Grain
YiSu Grain
Jul 29, 2026 · Fundamentals

45‑Minute Review: Master the Eight Architecture Types for the Soft Exam

This article guides readers through a 45‑minute review of the eight architecture categories—information‑system, layered, cloud‑native, SOA, embedded, communication, security, and big‑data—showing how to identify the relevant type from a problem statement, combine multiple architectures in a solution, and articulate the problem, solution, rationale, and cost with concrete examples and tables.

Cloud NativeSOAbig data
0 likes · 32 min read
45‑Minute Review: Master the Eight Architecture Types for the Soft Exam
Alibaba Cloud Native
Alibaba Cloud Native
Jul 27, 2026 · Artificial Intelligence

Why AgentScope 2.0 Is the Ideal Harness Runtime for Managed Agents

AgentScope 2.0 provides a stable, sandbox‑isolated runtime for Managed Agents by separating Brain (reasoning) and Hands (tool execution), offering multi‑tenant control, session persistence, flexible worker modes (local, cloud sandbox, self‑hosted), and a clear multi‑agent orchestration model for enterprise AI workloads.

AI RuntimeAgentScopeCloud Native
0 likes · 28 min read
Why AgentScope 2.0 Is the Ideal Harness Runtime for Managed Agents
dbaplus Community
dbaplus Community
Jul 26, 2026 · Cloud Native

Will AI Replace Kubernetes? Co‑Founder Brendan Burns on Its Rise and End

Brendan Burns recounts how he convinced Google to back Kubernetes, built the MVP in five days, navigated open‑source governance, tackled technical challenges like Etcd and declarative design, expanded the platform for AI workloads, and reflects on why even successful software like Kubernetes inevitably faces obsolescence.

AI workloadsCloud NativeKubernetes
0 likes · 31 min read
Will AI Replace Kubernetes? Co‑Founder Brendan Burns on Its Rise and End
IT Learning Made Simple
IT Learning Made Simple
Jul 25, 2026 · Industry Insights

What’s New in the 2026 System Architecture Designer Exam Syllabus?

The article explains why the exam syllabus is the first step for candidates, outlines the 2026 changes—including added cloud‑native, DevOps, and data‑architecture topics, refined microservice details, and reduced traditional software‑engineering content—shows the new weight distribution, and offers concrete study order, time‑allocation, learning methods, impact analysis, and resource recommendations.

Cloud NativeDevOpsdata architecture
0 likes · 8 min read
What’s New in the 2026 System Architecture Designer Exam Syllabus?
DataFunSummit
DataFunSummit
Jul 25, 2026 · Cloud Native

Evolution of Agent Infrastructure: Engineering Insights from Tencent Cloud Agent Runtime

The article analyzes how agents transition from demo to production, revealing that beyond model capabilities, stability, elasticity, security, and governance become critical, and explains the engineering challenges and solutions—including session management, state persistence, scheduling mismatches, sandbox isolation, and open‑source strategies—that underpin Tencent Cloud's Agent Runtime.

Agent RuntimeCloud NativeKubernetes
0 likes · 26 min read
Evolution of Agent Infrastructure: Engineering Insights from Tencent Cloud Agent Runtime
Ops Development Stories
Ops Development Stories
Jul 25, 2026 · Cloud Native

Practical Guide to Pyrra: The Kubernetes‑Native SLO Monitoring Tool

This comprehensive guide explains how Pyrra extends Sloth by providing a full SLO platform for Kubernetes, covering its architecture, four SLI types, rule generation, Web UI features, alert configuration, deployment options, Grafana integration, advanced usage, common pitfalls, and a detailed comparison to help you choose the right tool for reliable service monitoring.

Cloud NativeKubernetesPrometheus
0 likes · 24 min read
Practical Guide to Pyrra: The Kubernetes‑Native SLO Monitoring Tool
Alibaba Cloud Native
Alibaba Cloud Native
Jul 24, 2026 · Cloud Native

How Higress Serverless Enterprise Cuts Costs 90% and Boosts Auth Performance 30×

A SaaS platform’s consumer count surged from 200 to nearly 20,000, causing open‑source Higress authentication latency to jump 34‑fold and configuration size to balloon 8,457‑fold, while a local comparative test shows the Higress Enterprise Serverless edition maintains sub‑2 ms latency, 100 % success, tiny config footprints, and up to 90 % lower annual costs.

API GatewayCloud NativeHigress
0 likes · 10 min read
How Higress Serverless Enterprise Cuts Costs 90% and Boosts Auth Performance 30×
YiSu Grain
YiSu Grain
Jul 24, 2026 · Fundamentals

Enterprise Architecture vs Layered, SOA, Microservices & Cloud‑Native: Which to Pick?

The article explains how enterprise architecture, layered architecture, SOA, microservices and cloud‑native each address different scales of system design, shows how they can coexist in a retail modernization roadmap, and guides architects on selecting and integrating the appropriate approach for each problem domain.

Cloud NativeSOAarchitecture evolution
0 likes · 29 min read
Enterprise Architecture vs Layered, SOA, Microservices & Cloud‑Native: Which to Pick?
Su San Talks Tech
Su San Talks Tech
Jul 24, 2026 · Backend Development

Why More Teams Are Choosing Spring WebFlux for High‑Concurrency Applications

The article explains how Spring WebFlux replaces the thread‑per‑request model with an event‑loop, non‑blocking I/O and reactive streams, offering higher resource efficiency, built‑in back‑pressure, and better support for real‑time data, while also discussing its drawbacks and when virtual threads may be a viable alternative.

Cloud NativeR2DBCReactive Programming
0 likes · 18 min read
Why More Teams Are Choosing Spring WebFlux for High‑Concurrency Applications
Alibaba Cloud Native
Alibaba Cloud Native
Jul 23, 2026 · Backend Development

How an Agent Collaboration Failure Revealed RocketMQ’s AI‑Era Upgrade

The article dissects a multi‑Agent workflow that stalls for minutes, exposing why traditional message queues cannot handle AI‑driven long‑running, stateful sessions and how RocketMQ’s 5.x LiteTopic, event‑driven pull, and Suspend consumption model redesign the communication paradigm for AI workloads.

AICloud NativeEvent-Driven Pull
0 likes · 13 min read
How an Agent Collaboration Failure Revealed RocketMQ’s AI‑Era Upgrade
Top Architect
Top Architect
Jul 23, 2026 · Backend Development

How Taobao’s Backend Architecture Evolved Over a Decade

The article walks through Taobao’s backend architecture transformation from a single‑server setup to a cloud‑native, micro‑service ecosystem, detailing fourteen evolutionary stages—including separate Tomcat and DB, caching, load balancing, sharding, NoSQL, ESB, containerization, and cloud deployment—while highlighting key concepts, challenges, and design principles.

CachingCloud Nativebackend architecture
0 likes · 23 min read
How Taobao’s Backend Architecture Evolved Over a Decade
YiSu Grain
YiSu Grain
Jul 22, 2026 · Cloud Native

Day 31: Distinguishing Elasticity, Resilience, and Observability in Cloud‑Native Architecture

Moving an application to cloud VMs and Docker does not automatically grant cloud‑native capabilities; this article explains the seven cloud‑native principles—service‑orientation, elasticity, observability, resilience, full automation, zero‑trust, and continuous evolution—using concrete e‑commerce scenarios, tables, and step‑by‑step guidance to show how each principle solves specific problems and how they interrelate.

AutomationCloud NativeResilience
0 likes · 32 min read
Day 31: Distinguishing Elasticity, Resilience, and Observability in Cloud‑Native Architecture
IT Learning Made Simple
IT Learning Made Simple
Jul 22, 2026 · Fundamentals

Architect’s Reading List: From Beginner to Master

This article presents a curated reading list for software architects, organized by career stages and covering design patterns, code quality, architecture, distributed systems, cloud‑native topics, along with reading principles, recommendations, and a top‑10 book ranking to guide continuous learning.

Cloud NativeDesign Patternsbook list
0 likes · 10 min read
Architect’s Reading List: From Beginner to Master
JD Cloud Developers
JD Cloud Developers
Jul 22, 2026 · Cloud Native

How AI Quickly Reads Your Codebase: Three Evolutions of Joy-Code-Graph Cloud Service

The article explains how Joy-Code-Graph transforms AI code assistants from blind guesswork into globally aware tools by deploying a self‑hosted, cloud‑native code graph service that integrates directly with Joygen, offers zero‑install sandbox access, and persistently stores the graph in a dedicated repository branch.

AI programmingCloud NativeJoygen integration
0 likes · 10 min read
How AI Quickly Reads Your Codebase: Three Evolutions of Joy-Code-Graph Cloud Service
Cloud Architecture
Cloud Architecture
Jul 21, 2026 · Cloud Native

Stop Blindly Choosing Service Discovery: Deep Comparison of Eureka, Nacos & ZooKeeper

The article recounts a real‑world outage caused by Eureka’s self‑protection mode, then establishes five key evaluation dimensions for service discovery, provides a detailed side‑by‑side analysis of Eureka, ZooKeeper and Nacos—including design goals, failure modes, scalability and operational features—and offers concrete guidance on selecting, configuring and migrating to the most suitable registry for production micro‑service environments.

Cloud NativeEurekaNacos
0 likes · 35 min read
Stop Blindly Choosing Service Discovery: Deep Comparison of Eureka, Nacos & ZooKeeper