Tagged articles

cloud native

3345 articles · Page 1 of 34
Alibaba Cloud Native
Alibaba Cloud Native
Aug 23, 2026 · Artificial Intelligence

Why Is GPU Utilization Low? Try This Zero‑Intrusion AI Profiling Tool

The article introduces SysOM AI Profiling, a zero‑intrusive, cloud‑native performance observation and diagnosis solution for AI workloads that spans training to inference, single‑GPU to multi‑GPU clusters, and Python to GPU kernel layers, helping users pinpoint low GPU utilization, memory leaks, and communication bottlenecks.

AI profilingGPU utilizationcloud native
0 likes · 14 min read
Why Is GPU Utilization Low? Try This Zero‑Intrusion AI Profiling Tool
ITPUB
ITPUB
Aug 20, 2026 · Industry Insights

2026 China Database Technology Conference Launches: Data Fusion and AI Leadership

The 17th China Database Technology Conference (DTCC 2026) ran from August 20‑22 in Beijing, gathering top experts to discuss database kernel innovations, cloud‑native and distributed practices, AI‑driven data, vector databases, real‑time warehouses, and the emerging Agent era, while showcasing cutting‑edge solutions from Dameng, Tencent Cloud, Alibaba Cloud, OceanBase and GoldenDB.

AIAgentData Lake
0 likes · 15 min read
2026 China Database Technology Conference Launches: Data Fusion and AI Leadership
Airbnb Technology Team
Airbnb Technology Team
Aug 20, 2026 · Cloud Native

How Airbnb Built a Scalable, Reliable Kubernetes Sidecar for Dynamic Configuration

The article explains Airbnb's Sitar‑agent sidecar architecture, detailing the end‑to‑end configuration distribution lifecycle, key design choices such as sidecar versus in‑process deployment, pull‑model optimizations, and the migration from Sparkey to SQLite for robust, multi‑language support at massive scale.

Dynamic ConfigurationKubernetesRocksDB
0 likes · 15 min read
How Airbnb Built a Scalable, Reliable Kubernetes Sidecar for Dynamic Configuration
Tencent Cloud Middleware
Tencent Cloud Middleware
Aug 19, 2026 · Operations

How AI Gateway Makes Large-Model Calls Visible, Traceable, and Auditable

Enterprises deploying large-model APIs often struggle to see token usage, latency, and errors; the AI Gateway embeds metrics, structured logs, and distributed tracing at the gateway layer, providing token-level insights, request-level latency breakdowns, and full-chain auditability without code changes, as demonstrated in a real-world incident.

AI GatewayLLMLogging
0 likes · 17 min read
How AI Gateway Makes Large-Model Calls Visible, Traceable, and Auditable
java1234
java1234
Aug 18, 2026 · Backend Development

Why Is Jakarta EE Gaining Traction in Enterprise Java?

Jakarta EE, the rebranded Java EE now governed by the Eclipse Foundation, offers open standards, built‑in enterprise capabilities, cloud‑native improvements, a smooth upgrade path for legacy systems, and familiar APIs, making it an attractive alternative to Spring for long‑term, maintainable applications.

CDIEnterprise JavaJPA
0 likes · 9 min read
Why Is Jakarta EE Gaining Traction in Enterprise Java?
Smart Sea Tide
Smart Sea Tide
Aug 18, 2026 · Big Data

How to Choose and Architect a Data Lake Platform for Enterprise Digital Transformation

The article outlines the strategic need for a unified data lake in a digital‑focused enterprise, details functional and non‑functional requirements such as linear scalability, real‑time and batch processing, multi‑tenant support, security and governance, and presents a comprehensive architecture design that integrates storage, compute, and management components.

Big DataData Lakecloud native
0 likes · 19 min read
How to Choose and Architect a Data Lake Platform for Enterprise Digital Transformation
21CTO
21CTO
Aug 16, 2026 · Artificial Intelligence

Microsoft Joins Google in Adding Go Support for AI Agent Development

The article explains how Microsoft’s new Agent Framework for Go extends native AI agent capabilities—such as large‑model access, tool calls, and multi‑agent coordination—to the Go ecosystem, reflecting broader industry moves by Google and the growing demand for Go‑centric cloud‑native AI development.

AI AgentsAgent FrameworkAzure OpenAI
0 likes · 6 min read
Microsoft Joins Google in Adding Go Support for AI Agent Development
Cloud Architecture
Cloud Architecture
Aug 14, 2026 · Cloud Native

Complete Guide to Go Microservice Logging and Tracing with OpenTelemetry (Industrial‑Grade Solution)

When an alarm rang at 2:17 AM, a Go order service’s P99 latency surged from 220 ms to 4.6 s and its error rate climbed to 1.8 %; the article explains why many teams still see limited value after adopting OpenTelemetry, identifies three missing pieces—stable trace IDs, end‑to‑end context propagation, and production‑ready pipelines—and delivers a step‑by‑step, code‑first blueprint for building an industrial‑grade observability stack that scales in Kubernetes.

GoLoggingOpenTelemetry
0 likes · 40 min read
Complete Guide to Go Microservice Logging and Tracing with OpenTelemetry (Industrial‑Grade Solution)
TechVision Expert Circle
TechVision Expert Circle
Aug 13, 2026 · Cloud Native

Does a Three‑Hour Commute Actually Boost Anyone’s Efficiency?

A 2026 OPM survey of over 600,000 federal workers shows that mandatory office returns cut satisfaction and raise turnover, while modern cloud‑native collaboration tools and AI‑driven workflows enable remote teams to match or exceed on‑site productivity, revealing the hidden technical costs of forced commuting.

AI collaborationCRDTasynchronous workflow
0 likes · 13 min read
Does a Three‑Hour Commute Actually Boost Anyone’s Efficiency?
Alibaba Cloud Native
Alibaba Cloud Native
Aug 13, 2026 · Artificial Intelligence

Which Model Should Handle Your Request? Alibaba Cloud AI Gateway Intelligent Routing Goes Live

As the number of LLMs grows, developers face the dilemma of selecting the right model for each request; Alibaba Cloud’s AI Gateway introduces intelligent routing that evaluates model capabilities, cost, speed, and task suitability, making a single, optimal decision per request while preserving context in multi‑turn and Agent workflows.

AI GatewayAlibaba CloudIntelligent Routing
0 likes · 11 min read
Which Model Should Handle Your Request? Alibaba Cloud AI Gateway Intelligent Routing Goes Live
Woodpecker Software Testing
Woodpecker Software Testing
Aug 13, 2026 · Industry Insights

2026 Stress‑Testing ROI: When Is the Investment Worth It?

The article analyzes how AI‑assisted scenario generation, chaos‑as‑a‑service, and observability reshape stress‑testing costs in 2026, presenting ROI models, industry benchmarks, and critical thresholds that turn testing from a risk hedge into a growth lever.

AI-generated trafficROIchaos engineering
0 likes · 8 min read
2026 Stress‑Testing ROI: When Is the Investment Worth It?
Alibaba Cloud Native
Alibaba Cloud Native
Aug 12, 2026 · Cloud Native

Alibaba Cloud and Datadog Release OpenTelemetry Go Compile‑Time Instrumentation v1 for Zero‑Code Observability

The OpenTelemetry Go Compile‑Time Instrumentation project, jointly launched by Alibaba Cloud and Datadog, fills the last observability gap for Go by injecting tracing and metrics code at build time, offering zero‑code instrumentation, no runtime overhead, and seamless CI/CD integration while comparing it with manual and eBPF approaches.

Compile-Time InstrumentationGoOpenTelemetry
0 likes · 10 min read
Alibaba Cloud and Datadog Release OpenTelemetry Go Compile‑Time Instrumentation v1 for Zero‑Code Observability
Woodpecker Software Testing
Woodpecker Software Testing
Aug 12, 2026 · Cloud Native

Distributed vs Monolithic: Uncovering the Real Performance Trade‑offs

The article analyzes distributed and traditional monolithic architectures across response latency, throughput, scalability, and fault‑tolerance cost, revealing that distributed systems introduce network latency, serialization overhead, higher resource consumption, and complex failure modes that often offset their scalability benefits, and provides concrete case studies from Netflix, Alibaba, and LinkedIn.

Throughputcloud nativedistributed systems
0 likes · 7 min read
Distributed vs Monolithic: Uncovering the Real Performance Trade‑offs
samdeepthink
samdeepthink
Aug 12, 2026 · Operations

Why Observability Is More Than Monitoring: Finding the Root Cause Quickly

The article explains that observability goes beyond simple monitoring by combining metrics, logs, and traces to pinpoint where and why a system issue occurs, especially in microservice and cloud‑native environments, and stresses the importance of correlating data rather than merely collecting more.

AIOpsMicroservicescloud native
0 likes · 3 min read
Why Observability Is More Than Monitoring: Finding the Root Cause Quickly
Java Architecture Diary
Java Architecture Diary
Aug 12, 2026 · Cloud Native

Why Upgrading Your MCP Server to 2.0 Solves Stateless Session Issues

The article explains how MCP 1.x's stateful handshake caused node‑crash failures, sticky sessions, and serverless incompatibility, and how the 2.0 release removes the handshake, makes each request self‑describing via _meta and HTTP headers, introduces MRTR for multi‑round interactions, and provides a Java/TypeScript code walkthrough demonstrating the new stateless behavior.

JavaKubernetesMCP
0 likes · 8 min read
Why Upgrading Your MCP Server to 2.0 Solves Stateless Session Issues
Ray's Galactic Tech
Ray's Galactic Tech
Aug 10, 2026 · Cloud Native

Destruction and Rebirth: Deep Dive into ETCD Backup and Restore for Kubernetes Clusters

This article walks through a real‑world ETCD failure, explains why ETCD is the control‑plane brain, details the three‑layer ETCD architecture, exposes common backup pitfalls, and provides a production‑grade backup‑restore workflow—including snapshot API usage, Go implementation, verification steps, and post‑restore validation—for reliable Kubernetes disaster recovery.

GoKubernetesRestore
0 likes · 33 min read
Destruction and Rebirth: Deep Dive into ETCD Backup and Restore for Kubernetes Clusters
Java Architect Handbook
Java Architect Handbook
Aug 10, 2026 · Cloud Native

Tired of XXL‑Job? Try This Elegant Nacos‑Based Scheduling Solution

The article analyses why XXL‑Job’s separate registration, configuration, and weak sharding cause state inconsistency, observability gaps, and duplicate processing, then proposes JobFlow – a lightweight scheduler that removes redundant components, adds full‑traceId tracing, true sharding with distributed locks, exponential retry, and cloud‑native configuration managed by Nacos, all illustrated with concrete code snippets and deployment diagrams.

JavaNacosTask Scheduling
0 likes · 21 min read
Tired of XXL‑Job? Try This Elegant Nacos‑Based Scheduling Solution
Alibaba Cloud Infrastructure
Alibaba Cloud Infrastructure
Aug 9, 2026 · Cloud Native

How ACK One Fleet Transforms Agent Sandbox from Single-Cluster to Multi-Cluster

The article explains how ACK One Fleet upgrades the AI Agent Sandbox from a single‑cluster Kubernetes setup to a multi‑cluster architecture, addressing capacity limits, fault‑domain risks, and scheduling inefficiencies while providing global capacity control, water‑level balancing, fault‑tolerant failover, and faster sandbox startup through E2B and CRD integrations.

ACK OneAgent SandboxE2B
0 likes · 11 min read
How ACK One Fleet Transforms Agent Sandbox from Single-Cluster to Multi-Cluster
Woodpecker Software Testing
Woodpecker Software Testing
Aug 8, 2026 · Operations

Shift‑Left Performance Testing vs Traditional: A Must‑Read for Test Experts

The article examines how shift‑left performance testing moves responsibility earlier in the development lifecycle, contrasting it with traditional post‑release load testing, and demonstrates through real‑world cases how this approach improves efficiency, cost, quality, and collaboration in cloud‑native, high‑availability systems.

CI/CDDevOpsMicroservices
0 likes · 8 min read
Shift‑Left Performance Testing vs Traditional: A Must‑Read for Test Experts
Alibaba Cloud Native
Alibaba Cloud Native
Aug 8, 2026 · Backend Development

Designing AI‑Friendly Backend Architecture for 24/7 Unattended Development

The article outlines a comprehensive roadmap for transforming traditional backend systems into AI‑friendly architectures, introducing concepts such as Architecture Maps, Service Cards, SKILL packages, multi‑layered testing, permission tiers, and a Harness framework to enable reliable, 24/7 unattended AI‑driven development and operations.

AIBackend DevelopmentDevOps
0 likes · 39 min read
Designing AI‑Friendly Backend Architecture for 24/7 Unattended Development
Cloud Native Technology Community
Cloud Native Technology Community
Aug 6, 2026 · Cloud Native

5 Production Challenges for Running AI Workloads on Kubernetes: From GPU Scheduling to Observability

Running AI workloads on Kubernetes introduces five production‑grade challenges—complex GPU and accelerator management, workload‑aware scheduling, inference autoscaling beyond CPU metrics, multi‑layer observability, and Day 2 governance—requiring platform teams to extend their capabilities beyond traditional container operations.

AI workloadsDay 2 operationsGPU Scheduling
0 likes · 10 min read
5 Production Challenges for Running AI Workloads on Kubernetes: From GPU Scheduling to Observability
Alibaba Cloud Native
Alibaba Cloud Native
Aug 6, 2026 · Artificial Intelligence

OpenAgentPack: Managing and Migrating Cloud AI Agents Like Code

OpenAgentPack is an open‑source tool that lets you describe a cloud AI Agent’s entire workflow—including model, environment, skills, MCP, knowledge files and credentials—in a single agents.yaml file, version it in Git, preview changes with validate and plan commands, and redeploy the agent across providers while preserving reproducibility, collaboration and auditability.

AI AgentsCLIGitOps
0 likes · 7 min read
OpenAgentPack: Managing and Migrating Cloud AI Agents Like Code
System Architect Go
System Architect Go
Aug 5, 2026 · Cloud Native

Kubernetes Chronicle: From Borg to the Cloud‑Native Operating System

This article traces Kubernetes from its roots in Google’s Borg system through Docker’s rise, the open‑source launch, CNCF stewardship, key feature milestones like Deployments, CRDs, and Gateway API, and explains why it became the default cloud‑native orchestration platform.

BorgCRDContainer Orchestration
0 likes · 20 min read
Kubernetes Chronicle: From Borg to the Cloud‑Native Operating System
Alibaba Cloud Native
Alibaba Cloud Native
Aug 4, 2026 · Artificial Intelligence

AI Innovation Forum Shanghai: Key Takeaways, Multi‑Agent Architecture, and PPT Resources

The AI Innovation Practice Forum in Shanghai gathered over 70 tech professionals to present deep dives on multi‑agent governance, the Agent Native Cloud three‑layer model, AgentTeams collaboration platform, AgentLoop lifecycle flywheel, a cloud‑native network foundation, and next‑gen AIOps, with PPTs available for download.

AI AgentsAIOpsAgent Native Cloud
0 likes · 6 min read
AI Innovation Forum Shanghai: Key Takeaways, Multi‑Agent Architecture, and PPT Resources
Alibaba Cloud Native
Alibaba Cloud Native
Aug 4, 2026 · Operations

From Building Wheels to Embedding OpenAPI: Jingchen’s Choice of an Intelligent Ops Foundation

Facing exploding system complexity, unclear global topology, fragmented observability data, and noisy alerts, Jingchen migrated its full‑stack to the cloud and adopted Alibaba Cloud STAROps, a unified CMS 2.0 data base, UModel digital‑twin topology, and OpenAPI‑driven AI diagnostics to turn heavy‑lifting ops work into an automated, business‑focused capability.

Intelligent OperationsOpenAPISRE
0 likes · 10 min read
From Building Wheels to Embedding OpenAPI: Jingchen’s Choice of an Intelligent Ops Foundation
Alibaba Cloud Native
Alibaba Cloud Native
Aug 3, 2026 · Artificial Intelligence

Building a Financial‑Grade AI Agent Platform with AgentScope: A Practical Whitepaper

FinXScope, a financial‑grade AI‑native agent base built on AgentScope Java, serves as the core engine of the Agent Harness system, offering multi‑agent orchestration, dual‑mode execution, six‑layer architecture, high‑availability, security, observability and low‑code to high‑code pathways, and has already been adopted by dozens of leading financial institutions.

AI AgentsAgentScopeFinXScope
0 likes · 33 min read
Building a Financial‑Grade AI Agent Platform with AgentScope: A Practical Whitepaper
ITPUB
ITPUB
Aug 3, 2026 · Industry Insights

Why Docker’s Patchwork of Legacy Tech Still Dominates Cloud Computing

Despite being built on decades‑old components such as chroot, Linux namespaces, SLIRP and QEMU, Docker has become the de‑facto operating system of the cloud because it demands almost no changes to developers' existing workflows, offering a pragmatic, “good enough” solution that scales globally.

DockerEngineeringLegacy Technology
0 likes · 16 min read
Why Docker’s Patchwork of Legacy Tech Still Dominates Cloud Computing
Golang Shines
Golang Shines
Aug 3, 2026 · Cloud Native

How I Built a Production‑Ready HA Kubernetes Cluster in Minutes

When my manager suddenly demanded a production‑grade, highly available Kubernetes cluster integrated with a private Harbor registry, I followed a comprehensive step‑by‑step guide to finish the entire setup within a few hours, and now share the 83‑page manual for anyone to replicate.

Cluster DeploymentHarborKubernetes
0 likes · 3 min read
How I Built a Production‑Ready HA Kubernetes Cluster in Minutes
Architect Chen
Architect Chen
Aug 2, 2026 · Cloud Native

All Essential kubectl Commands for 2026: A Complete Guide

This article provides a concise, step‑by‑step reference of the most frequently used kubectl commands—including get, describe, logs, exec, apply, port‑forward, rollout, scale, and delete—showing exact syntax and typical use cases for managing Kubernetes resources.

Command LineContainer ManagementDevOps
0 likes · 4 min read
All Essential kubectl Commands for 2026: A Complete Guide
Alibaba Cloud Native
Alibaba Cloud Native
Aug 2, 2026 · Cloud Native

From Visibility to Self‑Healing: ChangjieTong’s Observability and AI‑Powered Ops Journey

ChangjieTong transformed its SaaS‑based, multi‑tenant finance cloud platform by building a five‑layer observability stack on Alibaba Cloud CloudMonitor 2.0, integrating a UModel digital‑twin, and deploying AI‑driven inspection, self‑healing and capacity‑prediction loops, which lifted SLA from 99.9% to 99.995% and cut average fault‑resolution time from over 10 minutes to under 30 seconds.

AIOpsCapacity PredictionDigital Twin
0 likes · 15 min read
From Visibility to Self‑Healing: ChangjieTong’s Observability and AI‑Powered Ops Journey
MaGe Linux Operations
MaGe Linux Operations
Jul 31, 2026 · Cloud Native

Choosing an Ingress Controller: Production Comparison of NGINX, Traefik, and APISIX

This article presents a production‑grade comparison of three Kubernetes Ingress controllers—NGINX, Traefik, and APISIX—by defining a four‑layer evaluation framework, detailing pre‑deployment checks, configuration examples, testing scripts, performance metrics, and rollout/rollback procedures to help teams select the most suitable solution.

APISIXKubernetesTraefik
0 likes · 24 min read
Choosing an Ingress Controller: Production Comparison of NGINX, Traefik, and APISIX
Alibaba Cloud Native
Alibaba Cloud Native
Jul 30, 2026 · Cloud Native

Replication‑Free Failover for RocketMQ: Achieving Second‑Level Takeover Without Data Copy

The ACM FSE‑2026 industry paper introduces a replication‑free failover mechanism for cloud‑native stateful services like Apache RocketMQ, using protocol‑level write isolation and multi‑attach storage to achieve second‑level recovery without extra data copies, while maintaining low cost and near‑native throughput.

Protocol FencingReplication-Free FailoverRocketMQ
0 likes · 8 min read
Replication‑Free Failover for RocketMQ: Achieving Second‑Level Takeover Without Data Copy
DevOps Operations Practice
DevOps Operations Practice
Jul 30, 2026 · Operations

Essential Velero Guide for Kubernetes Disaster Recovery

This article walks through using Velero to back up, restore, and migrate Kubernetes clusters, covering MinIO installation, Velero client and server setup, storage volume creation, backup location configuration, and execution of backup, restore, and scheduled backup commands.

KubernetesMinIOOperations
0 likes · 11 min read
Essential Velero Guide for Kubernetes Disaster Recovery
Alibaba Cloud Native
Alibaba Cloud Native
Jul 29, 2026 · Artificial Intelligence

Launching a Multi‑Agent AI Platform in One Week with Alibaba Cloud AgentTeams and AI Gateway

In just one week, XinYongZhongHe partnered with Alibaba Cloud to build an enterprise‑grade multi‑agent AI platform using AgentTeams and the AI Gateway, enabling coordinated agents, unified model governance, secure access, cost control, and a range of internal services that move AI from simple Q&A to task execution.

AI GatewayAI governanceAgentTeams
0 likes · 10 min read
Launching a Multi‑Agent AI Platform in One Week with Alibaba Cloud AgentTeams and AI Gateway
YiSu Grain
YiSu Grain
Jul 29, 2026 · Fundamentals

45‑Minute Review: Master the Eight Architecture Types for the Soft Exam

This article guides readers through a 45‑minute review of the eight architecture categories—information‑system, layered, cloud‑native, SOA, embedded, communication, security, and big‑data—showing how to identify the relevant type from a problem statement, combine multiple architectures in a solution, and articulate the problem, solution, rationale, and cost with concrete examples and tables.

Big DataSOASoftware Architecture
0 likes · 32 min read
45‑Minute Review: Master the Eight Architecture Types for the Soft Exam
Alibaba Cloud Native
Alibaba Cloud Native
Jul 27, 2026 · Artificial Intelligence

Why AgentScope 2.0 Is the Ideal Harness Runtime for Managed Agents

AgentScope 2.0 provides a stable, sandbox‑isolated runtime for Managed Agents by separating Brain (reasoning) and Hands (tool execution), offering multi‑tenant control, session persistence, flexible worker modes (local, cloud sandbox, self‑hosted), and a clear multi‑agent orchestration model for enterprise AI workloads.

AI RuntimeAgentScopeJava SDK
0 likes · 28 min read
Why AgentScope 2.0 Is the Ideal Harness Runtime for Managed Agents
dbaplus Community
dbaplus Community
Jul 26, 2026 · Cloud Native

Will AI Replace Kubernetes? Co‑Founder Brendan Burns on Its Rise and End

Brendan Burns recounts how he convinced Google to back Kubernetes, built the MVP in five days, navigated open‑source governance, tackled technical challenges like Etcd and declarative design, expanded the platform for AI workloads, and reflects on why even successful software like Kubernetes inevitably faces obsolescence.

AI workloadsKubernetesOpen Source
0 likes · 31 min read
Will AI Replace Kubernetes? Co‑Founder Brendan Burns on Its Rise and End
IT Learning Made Simple
IT Learning Made Simple
Jul 25, 2026 · Industry Insights

What’s New in the 2026 System Architecture Designer Exam Syllabus?

The article explains why the exam syllabus is the first step for candidates, outlines the 2026 changes—including added cloud‑native, DevOps, and data‑architecture topics, refined microservice details, and reduced traditional software‑engineering content—shows the new weight distribution, and offers concrete study order, time‑allocation, learning methods, impact analysis, and resource recommendations.

DevOpsMicroservicesSystem Architecture
0 likes · 8 min read
What’s New in the 2026 System Architecture Designer Exam Syllabus?
DataFunSummit
DataFunSummit
Jul 25, 2026 · Cloud Native

Evolution of Agent Infrastructure: Engineering Insights from Tencent Cloud Agent Runtime

The article analyzes how agents transition from demo to production, revealing that beyond model capabilities, stability, elasticity, security, and governance become critical, and explains the engineering challenges and solutions—including session management, state persistence, scheduling mismatches, sandbox isolation, and open‑source strategies—that underpin Tencent Cloud's Agent Runtime.

Agent RuntimeKubernetesRL Training
0 likes · 26 min read
Evolution of Agent Infrastructure: Engineering Insights from Tencent Cloud Agent Runtime
Ops Development Stories
Ops Development Stories
Jul 25, 2026 · Cloud Native

Practical Guide to Pyrra: The Kubernetes‑Native SLO Monitoring Tool

This comprehensive guide explains how Pyrra extends Sloth by providing a full SLO platform for Kubernetes, covering its architecture, four SLI types, rule generation, Web UI features, alert configuration, deployment options, Grafana integration, advanced usage, common pitfalls, and a detailed comparison to help you choose the right tool for reliable service monitoring.

KubernetesPyrraSLO
0 likes · 24 min read
Practical Guide to Pyrra: The Kubernetes‑Native SLO Monitoring Tool
Alibaba Cloud Native
Alibaba Cloud Native
Jul 24, 2026 · Cloud Native

How Higress Serverless Enterprise Cuts Costs 90% and Boosts Auth Performance 30×

A SaaS platform’s consumer count surged from 200 to nearly 20,000, causing open‑source Higress authentication latency to jump 34‑fold and configuration size to balloon 8,457‑fold, while a local comparative test shows the Higress Enterprise Serverless edition maintains sub‑2 ms latency, 100 % success, tiny config footprints, and up to 90 % lower annual costs.

API-gatewayHigressPerformance
0 likes · 10 min read
How Higress Serverless Enterprise Cuts Costs 90% and Boosts Auth Performance 30×
YiSu Grain
YiSu Grain
Jul 24, 2026 · Fundamentals

Enterprise Architecture vs Layered, SOA, Microservices & Cloud‑Native: Which to Pick?

The article explains how enterprise architecture, layered architecture, SOA, microservices and cloud‑native each address different scales of system design, shows how they can coexist in a retail modernization roadmap, and guides architects on selecting and integrating the appropriate approach for each problem domain.

MicroservicesSOAarchitecture evolution
0 likes · 29 min read
Enterprise Architecture vs Layered, SOA, Microservices & Cloud‑Native: Which to Pick?
Su San Talks Tech
Su San Talks Tech
Jul 24, 2026 · Backend Development

Why More Teams Are Choosing Spring WebFlux for High‑Concurrency Applications

The article explains how Spring WebFlux replaces the thread‑per‑request model with an event‑loop, non‑blocking I/O and reactive streams, offering higher resource efficiency, built‑in back‑pressure, and better support for real‑time data, while also discussing its drawbacks and when virtual threads may be a viable alternative.

R2DBCReactive ProgrammingSpring WebFlux
0 likes · 18 min read
Why More Teams Are Choosing Spring WebFlux for High‑Concurrency Applications
Alibaba Cloud Native
Alibaba Cloud Native
Jul 23, 2026 · Backend Development

How an Agent Collaboration Failure Revealed RocketMQ’s AI‑Era Upgrade

The article dissects a multi‑Agent workflow that stalls for minutes, exposing why traditional message queues cannot handle AI‑driven long‑running, stateful sessions and how RocketMQ’s 5.x LiteTopic, event‑driven pull, and Suspend consumption model redesign the communication paradigm for AI workloads.

AIEvent-Driven PullLiteTopic
0 likes · 13 min read
How an Agent Collaboration Failure Revealed RocketMQ’s AI‑Era Upgrade
Top Architect
Top Architect
Jul 23, 2026 · Backend Development

How Taobao’s Backend Architecture Evolved Over a Decade

The article walks through Taobao’s backend architecture transformation from a single‑server setup to a cloud‑native, micro‑service ecosystem, detailing fourteen evolutionary stages—including separate Tomcat and DB, caching, load balancing, sharding, NoSQL, ESB, containerization, and cloud deployment—while highlighting key concepts, challenges, and design principles.

Microservicesbackend architecturecaching
0 likes · 23 min read
How Taobao’s Backend Architecture Evolved Over a Decade
YiSu Grain
YiSu Grain
Jul 22, 2026 · Cloud Native

Day 31: Distinguishing Elasticity, Resilience, and Observability in Cloud‑Native Architecture

Moving an application to cloud VMs and Docker does not automatically grant cloud‑native capabilities; this article explains the seven cloud‑native principles—service‑orientation, elasticity, observability, resilience, full automation, zero‑trust, and continuous evolution—using concrete e‑commerce scenarios, tables, and step‑by‑step guidance to show how each principle solves specific problems and how they interrelate.

Microservicesautomationcloud native
0 likes · 32 min read
Day 31: Distinguishing Elasticity, Resilience, and Observability in Cloud‑Native Architecture
IT Learning Made Simple
IT Learning Made Simple
Jul 22, 2026 · Fundamentals

Architect’s Reading List: From Beginner to Master

This article presents a curated reading list for software architects, organized by career stages and covering design patterns, code quality, architecture, distributed systems, cloud‑native topics, along with reading principles, recommendations, and a top‑10 book ranking to guide continuous learning.

Design PatternsSoftware Architecturebook list
0 likes · 10 min read
Architect’s Reading List: From Beginner to Master
JD Cloud Developers
JD Cloud Developers
Jul 22, 2026 · Cloud Native

How AI Quickly Reads Your Codebase: Three Evolutions of Joy-Code-Graph Cloud Service

The article explains how Joy-Code-Graph transforms AI code assistants from blind guesswork into globally aware tools by deploying a self‑hosted, cloud‑native code graph service that integrates directly with Joygen, offers zero‑install sandbox access, and persistently stores the graph in a dedicated repository branch.

AI programmingJoygen integrationMCP
0 likes · 10 min read
How AI Quickly Reads Your Codebase: Three Evolutions of Joy-Code-Graph Cloud Service
Cloud Architecture
Cloud Architecture
Jul 21, 2026 · Cloud Native

Stop Blindly Choosing Service Discovery: Deep Comparison of Eureka, Nacos & ZooKeeper

The article recounts a real‑world outage caused by Eureka’s self‑protection mode, then establishes five key evaluation dimensions for service discovery, provides a detailed side‑by‑side analysis of Eureka, ZooKeeper and Nacos—including design goals, failure modes, scalability and operational features—and offers concrete guidance on selecting, configuring and migrating to the most suitable registry for production micro‑service environments.

EurekaMicroservicesNacos
0 likes · 35 min read
Stop Blindly Choosing Service Discovery: Deep Comparison of Eureka, Nacos & ZooKeeper
TechVision Expert Circle
TechVision Expert Circle
Jul 21, 2026 · Cloud Native

How to Build an Elastic Auto‑Scaling Cloud‑Native Application

After a 15‑fold traffic surge forced manual scaling of an e‑commerce platform, the team rebuilt the system with true elastic scaling—horizontal, vertical, and architectural—using Kubernetes, Envoy, KEDA, predictive autoscaling, and a comprehensive observability stack, achieving fully automated scaling from 12 to 80 pods in under 90 seconds and cutting peak resource costs by 60%.

Elastic ScalingKEDAKubernetes
0 likes · 13 min read
How to Build an Elastic Auto‑Scaling Cloud‑Native Application
Golang Shines
Golang Shines
Jul 20, 2026 · Cloud Native

7 Golden Rules for Building High‑Availability Cloud‑Native Go Services (Production‑Proven)

This article presents a step‑by‑step guide to building highly available cloud‑native Go systems, covering graceful error handling, structured logging, minimal dependencies, concurrency control, health checks, Raft‑based replication, timeout/retry strategies, circuit breaking, rate limiting, observability with Zap, Loki, Prometheus, OpenTelemetry, and future architectural directions.

GoLoggingMicroservices
0 likes · 18 min read
7 Golden Rules for Building High‑Availability Cloud‑Native Go Services (Production‑Proven)
Golang Shines
Golang Shines
Jul 19, 2026 · Cloud Native

10 Practical Tips to Quickly Build Cloud‑Native Apps with golang‑samples

This guide walks developers through ten hands‑on techniques for using Google Cloud's golang‑samples repository—covering environment setup, authentication methods, workflow orchestration, structured logging, AI integration, storage choices, serverless functions, secret management, observability, and automated testing and deployment—to accelerate production‑grade cloud‑native Go applications.

AI integrationGoGoogle Cloud
0 likes · 7 min read
10 Practical Tips to Quickly Build Cloud‑Native Apps with golang‑samples
Cloud Architecture
Cloud Architecture
Jul 17, 2026 · Cloud Native

Stop Hand‑Crafting ClusterRoles: Build a Production‑Grade Kubernetes RBAC Governance System with rbac‑manager

This article explains why manually managing ClusterRoles leads to governance chaos in Kubernetes, introduces rbac‑manager as a declarative controller that centralises binding creation, recycling and auditing, and provides a step‑by‑step guide with real‑world examples to build a scalable, production‑ready RBAC management workflow.

Access ControlKubernetesRBAC
0 likes · 23 min read
Stop Hand‑Crafting ClusterRoles: Build a Production‑Grade Kubernetes RBAC Governance System with rbac‑manager
Golang Shines
Golang Shines
Jul 17, 2026 · Cloud Native

Building a Scalable Go Service Mesh from Scratch: Core Cloud‑Native Practices

This article walks through why Go is ideal for cloud‑native development and demonstrates step‑by‑step how to build a scalable service mesh, covering static compilation, HTTP services, Go modules, Gin/Gorilla APIs, configuration, logging, health checks, service registration, load balancing, sidecar proxies, traffic interception, circuit breaking, rate limiting, retries, and distributed tracing with OpenTelemetry.

GoKubernetesMicroservices
0 likes · 16 min read
Building a Scalable Go Service Mesh from Scratch: Core Cloud‑Native Practices
dbaplus Community
dbaplus Community
Jul 15, 2026 · Backend Development

Why Are More Teams Switching from RabbitMQ to NATS?

The article compares RabbitMQ and NATS, outlining RabbitMQ's maturity and operational complexity versus NATS's lightweight, high‑performance, cloud‑native design, and explains when each solution is appropriate for modern microservice, edge, and AI‑driven architectures.

JetStreamMessage QueueMicroservices
0 likes · 8 min read
Why Are More Teams Switching from RabbitMQ to NATS?
MaGe Linux Operations
MaGe Linux Operations
Jul 15, 2026 · Cloud Native

How to Schedule, Isolate, and Allocate GPUs in a Kubernetes Cluster

Even when GPU nodes show up with nvidia‑smi, Pods can stay pending, see all devices, or suffer memory spikes; this guide walks through the full GPU resource chain in Kubernetes, from PCIe detection and driver loading to Device Plugin registration, node labeling, affinity, taints, isolation levels, MIG, time‑slicing, quotas, monitoring, and safe upgrade procedures.

Device PluginGPUKubernetes
0 likes · 34 min read
How to Schedule, Isolate, and Allocate GPUs in a Kubernetes Cluster
Su San Talks Tech
Su San Talks Tech
Jul 15, 2026 · Artificial Intelligence

How Codex Transforms Java Development: From Theory to Real-World Projects

Codex, OpenAI’s cloud‑native software‑engineering agent, replaces the traditional write‑test‑fix cycle with an automated loop that can pull repositories, modify multiple files, run tests in isolated sandboxes, and output merge‑ready diffs, delivering 60‑75% speed gains for Java backend tasks when used with well‑crafted prompts and proper governance.

AI code generationJavaOpenAI Codex
0 likes · 34 min read
How Codex Transforms Java Development: From Theory to Real-World Projects
Ray's Galactic Tech
Ray's Galactic Tech
Jul 14, 2026 · Cloud Native

Spring Boot + Netty MQTT Gateway for Million Connections & Millisecond Push

To support millions of persistent MQTT connections with sub‑millisecond latency, the article walks through a Spring Boot + Netty cloud‑native gateway design that separates connection, event, state and governance planes, details async authentication, back‑pressure handling, command state machines, and loss‑less Kubernetes roll‑outs.

KafkaKubernetesMQTT
0 likes · 38 min read
Spring Boot + Netty MQTT Gateway for Million Connections & Millisecond Push
Alibaba Cloud Native
Alibaba Cloud Native
Jul 14, 2026 · Cloud Native

How a 24/7 AI Community Admin Handles PRs at 2 AM with AgentTeams

In just three weeks, the AgentTeams‑powered AI digital employee "github‑manager" automatically reviewed 108 pull requests, processed 48 issues, and reduced first‑response time from days to under an hour for the LoongSuite open‑source project, while documenting the architecture, challenges, and lessons learned.

AI automationAgentTeamsGitHub PR review
0 likes · 19 min read
How a 24/7 AI Community Admin Handles PRs at 2 AM with AgentTeams
Long Ge's Treasure Box
Long Ge's Treasure Box
Jul 14, 2026 · Databases

Understanding Serverless Databases: PlanetScale, Neon, Supabase, and Turso

This article explains what serverless databases are, compares them with traditional databases, and provides detailed overviews, core concepts, and code examples for four major services—PlanetScale, Neon, Supabase, and Turso—highlighting features such as automatic scaling, pay‑as‑you‑go pricing, and global replication.

DatabaseNeonPlanetScale
0 likes · 10 min read
Understanding Serverless Databases: PlanetScale, Neon, Supabase, and Turso
Ray's Galactic Tech
Ray's Galactic Tech
Jul 13, 2026 · Artificial Intelligence

When AI Agents Meet Cloud‑Native: Practical Multi‑Agent Orchestration for High‑Concurrency Scenarios

The article explains why naïve multi‑agent demos fail in production, defines the core concepts of Task, Step, Agent Role and Event, proposes a four‑plane cloud‑native architecture, shows concrete Go and Python code, and provides detailed guidance on state machines, reliability, observability, security and budget governance for building scalable, production‑grade AI agent systems.

AI AgentsKubernetesMulti-agent orchestration
0 likes · 36 min read
When AI Agents Meet Cloud‑Native: Practical Multi‑Agent Orchestration for High‑Concurrency Scenarios
Alibaba Cloud Native
Alibaba Cloud Native
Jul 13, 2026 · Artificial Intelligence

How Alibaba Cloud AgentTeams Enables Enterprise-Scale Multi-Agent Operations

AgentTeams tackles the long‑term operation of enterprise AI agents by introducing a four‑layer architecture, a four‑defense security model, dynamic team hierarchies, sandboxed runtimes with elastic scaling, and a data‑driven evolution loop that continuously improves the system.

AI AgentsAgent evolutionSandbox runtime
0 likes · 16 min read
How Alibaba Cloud AgentTeams Enables Enterprise-Scale Multi-Agent Operations
Raymond Ops
Raymond Ops
Jul 11, 2026 · Cloud Native

Kubernetes HPA & VPA Auto-Scaling: Elastic Strategies for Traffic Spikes

An in‑depth comparison of Kubernetes Horizontal and Vertical Pod Autoscalers—including algorithms, configurations, performance benchmarks, mixed‑mode trade‑offs, custom‑metric integrations, and real‑world case studies—demonstrates how to choose and tune HPA, VPA, and KEDA for rapid traffic spikes while avoiding conflicts.

AutoscalingHPAKEDA
0 likes · 47 min read
Kubernetes HPA & VPA Auto-Scaling: Elastic Strategies for Traffic Spikes
Java Architect Handbook
Java Architect Handbook
Jul 10, 2026 · Artificial Intelligence

Spring AI 2.0 vs Spring AI Alibaba: Which One Should You Choose?

This article compares Spring AI 2.0 and Spring AI Alibaba, detailing their design philosophies, core architectures, recent upgrades, code examples, strengths, weaknesses, and ideal use‑cases, and explains how the two frameworks can be combined for enterprise AI solutions.

AI integrationGraph engineJava
0 likes · 19 min read
Spring AI 2.0 vs Spring AI Alibaba: Which One Should You Choose?
Architect Chen
Architect Chen
Jul 10, 2026 · Cloud Native

Comprehensive Guide to Docker Core Commands (2026 Edition)

This article provides a complete reference of essential Docker commands—including version, info, images, pull, run, ps, stop/start, exec, logs, stats, rm, rmi, prune, and inspect—along with example usages and typical scenarios such as environment verification, container management, and performance troubleshooting.

DevOpsDockerLinux
0 likes · 6 min read
Comprehensive Guide to Docker Core Commands (2026 Edition)
Java Backend Technology
Java Backend Technology
Jul 9, 2026 · Cloud Native

Quarkus 3.37 Beats Spring Boot: 50× Faster Startup and 70% Less Memory

Quarkus 3.37, released on June 24, 2026, introduces experimental JLink support, a full Hibernate ORM 7.4 upgrade, a reflection‑free Jackson serializer, AI‑native extensions, and numerous bug fixes, delivering up to 50‑fold faster startup, 70% lower memory usage, and measurable performance gains for Java cloud‑native applications.

AIJLinkJackson
0 likes · 11 min read
Quarkus 3.37 Beats Spring Boot: 50× Faster Startup and 70% Less Memory
Alibaba Cloud Observability
Alibaba Cloud Observability
Jul 6, 2026 · Cloud Native

Observe Every AI Agent Call Without Changing a Single Line of Code

OBI uses Linux kernel eBPF instrumentation to automatically capture and parse all AI‑related HTTP traffic—covering LLM, embedding, vector search, rerank and MCP tool calls—producing OpenTelemetry‑compatible traces and metrics without any code changes, enabling full‑stack observability of multi‑provider AI agents across languages with only ~1% CPU overhead.

AI ObservabilityGenAILinux kernel
0 likes · 21 min read
Observe Every AI Agent Call Without Changing a Single Line of Code
Alibaba Cloud Observability
Alibaba Cloud Observability
Jul 6, 2026 · Operations

How Qoder Embeds Ops Capability to Pinpoint Root Causes in One Sentence

The article shows how integrating Alibaba Cloud's STAROps plugin into Qoder lets developers diagnose production incidents with natural‑language queries, automatically gathering logs, metrics, topology and change data to deliver a structured root‑cause analysis and even generate fix code, cutting investigation time from tens of minutes to a few minutes.

AIDevOpsQoder
0 likes · 14 min read
How Qoder Embeds Ops Capability to Pinpoint Root Causes in One Sentence
ThinkingAgent
ThinkingAgent
Jul 4, 2026 · Cloud Native

Building the AI Infra Foundation: L0 Resource Layer for GPU Scheduling and Cloud‑Native Architecture

The article presents a detailed, step‑by‑step analysis of the L0 resource layer that underpins AI infrastructure, covering GPU scheduling, multi‑tier storage, low‑latency networking, core architectural components, key technologies such as MIG, Volcano, Kueue and RDMA, practical implementation patterns, quantitative acceptance criteria, and common pitfalls with best‑practice mitigations.

AI infrastructureGPU SchedulingJuiceFS
0 likes · 26 min read
Building the AI Infra Foundation: L0 Resource Layer for GPU Scheduling and Cloud‑Native Architecture
21CTO
21CTO
Jul 1, 2026 · Cloud Native

Microsoft Unveils Linux Containers for Native Windows Execution

Microsoft's preview adds a built‑in Linux container CLI and API, letting developers run Linux containers directly on Windows without third‑party tools, while introducing a new filesystem, network mode, and programmatic integration for Windows apps.

APICLILinux containers
0 likes · 5 min read
Microsoft Unveils Linux Containers for Native Windows Execution
dbaplus Community
dbaplus Community
Jun 29, 2026 · Cloud Computing

Why More Companies Are Dropping VMware for Proxmox

Since 2024, a growing number of enterprises—especially small‑to‑medium businesses and some large firms—are re‑evaluating the cost‑driven VMware licensing model and migrating to the open‑source Proxmox VE platform, which bundles KVM, LXC, Ceph, backup and clustering into a free, easy‑to‑manage solution that fits modern AI and Kubernetes workloads.

KubernetesOpen SourceProxmox
0 likes · 6 min read
Why More Companies Are Dropping VMware for Proxmox
Alibaba Cloud Big Data AI Platform
Alibaba Cloud Big Data AI Platform
Jun 29, 2026 · Big Data

How DataWorks Data Agent Evolved Across Three Stages and Its Cloud‑Native Engineering Practices

The article systematically outlines DataWorks Data Agent’s progression from a Copilot‑assisted tool to human‑AI collaboration and finally AI‑driven autonomy, details its four‑agent product matrix covering data development, operations diagnostics, autonomous governance and ChatBI, describes three architecture iterations (Dify, AgentScope, QwenCode/OpenClaw) and a cloud‑managed deployment, and cites real‑world efficiency gains such as cutting development cycles from hours to minutes.

AI AgentBig DataData Agent
0 likes · 15 min read
How DataWorks Data Agent Evolved Across Three Stages and Its Cloud‑Native Engineering Practices
Alibaba Cloud Infrastructure
Alibaba Cloud Infrastructure
Jun 29, 2026 · Cloud Native

How Argo Workflows and Alibaba Cloud ACS Redefine Gene Analysis Pipelines

By combining Alibaba Cloud's fully managed Argo Workflows with ACS's elastic compute, a gene bioinformatics platform boosted workflow efficiency by 70%, cut costs over 50% and reduced operational complexity 70%, delivering scalable, cost‑effective support for single‑cell, spatial transcriptomics and epigenomics research.

ACSAlibaba CloudArgo Workflows
0 likes · 8 min read
How Argo Workflows and Alibaba Cloud ACS Redefine Gene Analysis Pipelines
dbaplus Community
dbaplus Community
Jun 28, 2026 · Operations

Why Tencent Music Rejects AI Hype: Building an OpenClaw‑Powered Intelligent Ops Ecosystem

The article details Tencent Music's step‑by‑step evolution from manual alert handling to a three‑layer cloud‑native AIOps platform, describing data pipelines, dynamic 3‑sigma alerts, full‑link observability, and the OpenClaw sandbox with multi‑agent architecture that prioritises scenario‑driven, safe AI integration.

AIAIOpsData Engineering
0 likes · 17 min read
Why Tencent Music Rejects AI Hype: Building an OpenClaw‑Powered Intelligent Ops Ecosystem
Alibaba Cloud Native
Alibaba Cloud Native
Jun 26, 2026 · Cloud Native

One-Click Real-Time Stream Ingestion: Alibaba Cloud Kafka’s Native Data Lake Integration

Alibaba Cloud Message Queue for Kafka introduces a native message‑to‑lake capability that integrates Apache Iceberg with OSS Table Bucket, eliminating Spark/Flink/Kafka Connect, providing exactly‑once semantics, automatic schema management, dual write modes, smart partitioning, and up to ten‑fold performance gains across diverse real‑time analytics scenarios.

Apache IcebergData LakeKafka
0 likes · 12 min read
One-Click Real-Time Stream Ingestion: Alibaba Cloud Kafka’s Native Data Lake Integration
Golang Shines
Golang Shines
Jun 26, 2026 · Cloud Native

Why Every Ops Role Now Demands Kubernetes Skills (And a 100‑Question K8s Interview Guide)

After being laid off after five years in operations, the author realized that all job listings now require Docker and Kubernetes expertise, so they compiled a comprehensive "100 K8s Interview Questions" guide covering core concepts, architecture, resource management, networking, storage, security, troubleshooting, and ecosystem tools.

Container OrchestrationDevOpsDocker
0 likes · 7 min read
Why Every Ops Role Now Demands Kubernetes Skills (And a 100‑Question K8s Interview Guide)
360 Zhihui Cloud Developer
360 Zhihui Cloud Developer
Jun 26, 2026 · Databases

How ZestKV Redefines Cloud‑Native Serverless KV Storage: Design, Goals, and Use Cases

ZestKV, built on Pika, introduces a seven‑layer compute‑storage separation architecture that eliminates capacity limits, offers second‑level elastic scaling without data migration, maintains stable P99 latency, guarantees zero data loss, provides multi‑tenant isolation, and remains fully compatible with the Redis protocol for a wide range of cloud‑native workloads.

KV storeServerlesscloud native
0 likes · 9 min read
How ZestKV Redefines Cloud‑Native Serverless KV Storage: Design, Goals, and Use Cases
Architect Chen
Architect Chen
Jun 25, 2026 · Cloud Native

Four Key Ways to Deploy Microservices: From Bare Metal to Kubernetes

The article compares four microservice deployment approaches—physical servers, virtual machines, containerization with Docker, and Kubernetes clusters—detailing their implementation, advantages, drawbacks, and ideal scenarios, helping teams choose the most suitable strategy based on resource isolation, scalability, operational complexity, and team expertise.

KubernetesMicroservicescloud native
0 likes · 6 min read
Four Key Ways to Deploy Microservices: From Bare Metal to Kubernetes
Sohu Tech Products
Sohu Tech Products
Jun 24, 2026 · Cloud Native

Cloud‑Native Dynamic Routing & Session Persistence for AI Sandboxes via Web VNC

The article details how the team built a high‑performance, reliable cloud‑native gateway for millions of AI sandbox VNC sessions, addressing challenges of dynamic pod IPs, multi‑stage Web VNC traffic, session consistency, and security by using OpenResty, Lua scripts, Redis‑backed routing, cookie‑based state storage, and extensive Nginx tuning.

AI sandboxLuaOpenResty
0 likes · 26 min read
Cloud‑Native Dynamic Routing & Session Persistence for AI Sandboxes via Web VNC
Code of Duty
Code of Duty
Jun 22, 2026 · Cloud Native

Do Programmers Still Need a Personal Blog in 2026? An Honest Assessment

The article examines why, despite abundant content platforms, building a personal blog in 2026 remains valuable for programmers seeking long‑term technical archives, a controllable brand foundation, and hands‑on experience with servers, domains, Docker, Nginx, and HTTPS.

cloud nativedeveloper brandingpersonal blog
0 likes · 11 min read
Do Programmers Still Need a Personal Blog in 2026? An Honest Assessment
Alibaba Cloud Observability
Alibaba Cloud Observability
Jun 22, 2026 · Cloud Native

Zero‑Code Full‑Stack Observability with OpenTelemetry eBPF: CloudMonitor 2.0’s In‑Kernel “Lens”

OpenTelemetry eBPF Instrumentation (OBI) injects a kernel‑level, zero‑code probe that automatically captures OpenTelemetry‑compatible traces, metrics, and logs for over 15 protocols—including HTTP, gRPC, MySQL, Redis, Kafka, and CUDA—while handling cross‑language context propagation, GPU tracing, and seamless integration with CloudMonitor 2.0.

OpenTelemetryTracingZero-Code Monitoring
0 likes · 19 min read
Zero‑Code Full‑Stack Observability with OpenTelemetry eBPF: CloudMonitor 2.0’s In‑Kernel “Lens”
Ops Development Stories
Ops Development Stories
Jun 22, 2026 · Cloud Native

Design and Implementation of a Multi‑Cluster Arthas‑Based Online Diagnosis Platform

This article details the architecture, security mechanisms, and implementation of a unified Arthas online diagnosis platform that enables SSH‑free, audited access to Java applications across dozens of isolated Kubernetes clusters, covering control‑plane design, WebSocket tunneling, credential management, RBAC, and front‑end integration with Vue and xterm.js.

ArthasGoHMAC
0 likes · 23 min read
Design and Implementation of a Multi‑Cluster Arthas‑Based Online Diagnosis Platform
TechVision Expert Circle
TechVision Expert Circle
Jun 22, 2026 · Industry Insights

How CIOs Can Navigate the Deep‑Water Phase of Digital Transformation

The article analyzes why many enterprises now face entrenched legacy systems, data silos, and tightening security while AI delivers little ROI, and it offers CIOs practical, architecture‑driven strategies—including Strangler Fig migration, AI embedding, data‑fabric governance, and zero‑trust rollout—to break through these deep‑water challenges.

AI integrationCIOcloud native
0 likes · 12 min read
How CIOs Can Navigate the Deep‑Water Phase of Digital Transformation