Huawei Cloud Developer Alliance
Author

Huawei Cloud Developer Alliance

The Huawei Cloud Developer Alliance creates a tech sharing platform for developers and partners, gathering Huawei Cloud product knowledge, event updates, expert talks, and more. Together we continuously innovate to build the cloud foundation of an intelligent world.

490
Articles
0
Likes
1.7k
Views
0
Comments
Recent Articles

Latest from Huawei Cloud Developer Alliance

100 recent articles max
Huawei Cloud Developer Alliance
Huawei Cloud Developer Alliance
May 19, 2026 · Artificial Intelligence

How Cloud Agent Harness Grows Skills from Real Tasks: A Three‑Stage Self‑Evolution Mechanism

The article analyzes Huawei Cloud Agent Harness's three‑stage skill self‑evolution framework, detailing how agents automatically extract, evolve, and validate reusable skills from execution traces to overcome manual authoring bottlenecks and ensure continuous improvement.

AI agentsEvaluation PipelineLLM‑driven optimization
0 likes · 14 min read
How Cloud Agent Harness Grows Skills from Real Tasks: A Three‑Stage Self‑Evolution Mechanism
Huawei Cloud Developer Alliance
Huawei Cloud Developer Alliance
May 13, 2026 · Cloud Native

Why HPA Falls Short for LLMs and How Kthena Autoscaler Redefines Elastic Scaling

The article explains why traditional Kubernetes HPA cannot meet the unique demands of large‑language‑model inference, introduces Kthena Autoscaler’s model‑aware architecture, its dual stable/panic scaling modes, cost‑aware algorithms, flexible policy bindings, and provides practical configuration and observability guidance.

Kthena AutoscalerKubernetesLLM inference
0 likes · 10 min read
Why HPA Falls Short for LLMs and How Kthena Autoscaler Redefines Elastic Scaling
Huawei Cloud Developer Alliance
Huawei Cloud Developer Alliance
May 9, 2026 · Cloud Native

A Cloud‑Native Paradigm for Efficient Agent Hosting: Systematic Design of Agent Harness Infra

The article analyzes the challenges of deploying AI agents in cloud‑native environments—cold‑start latency, state persistence, and security isolation—and presents Huawei Cloud’s Agent Harness Infra, which uses capacity‑prediction, parallel scheduling, microVM‑based decoupling, and lightweight OS techniques to achieve up to 5× throughput, 100 ms startup and 80% warm‑start hit rates.

Capacity PredictionCloud NativeMicroVM
0 likes · 11 min read
A Cloud‑Native Paradigm for Efficient Agent Hosting: Systematic Design of Agent Harness Infra
Huawei Cloud Developer Alliance
Huawei Cloud Developer Alliance
Apr 29, 2026 · Artificial Intelligence

Deploy DeepSeek‑V4 on Ascend NPU with Kthena in 3 Minutes (Prefill‑Decode Separation)

This guide walks through deploying the DeepSeek‑V4‑Flash model on Ascend NPU using Kthena’s ModelRoute, detailing the Prefill‑Decode (P/D) separation architecture, KV cache transfer via Mooncake, configuration of ModelServing and ModelRoute resources, and flexible scaling of Prefill and Decode replicas for optimal performance.

Ascend NPUDeepSeek-V4KV Cache
0 likes · 22 min read
Deploy DeepSeek‑V4 on Ascend NPU with Kthena in 3 Minutes (Prefill‑Decode Separation)
Huawei Cloud Developer Alliance
Huawei Cloud Developer Alliance
Apr 23, 2026 · Artificial Intelligence

Why Agent Harness Is Central to AI Engineering: OfficeClaw Design & Implementation

The article explains how Agent Harness, defined by six core components (Execution Loop, Tool Registry, Context Manager, State Store, Lifecycle Hooks, Evaluation Interface), forms the operating system for AI agents, and details Huawei Cloud OfficeClaw’s layered architecture and real‑world deployment that boosts task reliability and efficiency.

AI engineeringContext ManagementMulti-Agent Systems
0 likes · 11 min read
Why Agent Harness Is Central to AI Engineering: OfficeClaw Design & Implementation
Huawei Cloud Developer Alliance
Huawei Cloud Developer Alliance
Apr 13, 2026 · Artificial Intelligence

How AReaL v1.0 Enables Scalable Agentic RL on Ascend NPU with AWEX Weight Sync

The new AReaL v1.0 release brings full Ascend NPU support, detailed installation guides, and a best‑practice example for training a 30B MoE model across four nodes, while the integrated AWEX weight‑sync mechanism dramatically reduces synchronization time, improving efficiency and stability for large‑scale Agentic RL workloads.

AWEXAgentic RLAscend NPU
0 likes · 12 min read
How AReaL v1.0 Enables Scalable Agentic RL on Ascend NPU with AWEX Weight Sync
Huawei Cloud Developer Alliance
Huawei Cloud Developer Alliance
Apr 10, 2026 · Cloud Computing

How Huawei’s Hybrid‑Cloud Claw Solution Secures and Localizes AI Skills

Huawei’s new hybrid‑cloud Claw solution addresses the security and accessibility challenges of AI Skills by providing an offline‑compatible, locally deployed ClawHub‑Lite that enables secure Skill acquisition, one‑click import, and instant invocation with RBAC controls, while supporting custom Skills and integration with popular IM platforms.

AI SkillsClawHubSecurity
0 likes · 5 min read
How Huawei’s Hybrid‑Cloud Claw Solution Secures and Localizes AI Skills
Huawei Cloud Developer Alliance
Huawei Cloud Developer Alliance
Apr 2, 2026 · Cloud Native

How Kthena Enables Production‑Grade LLM Inference on Kubernetes

This article analyzes the cloud‑native challenges of deploying large‑model inference on Kubernetes and presents Kthena’s architecture—ModelServing, Router, Autoscaler, and ModelBooster—along with Volcano integration, vLLM‑Ascend setup, and a real‑world Qwen3‑235B deployment case, highlighting performance gains and future directions.

Cloud NativeKthenaKubernetes
0 likes · 13 min read
How Kthena Enables Production‑Grade LLM Inference on Kubernetes
Huawei Cloud Developer Alliance
Huawei Cloud Developer Alliance
Mar 26, 2026 · Artificial Intelligence

How to Build a Full‑Stack RAG Chatbot Using LangChain, FAISS & Langfuse

This guide walks through an end‑to‑end RAG implementation with LangChain, covering multi‑format document loading, recursive text splitting, embedding selection, FAISS vector storage, ConversationalRetrievalChain setup, prompt engineering, source citation, Langfuse observability, and best‑practice configuration management.

FAISSLLMOpsLangChain
0 likes · 13 min read
How to Build a Full‑Stack RAG Chatbot Using LangChain, FAISS & Langfuse