All Articles

143517 articles · Page 415 of 7176
360 Tech Engineering
360 Tech Engineering
Apr 28, 2026 · Artificial Intelligence

How 360 AI Institute Boosted Airline Translation Accuracy from 70% to 96%

The 360 AI Research Institute tackled the zero‑tolerance translation demands of airline maintenance by building a specialized parallel corpus and applying RAG‑enhanced, SFT‑fine‑tuned, and RL‑reinforced models, raising Chinese‑to‑English translation accuracy from 70% to 96% and enabling a one‑month rollout.

AI translationRAGSFT
0 likes · 5 min read
How 360 AI Institute Boosted Airline Translation Accuracy from 70% to 96%
Old Zhang's AI Learning
Old Zhang's AI Learning
Apr 28, 2026 · Artificial Intelligence

vLLM 0.20 Arrives with DeepSeek V4 Support – What’s New?

The vLLM 0.20.0 release dramatically upgrades the inference engine with DeepSeek V4 support, default CUDA 13, PyTorch 2.11, Transformers v5 compatibility, FlashAttention 4 MLA prefill, TurboQuant 2‑bit KV cache, an online quantization front‑end, IR enhancements, Model Runner V2 features, and a slew of new models, while providing detailed installation and upgrade guidance.

CUDA 13DeepSeek V4FlashAttention
0 likes · 10 min read
vLLM 0.20 Arrives with DeepSeek V4 Support – What’s New?
James' Growth Diary
James' Growth Diary
Apr 28, 2026 · Artificial Intelligence

Mastering LangGraph Multi‑Agent Collaboration: The Supervisor Pattern Explained from Theory to Practice

The article examines why single‑agent setups fail, introduces the Supervisor pattern for clear responsibility separation, compares Tool‑Calling and Handoff approaches, provides a complete TypeScript implementation, explores hierarchical supervisors, and outlines five common pitfalls with concrete fixes.

HandoffLangGraphSupervisor Pattern
0 likes · 15 min read
Mastering LangGraph Multi‑Agent Collaboration: The Supervisor Pattern Explained from Theory to Practice
Code Mala Tang
Code Mala Tang
Apr 28, 2026 · Backend Development

Redis No Longer Dominates: Discover the Best Python Caching Alternatives

A benchmark of Redis, Memcached, DragonflyDB, and Cashews using the same FastAPI workload reveals that Redis falls behind on latency, throughput, and memory efficiency, while DragonflyDB and Cashews offer superior performance and developer experience for Python caching.

CachingCashewsDragonflyDB
0 likes · 11 min read
Redis No Longer Dominates: Discover the Best Python Caching Alternatives
Code Mala Tang
Code Mala Tang
Apr 28, 2026 · Operations

How Sub‑Agents Keep Claude Code Sessions Clean

Long Claude Code sessions quickly become noisy as every grep, find, and ls call stays in the main context, but using sub‑agents—including the built‑in Explore and Plan agents and the CLAUDE_CODE_FORK_SUBAGENT flag—isolates work, returns only concise summaries, and lets you monitor activity with a context‑timeline hook.

Claude CodeContext IsolationExplore
0 likes · 7 min read
How Sub‑Agents Keep Claude Code Sessions Clean
DataFunTalk
DataFunTalk
Apr 28, 2026 · Industry Insights

Why China Can’t Replicate Palantir – Not a Penguin in the Sahara, but a Different Beast

The article dissects Palantir’s rise—backed by In‑Q‑Tel, F‑class shares, and a subscription model—showing how the U.S. political‑capital ecosystem created a unique AI powerhouse that China’s project‑based procurement, legal constraints, and capital structure cannot emulate, and proposes a vertically‑focused, long‑term AI strategy suited to China’s own soil.

Chinese tech ecosystemIn-Q-TelPalantir
0 likes · 15 min read
Why China Can’t Replicate Palantir – Not a Penguin in the Sahara, but a Different Beast
DataFunTalk
DataFunTalk
Apr 28, 2026 · Artificial Intelligence

Manifold AI’s WorldScape 0.2 Tops WorldArena: How MoE Drives Superior Physics and 3D Understanding

Manifold AI’s WorldScape 0.2 achieved the highest overall score on the embodied world‑model benchmark WorldArena, outperforming giants like Google and Nvidia by excelling in comprehensive perception, physics compliance, and 3D accuracy while using only about 10 % of the parameters of competing models, thanks to a newly introduced MoE architecture.

MoEWorldArenaWorldScape
0 likes · 9 min read
Manifold AI’s WorldScape 0.2 Tops WorldArena: How MoE Drives Superior Physics and 3D Understanding
SuanNi
SuanNi
Apr 28, 2026 · Artificial Intelligence

Why Your AI Agent Fails and How Skills Can Fix It

The article argues that monolithic AI agents suffer from stability, extensibility, and knowledge‑retention problems, and proposes a modular "Skills" architecture—analogous to a microkernel OS—that turns expertise into reusable, version‑controlled assets, enabling cross‑platform deployment, better human‑AI collaboration, and reshaping the labor market.

AI Agentscross‑platform AIhuman‑AI collaboration
0 likes · 8 min read
Why Your AI Agent Fails and How Skills Can Fix It
Architect Chen
Architect Chen
Apr 28, 2026 · Backend Development

How High Must TPS Be for a Payment System to Be Considered High‑Throughput?

The article analyzes why payment systems require far higher transactions‑per‑second than typical web apps, outlines the challenges of distributed transactions under high load, and classifies TPS levels—from 1,000 daily to over 300,000—as benchmarks, citing Alipay’s Double 11 peak as an ultra‑high example.

TPSbackenddistributed transactions
0 likes · 3 min read
How High Must TPS Be for a Payment System to Be Considered High‑Throughput?
AI Waka
AI Waka
Apr 28, 2026 · Artificial Intelligence

Why Single-Agent AI Fails: Anthropic’s Multi-Agent Harness for Long-Running Tasks

The article explains that single‑agent AI collapses on long‑running tasks due to compound error probabilities, outlines four structural failure modes, and presents Anthropic’s three‑agent GAN‑style harness—Planner, Generator, Evaluator—detailing sprint contracts, primitives, token economics, and three real‑world case studies that demonstrate dramatically higher reliability and productivity.

AI HarnessAgentic OpsAnthropic
0 likes · 26 min read
Why Single-Agent AI Fails: Anthropic’s Multi-Agent Harness for Long-Running Tasks
Amazon Cloud Developers
Amazon Cloud Developers
Apr 28, 2026 · Cloud Computing

How AWS Achieved Day‑0 Adaptation of Xiaomi’s MiMo‑V2.5‑Pro on Trainium

AWS has completed a Day‑0 rapid adaptation of Xiaomi’s open‑source MiMo‑V2.5‑Pro model, enabling developers worldwide to run the 1‑trillion‑parameter, 1‑million‑token model on Amazon Trainium chips with high‑throughput, low‑latency inference via Neuron SDK integration, and offers three deployment paths—EC2, SageMaker, and EKS/ECS.

AI inferenceAWSAmazon Trainium
0 likes · 6 min read
How AWS Achieved Day‑0 Adaptation of Xiaomi’s MiMo‑V2.5‑Pro on Trainium
Digital Planet
Digital Planet
Apr 28, 2026 · Industry Insights

How to Eliminate Cap Fraud and Data Blind Spots in White‑Label Beverages for Under ¥100k

This article analyzes the twin challenges of cap fraud and opaque terminal data faced by low‑budget white‑label beverage brands, explains why heavyweight five‑code solutions are unsuitable, and presents a sub‑¥100,000 lightweight one‑code system with step‑by‑step deployment, flexible marketing modes, and a real‑world case study showing cost savings, data accuracy, and sales uplift.

Data AnalyticsOne-Codecap fraud
0 likes · 14 min read
How to Eliminate Cap Fraud and Data Blind Spots in White‑Label Beverages for Under ¥100k
Digital Planet
Digital Planet
Apr 28, 2026 · Industry Insights

What Is a Scenario Solution? A Structured Approach to Scenario Marketing

The article breaks down scenario marketing by applying Cartesian methodology to split consumption moments into physical, relational, and meaning fields, then maps them to functional, identity, and ritual solutions, offering a practical, enumerated framework for creating data‑driven, customer‑centric scenario solutions.

Cartesian methodologydigitalizationmeaning field
0 likes · 12 min read
What Is a Scenario Solution? A Structured Approach to Scenario Marketing
Digital Planet
Digital Planet
Apr 28, 2026 · Industry Insights

How Digital Channel Gaps Turned the Mengniu‑Yili Duopoly Upside‑Down

The 2025 dairy reports reveal a 400 billion‑yuan revenue gap between Mengniu and Yili, driven not by advertising or product quality but by a generational lag in channel digitalization that reshapes the competitive landscape, erodes Mengniu's marketing efficiency, and forces a strategic rethink for the whole industry.

Business Model AnalysisDairy IndustryMengniu
0 likes · 17 min read
How Digital Channel Gaps Turned the Mengniu‑Yili Duopoly Upside‑Down
Network Intelligence Research Center (NIRC)
Network Intelligence Research Center (NIRC)
Apr 28, 2026 · Artificial Intelligence

How AI Learned to Read the Genomic “Dialects” of 300,000 People for Precise Expression Prediction

This article reviews a study that overcomes the limitation of reference‑genome‑only models by pre‑training a genomic language model on 300,000 European individuals’ variants, creating UKBioBERT and the two‑stage UKBioFormer, which together deliver markedly better gene‑function representations and personalized expression predictions across cell lines and populations.

UKBioBERTUKBioFormerfunctional genomics
0 likes · 7 min read
How AI Learned to Read the Genomic “Dialects” of 300,000 People for Precise Expression Prediction
AntData
AntData
Apr 28, 2026 · Artificial Intelligence

Iterative Agent Evaluation Skill: Automating Bad‑Case Diagnosis with AI Pre‑Annotation

The article presents an end‑to‑end, eight‑phase automated evaluation pipeline for large‑model agents that replaces manual bad‑case inspection with AI‑assisted pre‑annotation, cutting analysis time from a full‑day to about 30 minutes and achieving over 90 % efficiency gain while enabling iterative knowledge‑base refinement.

AI Pre‑annotationAgent evaluationAutomated Pipeline
0 likes · 20 min read
Iterative Agent Evaluation Skill: Automating Bad‑Case Diagnosis with AI Pre‑Annotation
Old Zhang's AI Learning
Old Zhang's AI Learning
Apr 28, 2026 · Artificial Intelligence

OpenAI’s Latest Open‑Source Releases: Codex CLI, Plugins, Symphony, and Privacy‑Filter

OpenAI has recently open‑sourced three projects—Codex CLI, the openai/plugins repository, the engineering‑preview Symphony orchestration service, and the privacy‑filter model—detailing installation, plugin architecture, workflow orchestration design, and usage examples, while comparing them to competing agents and noting practical constraints.

AI AgentCodex CLIOpen Source
0 likes · 17 min read
OpenAI’s Latest Open‑Source Releases: Codex CLI, Plugins, Symphony, and Privacy‑Filter
Data STUDIO
Data STUDIO
Apr 28, 2026 · Backend Development

FastAPI in Production: Auth, Rate Limiting, and Zero‑Downtime with One Codebase

This article walks through a complete production‑ready FastAPI setup, covering secure OIDC/JWKS authentication, Redis‑backed token‑bucket rate limiting, zero‑downtime rolling deployments on Docker/Kubernetes, and observability best practices such as request‑ID middleware and structured JSON logging.

AuthenticationDockerFastAPI
0 likes · 20 min read
FastAPI in Production: Auth, Rate Limiting, and Zero‑Downtime with One Codebase
Wu Shixiong's Large Model Academy
Wu Shixiong's Large Model Academy
Apr 28, 2026 · Artificial Intelligence

Why Bigger Context Fails for Deep Research Agents and How IterResearch Fixes It

Interviewers point out that simply enlarging the LLM’s context window cannot prevent forgetting early conclusions in long‑step Deep Research tasks; the article explains the ReAct context issues, introduces the IterResearch framework with evolving reports, and compares its accuracy, cost, and scalability against ReAct and ReSum.

Context ManagementDeep ResearchIterResearch
0 likes · 17 min read
Why Bigger Context Fails for Deep Research Agents and How IterResearch Fixes It