Tagged articles

AACR-Bench

4 articles · Page 1 of 1
DataFunSummit
DataFunSummit
Sep 20, 2026 · Artificial Intelligence

Alibaba's OpenCodeReview: Why Production Agents Are Reclaiming Control from LLMs

Alibaba's OpenCodeReview adopts a hybrid deterministic-engineering-plus-agent architecture for code review, cutting token usage by ~9x versus generic coding agents by constraining agent autonomy with hard-coded filters, token guards, and line-resolution modules, trading lower recall for higher precision and reliability.

AACR-BenchAI code reviewAlibaba
0 likes · 16 min read
Alibaba's OpenCodeReview: Why Production Agents Are Reclaiming Control from LLMs
Architecture Digest
Architecture Digest
Sep 19, 2026 · Artificial Intelligence

Alibaba's Open Code Review: Deterministic Pipelines Slash Token Costs 9x

Alibaba open-sourced Open Code Review, an AI code review tool used internally for two years, which combines a deterministic rule engine for file selection and line positioning with LLMs for judgment only, achieving higher precision and 9x lower token consumption than Claude Code on a benchmark of 200 real PRs.

AACR-BenchAI code reviewAlibaba
0 likes · 8 min read
Alibaba's Open Code Review: Deterministic Pipelines Slash Token Costs 9x
Alibaba Cloud Developer
Alibaba Cloud Developer
Mar 9, 2026 · Artificial Intelligence

How Alibaba’s AI Code Review Assistant Cuts NPE Bugs with Context‑Aware Agents

This article explains Alibaba Group’s AI‑driven code review benchmark, the agent‑based assistant that understands repository context, its real‑world impact on reducing null‑pointer exceptions, and how the open‑source AACR‑Bench dataset provides a multi‑language, context‑aware evaluation standard for AI code review.

AACR-BenchAI code reviewAlibaba
0 likes · 19 min read
How Alibaba’s AI Code Review Assistant Cuts NPE Bugs with Context‑Aware Agents