PaperAgent
Author

PaperAgent

Daily updates, analyzing cutting-edge AI research papers

302
Articles
1
Likes
1.8k
Views
0
Comments
Recent Articles

Latest from PaperAgent

100 recent articles max
PaperAgent
PaperAgent
Sep 11, 2026 · Artificial Intelligence

Google's Procedural Graphs: Self-Evolving Execution Structures for LLM Agents

Google introduces Procedural Graphs, a self-evolving graph structure that explicitly encodes procedural knowledge for LLM agents, enabling them to locate, extract, and generate contextual guidance from a (procedure, relation, procedure) triplet graph, achieving 85% survival from 0% in CFO simulation and winning 21 of 24 model-benchmark combinations across 7 benchmarks and 4 LLMs.

Agent ReliabilityBenchmark EvaluationGraph-based Reasoning
0 likes · 9 min read
Google's Procedural Graphs: Self-Evolving Execution Structures for LLM Agents
PaperAgent
PaperAgent
Sep 9, 2026 · Artificial Intelligence

OpenAI Unveils AI Research Acceleration Metrics: A Three-Layer Measurement Framework

OpenAI publishes internal data on how AI agents accelerate research, introducing a three-layer measurement framework—usage, tasks, and results—showing median researchers spend $600/day on tokens, agents now handle 3.1 workdays per human day, task delegation spans six R&D stages but high-level planning remains human-led, and over half of successful 4-8 hour tasks still require human intervention.

AI agentsAI research methodologyOpenAI
0 likes · 8 min read
OpenAI Unveils AI Research Acceleration Metrics: A Three-Layer Measurement Framework
PaperAgent
PaperAgent
Sep 6, 2026 · Artificial Intelligence

Anthropic's Killer Multi-Agent Blueprint: One Loop, Skills, Harness & Snapshot Eval

Anthropic's production e-commerce and math-formalization agents share a unified architecture: a single-model loop with modular skills, tool calls to existing systems, code-enforced harness rules, and snapshot-based evaluation, enabling scalable, verifiable multi-agent systems.

Agent ArchitectureAnthropicLLM Engineering
0 likes · 17 min read
Anthropic's Killer Multi-Agent Blueprint: One Loop, Skills, Harness & Snapshot Eval
PaperAgent
PaperAgent
Sep 1, 2026 · Artificial Intelligence

Google Publishes Two Agent Skill Papers in One Day: WikiSkill and SKILL.state Break New Ground

Google released two Agent Skill papers—WikiSkill and SKILL.state—introducing a structured knowledge layer for skill evolution and a state‑machine execution model that dramatically reduces prompt length, improves accuracy, and demonstrates strong cross‑model transfer and robustness across a suite of benchmarks.

Agent SkillLLMPrompt Optimization
0 likes · 13 min read
Google Publishes Two Agent Skill Papers in One Day: WikiSkill and SKILL.state Break New Ground
PaperAgent
PaperAgent
Aug 31, 2026 · Artificial Intelligence

A First Systematic Review of Multimodal Agentic Frameworks

This article surveys multimodal agentic frameworks, proposing a taxonomy that maps modality‑fusion strategies to the five core agent modules (perception, reasoning, planning, memory, action), evaluates four application domains across five performance dimensions, and highlights architectural trade‑offs and benchmark results.

AI agentsAction PlanningAgentic Frameworks
0 likes · 15 min read
A First Systematic Review of Multimodal Agentic Frameworks
PaperAgent
PaperAgent
Aug 30, 2026 · Artificial Intelligence

Google Unveils How Gemini Supercharges AI Research

The article details Google's internal Co‑Scientist system that leverages Gemini to evolve hypotheses, generate and validate experimental code, and produce multi‑objective papers with safety checks, achieving superior results across chemistry, biology, and computer‑science benchmarks while dramatically cutting hallucinations and plagiarism.

AI agentsAI safetyCo-Scientist
0 likes · 9 min read
Google Unveils How Gemini Supercharges AI Research
PaperAgent
PaperAgent
Aug 29, 2026 · Artificial Intelligence

Automating Academic Literature Reviews with Doubao Work’s AI Agent

The author demonstrates how Doubao Work’s AI Agent can fully automate the academic literature review pipeline—searching arXiv, classifying papers, generating Excel tables, creating visualizations with matplotlib, and drafting a structured review—across desktop, web, and mobile, cutting hours of manual work to minutes.

AIAgentAutomation
0 likes · 8 min read
Automating Academic Literature Reviews with Doubao Work’s AI Agent
PaperAgent
PaperAgent
Aug 28, 2026 · Artificial Intelligence

How Tencent WorkBuddy’s TextIn xParse Turns PDFs into Structured Data for LLMs

The author evaluates Tencent WorkBuddy’s new TextIn xParse connector, showing how it converts complex, multi‑page PDFs—including mixed graphics and hierarchical headings—into accurate Markdown/JSON, enabling LLMs to answer detailed questions with preserved structure, and highlights the free 1,000‑page‑per‑day quota for developers.

AI agentLLMTextIn xParse
0 likes · 7 min read
How Tencent WorkBuddy’s TextIn xParse Turns PDFs into Structured Data for LLMs
PaperAgent
PaperAgent
Aug 27, 2026 · Artificial Intelligence

355 Recent Agent Papers + 821 Projects: A Free, Organized Resource

The article presents a curated, freely available collection of 355 recent Agent research papers and 821 implementation projects, categorized by research dimensions and conference venues, and offers a step‑by‑step reading plan to help researchers navigate the rapidly growing field.

AIAgentOpen Resources
0 likes · 4 min read
355 Recent Agent Papers + 821 Projects: A Free, Organized Resource