Tagged articles

scientific research

11 articles · Page 1 of 1
PaperAgent
PaperAgent
Oct 1, 2026 · Artificial Intelligence

ScienceBuddy: Open-Source AI Research Agent with Recursive Self-Improvement

ScienceBuddy is an open-source AI research agent that uses a recursive dual-loop architecture to continuously improve its harness and model weights from real user interactions, achieving 42.2% to 73.3% accuracy gains on scientific benchmarks across genomics, molecular biology, and pharmacology domains.

AI AgentLAB-BenchRecursive Self-Improvement
0 likes · 8 min read
ScienceBuddy: Open-Source AI Research Agent with Recursive Self-Improvement
Tech Architecture Stories
Tech Architecture Stories
Sep 9, 2026 · Industry Insights

AI Agents Go Vertical: 4 GitHub Projects Signal Shift Beyond Coding

This week's GitHub trending signals reveal AI agents moving beyond general coding into four vertical domains: verifiable architecture diagrams with archify, 3D immersive OSINT with gods-eye-view, 165 scientific research skills, and multi-agent classrooms with OpenMAIC, plus a Rust-native headless browser obscura for agent workflows.

AI agentsArchitecture DiagramsEducation Technology
0 likes · 9 min read
AI Agents Go Vertical: 4 GitHub Projects Signal Shift Beyond Coding
PaperAgent
PaperAgent
Aug 30, 2026 · Artificial Intelligence

Google Unveils How Gemini Supercharges AI Research

The article details Google's internal Co‑Scientist system that leverages Gemini to evolve hypotheses, generate and validate experimental code, and produce multi‑objective papers with safety checks, achieving superior results across chemistry, biology, and computer‑science benchmarks while dramatically cutting hallucinations and plagiarism.

AI agentsAI safetyCo-Scientist
0 likes · 9 min read
Google Unveils How Gemini Supercharges AI Research
Machine Learning Algorithms & Natural Language Processing
Machine Learning Algorithms & Natural Language Processing
Aug 26, 2026 · Artificial Intelligence

Can AI Do Independent Research? ASI‑Bench Measures Scientific Autonomy

ASI‑Bench, developed by Tsinghua and leading institutions, is a benchmark that evaluates AI’s scientific autonomy by progressively reducing method guidance across four levels, revealing that current models lose up to half their scientific score without detailed instructions, highlighting the gap to true independent research.

AI autonomyASI-BenchAgent Evaluation
0 likes · 14 min read
Can AI Do Independent Research? ASI‑Bench Measures Scientific Autonomy
Data Party THU
Data Party THU
Jul 10, 2026 · Artificial Intelligence

Anthropic Unveils Claude Science: An Integrated AI Workbench for Researchers

On June 30 2026 Anthropic introduced Claude Science, an AI‑powered workbench that unifies common research tools, generates auditable outputs, offers built‑in genomics, proteomics and cheminformatics skills, and manages scalable compute on HPC or cloud, enabling scientists to conduct end‑to‑end analyses from data acquisition to manuscript preparation.

AI workbenchClaude ScienceHPC
0 likes · 9 min read
Anthropic Unveils Claude Science: An Integrated AI Workbench for Researchers
ShiZhen AI
ShiZhen AI
Jun 10, 2026 · Artificial Intelligence

Claude Fable 5 Deep Dive: Coding Power Beats GPT‑5.5, Safety Trade‑off Explained

Anthropic’s newly released Claude Fable 5, the first publicly available Mythos‑level model, delivers SOTA performance across software engineering, coding, visual tasks and scientific research—outperforming GPT‑5.5 and Gemini on benchmarks—while offering a modest $10/$50 token pricing and a 5 % safety fallback that trades some flexibility for stronger safeguards.

AI benchmarksClaude Fable 5Mythos 5
0 likes · 14 min read
Claude Fable 5 Deep Dive: Coding Power Beats GPT‑5.5, Safety Trade‑off Explained
SuanNi
SuanNi
Apr 24, 2026 · Artificial Intelligence

Why GPT‑5.5 Beats Opus 4.7 and Sets a New Global SOTA

OpenAI’s newly released GPT‑5.5, marketed as a “next‑generation AI for real work,” outperforms competitors across coding, knowledge‑work, and scientific research benchmarks—achieving 82.7% accuracy on Terminal‑Bench 2.0, 58.6% on SWE‑Bench Pro, 84.9% on GDPval, and 98.0% on Tau2‑bench Telecom—while offering higher token efficiency and new pricing tiers.

AI AgentBenchmarkCoding
0 likes · 11 min read
Why GPT‑5.5 Beats Opus 4.7 and Sets a New Global SOTA
Smart Sea Tide
Smart Sea Tide
Apr 23, 2026 · Artificial Intelligence

UniScientist 30B: Open-Source Research Model Redefines Autonomous AI Science

UniScientist 30B, an open‑source large language model for scientific research, introduces a full hypothesis‑evidence‑reproducible‑iteration loop, achieves benchmark scores rivaling larger closed‑source models, and offers a modular architecture that turns research problems into verifiable unit tests.

AIBenchmarklarge language model
0 likes · 5 min read
UniScientist 30B: Open-Source Research Model Redefines Autonomous AI Science
Data Party THU
Data Party THU
Mar 12, 2026 · Artificial Intelligence

Can a 30B LLM Truly Conduct Autonomous Scientific Research? Inside UniScientist

UniScientist, a 30‑billion‑parameter open‑source model from UniPat AI, demonstrates a closed‑loop scientific research workflow—generating hypotheses, gathering evidence, performing reproducible derivations, and iteratively refining conclusions—while achieving benchmark scores comparable to much larger proprietary systems across multiple scientific evaluation suites.

benchmarkinglarge language modelscientific research
0 likes · 10 min read
Can a 30B LLM Truly Conduct Autonomous Scientific Research? Inside UniScientist
SuanNi
SuanNi
Mar 9, 2026 · Artificial Intelligence

How UniScientist Beats GPT‑5.4 on FrontierScience Benchmarks

UniScientist, a 30B‑parameter AI model co‑developed by UniPat AI and Peking University, leverages a meticulously curated scientific dataset and a powerful code interpreter to achieve 33.3% success on the FrontierScience‑Research benchmark, surpassing the newly released GPT‑5.4 and demonstrating superior multi‑disciplinary research capabilities.

AIdatasetlarge language model
0 likes · 12 min read
How UniScientist Beats GPT‑5.4 on FrontierScience Benchmarks
Baidu Tech Salon
Baidu Tech Salon
Aug 30, 2024 · Artificial Intelligence

AI4R&D: AI-Driven R&D New Paradigm

The AI4R&D report released on July 25 2024 outlines how artificial‑intelligence technologies are reshaping scientific research and industrial development by boosting computing power, data utilization and toolchains, delivering breakthroughs in drug discovery, materials and simulation, and mapping an emerging ecosystem that promises faster translation of innovation into practical applications.

AI ecosystemAI4R&DArtificial Intelligence
0 likes · 5 min read
AI4R&D: AI-Driven R&D New Paradigm