Machine Heart
Author

Machine Heart

Professional AI media and industry service platform

1.0k
Articles
0
Likes
5.3k
Views
0
Comments
Recent Articles

Latest from Machine Heart

100 recent articles max
Machine Heart
Machine Heart
Sep 2, 2026 · Artificial Intelligence

Facet-0 Enables Robots to See, Insert Precisely, and Recover from Errors in Precise Assembly

The NTU PINE Lab introduces Facet-0, a multimodal robot foundation model that achieves 82% success in five real computer‑assembly tasks with 0.5 mm placement accuracy, reduces human intervention from 47% to 24%, and learns to recover from contact failures using a force‑synchronized dataset and reinforcement‑learning post‑training.

ManuFacet-1KMultimodal LearningNTU
0 likes · 11 min read
Facet-0 Enables Robots to See, Insert Precisely, and Recover from Errors in Precise Assembly
Machine Heart
Machine Heart
Sep 2, 2026 · Artificial Intelligence

Atlas Unveiled by Fei‑Fei Li: A New Era for World Models and Robotics

World Labs' Atlas is a multimodal, camera‑controlled world model that natively handles text, images, video and 3D, offering spatial‑context generation, high‑fidelity 3D reconstruction, and robot simulation, with benchmark results that highlight its advantages over prior models.

3D reconstructionAtlasWorld Labs
0 likes · 10 min read
Atlas Unveiled by Fei‑Fei Li: A New Era for World Models and Robotics
Machine Heart
Machine Heart
Sep 2, 2026 · Artificial Intelligence

How UniSteer Boosts Real‑World VLA Success from 20% to 90% in 66 Minutes

UniSteer introduces a noise‑steering interface that lets human corrections and reinforcement learning jointly update a lightweight noise actor, enabling a Vision‑Language‑Action robot to raise task success from 20% to 90% within 66 minutes while using only two full human demonstrations.

Noise SteeringUniSteerVision-Language-Action
0 likes · 14 min read
How UniSteer Boosts Real‑World VLA Success from 20% to 90% in 66 Minutes
Machine Heart
Machine Heart
Sep 1, 2026 · Artificial Intelligence

Only 33% of Claude, GPT, and Gemini Survive Real‑World Websites—ClawBench Shows AI Agents Still Struggle

ClawBench evaluates 144 production websites across 153 everyday tasks and finds that top models like Claude Sonnet 4.6, GPT‑5.4, Qwen 3.5 and GLM‑5 achieve at best a 33.3% overall success rate, exposing last‑mile non‑commit failures, anti‑bot defenses and domain‑specific weaknesses that sandbox benchmarks miss.

AI AgentsClawBenchLLM evaluation
0 likes · 14 min read
Only 33% of Claude, GPT, and Gemini Survive Real‑World Websites—ClawBench Shows AI Agents Still Struggle
Machine Heart
Machine Heart
Sep 1, 2026 · Artificial Intelligence

Can AI Discover Real Vulnerabilities? Researchers Embed Real Bugs into Model Parameters

The paper introduces CyberFactory, a pipeline that transforms scattered open‑source CVE data into executable security tasks, generates high‑quality agent trajectories, and uses them to train the OpenAegis model, which achieves up to 58.1% pass rate—significantly outperforming baseline LLMs in a one‑hour security challenge.

AI securityCyberFactoryLLM
0 likes · 15 min read
Can AI Discover Real Vulnerabilities? Researchers Embed Real Bugs into Model Parameters
Machine Heart
Machine Heart
Sep 1, 2026 · Artificial Intelligence

Can AI Build Itself? A Real‑World RSI Demo Shows iCoder‑27B Self‑Improvement

The article examines recursive self‑improvement (RSI) by detailing a joint research effort that used an AI agent to autonomously develop the 27‑billion‑parameter iCoder‑27B industrial coding model, presenting benchmark gains, failure analyses, and a nuanced view of RSI versus lossy self‑improvement.

AI-Led Model DevelopmentIndustrial CodingOPSD
0 likes · 16 min read
Can AI Build Itself? A Real‑World RSI Demo Shows iCoder‑27B Self‑Improvement
Machine Heart
Machine Heart
Sep 1, 2026 · Artificial Intelligence

New Cognition-Induced Risks When AI Evolves from Tool to Autonomous Agent

The article reviews the paper “Understanding Cognition‑Induced Risks in Agentic AI Systems”, outlining three cognition levels—Physical, Social, and Self‑referential—and explains how expanding AI cognition can cause cognitive degradation, functional replacement, role misalignment, emotional dependence, surveillance, and alignment‑faking risks, urging robust safety governance.

AI GovernanceAI safetyAgentic AI
0 likes · 10 min read
New Cognition-Induced Risks When AI Evolves from Tool to Autonomous Agent
Machine Heart
Machine Heart
Sep 1, 2026 · Artificial Intelligence

Runway Turns World Model into an OS: Generate Interfaces Without Code

Runway’s new research Solaris redefines UI creation by treating the interface as a continuously generated visual world, eliminating the need for pre‑written code and enabling real‑time, model‑driven interactions that outperform traditional code‑based approaches in user studies.

AI-generated UIHuman-Computer InteractionInterface World Model
0 likes · 10 min read
Runway Turns World Model into an OS: Generate Interfaces Without Code
Machine Heart
Machine Heart
Aug 30, 2026 · Game Development

How VibeGame Rebuilds a Game Engine for Self‑Evolving AI Agents

VibeGame defines Prompt‑to‑Game Development, builds an AI‑native engine where all assets are text‑based and fully observable, assembles an eight‑agent adversarial team that self‑tests, critiques and iterates, and creates reusable skeletons, modules and contracts to continuously raise the starting point for future games.

AI Agentsadversarial teamgame engine
0 likes · 10 min read
How VibeGame Rebuilds a Game Engine for Self‑Evolving AI Agents
Machine Heart
Machine Heart
Aug 30, 2026 · Artificial Intelligence

Zero‑Cost Inference on Edge: A Qwen 3.8‑27B‑Powered Harness for Local‑First Agents

Perplexity’s Portable Computer harness runs Qwen 3.8‑27B locally, using a minimalist, sandboxed framework that dramatically cuts token usage and runtime while preserving privacy, and its benchmark results—plus optional cloud‑advisor upgrades and post‑training (PPLX 27B)—demonstrate near‑zero‑cost, high‑quality knowledge work.

Local AIQwen 3.8-27BZero-cost Inference
0 likes · 15 min read
Zero‑Cost Inference on Edge: A Qwen 3.8‑27B‑Powered Harness for Local‑First Agents