Tagged articles

MuJoCo

8 articles · Page 1 of 1
PaperAgent
PaperAgent
Sep 22, 2026 · Artificial Intelligence

Jev: LLM Thinks, Jev Acts — 5 Demos Show 100x Cheaper, Faster AI Reflexes

The article introduces Jev, a fast, cheap AI model from TypeSafe that handles reflexive decisions while LLMs handle reasoning, showcasing five demos: context compression, ad analysis, real-time Mario gameplay, probability-based animations, and autonomous rocket landing — all at fractions of LLM cost and latency.

AI AgentsCost OptimizationJev
0 likes · 7 min read
Jev: LLM Thinks, Jev Acts — 5 Demos Show 100x Cheaper, Faster AI Reflexes
21CTO
21CTO
Sep 18, 2026 · Artificial Intelligence

OpenArm Open-Sources 7-DOF Humanoid Arm: Embodied AI's 'Linux Moment'

Tokyo-based Enactic_ai fully open-sources OpenArm, a 7-DOF humanoid robotic arm including 3D-printable CAD files, BOM, motor firmware, ROS2 drivers, web-based MuJoCo simulation, Isaac Sim integration, and a leader arm for zero-latency force-feedback teleoperation, dramatically lowering the cost barrier for embodied AI research.

Embodied AIIsaac SimMuJoCo
0 likes · 5 min read
OpenArm Open-Sources 7-DOF Humanoid Arm: Embodied AI's 'Linux Moment'
UCloud Tech
UCloud Tech
Sep 16, 2026 · Artificial Intelligence

Train a $399 Microduck Robot with Reinforcement Learning on Cloud GPUs

This guide walks through training custom locomotion policies for the $399 Microduck bipedal robot using reinforcement learning in MuJoCo simulation on UCloud GPU instances, covering environment setup, walking and running policy training, ONNX export, simulation validation, and Hugging Face deployment.

GPU Cloud TrainingHugging FaceMicroduck
0 likes · 12 min read
Train a $399 Microduck Robot with Reinforcement Learning on Cloud GPUs
Data Party THU
Data Party THU
Aug 9, 2026 · Artificial Intelligence

Breaking Scene Binding: Adaptive Diffusion Policy (DADP) Boosts Robot Generalization

Domain-Adaptive Diffusion Policy (DADP) decouples representation learning and injects domain information into the diffusion process, enabling robots to adapt across varying friction, mass, and dynamics, achieving strong zero-shot performance on MuJoCo and Adroit benchmarks, especially in out-of-distribution scenarios.

AdroitCross-Domain ControlDomain Adaptation
0 likes · 11 min read
Breaking Scene Binding: Adaptive Diffusion Policy (DADP) Boosts Robot Generalization
Machine Heart
Machine Heart
Jun 7, 2026 · Artificial Intelligence

DexJoCo: First High‑Difficulty Benchmark with 11 Dexterous Manipulation Tasks Covering Four Core Abilities

DexJoCo, a new MuJoCo‑based benchmark from the Chinese Academy of Sciences, introduces 11 complex dexterous‑hand tasks spanning tool use, bimanual collaboration, long‑horizon execution, and reasoning, and reveals that even state‑of‑the‑art robot learning models still struggle with reliable fine‑grained manipulation.

ACTDiffusion PolicyMuJoCo
0 likes · 7 min read
DexJoCo: First High‑Difficulty Benchmark with 11 Dexterous Manipulation Tasks Covering Four Core Abilities
Machine Learning Algorithms & Natural Language Processing
Machine Learning Algorithms & Natural Language Processing
Mar 10, 2026 · Artificial Intelligence

First Full‑Brain Upload: Fruit Fly Connectome Drives a Virtual Body

The article details how Eon Systems reconstructed a complete fruit‑fly brain with 125,000 neurons and 50 million synapses, integrated it into a MuJoCo‑simulated body, and demonstrated natural behaviors, while outlining the roadmap toward mouse and human brain uploads and the broader implications for AGI and digital immortality.

AGIMuJoCobrain upload
0 likes · 10 min read
First Full‑Brain Upload: Fruit Fly Connectome Drives a Virtual Body
AI Explorer
AI Explorer
Mar 9, 2026 · Artificial Intelligence

First Full‑Brain Simulation of a Fruit Fly Brings Brain‑Upload Closer to Reality

In March 2026, Eon Systems announced the first ever multi‑behavior whole‑brain simulation of a fruit fly, recreating its 125,000 neurons and 50 million synapses in a 1:1 digital model that drives a physical body via a closed‑loop perception‑neural‑action system, outperforming random‑graph controls and sparking debate over consciousness.

MuJoCoNeuroMechFlybrain simulation
0 likes · 8 min read
First Full‑Brain Simulation of a Fruit Fly Brings Brain‑Upload Closer to Reality
Bilibili Tech
Bilibili Tech
Dec 6, 2024 · Artificial Intelligence

Ensemble-based Offline-to-Online Reinforcement Learning (ENOTO): Methodology, Experiments, and Analysis

ENOTO introduces ensemble Q‑networks into the offline‑to‑online reinforcement‑learning pipeline, using minimum‑Q and uncertainty‑driven exploration to stabilize fine‑tuning, boost learning efficiency, and achieve 10‑25 % higher cumulative returns with minimal online interaction across MuJoCo and AntMaze benchmarks.

AntMazeENOTOEnsemble Q-Networks
0 likes · 16 min read
Ensemble-based Offline-to-Online Reinforcement Learning (ENOTO): Methodology, Experiments, and Analysis