arXiv · Computation and Language · 13 Aug 2026 · paper
Embodied agents are increasingly built as systems around foundation models, where performance depends not only on model weights but also on the skills, context, action interfaces, and execu…
arXiv · Machine Learning · 13 Aug 2026 · paper
Autonomous rehabilitation systems must not only recognize human motion but also provide structured feedback to support users without continuous therapist supervision. This paper presents a…
arXiv · Machine Learning · 13 Aug 2026 · paper
A central goal in robot learning is to move beyond task-specific human data collection toward robots that improve through autonomous interaction. Yet fully autonomous learning remains diffi…
arXiv · Artificial Intelligence · 13 Aug 2026 · paper
Rather than providing an exhaustive survey, this paper presents a concise tutorial on world models and world action models for robotics. After reading the tutorial, readers should have a cl…
arXiv · Machine Learning · 13 Aug 2026 · paper
Contact-rich manipulation requires force sensitivity, but many robot arms lack dedicated force sensors due to their high cost. We present Neural External Torque Estimation (NEXT), a data-dr…
arXiv · Machine Learning · 13 Aug 2026 · paper
Fine-tuning pre-trained robot policies with reinforcement learning (RL) often inherits the bottlenecks introduced by pre-training with behavioral cloning (BC), which produces narrow action…
arXiv · Artificial Intelligence · 13 Aug 2026 · paper
After an older adult misses a medication dose, a software agent can send another reminder and an embodied agent can bring the medication. Yet neither explains whether the person forgot, is…
arXiv · Machine Learning · 13 Aug 2026 · paper
Safe offline RL typically assumes access to dense per-step cost annotations, but in practice supervisors provide only trajectory-level stop-feedback: a binary signal at the first unsafe tra…
arXiv · Artificial Intelligence · 13 Aug 2026 · paper
Integrating locomotion and manipulation is essential for robot autonomy, but scaling standard Reinforcement Learning (RL) to complex tasks is severely bottlenecked by the slow, manual proce…
arXiv · Artificial Intelligence · 13 Aug 2026 · paper
The prevailing recipe for Vision-Language-Action (VLA) models couples a pretrained VLM with a separately trained flow-matching action expert. This makes the VLM a context encoder rather tha…
arXiv · Artificial Intelligence · 13 Aug 2026 · paper
Motion-centric video reasoning is fundamental to interactive applications such as robotic manipulation and autonomous navigation. However, multimodal large language models (MLLMs) typically…
arXiv · Artificial Intelligence · 13 Aug 2026 · paper
The convergence of artificial intelligence (AI), Industrial Internet of Things, cyber-physical systems, and advanced robotics is reshaping manufacturing faster than engineering curricula ca…
arXiv · Artificial Intelligence · 13 Aug 2026 · paper
World action models (WAMs) condition robot actions on predicted futures, but iterative video rollout increases deployment latency. We ask whether action generation requires the evolving rol…
arXiv · Artificial Intelligence · 13 Aug 2026 · paper
World Action Models (WAMs) couple future visual prediction with robot action generation, enabling policies to model how the physical world evolves during interaction. Existing WAMs differ i…
arXiv · Machine Learning · 13 Aug 2026 · paper
The success of coding agents has established the harness as a paradigm: what an agent achieves depends not on the model alone, but on the infrastructure around it. We ask whether the same p…
arXiv · Artificial Intelligence · 13 Aug 2026 · paper
Comparative feedback, asking people which of two behaviors they prefer, has become a standard way to align robot and agent behavior with human intent when the reward itself cannot be specif…
arXiv · Artificial Intelligence · 13 Aug 2026 · paper
Cyber-physical systems (CPS) are typically developed by multiple stakeholders who produce artefacts tailored to their specific domains of expertise. The behaviour of these systems emerges f…
NVIDIA Developer · 11 Aug 2026
Video is a core data path across NVIDIA Jetson applications, from robotics and intelligent video analytics to industrial automation, healthcare, media...
NVIDIA Developer · 04 Aug 2026
A central challenge in robotics is building policies that generalize beyond the demonstrations they’re trained on. A policy that succeeds in a training scene...
Google · 30 Jul 2026 · release software
Two robots: Duo and Apollo
NVIDIA Developer · 28 Jul 2026
Unlike autonomous driving or industrial robotics, healthcare robotics can’t rely on internet-scale data collection or unlimited real-world experimentation....
Hugging Face · 27 Jul 2026
Hugging Face · 21 Jul 2026
NVIDIA Developer · 20 Jul 2026
Developers building 3D, design, simulation, robotics, and industrial digital twin applications need ways to bring physical AI capabilities into the tools and...
NVIDIA Developer · 12 Jul 2026 · tutorial
Robotics foundation models have made remarkable progress. Today's best systems can follow natural language instructions to pick, place, sort, and manipulate a...
NVIDIA Developer · 07 Jul 2026
As more teams move from humanoid robot bring-up to task-specific skill development, the need for repeatable development workflows is growing. Building humanoids...
Hugging Face · 07 Jul 2026
Berkeley AI Research · 01 Jul 2026 · paper
Congratulations to the Berkeley Artificial Intelligence Research (BAIR) Lab class of 2026! This year, BAIR celebrates another remarkable group of Ph.D. graduates whose curiosity, creativity…
NVIDIA Developer · 24 Jun 2026
An increasingly common design pattern for autonomous vehicles (AVs), robotics, and spatial AI systems is bird's-eye-view (BEV) perception. BEV models project...
NVIDIA Developer · 22 Jun 2026
Physical AI—robots working autonomously alongside people in factories, warehouses, hospitals, and homes—is arriving faster than most expected. Traditional...
Hugging Face · 17 Jun 2026
NVIDIA Developer · 15 Jun 2026
Quick glossary for readers new to VLA/WAM terminology VLA Vision-Language-Action model: a robot policy that starts from a pretrained VLM backbone and adapts it...
NVIDIA Developer · 01 Jun 2026
Physical AI systems must understand the real world before they can act within it. Robots, autonomous vehicles, and smart spaces need to understand what's...
Berkeley AI Research · 20 Apr 2026 · paper
GRASP is a new gradient-based planner for learned dynamics (a “world model”) that makes long-horizon planning practical by (1) lifting the trajectory into virtual states so optimization is…
Hugging Face · 09 Mar 2026
Hugging Face · 05 Mar 2026
Hugging Face · 13 Nov 2025
Berkeley AI Research · 01 Nov 2025 · paper
In this post, I’ll introduce a reinforcement learning (RL) algorithm based on an “alternative” paradigm: divide and conquer . Unlike traditional methods, this algorithm is not based on temp…
Hugging Face · 29 Oct 2025
Hugging Face · 28 Oct 2025 · tutorial
Hugging Face · 24 Oct 2025
Hugging Face · 16 Sep 2025
Hugging Face · 10 Jul 2025
Hugging Face · 09 Jul 2025
Hugging Face · 11 Jun 2025
Hugging Face · 03 Jun 2025
Hugging Face · 11 May 2025
Hugging Face · 14 Apr 2025
Hugging Face · 11 Mar 2025
Hugging Face · 04 Feb 2025
Hugging Face · 27 Aug 2024
Hugging Face · 25 Mar 2024
OpenAI · 15 Oct 2019 · comunicato aziendale
We’ve trained a pair of neural networks to solve the Rubik’s Cube with a human-like robot hand. The neural networks are trained entirely in simulation, using the same reinforcement learning…
OpenAI · 05 Jun 2019 · comunicato aziendale
We hosted the first OpenAI Robotics Symposium on April 27, 2019.
OpenAI · 07 Nov 2018 · comunicato aziendale
We’ve developed an energy-based model that can quickly learn to identify and generate instances of concepts, such as near, above, between, closest, and furthest, expressed as sets of 2d poi…
OpenAI · 30 Jul 2018 · comunicato aziendale
We’ve trained a human-like robot hand to manipulate physical objects with unprecedented dexterity.
OpenAI · 26 Feb 2018 · comunicato aziendale
We’re releasing eight simulated robotics environments and a Baselines implementation of Hindsight Experience Replay, all developed for our research over the past year. We’ve used these envi…
OpenAI · 26 Feb 2018 · comunicato aziendale
OpenAI · 19 Oct 2017 · comunicato aziendale
Our latest robotics techniques allow robot controllers, trained entirely in simulation and deployed on physical robots, to react to unplanned changes in the environment as they solve simple…
OpenAI · 18 Oct 2017 · comunicato aziendale