# world-model

Published articles for world-model.

This is one page of public article previews, not the complete archive. Follow Next page to continue. Summaries are not the original full articles.

## REFACTOR-VLA: Unsupervised Library Learning of Typed Motor Programs

DevFeed: [REFACTOR-VLA: Unsupervised Library Learning of Typed Motor Programs](<https://devfeed.tech/articles/refactor-vla-unsupervised-library-learning-of-typed-motor-programs-6732.md>)

Original publisher: [Read original article](<https://machinelearning.apple.com/research/refactor-vla-motor-programs>)

Published: 2026-09-02T00:00:00Z

Content type: article

Language: en

Sources: [Apple Machine Learning Research](<https://devfeed.tech/sources/apple-machine-learning-research.md>)

Topics: [World models](<https://devfeed.tech/topics/world-models.md>), [Large Language Model](<https://devfeed.tech/topics/llm.md>)

Tags: [models](<https://devfeed.tech/tags/models.md>), [skills](<https://devfeed.tech/tags/skills.md>), [training](<https://devfeed.tech/tags/training.md>), [world-model](<https://devfeed.tech/tags/world-model.md>)

### AI overview

REFACTOR-VLA is a vision-language-action system that learns reusable typed motor-program skills through alternating world-model-based clustering and policy optimization. On LIBERO, the article reports that scaling the world model reduced performance across all four benchmark suites, while an InfoNCE auxiliary loss improved skill clustering.

### Source excerpt

Most current vision-language-action (VLA) models--such as OpenVLA, π0, RT-2, and RDT-1B--are "monolithic." This means they generate raw motor commands or very short sequences of actions, without organizing behaviors into reusable, well-defined abstractions. As a result, these models perform poorly on long-horizon (multi-step) tasks, and it's difficult to interpret what they have learned. Existing approaches for discovering skills often avoid the core problem of deciding when two action sequences are "behaviorally equivalent." For example, AtomicVLA and AtomSkill group action sequences by...

## Beyond VLAs: How World Action Models Reshape Robot Manipulation

DevFeed: [Beyond VLAs: How World Action Models Reshape Robot Manipulation](<https://devfeed.tech/articles/beyond-vlas-how-world-action-models-reshape-robot-manipulation-6764.md>)

Original publisher: [Read original article](<https://developer.nvidia.com/blog/beyond-vlas-how-world-action-models-reshape-robot-manipulation/>)

Author: Michelle Horton

Published: 2026-08-04T16:00:00Z

Content type: article

Language: en

Sources: [NVIDIA Developer](<https://devfeed.tech/sources/nvidia-developer.md>), [NVIDIA Technical Blog](<https://devfeed.tech/sources/nvidia-technical-blog.md>)

Topics: [Robotics](<https://devfeed.tech/topics/robotics.md>), [World models](<https://devfeed.tech/topics/world-models.md>), [vlm](<https://devfeed.tech/topics/vlm.md>), [post-training](<https://devfeed.tech/topics/post-training.md>), [Cosmos](<https://devfeed.tech/topics/cosmos.md>), [Nvidia](<https://devfeed.tech/topics/nvidia.md>), [NVIDIA Research](<https://devfeed.tech/topics/nvidia-research.md>)

Tags: [agentic-ai-generative-ai](<https://devfeed.tech/tags/agentic-ai-generative-ai.md>), [cosmos](<https://devfeed.tech/tags/cosmos.md>), [featured](<https://devfeed.tech/tags/featured.md>), [nvidia](<https://devfeed.tech/tags/nvidia.md>), [nvidia-research](<https://devfeed.tech/tags/nvidia-research.md>), [physical-ai](<https://devfeed.tech/tags/physical-ai.md>), [post-training](<https://devfeed.tech/tags/post-training.md>), [robot-manipulation](<https://devfeed.tech/tags/robot-manipulation.md>), [robotics](<https://devfeed.tech/tags/robotics.md>), [simulation-modeling-design](<https://devfeed.tech/tags/simulation-modeling-design.md>), [thor](<https://devfeed.tech/tags/thor.md>), [vlm](<https://devfeed.tech/tags/vlm.md>), [world-model](<https://devfeed.tech/tags/world-model.md>), [zero-shot](<https://devfeed.tech/tags/zero-shot.md>)

### AI overview

The article explains how World Action Models (WAMs) use video world models as backbones for robot policies, addressing the physical-generalization limitations of vision-language-action models. It discusses post-training WAMs into specialized policies and presents NVIDIA Cosmos 3 as a foundation for building them.

### Source excerpt

A central challenge in robotics is building policies that generalize beyond the demonstrations they're trained on. A policy that succeeds in a training scene...

## Developing Healthcare Robotics with GPU-Native Medical Physics Simulation

DevFeed: [Developing Healthcare Robotics with GPU-Native Medical Physics Simulation](<https://devfeed.tech/articles/developing-healthcare-robotics-with-gpu-native-medical-physics-simulation-6808.md>)

Original publisher: [Read original article](<https://developer.nvidia.com/blog/developing-healthcare-robotics-with-gpu-native-medical-physics-simulation/>)

Author: Michelle Horton

Published: 2026-07-28T20:49:21Z

Content type: article

Language: en

Sources: [NVIDIA Developer](<https://devfeed.tech/sources/nvidia-developer.md>), [NVIDIA Technical Blog](<https://devfeed.tech/sources/nvidia-technical-blog.md>)

Topics: [Isaac for Healthcare](<https://devfeed.tech/topics/isaac-for-healthcare.md>), [Robotics](<https://devfeed.tech/topics/robotics.md>), [Reinforcement learning](<https://devfeed.tech/topics/reinforcement-learning.md>), [GPU](<https://devfeed.tech/topics/gpu.md>), [datasets](<https://devfeed.tech/topics/datasets.md>), [Medical imaging](<https://devfeed.tech/topics/medical-imaging.md>), [Cosmos](<https://devfeed.tech/topics/cosmos.md>), [Open Source](<https://devfeed.tech/topics/open-source.md>), [real-time](<https://devfeed.tech/topics/real-time.md>), [Nvidia](<https://devfeed.tech/topics/nvidia.md>)

Tags: [agentic-ai-generative-ai](<https://devfeed.tech/tags/agentic-ai-generative-ai.md>), [data](<https://devfeed.tech/tags/data.md>), [datasets](<https://devfeed.tech/tags/datasets.md>), [featured](<https://devfeed.tech/tags/featured.md>), [gpu](<https://devfeed.tech/tags/gpu.md>), [healthcare](<https://devfeed.tech/tags/healthcare.md>), [isaac](<https://devfeed.tech/tags/isaac.md>), [isaac-for-healthcare](<https://devfeed.tech/tags/isaac-for-healthcare.md>), [isaac-sim](<https://devfeed.tech/tags/isaac-sim.md>), [medical-imaging](<https://devfeed.tech/tags/medical-imaging.md>), [nvidia](<https://devfeed.tech/tags/nvidia.md>), [open-source](<https://devfeed.tech/tags/open-source.md>), [physical-ai](<https://devfeed.tech/tags/physical-ai.md>), [physics](<https://devfeed.tech/tags/physics.md>), [real-time](<https://devfeed.tech/tags/real-time.md>), [reinforcement-learning](<https://devfeed.tech/tags/reinforcement-learning.md>), [robotics](<https://devfeed.tech/tags/robotics.md>), [simulation-modeling-design](<https://devfeed.tech/tags/simulation-modeling-design.md>), [warp](<https://devfeed.tech/tags/warp.md>), [world-model](<https://devfeed.tech/tags/world-model.md>)

### AI overview

The article presents NVIDIA's open source, GPU-accelerated Medical Physics Simulation framework for healthcare robotics. It addresses limited medical robotics data, poor generalization, and slow development by enabling anatomical digital twins, device-anatomy and medical imaging simulation, and GPU-scale reinforcement learning within Isaac for Healthcare, Isaac Sim, and Isaac Lab.

### Source excerpt

Unlike autonomous driving or industrial robotics, healthcare robotics can't rely on internet-scale data collection or unlimited real-world experimentation....

## LeRobot v0.6.0: Imagine, Evaluate, Improve

DevFeed: [LeRobot v0.6.0: Imagine, Evaluate, Improve](<https://devfeed.tech/articles/lerobot-v0-6-0-imagine-evaluate-improve-7329.md>)

Original publisher: [Read original article](<https://huggingface.co/blog/lerobot-release-v060>)

Author: Steven Palma; Pepijn Kooijmans; Caroline Pascal; Khalil Meftah; Maxime Ellerbach; Martino Russi; Nikodem Bartnik; Nicolas Rabault; Thomas Wolf

Published: 2026-07-07T00:00:00Z

Content type: article

Language: en

Sources: [Hugging Face - Blog](<https://devfeed.tech/sources/hugging-face-blog.md>)

Topics: [lerobot](<https://devfeed.tech/topics/lerobot.md>), [World models](<https://devfeed.tech/topics/world-models.md>), [Robotics](<https://devfeed.tech/topics/robotics.md>), [vlm](<https://devfeed.tech/topics/vlm.md>), [Cosmos](<https://devfeed.tech/topics/cosmos.md>), [foundation-models](<https://devfeed.tech/topics/foundation-models.md>), [datasets](<https://devfeed.tech/topics/datasets.md>), [Open Source](<https://devfeed.tech/topics/open-source.md>)

Tags: [artificial-intelligence](<https://devfeed.tech/tags/artificial-intelligence.md>), [cli](<https://devfeed.tech/tags/cli.md>), [cosmos](<https://devfeed.tech/tags/cosmos.md>), [datasets](<https://devfeed.tech/tags/datasets.md>), [gpu](<https://devfeed.tech/tags/gpu.md>), [lerobot](<https://devfeed.tech/tags/lerobot.md>), [models](<https://devfeed.tech/tags/models.md>), [open-source](<https://devfeed.tech/tags/open-source.md>), [robotics](<https://devfeed.tech/tags/robotics.md>), [science](<https://devfeed.tech/tags/science.md>), [training](<https://devfeed.tech/tags/training.md>), [vlm](<https://devfeed.tech/tags/vlm.md>), [world-model](<https://devfeed.tech/tags/world-model.md>)

### AI overview

LeRobot v0.6.0 adds world-model policies, new vision-language-action models, reward-model APIs, simulation benchmarks, human-in-the-loop CLI corrections, FSDP and cloud training. It also expands dataset capabilities with depth support, automated language annotation, custom video encoding, and faster data loading.

### Source excerpt

We're on a journey to advance and democratize artificial intelligence through open source and open science.

## How Jua delivers the world's most accurate physics simulations 3x faster with ClickHouse Cloud

DevFeed: [How Jua delivers the world's most accurate physics simulations 3x faster with ClickHouse Cloud](<https://devfeed.tech/articles/how-jua-delivers-the-world-s-most-accurate-physics-simulations-3x-faster-with-clickhouse-cloud-5361.md>)

Original publisher: [Read original article](<https://clickhouse.com/blog/jua-physics-foundation-model>)

Author: ClickHouse

Published: 2026-06-30T00:00:00Z

Content type: article

Language: en

Sources: [ClickHouse Blog](<https://devfeed.tech/sources/clickhouse-blog.md>)

Topics: [clickhouse](<https://devfeed.tech/topics/clickhouse.md>), [Cloud](<https://devfeed.tech/topics/cloud.md>), [data](<https://devfeed.tech/topics/data.md>), [foundation-models](<https://devfeed.tech/topics/foundation-models.md>), [World models](<https://devfeed.tech/topics/world-models.md>), [Artificial Intelligence](<https://devfeed.tech/topics/ai.md>), [Compression](<https://devfeed.tech/topics/compression.md>), [Fine-tuning](<https://devfeed.tech/topics/fine-tuning.md>), [Google](<https://devfeed.tech/topics/google.md>), [Microsoft](<https://devfeed.tech/topics/microsoft.md>), [Nvidia](<https://devfeed.tech/topics/nvidia.md>)

Tags: [ai](<https://devfeed.tech/tags/ai.md>), [architecture](<https://devfeed.tech/tags/architecture.md>), [benchmarks](<https://devfeed.tech/tags/benchmarks.md>), [clickhouse](<https://devfeed.tech/tags/clickhouse.md>), [cloud](<https://devfeed.tech/tags/cloud.md>), [compression](<https://devfeed.tech/tags/compression.md>), [compute](<https://devfeed.tech/tags/compute.md>), [cost](<https://devfeed.tech/tags/cost.md>), [data](<https://devfeed.tech/tags/data.md>), [efficiency](<https://devfeed.tech/tags/efficiency.md>), [fine-tuning](<https://devfeed.tech/tags/fine-tuning.md>), [global](<https://devfeed.tech/tags/global.md>), [google](<https://devfeed.tech/tags/google.md>), [microsoft](<https://devfeed.tech/tags/microsoft.md>), [net-11](<https://devfeed.tech/tags/net-11.md>), [nvidia](<https://devfeed.tech/tags/nvidia.md>), [renewables](<https://devfeed.tech/tags/renewables.md>), [speed](<https://devfeed.tech/tags/speed.md>), [world-model](<https://devfeed.tech/tags/world-model.md>)

### AI overview

Jua uses ClickHouse Cloud to deliver physics simulation data for energy forecasting faster and more efficiently. The platform reduced forecast delivery time from one hour to 20 minutes, cut compute costs by a third, and reduced historical query times from hours to seconds. Jua's EPT-2 physics foundation model learns atmospheric physics from observational data and can transfer to other fluid-dynamics problems with minimal fine-tuning.

### Source excerpt

Jua replaced a file-based forecast pipeline with ClickHouse Cloud, cutting data delivery time from one hour to 20 minutes and historical query times from hours to seconds -- giving energy traders a faster edge.

## Waypoint-1.5: Higher-Fidelity Interactive Worlds for Everyday GPUs

DevFeed: [Waypoint-1.5: Higher-Fidelity Interactive Worlds for Everyday GPUs](<https://devfeed.tech/articles/waypoint-1-5-higher-fidelity-interactive-worlds-for-everyday-gpus-7565.md>)

Original publisher: [Read original article](<https://huggingface.co/blog/waypoint-1-5>)

Author: Andrew Lapp; Louis Castricato; Scott Fox; Shahbuland Matiana; David Rossi

Published: 2026-04-09T00:00:00Z

Content type: article

Language: en

Sources: [Hugging Face - Blog](<https://devfeed.tech/sources/hugging-face-blog.md>)

Topics: [World models](<https://devfeed.tech/topics/world-models.md>), [Hardware](<https://devfeed.tech/topics/hardware.md>), [real-time](<https://devfeed.tech/topics/real-time.md>), [data](<https://devfeed.tech/topics/data.md>), [AI, ML & Data Engineering](<https://devfeed.tech/topics/ai-ml-data-engineering.md>), [systems](<https://devfeed.tech/topics/systems.md>), [Artificial Intelligence](<https://devfeed.tech/topics/ai.md>)

Tags: [announcement](<https://devfeed.tech/tags/announcement.md>), [artificial-intelligence](<https://devfeed.tech/tags/artificial-intelligence.md>), [compute](<https://devfeed.tech/tags/compute.md>), [data](<https://devfeed.tech/tags/data.md>), [diffusion](<https://devfeed.tech/tags/diffusion.md>), [game-dev](<https://devfeed.tech/tags/game-dev.md>), [hardware](<https://devfeed.tech/tags/hardware.md>), [local](<https://devfeed.tech/tags/local.md>), [models](<https://devfeed.tech/tags/models.md>), [performance](<https://devfeed.tech/tags/performance.md>), [real-time](<https://devfeed.tech/tags/real-time.md>), [training](<https://devfeed.tech/tags/training.md>), [world-model](<https://devfeed.tech/tags/world-model.md>)

### AI overview

Waypoint-1.5 is Overworld's real-time video world model, designed to generate interactive environments locally on consumer hardware. The release adds 720p and 360p model tiers, supports up to 60 FPS on higher-end desktop GPUs, broadens hardware accessibility, uses nearly 100 times more training data than Waypoint-1, and applies more efficient video modeling to improve coherence and responsiveness.

### Source excerpt

We're on a journey to advance and democratize artificial intelligence through open source and open science.

## Project Genie: Experimenting with infinite, interactive worlds

DevFeed: [Project Genie: Experimenting with infinite, interactive worlds](<https://devfeed.tech/articles/project-genie-experimenting-with-infinite-interactive-worlds-6233.md>)

Original publisher: [Read original article](<https://deepmind.google/blog/project-genie-experimenting-with-infinite-interactive-worlds/>)

Author: Diego Rivas

Published: 2026-01-29T17:01:05Z

Content type: release

Language: en

Sources: [Google DeepMind News](<https://devfeed.tech/sources/google-deepmind-news.md>)

Topics: [World models](<https://devfeed.tech/topics/world-models.md>), [Artificial Intelligence](<https://devfeed.tech/topics/ai.md>), [Google AI](<https://devfeed.tech/topics/google-ai.md>), [Simulation](<https://devfeed.tech/topics/simulation.md>), [Web app](<https://devfeed.tech/topics/webapp.md>)

Tags: [3d](<https://devfeed.tech/tags/3d.md>), [ai](<https://devfeed.tech/tags/ai.md>), [experimental](<https://devfeed.tech/tags/experimental.md>), [google](<https://devfeed.tech/tags/google.md>), [google-ai](<https://devfeed.tech/tags/google-ai.md>), [none](<https://devfeed.tech/tags/none.md>), [prototype](<https://devfeed.tech/tags/prototype.md>), [research](<https://devfeed.tech/tags/research.md>), [research-prototype](<https://devfeed.tech/tags/research-prototype.md>), [simulation](<https://devfeed.tech/tags/simulation.md>), [web](<https://devfeed.tech/tags/web.md>), [web-app](<https://devfeed.tech/tags/web-app.md>), [world-model](<https://devfeed.tech/tags/world-model.md>)

### AI overview

Google is opening Project Genie to Google AI Ultra subscribers in the U.S. The experimental research prototype, powered by Genie 3, Nano Banana Pro and Gemini, lets users create, explore and remix interactive worlds.

### Source excerpt

Google AI Ultra subscribers in the U.S. can try out Project Genie, an experimental research prototype that lets you create and explore worlds.

## Introducing Waypoint-1: Real-time interactive video diffusion from Overworld

DevFeed: [Introducing Waypoint-1: Real-time interactive video diffusion from Overworld](<https://devfeed.tech/articles/introducing-waypoint-1-real-time-interactive-video-diffusion-from-overworld-7564.md>)

Original publisher: [Read original article](<https://huggingface.co/blog/waypoint-1>)

Author: Andrew Lapp; Louis Castricato; Scott Fox; Shahbuland Matiana; David Rossi

Published: 2026-01-20T00:00:00Z

Content type: article

Language: en

Sources: [Hugging Face - Blog](<https://devfeed.tech/sources/hugging-face-blog.md>)

Topics: [World models](<https://devfeed.tech/topics/world-models.md>), [real-time](<https://devfeed.tech/topics/real-time.md>), [Artificial Intelligence](<https://devfeed.tech/topics/ai.md>), [Inference](<https://devfeed.tech/topics/inference.md>), [Latency](<https://devfeed.tech/topics/latency.md>), [Streaming](<https://devfeed.tech/topics/streaming.md>), [Transformer](<https://devfeed.tech/topics/transformer.md>), [Tooling](<https://devfeed.tech/topics/tooling.md>), [AI Development](<https://devfeed.tech/topics/ai-development.md>), [Open Source](<https://devfeed.tech/topics/open-source.md>)

Tags: [announcement](<https://devfeed.tech/tags/announcement.md>), [artificial-intelligence](<https://devfeed.tech/tags/artificial-intelligence.md>), [diffusion](<https://devfeed.tech/tags/diffusion.md>), [inference](<https://devfeed.tech/tags/inference.md>), [latency](<https://devfeed.tech/tags/latency.md>), [model](<https://devfeed.tech/tags/model.md>), [models](<https://devfeed.tech/tags/models.md>), [open-source](<https://devfeed.tech/tags/open-source.md>), [performance](<https://devfeed.tech/tags/performance.md>), [procedural](<https://devfeed.tech/tags/procedural.md>), [real-time](<https://devfeed.tech/tags/real-time.md>), [streaming](<https://devfeed.tech/tags/streaming.md>), [tooling](<https://devfeed.tech/tags/tooling.md>), [train](<https://devfeed.tech/tags/train.md>), [world-model](<https://devfeed.tech/tags/world-model.md>)

### AI overview

Overworld introduces Waypoint-1, a real-time interactive video diffusion model that generates explorable worlds from frames and responds to text, mouse, and keyboard controls. The article describes its frame-causal rectified flow transformer, training on diverse video game footage, diffusion forcing, self-forcing, and the WorldEngine inference library.

### Source excerpt

We're on a journey to advance and democratize artificial intelligence through open source and open science.