What is the best open-source implementation of "Cosmos World Foundation Model Platform for Physical AI"?

The best maintained implementation is nvidia-cosmos/cosmos-predict1 with 448 stars on GitHub. Confidence: high. Reproducibility: Strong.

How reproducible is "Cosmos World Foundation Model Platform for Physical AI"?

Estimated time to first reproduction: a few hours. No risk flags identified. Start with nvidia-cosmos/cosmos-predict1 and validate setup instructions in README.

Are there pretrained models available for "Cosmos World Foundation Model Platform for Physical AI"?

Yes, 3 Hugging Face models found. The top result is nvidia/Cosmos-1.0-Diffusion-7B-Text2World with 6,326 downloads.

What framework is used to implement "Cosmos World Foundation Model Platform for Physical AI"?

The primary implementation uses pytorch.

Cosmos World Foundation Model Platform for Physical AI

NVIDIA, :, Niket Agarwal, Arslan Ali, Maciej Bala, Yogesh Balaji, Erik Barker, Tiffany Cai, Prithvijit Chattopadhyay, Yongxin Chen, Yin Cui, Yifan Ding, Daniel Dworakowski, Jiaojiao Fan, Michele Fenzi, Francesco Ferroni, Sanja Fidler, Dieter Fox, Songwei Ge, Yunhao Ge, Jinwei Gu, Siddharth Gururani, Ethan He, Jiahui Huang, Jacob Huffman, Pooya Jannaty, Jingyi Jin, Seung Wook Kim, Gergely Klár, Grace Lam, Shiyi Lan, Laura Leal-Taixe, Anqi Li, Zhaoshuo Li, Chen-Hsuan Lin, Tsung-Yi Lin, Huan Ling, Ming-Yu Liu, Xian Liu, Alice Luo, Qianli Ma, Hanzi Mao, Kaichun Mo, Arsalan Mousavian, Seungjun Nah, Sriharsha Niverty, David Page, Despoina Paschalidou, Zeeshan Patel, Lindsey Pavao, Morteza Ramezanali, Fitsum Reda, Xiaowei Ren, Vasanth Rao Naik Sabavat, Ed Schmerling, Stella Shi, Bartosz Stefaniak, Shitao Tang, Lyne Tchapmi, Przemek Tredak, Wei-Cheng Tseng, Jibin Varghese, Hao Wang, Haoxiang Wang, Heng Wang, Ting-Chun Wang, Fangyin Wei, Xinyue Wei, Jay Zhangjie Wu, Jiashu Xu, Wei Yang, Lin Yen-Chen, Xiaohui Zeng, Yu Zeng, Jing Zhang, Qinsheng Zhang, Yuxuan Zhang, Qingqing Zhao, Artur Zolkowski

Published: Jan 7, 2025

Best maintained implementation now

Evidence: Direct

Domain fit: AI-adjacent

Verified repos: 4

Top repo stars: 448

Paper appears method- or tooling-adjacent to AI workflows with partial ecosystem coverage.

Framework: pytorch

Time to first repro: a few hours

No risk flags

arXiv PDF

Physical AI needs to be trained digitally first. It needs a digital twin of itself, the policy model, and a digital twin of the world, the world model. In this paper, we present the Cosmos World Foundation Model Platform to help developers build customized world models for their Physical AI setups. We position a world foundation model as a general-purpose world model that can be fine-tuned into customized world model ...

Read full abstract

s for downstream applications. Our platform covers a video curation pipeline, pre-trained world foundation models, examples of post-training of pre-trained world foundation models, and video tokenizers. To help Physical AI builders solve the most critical problems of our society, we make Cosmos open-source and our models open-weight with permissive licenses available via https://github.com/nvidia-cosmos/cosmos-predict1.

Technical details

Canonical key: arxiv-2501.03575

Cache status: Fresh

Generated at: May 27, 2026, 1:52 PM

Artifact coverage: direct

HF provider: ok (token)

PWC source used: Yes

LLM status: not_generated

LLM model: n/a

LLM generated: Unknown

LLM content type: n/a

HF policy: hf-relevance-v27

implementation starting point

Benchmarks: thin evidence

Time to repro: a few hours

pytorch

Results & Benchmarks

Freshness tier: hot

Direct + Inferred Evidence

Some benchmark signal exists in the extracted evidence, but it is not structured strongly enough yet for a confident benchmark decision.

Physical AI needs to be trained digitally first.

Use This Implementation Because…

Confidence: high

nvidia-cosmos/cosmos-predict1 is the strongest maintained implementation based on ranking signals. CI workflows are present. License is declared (Apache-2.0).

Open nvidia-cosmos/cosmos-predict1

Reproduction Risks

No repository-level red flags were detected, but paper-specific preprocessing and hyperparameter details may still be under-specified.

Evidence disclosure

Evidence graph: 4 refs, 4 links.

Utility signals: depth 70/100, grounding 95/100, status medium.

Implementation Comparison

Top 3 paths

Compare maintenance quality, reproducibility coverage, and evidence confidence before choosing a reproduction baseline.

nvidia-cosmos/cosmos-predict1

best maintained

Maintenance: Recently updated

Confidence: High

Reproducibility: Strong

Official implementation from Papers with Code · Repository link is mentioned in the paper metadata

Stars: 448
Last push: Jan 6, 2026 (142d ago)

CIDockerfileReleasesDependencies

Risk flags

No obvious maintenance or reproducibility risks detected.

nvidia/cosmos-tokenizer

historical official

Maintenance: Archived

Confidence: High

Reproducibility: Limited

Official implementation from Papers with Code · Repository link is mentioned in the paper metadata

Stars: 1,724
Last push: Feb 11, 2025 (470d ago)

DockerfileDependencies

Risk flags

Repository archived
No push in 12+ months
No CI pipeline detected

nvlabs/tokenbench

alternative

Maintenance: Stale

Confidence: High

Reproducibility: Moderate

Official implementation from Papers with Code · Repository link is mentioned in the paper metadata

Stars: 159
Last push: Jan 13, 2025 (499d ago)

DockerfileDependencies

Risk flags

No push in 12+ months
No CI pipeline detected
No tagged releases

Best implementation now

nvidia-cosmos/cosmos-predict1

Confidence: High

Reproducibility: Strong

Cosmos-Predict1 is a collection of general-purpose world foundation models for Physical AI that can be fine-tuned into customized world models for downstream applications.

Stars: 448

Forks: 79

Last push: Jan 6, 2026

License: Apache-2.0

Official implementation from Papers with Code

Repository link is mentioned in the paper metadata

Strong overlap with paper title keywords

Community adoption signal (448 stars)

License ✓

CI ✓

Deps ✓

Docker ✓

Selected nvidia-cosmos/cosmos-predict1 as the strongest maintained implementation for new work.
Includes CI workflow signals.
Includes dependency/environment manifest signals.
Repository activity is within the last 24 months.

Historical official implementation

Preserved for provenance. Not recommended as the default path for new builds.

nvidia/cosmos-tokenizer

Stars: 1,724

Last push: Feb 11, 2025

Archived

Reproduction readiness

Ready to Run

Time to first repro: hours

Last checked: May 27, 2026

Ready to reproduce

· Clone nvidia-cosmos/cosmos-predict1 and install dependencies from pyproject.toml.
· Dockerfile available for containerized reproduction.
· CI pipeline detected — automated tests are in place.
· Last updated 142 days ago.

Open nvidia-cosmos/cosmos-predict1

Quick start

git clone https://github.com/nvidia-cosmos/cosmos-predict1.git
pip install -e .

Additional implementations

Official

nvlabs/tokenbench
Confidence: High

A Video Tokenizer Evaluation Dataset

Stars: 159

Forks: 12

Last push: Jan 13, 2025

License: Apache-2.0

Community

pham-tuan-binh/learning-world-model-learning
Confidence: Medium

Learning world model learning from scratch

Stars: 57

Last push: May 5, 2026

Possible but unverified matches (1)

These repositories had low-confidence matching signals and are hidden by default.

nvidia-cosmos/cosmos-transfer1

Confidence: Low

Stars: 805

Hugging Face artifacts

No direct paper-linked artifacts were found. Showing strongest curated related artifacts for faster exploration.

Models

nvidia/Cosmos-1.0-Diffusion-7B-Text2World

Curated Related

Downloads: 6,326

Likes: 234
nvidia/Cosmos-Guardrail1

Curated Related

Downloads: 4,288

Likes: 21
nvidia/Cosmos-1.0-Diffusion-7B-Video2World

Curated Related

Downloads: 1,599

Likes: 40

Broaden model search

Computer vision cosmos world foundation model

Datasets

No trustworthy dataset matches right now.

Search datasets on Hugging Face

Spaces

No trustworthy demo spaces right now.

Search spaces on Hugging Face

Explore on Hugging Face

Search models Search datasets Search spaces

Research context

Tasks

Computer vision

Methods

None detected

Domains

Computer vision

Evaluation & Human Feedback Data

Open this paper in HFEPX to review benchmark signals, evaluation modes, and human-feedback protocol context.

Open in HFEPX

Explore Similar Papers

Jump to Paper2Code search queries derived from this paper's research context.

Computer vision

Need human evaluators for your AI research? Scale annotation with expert AI Trainers.

Post a Job Get a Quote