Cosmos Predict 2.5: what the repository publishes
Cosmos Predict 2.5 fills every cell: the NVIDIA Open Model License with a sentence saying models are commercially usable, 2,059,174,912 parameters, a requirement of 32.54 GB of GPU VRAM, and 720P video at 16FPS delivered as a 5-second clip. As of 2026-09-22.
| Field | What is stated | Detail |
|---|---|---|
| Licence | NVIDIA Open Model License | The card states that models are commercially usable |
| Size | 2,059,174,912 parameters | Published as an exact count, not a rounded figure |
| Hardware | 32.54 GB of GPU VRAM | Ampere, Hopper and Blackwell named as supported |
| Resolution | 720P | Stated with a 16FPS frame rate |
| Length | A 5-second clip | Seconds, stated alongside the frame rate |
Inclusion rule. The same five fields for every family here. A field the repository does not state is recorded as not stated rather than estimated from the parameter count. Order. Fixed field order, identical on every family page.
1The fullest row here, and the precision is the interesting part
A parameter count to the digit, a memory figure to two decimal places, a frame rate, a clip length and a licence that addresses commercial use inside the file. Every other entry in this register leaves at least one of those for the reader to find elsewhere or to guess.
Precision is not accuracy, and the card is careful about which it is offering: the VRAM figure is written as a requirement for this model rather than as a measurement of one particular run. A reader still has to ask what settings it assumes. They do not have to ask what the number is.
2A licence that says commercially usable inside its own text
The card states the model is released under the NVIDIA Open Model License and that models are commercially usable. That is a vendor-written licence answering, in its own words, the question most licences in this register hand off to a separate page or leave alone.
The register still records the name rather than the permission, for the same reason it does everywhere else: a sentence on a model card summarises a document, and the document governs. But the presence of that sentence is itself a fact, and it is the reason this row reads differently from the other vendor-written licences here.
3Architectures named, not just one card
Ampere, Hopper and Blackwell are listed as supported, which is a statement about card generations rather than about a single product. Anybody holding a card can place it in a generation without doing any memory arithmetic first.
Naming generations and stating 32.54 GB together is what makes a hardware plan checkable from published material: the generation settles whether the numeric formats are supported, the figure settles whether the memory is. Most rows in this register carry one of those two and leave the other open.
4The same field, across every family
| Family | Licence |
|---|---|
| Allegro | apache-2.0 |
| CogVideoX | Split by model size |
| Cosmos Predict 2.5 | NVIDIA Open Model License |
| FramePack | Apache 2.0 |
| HunyuanVideo | tencent-hunyuan-community |
| LongCat-Video | MIT License |
| LTX-Video | Apache 2.0 |
| MAGI-1 | Apache License 2.0 |
| Mochi 1 | Apache 2.0 |
| Open-Sora | Apache 2.0 |
| Pyramid Flow | stabilityai-ai-community |
| SkyReels-V2 | skywork-license |
| Stable Video Diffusion | stable-video-diffusion-community |
| Step-Video-T2V | MIT |
| Wan | Apache 2.0 |
Inclusion rule. Families whose vendor has published weights openly. A family whose repository does not state this particular field keeps its row, marked hollow. Order. Alphabetical by family name.
| Family | Size |
|---|---|
| Allegro | VAE 175M and DiT 2.8B |
| CogVideoX | 2B and 5B, carried in the model names |
| Cosmos Predict 2.5 | 2,059,174,912 parameters |
| FramePack | 13B models |
| HunyuanVideo | Over 13 billion parameters |
| LongCat-Video | 13.6B parameters |
| LTX-Video | A 13B model |
| MAGI-1 | 24B and 4.5B models |
| Mochi 1 | 10 billion parameters |
| Open-Sora | An 11B model |
| Pyramid Flow | No parameter count recorded |
| SkyReels-V2 | 1.3B, 5B and 14B variants |
| Stable Video Diffusion | 2B params |
| Step-Video-T2V | 30 billion parameters |
| Wan | A 5B text-to-video and image-to-video model |
Inclusion rule. Families whose vendor has published weights openly. A family whose repository does not state this particular field keeps its row, marked hollow. Order. Alphabetical by family name.
| Family | Hardware |
|---|---|
| Allegro | 9.3G BF16 with cpu_offload |
| CogVideoX | From 4GB in FP16 on the 2B |
| Cosmos Predict 2.5 | 32.54 GB of GPU VRAM |
| FramePack | 6GB minimum for 1800 frames at 30fps |
| HunyuanVideo | 60GB peak at 720px1280px, 129 frames |
| LongCat-Video | No memory figure recorded |
| LTX-Video | No minimum stated |
| MAGI-1 | H100 or H800 x 8 for the 24B |
| Mochi 1 | About 60GB VRAM on a single GPU |
| Open-Sora | 60.3GB peak at 768px on one GPU |
| Pyramid Flow | No memory figure recorded |
| SkyReels-V2 | About 14.7GB peak on the 1.3B at 540P |
| Stable Video Diffusion | About 180s on an A100 80GB card |
| Step-Video-T2V | 78.55 GB peak at 768px768px204f |
| Wan | A consumer 4090 on the 5B release |
Inclusion rule. Families whose vendor has published weights openly. A family whose repository does not state this particular field keeps its row, marked hollow. Order. Alphabetical by family name.
- LicenceNVIDIA Open Model License, with the card stating that models are commercially usablea vendor-written licence answering the commercial question in its own text
- Model sizeNumber of model parameters: 2,059,174,912published as an exact count
- HardwareThis model requires 32.54 GB of GPU VRAM, on Ampere, Hopper or Blackwell hardwarea figure and the supported architectures
- Resolution and lengthProduces 720P video with 16FPS, as a 5-second clipboth stated on the card
5Sources
Read from huggingface.co/nvidia on 2026-09-22. Hardware figures across every family are collected on requirements; the licence column is set out on licence and what counts as a published weight is on how read. Previous family: MAGI-1. Next family: LongCat-Video.