tangyue0820 commited on
Commit
daf3d37
·
verified ·
1 Parent(s): da579b9

docs: scope model versions and parameter list to Cosmos3-Super-Text2Image

Browse files
Files changed (1) hide show
  1. README.md +0 -16
README.md CHANGED
@@ -32,18 +32,6 @@ This model is ready for commercial and non-commercial use.
32
  **Model Developer:** NVIDIA
33
 
34
  ### Model Versions
35
- - Cosmos3-Nano:
36
- - Given multimodal inputs including text, images, video, audio, and action trajectories, generate coherent text, images, video, audio, and action outputs for multimodal understanding, world simulation, future prediction, action reasoning, and Physical AI applications.
37
-
38
- - Cosmos3-Super:
39
- - Given multimodal inputs including text, images, video, audio, and action trajectories, generate coherent text, images, video, audio, and action outputs for multimodal understanding, world simulation, future prediction, action reasoning, and Physical AI applications.
40
-
41
- - Cosmos3-Nano-Policy-DROID:
42
- - Given language instructions and visual observations from the DROID robot platform, generate robot action trajectories for manipulation and control tasks.
43
-
44
- - Cosmos3-Super-Image2Video:
45
- - Given one input image and text instructions, generate temporally coherent video sequences that are consistent with the provided visual content.
46
-
47
  - Cosmos3-Super-Text2Image:
48
  - Given text input, generate high-fidelity images that are consistent with the provided description.
49
 
@@ -76,10 +64,6 @@ Cosmos3 is an Omni-modal foundation model built on a Mixture-of-Transformers (Mo
76
 
77
  **Number of trainable model parameters:**
78
 
79
- - Cosmos3-Nano: 16B
80
- - Cosmos3-Super: 64B
81
- - Cosmos3-Nano-Policy-DROID: 16B
82
- - Cosmos3-Super-Image2Video: 64B
83
  - Cosmos3-Super-Text2Image: 64B
84
 
85
  ## Input/Output Specifications
 
32
  **Model Developer:** NVIDIA
33
 
34
  ### Model Versions
 
 
 
 
 
 
 
 
 
 
 
 
35
  - Cosmos3-Super-Text2Image:
36
  - Given text input, generate high-fidelity images that are consistent with the provided description.
37
 
 
64
 
65
  **Number of trainable model parameters:**
66
 
 
 
 
 
67
  - Cosmos3-Super-Text2Image: 64B
68
 
69
  ## Input/Output Specifications