Skip to content
Explore models

Image

Cosmos 3 Super Image

Cosmos3 is a collection of Omnimodal world models capable of generating dynamic, high-quality video, image, audio, and action commands from combinations of text, image, video, and action trajectory inputs.

  • Text to image
Create with Cosmos 3 Super Image
Preview for Cosmos 3 Super Image

Catalog preview for Cosmos 3 Super Image. Review your own result before publishing.

Use Cosmos 3 Super Image through the API.

Start with this model ID for the endpoint featured on this page:

nvidia/cosmos-3-super/text-to-image

Check the current model schema for supported inputs and options. Other tasks may use different model IDs.

From quote to result.

Request a credit quote for your input and settings before submitting a generation. A sample output or the price of another model is not a quote for your request.

Follow the generation workflow, then poll the job for its status and result. The API and studio use the same balance.