Image
Cosmos 3 Super Image
Cosmos3 is a collection of Omnimodal world models capable of generating dynamic, high-quality video, image, audio, and action commands from combinations of text, image, video, and action trajectory inputs.
- Text to image
Sign in to create. Review the credit cost in the studio before running.
Catalog preview for Cosmos 3 Super Image. Review your own result before publishing.
Use Cosmos 3 Super Image through the API.
Start with this model ID for the endpoint featured on this page:
nvidia/cosmos-3-super/text-to-image
Check the current model schema for supported inputs and options. Other tasks may use different model IDs.
From quote to result.
Request a credit quote for your input and settings before submitting a generation. A sample output or the price of another model is not a quote for your request.
Follow the generation workflow, then poll the job for its status and result. The API and studio use the same balance.