Triton AI Docs
Developer API

Models

Compare current cloud and UC San Diego-hosted models by task, limits, and price.

This catalog reads https://tritonai-api.ucsd.edu/public/model_hub when the page loads. It includes only records where is_public_model_group is exactly true.

The site caches a successful response for five minutes. If a later request fails, the site can show the last successful in-memory response.

Hosting and data use

  • UC San Diego-hosted: These on-prem routes run on infrastructure managed by UC San Diego.
  • Enterprise cloud: These routes use approved enterprise cloud providers.

Read the Triton AI trust, privacy, and hosting guidance before you select a route for protected data.

Access can differ

The table shows public model groups. Your Developer API approval controls the routes that your key can use.

How models are ordered

The main lists favor the current model generation and distinct task-specific models. Earlier versions remain available in a collapsed section. New aliases stay visible while task guidance is pending.

Loading the public model catalog.

How to choose

Start with the task guidance for each model. Then compare quality, latency, and cost with representative inputs from your application.

The capability fields come from the live Model Hub. The task guidance comes from provider documentation and model cards. A provider can change a model without changing its API alias.

Field notes

Prop

Type

On this page