CH·02CLI reference

gpu models get

Show details for one model (id or alias, optionally :economy)

Show details for one model (id or alias, optionally :economy)

Synopsis

Shows the full record for a single model — id, modality, tier, context length, per-1M-token pricing (input, cached input where the model publishes a cache-hit rate, and output), aliases, supported parameters, and status. Accepts a canonical id or any alias, optionally with a :economy suffix. Requires a key with serverless:read (or full_access).

gpu models get <model> [flags]

Examples

gpu models get gpuai/bge-large-en-v1.5
gpu models get bge-large                          # alias resolves to the canonical id
gpu models get gpuai/qwen2.5-7b-instruct:economy  # economy tier

Options

  -h, --help   help for get

Options inherited from parent commands

      --api-base string   API base URL (override with GPUAI_API_BASE env) (default "https://api.gpu.ai/v1")
      --debug             Enable debug logging to stderr
  -o, --output string     Output format: table|json (default table on TTY, json otherwise)

SEE ALSO

  • gpu models - Browse the serverless inference model catalog

Time-of-day pricing

A model priced by time of day also shows:

Pricing period                 off-peak
Peak windows (UTC)             01:00-04:00, 06:00-10:00
Peak $/1M (in/cached/out)      $0.3018 / $0.00114 / $0.8862
Off-peak $/1M (in/cached/out)  $0.1509 / $0.00057 / $0.4431

The Input $/1M / Output $/1M rows above them are the rates in effect at the moment the command ran (Pricing period says which). A request is billed at the rate in effect when it arrives, in UTC. Models without time-of-day pricing show none of these rows.

← The gpu CLI