nodus create inferenceendpoint
View Markdownnodus create inferenceendpoint
Section titled “nodus create inferenceendpoint”Create a inferenceendpoint
nodus create inferenceendpoint NAME [flags]Examples
Section titled “Examples” nodus create inferenceendpoint support-bot --model nodus/gpt-oss-120b --rpm 120 --max-cost 5Options
Section titled “Options” --allowed-key stringArray API key name allowed to call it (repeatable) --dry-run string[="server"] server returns the estimate without creating; client prints the manifest (default "none") -h, --help help for inferenceendpoint --max-concurrent int Open requests --max-cost string Spend limit in USD --model string Catalog model, as 'nodus inference models' lists them -o, --output string Output format: name, json, yaml, or estimate (with --dry-run=server) --rpm int Requests per minute --tpm int Tokens per minuteOptions inherited from parent commands
Section titled “Options inherited from parent commands” --context string Context from the config file to use --org string Organization (selects the context for that org) -p, --project string Project to work in -v, --verbose Log each API request (never credentials)SEE ALSO
Section titled “SEE ALSO”- nodus create - Create resources from manifests or with a typed generator