Instance Type Service

List inference instance types

GET
/public/inference-instance-types

Lists hardware instance types currently available to inference deployments, including GPU resources, pricing, regions, and best-effort capacity headroom.

Authorization

bearerAuth
AuthorizationBearer <token>

In: header

Response Body

application/json

application/json

curl -X GET "https://example.com/public/inference-instance-types"
{  "data": [    {      "id": "string",      "name": "string",      "description": "string",      "gpuType": "string",      "gpuCount": 0,      "gpuMemoryGib": 0,      "priceCentsPerHour": 0,      "regions": [        {          "name": "string",          "headroom": {            "value": 0,            "relation": "RELATION_EQ"          },          "compliance": [            {              "policy": {                "hipaa": true              },              "headroom": {                "value": 0,                "relation": "RELATION_EQ"              }            }          ]        }      ]    }  ],  "next_cursor": "string",  "object": "list"}