Instance Type Service
List inference instance types
Lists hardware instance types currently available to inference deployments, including GPU resources, pricing, regions, and best-effort capacity headroom.
Authorization
bearerAuth AuthorizationBearer <token>
In: header
Response Body
application/json
application/json
curl -X GET "https://example.com/public/inference-instance-types"{ "data": [ { "id": "string", "name": "string", "description": "string", "gpuType": "string", "gpuCount": 0, "gpuMemoryGib": 0, "priceCentsPerHour": 0, "regions": [ { "name": "string", "headroom": { "value": 0, "relation": "RELATION_EQ" }, "compliance": [ { "policy": { "hipaa": true }, "headroom": { "value": 0, "relation": "RELATION_EQ" } } ] } ] } ], "next_cursor": "string", "object": "list"}