List accelerators
Returns items and a total for the accelerators held by the tenant selected by the X-CostGraph-Tenant-ID header. An accelerator that has reported gets its own item carrying its utilisation band, its 24h, 7d and 30d utilisation and memory, its hourly and monthly cost where the catalog prices it, and its MIG partitions with the ones that ran no compute marked idle. The accelerators on a machine that have not reported are counted in one further item per machine, with a null gpu_uuid, unreported_card_count set, and every measured field null rather than zero. A machine that reports its accelerators through the host agent rather than through Kubernetes gets its own item in the same shape, counted in unreported_card_count with every measured field null, since the agent reports what the hardware is and never how busy it was; such a machine is listed separately from any Kubernetes node and is never merged with one. model_source names what answered for the model, and reported_by names which path answered for the accelerator: operator for a card reporting through Kubernetes, agent for one the host agent inventoried, and catalog for cards the machine is priced for but none has reported. placement says where the accelerator sits and carries the ids a caller links on: placement.instance_id always, and placement.cluster_id, placement.node_id and placement.node_name where the machine is a Kubernetes node we hold, which is also what makes placement.kind workload rather than instance. placement.namespace names the namespace of the workload that asked for the accelerator, and stays null where no workload on that node asks for one or where more than one namespace does. gpu_availability states every signal the card could report, read from the measurements stored against it: supported where it was measured, capacity_unknown where it was measured but nothing expresses it as a percentage, unsupported where the card reports other signals but not this one, and no_data where the card reported nothing at all; no signal is ever omitted or returned as zero. A report whose 7-day or 30-day window has not been computed carries window_7d or window_30d as null rather than as a zero measurement. cost carries the currency every amount is in, the hourly and projected monthly rate the catalog prices the card at, and mtd_reason saying why mtd and change_percent are null; spend to date is billed for the whole machine, so it is never attributed to one accelerator and never rendered as zero. utilization_sparkline and memory_utilization_sparkline each carry the trailing 24 hours as 12 fixed-width 2 hour buckets, oldest first, each holding the highest percentage the card reached in that bucket; a bucket the card took no sample in is null, a card that has never reported carries twelve nulls rather than zeros, and a card carved into MIG partitions is read once for the whole card rather than once per partition. total_count is the filtered count before paging.
Authorizations
Enter "Bearer {token}"
Headers
Tenant ID
Query Parameters
Filter to one accelerator model
Filter to one utilisation band
Filter to accelerators held by one instance
Filter to accelerators on the machines of one Kubernetes cluster
Filter to one or more cloud providers
Filter to where the accelerator sits: instance or workload
Case-insensitive match over the accelerator model and the name of the machine holding it
Filter to partitioned or whole cards
Only accelerators that reported within this window, as a duration such as 24h, 7d or 30d. Omit to include ones that stopped reporting.
Number of rows to return (default 50, max 200)
Rows to skip