Are you an LLM? Read llms.txt for a summary of the docs, or llms-full.txt for the full context.
Skip to content
Endpoints / Usage
GEThttps://api.lazu.ai/api/usage/requests/{request_id}

Request details

Fetch the customer receipt for one completed request, including normalized usage dimensions, billing line items and public timing.

BillingRead-only reference endpoint; this is not a model generation request. Pricing & lanes ↗

Path parameters

request_idstringrequired

Request ID for a request made by the same API key. Lazu rejects attempts to read another key's request details.

Response fields

idstring

The requested request ID.

modelstring

Requested model or route name.

usage.measurementobjectnullable

Input and output provenance: provider, local_estimate, or unavailable. This optional field is absent on historical or unclassified records; absence does not prove provider measurement.

usage.dimensionsobject[]

Normalized dimensions such as input, output, cache_read, cache_write_5m, cache_write_1h, cache_miss, audio and image dimensions.

billing.line_itemsobject[]

Billable usage lines with price, quantity and total charge.

routingobjectnullable

Public status and timing only. Internal routing decisions and provider identities are not exposed.

Cache field guidance

OpenAI-compatible responses may show cache reads as

usage.prompt_tokens_details.cached_tokens or usage.input_tokens_details.cached_tokens. Cache writes may appear as cache_write_tokens, cache_write_5m_tokens or cache_write_1h_tokens only when the upstream reports them.

For billing and reconciliation, this endpoint is the more complete source.

See also

RequestGET /api/usage/requests/{request_id}
curl https://api.lazu.ai/api/usage/requests/req_lazu_01ABCDEF \
  -H "Authorization: Bearer $LAZU_API_KEY"
ResponseExample
{
  "object": "usage.request",
  "id": "req_lazu_01ABCDEF",
  "created_at": 1765980000,
  "model": "gpt-6-luna",
  "endpoint": "/v1/chat/completions",
  "usage": {
    "prompt_tokens": 1200,
    "completion_tokens": 300,
    "total_tokens": 1500,
    "dimensions": [
      { "name": "input", "quantity": 1200, "unit": "token" },
      { "name": "cache_read", "quantity": 300, "unit": "token" },
      { "name": "output", "quantity": 300, "unit": "token" }
    ]
  },
  "billing": {
    "currency": "USD",
    "amount_microusd": 480,
    "pricing_version": 3,
    "line_items": [
      {
        "dimension": "input",
        "quantity": 1200,
        "unit_microusd": 150,
        "amount_microusd": 180
      }
    ]
  },
  "routing": {
    "status_code": 200,
    "total_duration_ms": 1840,
    "ttft_ms": 420,
    "was_streaming": true
  },
  "provider_usage": {
    "family": "openai",
    "raw_fields": { "prompt_tokens": 1200, "completion_tokens": 300 }
  },
  "log_id": "log_01ABCDEF"
}
Request receipt

After a billable request, use its gateway request ID to inspect usage dimensions, billing line items and timing.

Read a real request receipt →

Where request_id comes from

Read X-Lazu-Request-Id, an error payloadrequest_id, or the gateway response request ID.