LLMInferenceProbeResult
Outcome of the optional generation round-trip, reported separately from the model-list probe so "the list opens but inference fails" is legible.
the completion request returned 2xx
ok | auth_failed | http_error | unreachable | invalid_url | model_required
Human-readable inference result in English. Clients may localize it using the structured reason value.
status_code object
upstream HTTP status when a response was received
- integer
- null
probed_url object
completion URL actually requested
- string
- null
body object
Upstream response body, masked and length-capped. It is also returned when the request fails, and is null when no response was received.
- string
- null
content object
generated text, masked and length-capped
- string
- null
reasoning_content object
chain of thought when the endpoint reports one (Anthropic thinking blocks, OpenAI-compatible reasoning_content/reasoning)
- string
- null
usage object
token usage block the endpoint reported, as-is
- object
- null
finish_reason object
OpenAI finish_reason / Anthropic stop_reason, when reported
- string
- null
ttft_ms object
Time to the first token in milliseconds, measured from request start to the first streamed text or thinking delta. It is null unless measure_ttft was requested, and also when the stream carried no delta.
- number
- null
{
"success": true,
"reason": "string",
"message": "string",
"latency_ms": 0,
"status_code": 0,
"probed_url": "string",
"body": "string",
"content": "string",
"reasoning_content": "string",
"usage": {},
"finish_reason": "string",
"ttft_ms": 0
}