OpenAI API Check | Latency, Cache, Model Purity and Relay Test
Check OpenAI API and OpenAI-compatible relays online: connectivity, P95/TTFT latency, model purity, prompt cache, rate limits and token cost.
Quick Answer
An OpenAI API check validates connectivity, response model fields, P95 latency, TTFT, cached_tokens, rate limit behavior and token usage to determine whether an OpenAI channel is usable, stable and free of relay anomalies.
How It Works
- 1Choose the OpenAI-compatible protocol.
- 2Enter API base URL, model id and a temporary API key.
- 3Run streaming and non-streaming probes to collect latency, TTFT and usage fields.
- 4Check cached_tokens and model identity consistency.
Use Cases
- Verify official OpenAI API connectivity.
- Check whether an OpenAI-compatible relay strips cache fields.
- Validate GPT model latency before production.
- Debug 401, 429, 5xx or invalid model errors.
Core Metrics
Connectivity
Auth, model id and basic response validity
cached_tokens
OpenAI usage field for prompt cache hits
TTFT
Streaming time to first token
Rate limits
429 and x-ratelimit response headers
Why OpenAI-compatible APIs need testing
Many gateways claim OpenAI compatibility, but usage, streaming, tool calls, cache fields and error codes may differ. A check exposes these differences before launch.
How to use the report
Use the report for initial channel screening and pre-rollout acceptance. For procurement, repeat with fixed regions, business prompts and larger samples.
Related Search Terms
FAQ
Does AIBench store my OpenAI API key?
No. The key is only used for the current probe request and is not written to reports, logs or databases. Use a temporary test key whenever possible.
Can it test OpenAI-compatible relays?
Yes. Enter the relay base URL and model id. The report focuses on latency, cache fields, model purity and rate limit behavior.
Related Guides and Comparisons
Start a Real API Check
Paste a temporary test key and run the same probes for latency, model purity, cache, rate limits and cost.