RunLog Benchmark

Your prompt. Real model responses.

Compare response text, elapsed request time, token usage and provider-reported cost. Review outputs yourself: RunLog does not assign an automated quality score.

Run a comparison

Requires an active Benchmark entitlement and a RunLog customer API key. Requests send your prompt to the configured model gateway and selected model providers. Model availability is checked by the service; unavailable models return an error. Manage your key · Check subscription availability

Held only in this page; never saved to browser storage.Supported model IDs are loaded from the service configuration. Maximum 4,000 prompt characters and 512 output tokens per model.

No benchmark has been run.