Skip to content

Repository files navigation

OpenRouter Benchmark Harness

OpenRouter's internal benchmarking harness, externalized for transparency. We port benchmarks here so we can run them scalably on our infrastructure and iterate quickly.

bun install
OPENROUTER_API_KEY=... bun run bench -- --benchmark gpqa_diamond --model openai/gpt-4o-mini --limit 5

See CONTRIBUTING.md before proposing changes. Report security issues privately as described in SECURITY.md.

About

OpenRouter TypeScript harness for reproducible LLM benchmarks and evaluations.

Topics

Resources

Contributing

Security policy

Stars

1 star

Watchers

0 watching

Forks

Releases

Packages

Contributors

Languages