Overview

The open-model control room for SMB & mid-market teams.

Stop guessing which model to use.

LLM Switchboard catalogs every open model on Groq and NVIDIA build.nvidia.com β€” plus 303+ models you can run on your own hardware β€” scores them on the dimensions that matter for your job, and routes each request to the best fit. Cloud or local. One API. No lock-in.

⚑ 53 cloud models Groq Β· 18 NVIDIA Β· 27 ⬇ 303 run locally < 25GB
53
Cloud models
18 Groq Β· 27 NVIDIA
303
Local models < 25GB
run on your own hardware
1.0k t/s
Fastest (cloud)
GPT-OSS 20B
15
Benchmarks graded
by trustworthiness (BQS)
⬇ Run it on your own hardware no API key, $0
Browse 303 models β†’

303 open models under 25 GB with one-command Ollama/Docker setup, picked by benchmark β€” and a built-in sandbox to test them in-browser. reasoning 80 Β· coding 35 Β· vision 46 Β· STT 45 Β· TTS 38 Β· embeddings 59

What's inside
Talk to the SDR live demo
✸ icompaas SDR · routed to fastest Groq model
This is use case #3, embedded

The homepage SDR chat is just a LLM Switchboard recipe: job type chat, policy fastest that passes, pinned to Groq for sub-second replies. Drop a key in .env and it answers live; without one it shows which model it would call.

See all business recipes β†’
Sign in to continue

LLM Switchboard is private β€” sign in with Authlee to access the control room.

Sign in with Authlee
← Back to home