▸ surp.ivc.lol usage
new: genuinely free AI models are live — try surp/free + see live budgets · token-gating prototype · vote on SRP

Surp Value Index (SVI)

the one number: cost × intelligence × speed

A single composite score per model — weighted geometric mean of three normalized sub-scores. Buyers optimize for the highest SVI in class; suppliers climb the leaderboard by submitting verified benchmark results or improving real served TPS.

## weights

cost 45% · intelligence 40% · speed 15%

## leaderboard

#modelSVIcostintelspeedpricetps
1gpt-5.6-luna83.770.595.0100.0$0.0184/M271.8 tps
2deepseek-v4-flash66.8100.060.026.5$0.0091/M71.9 tps
3glm-5.266.473.270.042.9$0.0171/M116.6 tps
4deepseek-v4-pro59.168.170.024.7$0.0197/M67.0 tps
5deepseek-v4-flash-073150.350.460.031.3$0.0360/M85.0 tps
6llama-3.3-70b-instruct47.644.470.021.1$0.0464/M57.4 tps

## route by your own lens

You don't have to use the default weights. Pass surp_mode to any combo to route by the lens that fits the job:

modeweights (cost·intel·speed)use for
costpure cheapestovernight batch, agents that can wait
value45·40·15default SVI — best all-round value
balanced33·33·33no strong preference
speed15·15·70interactive work, pair programming
intel20·60·20hard reasoning problems
curl -X POST https://surp.ivc.lol/v1/chat/completions \
  -H "Content-Type: application/json" \
  -d '{"model":"surp/best-chat","surp_mode":"speed",
       "messages":[{"role":"user","content":"hi"}]}'

Or go fully custom with surp_weights — a cost:intel:speed triple, e.g. "surp_weights":"0.3:0.4:0.3". Models without a verified TPS fall back to cheapest so routing never fails. The 402 response includes routing_reason so you can see which lens won.

## submit a benchmark

curl -X POST https://surp.ivc.lol/api/svi/benchmark \
  -H "Content-Type: application/json" \
  -d '{"model":"my-quantized-model","mmlu":88,"humaneval":92,"submitter":"you"}'

Verified submissions move the leaderboard. Missing axes fall back to the model class default so partial submissions still count.