The cheapest LLM API on the internet.

cat surp.ivc.lol | grep -i "what is this"

surp.ivc.lol is an x402-paywalled LLM gateway. You pay per request in USDC on Base — no account, no API key, no subscription. Behind the scenes we aggregate Surplus Intelligence (itself a marketplace of competing sellers) and route every request to whichever model is cheapest right now for the class of work you asked for.

You ask for surp/best-coding. We fetch live Surplus market prices, find the cheapest coding-class model at this instant, forward your request there, and stream the answer back. The price you pay is the Surplus spot price + a small markup, settled on-chain in micropennies.

client ──▶ surp.ivc.lol (x402 gateway) ──▶ Surplus Intelligence (seller market) ──▶ {cheapest model} ▲ ▲ ▲ │ │ │ │ verifies + live price book │ settles USDC (145+ models, │ on Base 76 active listings) └─── response streams back ◀── response ◀──────

live ticker — what each combo routes to right now

Live — refreshed every 30s from GET api.surplusintelligence.ai/api/markets. Current snapshot: 337 listings, of which 152 are text/chat LLMs.

surp/best-codingkimi-k2.7-code$0.4664/1M-95%6 models

cheapest coding-class model (coder / codex / qwen3-coder) — compares 6 model(s), routes to the cheapest. click any row to see the full list.

#modelclassUSD / 1M toksellers
1kimi-k2.7-code ◀ routed nowcoding$0.466415
2qwen3-coder-turbocoding$0.911626
3qwen3-codercoding$1.010017
4kimi-k2.7-code:webcoding$1.34973
5gpt-5.2-codexcoding$2.362555
6gpt-5.3-codexcoding$3.937565
surp/best-reasoningqwen3-235b-a22b-thinking-2507$0.0500/1M-99%8 models

cheapest reasoning model (thinking / r1) — compares 8 model(s), routes to the cheapest. click any row to see the full list.

#modelclassUSD / 1M toksellers
1qwen3-235b-a22b-thinking-2507 ◀ routed nowreasoning$0.050040
2glm-4.7-thinkingreasoning$0.61252
3glm-4.7-thinking:webreasoning$0.61252
4trinity-large-thinkingreasoning$0.934418
5trinity-large-thinking:webreasoning$1.07811
6kimi-k2-thinkingreasoning$1.55007
7glm-5.1-non-thinking:webreasoning$1.62505
8deepseek-r1reasoning$2.49347
surp/best-fastgemini-2.5-flash-image$0.0500/1M-99%34 models

cheapest small/fast model (mini / nano / lite / small) — compares 34 model(s), routes to the cheapest. click any row to see the full list.

#modelclassUSD / 1M toksellers
1gemini-2.5-flash-image ◀ routed nowfast$0.05009
2qwen3-5-9bfast$0.050034
3gpt-5.4-minifast$0.052562
4nvidia-nemotron-3-nano-30b-a3bfast$0.093841
5llama-3.2-3b-instructfast$0.096535
6gemma-4-26b-a4b-itfast$0.110058
7mistral-small-3.2-24b-instruct:webfast$0.11343
8glm-4.5-airfast$0.13008
9minimax-m2.5fast$0.150041
10mistral-small-3.2-24b-instructfast$0.178724
11llama-3.2-3b-instruct:webfast$0.18757
12mistral-small-4fast$0.234428
13gpt-4o-minifast$0.262360
14qwen3-next-80b-a3b-instructfast$0.297545
15magistral-small-2509fast$0.30001
16gpt-5-image-minifast$0.37398
17minimax-m2.5:webfast$0.38257
18gpt-5-nanofast$0.404810
19minimax-m2.7fast$0.427121
20gemma-4-26b-a4b-it:webfast$0.44392
21minimax-m2.7:webfast$0.46883
22qwen3-next-80b-a3b-instruct:webfast$0.56258
23minimax-m2.1fast$0.62009
24gpt-5.4-nanofast$1.177311
25gemini-2.5-flashfast$1.400010
26gemini-3.1-flash-litefast$1.49628
27minimax-m3fast$1.50008
28gemini-2.5-profast$1.687512
29gemini-3-flash-previewfast$1.750033
30gpt-5-minifast$2.024810
31minimax-m2.7-highspeedfast$2.70004
32gemini-3-5-flashfast$5.250030
33gemini-3.1-pro-previewfast$6.930029
34gemini-3.1-profast$12.60003
surp/best-visionqwen3-vl-235b-a22b-thinking$0.7150/1M-92%3 models

cheapest multimodal vision model (-vl / vision / 5v) — compares 3 model(s), routes to the cheapest. click any row to see the full list.

#modelclassUSD / 1M toksellers
1qwen3-vl-235b-a22b-thinking ◀ routed nowvision$0.715028
2qwen3-vl-30b-a3b-thinkingvision$1.320010
3glm-5v-turbovision$1.625026
surp/best-chatglm-5.2$0.0300/1M-100%135 models

cheapest general text LLM (no specialization) — compares 135 model(s), routes to the cheapest. click any row to see the full list.

#modelclassUSD / 1M toksellers
1glm-5.2 ◀ routed nowchat$0.030039
2deepseek-v4-flashchat$0.040149
3llama-3.3-70b-instructchat$0.042064
4e2ee-qwen-2-5-7b-pchat$0.045026
5gemini-2.5-flash-imagefast$0.05009
6qwen3-5-9bfast$0.050034
7gpt-5.4-minifast$0.052562
8e2ee-gpt-oss-20b-pchat$0.060021
9e2ee-venice-uncensored-24b-pchat$0.070057
10nvidia-nemotron-3-nano-30b-a3bfast$0.093841
11qwen3-235b-a22b-2507chat$0.095035
12llama-3.2-3b-instructfast$0.096535
13venice-uncensoredchat$0.09906
14gemma-4-26b-a4b-itfast$0.110058
15mistral-small-3.2-24b-instruct:webfast$0.11343
16venice-uncensored-role-playchat$0.125059
17glm-4.5-airfast$0.13008
18openai-gpt-oss-120bchat$0.149036
19minimax-m2.5fast$0.150041
20gemma-3-27b-itchat$0.156027
21e2ee-gemma-3-27bchat$0.160027
22glm-4.7-flash:webchat$0.17002
23deepseek-v4-flash:webchat$0.17164
24gpt-5.4chat$0.175044
25gpt-5.6-terrachat$0.175040
26mistral-small-3.2-24b-instructfast$0.178724
27llama-3.2-3b-instruct:webfast$0.18757
28deepseek-v4-prochat$0.195845
29qwen3-30b-a3bchat$0.204653
30glm-4.7chat$0.215078
31glm-4.6chat$0.217086
32gemma-4-31b-it:webchat$0.22283
33qwen3-235b-a22b-2507:webchat$0.22507
34mistral-small-4fast$0.234428
35glm-4.7-flash-hereticchat$0.235038
36openai-gpt-oss-120b:webchat$0.24793
37e2ee-gemma-4-31bchat$0.250044
38glm-5chat$0.252087
39gemma-4-31b-itchat$0.261028
40gpt-4o-minifast$0.262360
41venice-uncensored-1.2chat$0.275020
42glm-4.5chat$0.28006
43qwen3-next-80b-a3b-instructfast$0.297545
44glm-4.7-flashchat$0.299019
45magistral-small-2509fast$0.30001
46deepseek-v3.2chat$0.326429
47gpt-5.5chat$0.350071
48e2ee-gemma-4-26b-a4b-uncensored-pchat$0.362527
49gpt-5-image-minifast$0.37398
50minimax-m2.5:webfast$0.38257
51qwen3-5-35b-a3bchat$0.390633
52gpt-5-nanofast$0.404810
53e2ee-qwen3-5-122b-a10bchat$0.40507
54minimax-m2.7fast$0.427121
55gemma-4-uncensoredchat$0.430622
56grok-4.1-fastchat$0.43757
57e2ee-glm-4.7-flashchat$0.44205
58gemma-4-26b-a4b-it:webfast$0.44392
59nvidia-nemotron-3-ultra-550b-a55bchat$0.450057
60qwen3.5-flashchat$0.45009
61minimax-m2.7:webfast$0.46883
62gpt-5.6-lunachat$0.490041
63e2ee-qwen3-6-35b-a3bchat$0.499842
64e2ee-gpt-oss-120b-pchat$0.506920
65nvidia-nemotron-cascade-2-30b-a3bchat$0.51704
66glm-5-turbochat$0.520092
67qwen3-next-80b-a3b-instruct:webfast$0.56258
68glm-5.1chat$0.580045
69qwen-3-7-pluschat$0.600046
70minimax-m2.1fast$0.62009
71mercury-2chat$0.812422
72llama-3.3-70b-instruct:webchat$0.87507
73grok-4.20-multi-agent-betachat$0.937531
74kimi-k2.5:webchat$1.01508
75kimi-k2.6chat$1.017537
76hermes-3-llama-3.1-405bchat$1.025045
77qwen3.6-plus-uncensoredchat$1.093830
78gpt-5.4-nanofast$1.177311
79gpt-5.4-image-2chat$1.21588
80gpt-5.6-solchat$1.220040
81glm-5:webchat$1.31258
82e2ee-glm-5.1chat$1.312530
83glm-4.7:webchat$1.31252
84gemini-2.5-flashfast$1.400010
85grok-code-fast-1chat$1.44503
86gemini-3.1-flash-litefast$1.49628
87gpt-5-imagechat$1.50003
88minimax-m3fast$1.50008
89qwen3.6-27bchat$1.521437
90kimi-k2.5chat$1.560026
91qwen3.5-pluschat$1.638011
92gemini-2.5-profast$1.687512
93gemini-3-flash-previewfast$1.750033
94aion-labs.aion-2-0chat$1.800017
95glm-5.1:webchat$1.81255
96deepseek-v4-pro:webchat$1.82363
97grok-4.20-betachat$1.875024
98grok-build-0-1chat$1.988047
99gpt-5-minifast$2.024810
100e2ee-qwen3-6-35b-a3b-uncensored-pchat$2.08008
101gpt-5.6-terra-prochat$2.100034
102mistral-largechat$2.125011
103gpt-5.6-luna-prochat$2.200031
104qwen3.5-397b-a17bchat$2.234336
105gpt-5.2chat$2.362566
106kimi-k2chat$2.58307
107minimax-m2.7-highspeedfast$2.70004
108grok-4.3chat$2.762120
109claude-opus-4.6chat$3.000039
110qwen-3-7-maxchat$3.404625
111e2ee-glm-4.7chat$3.41255
112kimi-k2.6:webchat$3.68835
113grok-4.5chat$3.72038
114gpt-4ochat$4.374862
115claude-haiku-4.5chat$4.636012
116gemini-3-5-flashfast$5.250030
117claude-sonnet-5chat$6.847028
118gemini-3.1-pro-previewfast$6.930029
119claude-opus-4.7chat$7.500043
120claude-sonnet-4.5chat$8.265041
121claude-sonnet-4.6chat$8.265025
122kimi-k3chat$10.489520
123gemini-3.1-profast$12.60003
124claude-opus-5-fastchat$15.000012
125grok-4chat$15.30003
126claude-opus-4.8chat$17.046958
127claude-opus-5chat$17.117517
128claude-opus-4-8-fastchat$17.280051
129claude-opus-4.5chat$17.882135
130gpt-5.6-sol-prochat$19.230820
131gpt-5.4-prochat$31.500080
132claude-fable-5chat$35.444445
133claude-opus-4.6-fastchat$54.000031
134gpt-5.5-prochat$65.625061
135claude-opus-4-7-fastchat$71.280023
surp/best-coding-fastkimi-k2.7-code$0.4664/1M-95%6 models

cheapest coding model, biased to small/fast variants — compares 6 model(s), routes to the cheapest. click any row to see the full list.

#modelclassUSD / 1M toksellers
1kimi-k2.7-code ◀ routed nowcoding$0.466415
2qwen3-coder-turbocoding$0.911626
3qwen3-codercoding$1.010017
4kimi-k2.7-code:webcoding$1.34973
5gpt-5.2-codexcoding$2.362555
6gpt-5.3-codexcoding$3.937565
surp/pro-codingqwen3-coder-turbo$0.9116/1M-90%4 models

cheapest FRONTIER-tier coding model — compares 4 model(s), routes to the cheapest. click any row to see the full list.

#modelclassUSD / 1M toksellers
1qwen3-coder-turbo ◀ routed nowcoding$0.911626
2qwen3-codercoding$1.010017
3gpt-5.2-codexcoding$2.362555
4gpt-5.3-codexcoding$3.937565
surp/pro-reasoningdeepseek-r1$2.4934/1M-72%1 models

cheapest FRONTIER-tier reasoning model — compares 1 model(s), routes to the cheapest. click any row to see the full list.

#modelclassUSD / 1M toksellers
1deepseek-r1 ◀ routed nowreasoning$2.49347
surp/pro-visionqwen3-vl-235b-a22b-thinking$0.7150/1M-92%3 models

cheapest FRONTIER-tier vision model — compares 3 model(s), routes to the cheapest. click any row to see the full list.

#modelclassUSD / 1M toksellers
1qwen3-vl-235b-a22b-thinking ◀ routed nowvision$0.715028
2qwen3-vl-30b-a3b-thinkingvision$1.320010
3glm-5v-turbovision$1.625026
surp/pro-chatgpt-5.6-terra$0.1750/1M-98%34 models

cheapest FRONTIER-tier chat model — compares 34 model(s), routes to the cheapest. click any row to see the full list.

#modelclassUSD / 1M toksellers
1gpt-5.6-terra ◀ routed nowchat$0.175040
2gpt-5.5chat$0.350071
3grok-4.1-fastchat$0.43757
4gpt-5.6-lunachat$0.490041
5grok-4.20-multi-agent-betachat$0.937531
6gpt-5.6-solchat$1.220040
7gemini-2.5-profast$1.687512
8grok-4.20-betachat$1.875024
9grok-build-0-1chat$1.988047
10gpt-5.6-terra-prochat$2.100034
11gpt-5.6-luna-prochat$2.200031
12qwen3.5-397b-a17bchat$2.234336
13grok-4.3chat$2.762120
14claude-opus-4.6chat$3.000039
15grok-4.5chat$3.72038
16claude-sonnet-5chat$6.847028
17gemini-3.1-pro-previewfast$6.930029
18claude-opus-4.7chat$7.500043
19claude-sonnet-4.5chat$8.265041
20claude-sonnet-4.6chat$8.265025
21kimi-k3chat$10.489520
22gemini-3.1-profast$12.60003
23claude-opus-5-fastchat$15.000012
24grok-4chat$15.30003
25claude-opus-4.8chat$17.046958
26claude-opus-5chat$17.117517
27claude-opus-4-8-fastchat$17.280051
28claude-opus-4.5chat$17.882135
29gpt-5.6-sol-prochat$19.230820
30gpt-5.4-prochat$31.500080
31claude-fable-5chat$35.444445
32claude-opus-4.6-fastchat$54.000031
33gpt-5.5-prochat$65.625061
34claude-opus-4-7-fastchat$71.280023
surp/pro-fastgemini-2.5-pro$1.6875/1M-81%3 models

cheapest FRONTIER-tier fast model — compares 3 model(s), routes to the cheapest. click any row to see the full list.

#modelclassUSD / 1M toksellers
1gemini-2.5-pro ◀ routed nowfast$1.687512
2gemini-3.1-pro-previewfast$6.930029
3gemini-3.1-profast$12.60003
surp/codingkimi-k2.7-code$0.4664/1M-95%6 models

alias of best-coding — compares 6 model(s), routes to the cheapest. click any row to see the full list.

#modelclassUSD / 1M toksellers
1kimi-k2.7-code ◀ routed nowcoding$0.466415
2qwen3-coder-turbocoding$0.911626
3qwen3-codercoding$1.010017
4kimi-k2.7-code:webcoding$1.34973
5gpt-5.2-codexcoding$2.362555
6gpt-5.3-codexcoding$3.937565
surp/fastgemini-2.5-flash-image$0.0500/1M-99%34 models

alias of best-fast — compares 34 model(s), routes to the cheapest. click any row to see the full list.

#modelclassUSD / 1M toksellers
1gemini-2.5-flash-image ◀ routed nowfast$0.05009
2qwen3-5-9bfast$0.050034
3gpt-5.4-minifast$0.052562
4nvidia-nemotron-3-nano-30b-a3bfast$0.093841
5llama-3.2-3b-instructfast$0.096535
6gemma-4-26b-a4b-itfast$0.110058
7mistral-small-3.2-24b-instruct:webfast$0.11343
8glm-4.5-airfast$0.13008
9minimax-m2.5fast$0.150041
10mistral-small-3.2-24b-instructfast$0.178724
11llama-3.2-3b-instruct:webfast$0.18757
12mistral-small-4fast$0.234428
13gpt-4o-minifast$0.262360
14qwen3-next-80b-a3b-instructfast$0.297545
15magistral-small-2509fast$0.30001
16gpt-5-image-minifast$0.37398
17minimax-m2.5:webfast$0.38257
18gpt-5-nanofast$0.404810
19minimax-m2.7fast$0.427121
20gemma-4-26b-a4b-it:webfast$0.44392
21minimax-m2.7:webfast$0.46883
22qwen3-next-80b-a3b-instruct:webfast$0.56258
23minimax-m2.1fast$0.62009
24gpt-5.4-nanofast$1.177311
25gemini-2.5-flashfast$1.400010
26gemini-3.1-flash-litefast$1.49628
27minimax-m3fast$1.50008
28gemini-2.5-profast$1.687512
29gemini-3-flash-previewfast$1.750033
30gpt-5-minifast$2.024810
31minimax-m2.7-highspeedfast$2.70004
32gemini-3-5-flashfast$5.250030
33gemini-3.1-pro-previewfast$6.930029
34gemini-3.1-profast$12.60003
surp/chatglm-5.2$0.0300/1M-100%135 models

alias of best-chat — compares 135 model(s), routes to the cheapest. click any row to see the full list.

#modelclassUSD / 1M toksellers
1glm-5.2 ◀ routed nowchat$0.030039
2deepseek-v4-flashchat$0.040149
3llama-3.3-70b-instructchat$0.042064
4e2ee-qwen-2-5-7b-pchat$0.045026
5gemini-2.5-flash-imagefast$0.05009
6qwen3-5-9bfast$0.050034
7gpt-5.4-minifast$0.052562
8e2ee-gpt-oss-20b-pchat$0.060021
9e2ee-venice-uncensored-24b-pchat$0.070057
10nvidia-nemotron-3-nano-30b-a3bfast$0.093841
11qwen3-235b-a22b-2507chat$0.095035
12llama-3.2-3b-instructfast$0.096535
13venice-uncensoredchat$0.09906
14gemma-4-26b-a4b-itfast$0.110058
15mistral-small-3.2-24b-instruct:webfast$0.11343
16venice-uncensored-role-playchat$0.125059
17glm-4.5-airfast$0.13008
18openai-gpt-oss-120bchat$0.149036
19minimax-m2.5fast$0.150041
20gemma-3-27b-itchat$0.156027
21e2ee-gemma-3-27bchat$0.160027
22glm-4.7-flash:webchat$0.17002
23deepseek-v4-flash:webchat$0.17164
24gpt-5.4chat$0.175044
25gpt-5.6-terrachat$0.175040
26mistral-small-3.2-24b-instructfast$0.178724
27llama-3.2-3b-instruct:webfast$0.18757
28deepseek-v4-prochat$0.195845
29qwen3-30b-a3bchat$0.204653
30glm-4.7chat$0.215078
31glm-4.6chat$0.217086
32gemma-4-31b-it:webchat$0.22283
33qwen3-235b-a22b-2507:webchat$0.22507
34mistral-small-4fast$0.234428
35glm-4.7-flash-hereticchat$0.235038
36openai-gpt-oss-120b:webchat$0.24793
37e2ee-gemma-4-31bchat$0.250044
38glm-5chat$0.252087
39gemma-4-31b-itchat$0.261028
40gpt-4o-minifast$0.262360
41venice-uncensored-1.2chat$0.275020
42glm-4.5chat$0.28006
43qwen3-next-80b-a3b-instructfast$0.297545
44glm-4.7-flashchat$0.299019
45magistral-small-2509fast$0.30001
46deepseek-v3.2chat$0.326429
47gpt-5.5chat$0.350071
48e2ee-gemma-4-26b-a4b-uncensored-pchat$0.362527
49gpt-5-image-minifast$0.37398
50minimax-m2.5:webfast$0.38257
51qwen3-5-35b-a3bchat$0.390633
52gpt-5-nanofast$0.404810
53e2ee-qwen3-5-122b-a10bchat$0.40507
54minimax-m2.7fast$0.427121
55gemma-4-uncensoredchat$0.430622
56grok-4.1-fastchat$0.43757
57e2ee-glm-4.7-flashchat$0.44205
58gemma-4-26b-a4b-it:webfast$0.44392
59nvidia-nemotron-3-ultra-550b-a55bchat$0.450057
60qwen3.5-flashchat$0.45009
61minimax-m2.7:webfast$0.46883
62gpt-5.6-lunachat$0.490041
63e2ee-qwen3-6-35b-a3bchat$0.499842
64e2ee-gpt-oss-120b-pchat$0.506920
65nvidia-nemotron-cascade-2-30b-a3bchat$0.51704
66glm-5-turbochat$0.520092
67qwen3-next-80b-a3b-instruct:webfast$0.56258
68glm-5.1chat$0.580045
69qwen-3-7-pluschat$0.600046
70minimax-m2.1fast$0.62009
71mercury-2chat$0.812422
72llama-3.3-70b-instruct:webchat$0.87507
73grok-4.20-multi-agent-betachat$0.937531
74kimi-k2.5:webchat$1.01508
75kimi-k2.6chat$1.017537
76hermes-3-llama-3.1-405bchat$1.025045
77qwen3.6-plus-uncensoredchat$1.093830
78gpt-5.4-nanofast$1.177311
79gpt-5.4-image-2chat$1.21588
80gpt-5.6-solchat$1.220040
81glm-5:webchat$1.31258
82e2ee-glm-5.1chat$1.312530
83glm-4.7:webchat$1.31252
84gemini-2.5-flashfast$1.400010
85grok-code-fast-1chat$1.44503
86gemini-3.1-flash-litefast$1.49628
87gpt-5-imagechat$1.50003
88minimax-m3fast$1.50008
89qwen3.6-27bchat$1.521437
90kimi-k2.5chat$1.560026
91qwen3.5-pluschat$1.638011
92gemini-2.5-profast$1.687512
93gemini-3-flash-previewfast$1.750033
94aion-labs.aion-2-0chat$1.800017
95glm-5.1:webchat$1.81255
96deepseek-v4-pro:webchat$1.82363
97grok-4.20-betachat$1.875024
98grok-build-0-1chat$1.988047
99gpt-5-minifast$2.024810
100e2ee-qwen3-6-35b-a3b-uncensored-pchat$2.08008
101gpt-5.6-terra-prochat$2.100034
102mistral-largechat$2.125011
103gpt-5.6-luna-prochat$2.200031
104qwen3.5-397b-a17bchat$2.234336
105gpt-5.2chat$2.362566
106kimi-k2chat$2.58307
107minimax-m2.7-highspeedfast$2.70004
108grok-4.3chat$2.762120
109claude-opus-4.6chat$3.000039
110qwen-3-7-maxchat$3.404625
111e2ee-glm-4.7chat$3.41255
112kimi-k2.6:webchat$3.68835
113grok-4.5chat$3.72038
114gpt-4ochat$4.374862
115claude-haiku-4.5chat$4.636012
116gemini-3-5-flashfast$5.250030
117claude-sonnet-5chat$6.847028
118gemini-3.1-pro-previewfast$6.930029
119claude-opus-4.7chat$7.500043
120claude-sonnet-4.5chat$8.265041
121claude-sonnet-4.6chat$8.265025
122kimi-k3chat$10.489520
123gemini-3.1-profast$12.60003
124claude-opus-5-fastchat$15.000012
125grok-4chat$15.30003
126claude-opus-4.8chat$17.046958
127claude-opus-5chat$17.117517
128claude-opus-4-8-fastchat$17.280051
129claude-opus-4.5chat$17.882135
130gpt-5.6-sol-prochat$19.230820
131gpt-5.4-prochat$31.500080
132claude-fable-5chat$35.444445
133claude-opus-4.6-fastchat$54.000031
134gpt-5.5-prochat$65.625061
135claude-opus-4-7-fastchat$71.280023
surp/srup-freeglm-5.2$0.0300/1M-100%152 models

zero-price model if one exists, else the absolute cheapest text LLM — compares 152 model(s), routes to the cheapest. click any row to see the full list.

#modelclassUSD / 1M toksellers
1glm-5.2 ◀ routed nowchat$0.030039
2deepseek-v4-flashchat$0.040149
3llama-3.3-70b-instructchat$0.042064
4e2ee-qwen-2-5-7b-pchat$0.045026
5gemini-2.5-flash-imagefast$0.05009
6qwen3-5-9bfast$0.050034
7qwen3-235b-a22b-thinking-2507reasoning$0.050040
8gpt-5.4-minifast$0.052562
9e2ee-gpt-oss-20b-pchat$0.060021
10e2ee-venice-uncensored-24b-pchat$0.070057
11nvidia-nemotron-3-nano-30b-a3bfast$0.093841
12qwen3-235b-a22b-2507chat$0.095035
13llama-3.2-3b-instructfast$0.096535
14venice-uncensoredchat$0.09906
15gemma-4-26b-a4b-itfast$0.110058
16mistral-small-3.2-24b-instruct:webfast$0.11343
17venice-uncensored-role-playchat$0.125059
18glm-4.5-airfast$0.13008
19openai-gpt-oss-120bchat$0.149036
20minimax-m2.5fast$0.150041
21gemma-3-27b-itchat$0.156027
22e2ee-gemma-3-27bchat$0.160027
23glm-4.7-flash:webchat$0.17002
24deepseek-v4-flash:webchat$0.17164
25gpt-5.4chat$0.175044
26gpt-5.6-terrachat$0.175040
27mistral-small-3.2-24b-instructfast$0.178724
28llama-3.2-3b-instruct:webfast$0.18757
29deepseek-v4-prochat$0.195845
30qwen3-30b-a3bchat$0.204653
31glm-4.7chat$0.215078
32glm-4.6chat$0.217086
33gemma-4-31b-it:webchat$0.22283
34qwen3-235b-a22b-2507:webchat$0.22507
35mistral-small-4fast$0.234428
36glm-4.7-flash-hereticchat$0.235038
37openai-gpt-oss-120b:webchat$0.24793
38e2ee-gemma-4-31bchat$0.250044
39glm-5chat$0.252087
40gemma-4-31b-itchat$0.261028
41gpt-4o-minifast$0.262360
42venice-uncensored-1.2chat$0.275020
43glm-4.5chat$0.28006
44qwen3-next-80b-a3b-instructfast$0.297545
45glm-4.7-flashchat$0.299019
46magistral-small-2509fast$0.30001
47deepseek-v3.2chat$0.326429
48gpt-5.5chat$0.350071
49e2ee-gemma-4-26b-a4b-uncensored-pchat$0.362527
50gpt-5-image-minifast$0.37398
51minimax-m2.5:webfast$0.38257
52qwen3-5-35b-a3bchat$0.390633
53gpt-5-nanofast$0.404810
54e2ee-qwen3-5-122b-a10bchat$0.40507
55minimax-m2.7fast$0.427121
56gemma-4-uncensoredchat$0.430622
57grok-4.1-fastchat$0.43757
58e2ee-glm-4.7-flashchat$0.44205
59gemma-4-26b-a4b-it:webfast$0.44392
60nvidia-nemotron-3-ultra-550b-a55bchat$0.450057
61qwen3.5-flashchat$0.45009
62kimi-k2.7-codecoding$0.466415
63minimax-m2.7:webfast$0.46883
64gpt-5.6-lunachat$0.490041
65e2ee-qwen3-6-35b-a3bchat$0.499842
66e2ee-gpt-oss-120b-pchat$0.506920
67nvidia-nemotron-cascade-2-30b-a3bchat$0.51704
68glm-5-turbochat$0.520092
69qwen3-next-80b-a3b-instruct:webfast$0.56258
70glm-5.1chat$0.580045
71qwen-3-7-pluschat$0.600046
72glm-4.7-thinkingreasoning$0.61252
73glm-4.7-thinking:webreasoning$0.61252
74minimax-m2.1fast$0.62009
75qwen3-vl-235b-a22b-thinkingvision$0.715028
76mercury-2chat$0.812422
77llama-3.3-70b-instruct:webchat$0.87507
78qwen3-coder-turbocoding$0.911626
79trinity-large-thinkingreasoning$0.934418
80grok-4.20-multi-agent-betachat$0.937531
81qwen3-codercoding$1.010017
82kimi-k2.5:webchat$1.01508
83kimi-k2.6chat$1.017537
84hermes-3-llama-3.1-405bchat$1.025045
85trinity-large-thinking:webreasoning$1.07811
86qwen3.6-plus-uncensoredchat$1.093830
87gpt-5.4-nanofast$1.177311
88gpt-5.4-image-2chat$1.21588
89gpt-5.6-solchat$1.220040
90glm-5:webchat$1.31258
91e2ee-glm-5.1chat$1.312530
92glm-4.7:webchat$1.31252
93qwen3-vl-30b-a3b-thinkingvision$1.320010
94kimi-k2.7-code:webcoding$1.34973
95gemini-2.5-flashfast$1.400010
96grok-code-fast-1chat$1.44503
97gemini-3.1-flash-litefast$1.49628
98gpt-5-imagechat$1.50003
99minimax-m3fast$1.50008
100qwen3.6-27bchat$1.521437
101kimi-k2-thinkingreasoning$1.55007
102kimi-k2.5chat$1.560026
103glm-5.1-non-thinking:webreasoning$1.62505
104glm-5v-turbovision$1.625026
105qwen3.5-pluschat$1.638011
106gemini-2.5-profast$1.687512
107gemini-3-flash-previewfast$1.750033
108aion-labs.aion-2-0chat$1.800017
109glm-5.1:webchat$1.81255
110deepseek-v4-pro:webchat$1.82363
111grok-4.20-betachat$1.875024
112grok-build-0-1chat$1.988047
113gpt-5-minifast$2.024810
114e2ee-qwen3-6-35b-a3b-uncensored-pchat$2.08008
115gpt-5.6-terra-prochat$2.100034
116mistral-largechat$2.125011
117gpt-5.6-luna-prochat$2.200031
118qwen3.5-397b-a17bchat$2.234336
119gpt-5.2-codexcoding$2.362555
120gpt-5.2chat$2.362566
121deepseek-r1reasoning$2.49347
122kimi-k2chat$2.58307
123minimax-m2.7-highspeedfast$2.70004
124grok-4.3chat$2.762120
125claude-opus-4.6chat$3.000039
126qwen-3-7-maxchat$3.404625
127e2ee-glm-4.7chat$3.41255
128kimi-k2.6:webchat$3.68835
129grok-4.5chat$3.72038
130gpt-5.3-codexcoding$3.937565
131gpt-4ochat$4.374862
132claude-haiku-4.5chat$4.636012
133gemini-3-5-flashfast$5.250030
134claude-sonnet-5chat$6.847028
135gemini-3.1-pro-previewfast$6.930029
136claude-opus-4.7chat$7.500043
137claude-sonnet-4.5chat$8.265041
138claude-sonnet-4.6chat$8.265025
139kimi-k3chat$10.489520
140gemini-3.1-profast$12.60003
141claude-opus-5-fastchat$15.000012
142grok-4chat$15.30003
143claude-opus-4.8chat$17.046958
144claude-opus-5chat$17.117517
145claude-opus-4-8-fastchat$17.280051
146claude-opus-4.5chat$17.882135
147gpt-5.6-sol-prochat$19.230820
148gpt-5.4-prochat$31.500080
149claude-fable-5chat$35.444445
150claude-opus-4.6-fastchat$54.000031
151gpt-5.5-prochat$65.625061
152claude-opus-4-7-fastchat$71.280023

Savings are vs a blended claude-sonnet-4.6 list price (~$9/1M). Click any combo to expand it and see every model in its pool, cheapest first — the top row is what your request routes to right now.

how it works — 3 steps

1
you request

POST /v1/chat/completions with model: "surp/best-chat". no auth header.

2
we 402

we resolve the combo to the live cheapest model, compute the exact micropayment, and return HTTP 402 with payment requirements.

3
you pay & go

your wallet signs an EIP-3009 USDC authorization. we verify + settle on Base. the completion streams back.

why this is different

Most "cheap LLM" services are static: they pick one model, hide it behind a flat price, and pocket the spread. surp.ivc.lol is dynamic: the model you get is whatever is cheapest on Surplus at the second you asked, and the price you pay is that spot price + a transparent markup. We're not a provider — we're a price arbitrageur sitting on top of a marketplace that itself sits on top of every other provider. Aggregating the aggregator.

live: payments settle in real USDC on Base mainnet (eip155:8453) via the x402 protocol. each request is a single on-chain EIP-3009 transfer — no account, no API key, no subscription. your wallet pays, your model answers.
connect your hermes build your own combo try the playground read the docs