We buy unused AI compute commitments at a discount - before they go to waste - and pass the savings on to you. Made by @keak_ai

Cheaper Inference retweeted
From prompt optimization using genetic pareto to benchmarking for field performance intel, #CheaperInference @CheaperInfer has it all if not more ! See for yourself at cheaperinference.com/signup?…
4
4
743
Live and discounted on cheaper inference
Introducing Claude Opus 5.5, the first model in our new Claude 5.5 family. It performs at the level of Claude Fable 5.1 for most tasks, and costs 40% less to run than Opus 5.
3
1
8
6,232
Replying to @fabrice_mayrand
At least change the privacy policy @LeoYe_AI 😂
2
177
Astra is live and discounted!
19
4
103
546,627
Cut the AI bill. Keep the models. Keep your messages, tools, streaming, and response handling. Change two configuration values and access discounted model capacity across leading providers.
102
117
1,183
14,402,999
You folks, are straight to the point, I subscribed should give it a try sometime tomorrow :)
2
9
6,069
Let us know if you need help with anything!
1
5
4,203
Cheaper Inference retweeted
Cheaper Inference is up!
All the AI is down
15
7
67
17,628
We've acquired omniroute.online ! Try the #1 open source router and stop paying routing fees!
14
6
63
40,907
Omniroute has joined cheaper inference!
We've acquired omniroute.online ! Try the #1 open source router and stop paying routing fees!
10
3
26
25,114
GPT 5.6 Luna is 60% off!
3
1
16
4,460
Time to use cheaper inference in Cursor
We’re ending our partnership with Cursor following its acquisition by SpaceX. Under our proposal, Cursor’s direct access to our models would end on November 12. We know that the people most affected by this decision are the developers who rely on OpenAI models in Cursor. We care about their experience in this transition and we’re ready to go above and beyond to support them. openai.com/index/our-decisio…
3
2
15
6,603
Hey @CheaperInfer - service down for much of today... Any updates?
2
2
126
Hi Andy, we're dealing with a surge in demand that is pushing our infrastructure to its limits. We're working on adding capacity and should be back to normal in ~2h
1
2
102
Cheaper Inference retweeted
We’re experiencing a surge in demand that’s pushing @CheaperInfer’s infrastructure to its limits. You may encounter some issues over the next few hours while we scale up capacity. We apologize for the disruption. We expect everything to be back to normal within approximately two hours.
7
4
29
9,138
GLM 5.3 is live and discounted 🫡
3
3
13
5,241
service down. rug pull?
2
3
174
Huge spike in demand caused us to be temporarily down today, we're sorry about that!
2
169
3.7 Flash is live and discounted by 30% 🫡
1
3
3,694
deepseek v4 pro 0813 is now live and discounted 🫡
1
1
7
22,461
Prices will come down even more over the next couple days as we get more capacity
2
2,720
This is true, DeepSeek is completely dominating our rankings
Prediction: Deepseek v4 flash will take over Claude as the no.1 winner in market share. It's the biggest story of 2026! Three reasons: 1. It's so cheap. 100x cheaper in unit token pricing. 2. It's so fast. Much much faster than Claude. My vibe check is around 2-3x faster. 3. It's so cache-efficient: most of my tokens spent are in cache read, and cache input pricing are $0.5/Mt Opus 5 v.s. $0.0028/Mt Deepseek. Deepseek has a much higher cache hit rate, which means the overall effective cost per task could be upwards of 500x cheaper. For something that's 500x cheaper AND 2-3x faster per task, why wouldn't Deepseek V4 Flash win.
1
1
10
7,283
I’d pay $1,000 a month for any service that gives me unlimited Kimi K3.
100
6
660
82,710
Happened later than we thought but Opus 5 is now the top model by usage 👑
6
5,054