@petergyang@nateherk I think that they dropped sol, are serving what would have been a 5.6 terra model, and when they drop the "MORE EFFICIENT ASTRA!!!!!"🙄 it's going to be a sol class model.
If you route through OpenRouter:
• Check the cache-read column, not just input/output
• Use ignore or an only allowlist
• Don't trust "cheapest" on cache-heavy workloads
@OpenRouter: does price routing account for cache read? It should.
What we did @vxbeai : @wafer_ai is banned from every model and every lane in our routing.
We're also adding a guard: a provider can't win on one meter while charging far more on another. Every meter gets priced before a provider is admitted.
To be fair, nobody's hiding anything. It's all on the price page.
But price routing optimizes the number everyone looks at, and cache read is the number nobody looks at.
That gap is the whole business model here.
Why this matters: agents resend their whole context every turn, and most of it is cache hits.
Say 100k context, 90% cached:
Wafer saves you $0.00001 on input
Wafer charges you $0.0045 more on cache
Per million turns: $10 saved, $4,500 lost.
Straight from the provider table:
Input: $0.099/M (Wafer) vs $0.100/M (everyone else)
Cache read: $0.06/M (@wafer_ai ) vs ~$0.01/M (everyone else @OpenRouter, @stripe)
One of those numbers gets you routed there.
The other one is your bill.
And it didn't start there. Highest cache-read price on this route:
9/17: $0.03/M
9/22: $0.05/M, set by Wafer
Today: $0.06/M, still Wafer
A 2x climb in a week, while holding the input undercut that keeps traffic flowing in.
The cheapest provider on @OpenRouter isn't the cheapest.
@wafer_ai wins price routing on DeepSeek V4.1 Flash by pricing input $0.001/M under everyone else.
Their cache-read price is 6x the field. And it keeps going up. 🧵
GUYS. IM GOING TO OFFICIALLY BE A CONTRIBUTOR TO @Cloudflare Think on GitHub. It's the first time any of my work has been accepted anywhere and I'm so excited! Huge thanks to @threepointone for keeping my commits in.
@IndependentEco I've gotten SO much work done since the release. I used my reset already and now I've used 86% of my weekly limit, but the amount of work I got done for that is INSANE.
I used my entire weekly usage limits on Opus 5.5 in a day, but here's the thing, I feel like the amount of work that got accomplished in that one day, including verification and integration tests, was total what I had been getting done in a week before. This model is GOOD.
I get where you are coming from, and I appreciate you going out of your way so that users can address issues. It's something that I wish I saw more of, and a solution to a problem now is always better than nothing. But I just wanted to voice my opinion so that ideally these types of things can be corrected in the next model release.
You guys are cooking hard and I know that what's coming next will be great. I just get tired of all of the marketing games. Like, Sol, Terra and Luna were supposed to be permanent representations of capability, but I don't feel like they stayed that "permanent." Anywho, thanks for your comment back, I appreciate being heard.
732 Followers 6K FollowingI help ecommerce brands increase sales with high converting Shopify product pages.
Designed in Figma. Built in Replo & Instant.
DM for a free 1-minute page idea
593 Followers 668 Following10+ Years Android Dev | Kotlin Multiplatform
Building mobile SaaS and sharing the real progress, data and lessons.
DM for collabs AND projects
766 Followers 2K FollowingAhead AI ships the future. The apology is already on the calendar. Web comic.
A workplace satire about artificial confidence
Author: @realdavidstorm
21K Followers 752 FollowingHelping you build your remote empire | Follow to grow your online store + achieve your goals | CEO at Pure Private Label | Co-Founder RevMulti Marketing
1K Followers 403 FollowingFounder of @fewerapp & https://t.co/yUtnAPjt5I. Helping companies safely remove organizational bloat. AI should help decide what not to do.
2K Followers 362 Followinggrowth at @jitsucom (yc s20) · prev @GooseworksAI (yc w23), @lyzr__ai · bootstrapped my own saas to $120k arr · book a call 👇
527 Followers 326 FollowingNeed an automation, a web or mobile app built?
6+ yrs • 40+ apps • 50+ systems
Fast, safe, and reliable!
https://t.co/lP7np3G0JU
48K Followers 2K Followinghe/him • 🏳️🌈 • software engineer - hci/ux • texas ex • humans were not meant to be in constant communication • https://t.co/BocaEXmCBk
144K Followers 94 FollowingSane + 🌶️ takes in an insane AI world... AI capabilities researcher: co-created RLHF/ChatGPT @ @openai now trying to right the wrong 🤭 (ceo @typesafeai)
606K Followers 59K FollowingThe automated life | San Francisco/Silicon Valley AI, robotics, BCIs | Ex-Microsoft, Rackspace, Fast Company | Wrote eight books about the future.
11K Followers 987 FollowingChasing the horizon of tomorrow's ✨ | Business, AI, Apps & Open Source | Tech Expert | Ambassador @cognition | DMs open for collabs and opportunities.
593 Followers 668 Following10+ Years Android Dev | Kotlin Multiplatform
Building mobile SaaS and sharing the real progress, data and lessons.
DM for collabs AND projects
84K Followers 1 FollowingWorkflow automation for technical teams to build AI solutions that integrate with any app or API at no-code speed and code flexibility. Open and self-hostable
1K Followers 766 FollowingBuilding @researchpodapp with the goal of making science more accessible | Previously at CODE & @tesla | Author of Musk’s Memos
338K Followers 299 FollowingA little bit geek, wonk, and nerd. Repeat entrepreneur, recovering lawyer, and former ski instructor. Co-founder & CEO of Cloudflare (NYSE: NET).
21K Followers 752 FollowingHelping you build your remote empire | Follow to grow your online store + achieve your goals | CEO at Pure Private Label | Co-Founder RevMulti Marketing