80 points joebuckwilliams 4 hours ago 41 comments
bix6 1 hour ago | parent
Edit: miss me with the downvotes. These guys are clowns. I want alternatives. Thank you to everyone offering help!
marcuskaz 1 hour ago | parent
bix6 1 hour ago | parent
tolugenius 1 hour ago | parent
bix6 1 hour ago | parent
tolugenius 50 minutes ago | parent
jmtulloss 1 hour ago | parent
OpenRouter isn’t a provider, they route to other providers, so you would need to specify which ones you’re comfortable with anyway.
[1]:https://docs.fireworks.ai/guides/security_compliance/data_ha...
bix6 1 hour ago | parent
SkyBelow 51 minutes ago | parent
For OpenRouter, you can setup an API key and limit it to only models that claim to not train on data, but that is just a claim. You can then use trust to judge which providers will honor that claim.
But if the data is really sensitive, you might want either a local model or a business subscription with some big name in the US that legally promises no data training.
So are we talking some app idea you are playing around with, or files filled with PHI/PII that you have legal mandates to safeguard? If the latter, I would stick to only provider with enterprise agreements to not store/train on the data. Even the ones who promise no training are likely storing the data for monitoring for abuse or such short term.
bix6 44 minutes ago | parent
cmiles8 1 hour ago | parent
myaccountonhn 1 hour ago | parent
Roark66 1 hour ago | parent
I'm very happy with Deepinfra. Less model coverage, but good prices and quite fast.
However I have to caution you about one thing.
No one will give you as many input tokens for so little money as Claude Max x5 (maybe x20 too, I use x5).
I tend to use 1.1B to 1.4B a week about 0.8-1B cached. Even with cache were talking thousands of $ in API prices a week. Hundreds if we're talking cheap cloud like Deepinfra.
However, local AI well setup is actually a good alternative for this if Claude Max was unavailable.
For example my system a ryzen 7950x 192GB ram, 5x rtx3090 plus an rtx5060 ti 16gb. (3 rtx3090 cards via usb4 egpu dock). Let's me run Qwen3.8-Flash-Next with 3slots (no rtx5060 used) at 55tok/s decode dropping to 50 at the end of a 260k context, 1200tok/s refill dropping to 950 at the end of context.
With RAM and ssd cashing and 80% cache were talking on the order of 4B a week could be ingested by this setup (roughly) if it was running 24/7. I found 6 interactive cloud code sessions are fairly pleasant with this 3 user setup.
If I include the rtx5060 in the mix I can bump to 5 users, but it slows down by about 15% (note the speeds are give are for one active user, multiple users at once see maybe 70% of tgat per user so aggregate is much higher in multi user setup).
So in theory I should be able to run 10 cloud code sessions. Although I'm testing CC alternative now (pi with own plugins) because this model, while multimodal has only 260k context 30k of which CC eats on the getgo.
Many people say local AI makes no sense financially. But in the event you process huge inputs that are often cached it does make sense.
bix6 55 minutes ago | parent
I think your local setup would be a bit much for me capability / price wise. But maybe I can find a scaled down version. I don’t need insane tok/s. Oftentimes I just let things run and come back later.
Ciantic 27 minutes ago | parent
OpenRouter allows to make an API key locked to certain provider, I do that myself.
badsectoracula 1 hour ago | parent
You have to be the embodiment of hubris to believe the only reason a technologically advanced country of 1.4 billion people with several AI labs and government support can make competitive models is because they all distill yours.
bpodgursky 1 hour ago | parent
benxh 53 minutes ago | parent
toasty228 53 minutes ago | parent
https://www.heise.de/en/news/DeepSeek-orders-160-000-Huawei-...
bpodgursky 47 minutes ago | parent
Like I said, it will delay the Chinese labs for a couple years. These are not even top-line chips.
Frankly, China was not going to allow them to depend on NVIDIA forever, I don't think this motivates domestic production that much over the counterfactual. If NVIDIA didn't have export controls, China was going to set up formal import controls. They want to own their entire supply chain.
lightedman 27 minutes ago | parent
Not even that long. China can simply order its industries to stop supplying externally and go full-domestic. It has happened before in others parts of Chinese industry; it will happen again. As it stands China can easily build computational nodes at scale, and their LineShine supercomputer holds top place in the Supercomputer Top 500 at almost 2.2 exaflops. They have zero issues building performant hardware.
Couple of years? Their current technology level can make that less than a month.
giantrobot 20 minutes ago | parent
Also as we have seen, the Chinese labs will just buy their kit through cut outs or do training in locations where there's no embargo.
toasty228 54 minutes ago | parent
scottLobster 44 minutes ago | parent
They aren't patriotic and they don't really care about competing with China, they fancy themselves princes and are playing on the silents/boomers in charge being irreparably stuck in the Cold War.
rglover 39 minutes ago | parent
I don't agree with all of their means for getting here, but to accomplish what they have in ~30 or so years [1] should blow minds far more than it does in the West. And they did it by convincing us to give them control of our manufacturing and selling us (literal) boatloads of cheap junk. Historians will look back on this era as one of the greatest demonstrations of "winning without firing a single shot."
This whole insular POV is tired, ignorant, and frankly just another indicator that America has lost the plot, shit-faced on its own arrogance.
altcognito 25 minutes ago | parent
I'm confused, what has America and the west lost the plot on? You would have preferred the west to exploit China more?
This framing of a zero sum game is counterproductive.
black6 21 minutes ago | parent
rayiner 22 minutes ago | parent
oceanplexian 8 minutes ago | parent
There’s a reason China has a GDP per capita somewhere around that of Argentina.
Of course, they also have futuristic cities, maglev, and plenty of brilliant engineers and scientists, therefore it makes sense to be pragmatic. But there are a lot people talking up China because they hate the US and are bought into a massive propaganda / influencer campaign.
amluto 33 minutes ago | parent
(a) Sell many billions of dollars of fancy chips to China, thus bringing in billions of dollars of money and billions of dollars of trade balance improvement. And let China continue to build models quickly when they seem oddly happy to export the trained weights for free.
Or
(b) Decline to do so, thus reducing exports and very very strongly encouraging China to do everything in their power to avoid needing to depend on our chips.
In exchange for (b), what do we gain? A temporary advantage in model training and availability?
sschueller 27 minutes ago | parent
rayiner 25 minutes ago | parent
dfxm12 11 minutes ago | parent
mcshicks 15 minutes ago | parent
ultrarunner 7 minutes ago | parent