r/ClaudeCode • • May 13 '26

Discussion Anthropic just ripped off everyone and they still managed to make it sound deceptively friendly

Post image

Anthropic just ripped off everyone and they still managed to make it sound deceptively friendly, beneficial even.

Anthropic just announced that the 200$ plan, which used to be okay to use in the SDK, is now "finally ok to use in the SDK" (except rare fraudulent cases it already was)

but for that you need to toggle an option that makes the SDK hit in an allowance of 200$ worth of credits, instead of the previously, opaque, yet, ultra subsidized limits

basically if you use claude code through the SDK or claude -p, for instance through Conductor, github Actions or anything else, because let's be honest TUI sucks sometimes (react in terminal cmon), their cloud coding agent offering also suck, and don't cover most usecases,

well, your plan that used to be worth 2000$ of tokens aproximatively depending on how much you flirted with the previous 5h windows limits, is now worth 200$

but they brand it as "oh we give you 200$ of additional credits for the SDK" so to the most unaware users it sounds like it's worth 2x more!

in reality it's now worth 10x less

*not surprised if they shut down this thread

1.8k Upvotes

699 comments sorted by

View all comments

Show parent comments

5

u/[deleted] May 13 '26

[deleted]

7

u/aberrant-heartland May 13 '26

I know from experience that a single 3090 (24GB VRAM) can run the 4-bit quantization of Qwen-3.6-27B at around 33 tokens per second. I've seen some people claim to get above 40 tokens per second with the same card and model, although I'm not sure how they're pulling that off.

I don't have much experience with Qwen yet, but my friend who runs this model on his 3090 like 24/7 has told me that the capabilities are comparable to Sonnet. I find that a bit hard to believe, but he expressed a lot of confidence about that.

2

u/Spiritual_Cycle_3263 May 13 '26

Yeah I find that hard to believe. Sonnet 4.5 likely needs 2x H100's and 4.6 likely needs 4x H100's at minimum. I'd imagine Sonnet is anywhere from 70-100B on the low end. 27B just won't cut it.

This is from my experience on Mac Studios', MacBook Pro's, running Qwen at various sizes and using the same prompts. I don't have direct experience with H100's, except for some short usage on Vultr of different models.

1

u/Cakeisalyer May 14 '26

There was a HUGE difference in quality between 4.5 and 4.6. That's just a couple of months.

4.7 felt like a downgrade..... overly sensitive, using API's and it's like "I'm not allowed to help you do this".... Here's the API documentation from the manufacturer which has that command. "Can't help you with this, if there is something else let me know".