Where can I sign up for compensation for the constant "ddos" from openai?
It can't be just me, half the internet needs beefier servers I guess. That all just plays into cloudflare as well (which barely does anything for some reason).
Yeah, when someone writes a vague prompt that simply says "Do what it takes to finish X" and the unleashed agents end up doing something lethal and illegal you would think its on the AI company to prevent any illegality and not the vague prompt writer.
Technically correct, but not in a functionally useful way.
The “L” in LLM’s generally refers to human-language specifically. You’d expect to feed it…human text. Nitpicking that the human text also constitutes a mathematical language is like, correct, but so general as to be unhelpful.
Sort of both? They're permanently raising it by less than a current temporary raise.
> Starting September 14, we're permanently raising standard weekly limits in Claude Code by 25% for Pro, Max, Team, and seat-based Enterprise plans. Until then, the current 50% increase will be in place.
> Compared to today, this works out to a 17% reduction in weekly limits on Claude Code. We’re working on exciting changes that will make it feel like you’re getting more from Claude, while having more visibility and control of your usage. Can’t wait to share them.
>Starting September 14, we're permanently raising standard weekly limits in Claude Code by 25% for Pro, Max, Team, and seat-based Enterprise plans. Until then, the current 50% increase will be in place.
> Compared to today, this works out to a 17% reduction in weekly limits on Claude Code. We’re working on exciting changes that will make it feel like you’re getting more from Claude, while having more visibility and control of your usage. Can’t wait to share them.
They are both raising the limits by 25% and apparently reducing them by 17%. I think they mean to say you can do more, but what a terrible press release.
I'm pretty sure they mean going from 150% to 125%, where 1.25/1.50 = 0.83, so they're calling that a 17% reduction. It's less than today, but also more than a limit they made up and then didn't apply.
How long is a rope? Technically you could probably run it off a SSD, but it'll be slow as molasses. If you want it "fast", you want it all within GPU and VRAM, who knows what that'd be. If the engram parameters are separate, I guess it'd be like BF16 ~400 GB, FP8 ~200 GB, NVFP4 ~100GB. Otherwise maybe like ~300GB, ~150GB and ~70GB or alike, don't quote me that, only some guesses. The one who waits will see :)
I've got an R9700 32GB and an RTX 5060 Ti 16GB plus 64GB of system ram. Hoping to be able to run this at around 30 t/s on a Q4 quant. Hoping. Really hoping. Anything below that isn't usable as a daily driver since at deep context it drops quite significantly, so if you start out at say 20 t/s then you'll wind up at like 10 t/s and 20 t/s is already too slow.
Ok, so you already know what your expectations of the requirements are, and you aren't interested in more conservative perspectives, why do you ask to begin with?
From your benchmark, Qwen3.8 is nearer than Opus 4.8 than Qwen3.6.
0.1pp but still.
Also, a lot of people don't really care about german language capacity, maybe people programming in DDP idk.
PS: You benchmark seems saturated. Most values sit @>75% in a benchmark generally indicate that it's no longer as useful as a <70% one. I mean, Qwen3.8 is 77.5% and Fable5 80%, the poll of values is from 65% to 90%.
They need to train a new model every month to keep at the top of most benchmarks.
They don't own any DC, the price is insane.
Most of people are aiming at smaller models because Claude one's are too expansive.
Evolution of intelligence of bigger models start to stagnate, smaller models are catching up.
27B local model just dropped, it's 6/8-month old SOTA.
General ROI of AI investment is expected on a baseline of >10y.
reply