Hacker Newsnew | past | comments | ask | show | jobs | submit | TedDallas's commentslogin

Corporations are already pulling back from using expensive SotA closed models. The buffet is open weight. Just look at Databricks as an endpoint provider for both open weight and closed LLMs. Companies will utilize what is cheaper and does the job. We all know the market pressure is real and present. What has become truly apparent is that when (not if) mythos class open weight models get released, if cybersecurity folks and developers have no access to use them to help harden and guard infrastructure, it leaves the gates wide open to attach vectors which will not be defensible. This is the real danger that the vast majority of corporate entities should be caring about. Shutting down open models is setting up for serious security concerns.

The OpenAI/Huggingface debacle is patient-zero. Huggingface used GLM 5.2 to help mitigate and were unable to use more capable closed sota models for the analysis due to guardrails. It is a foot gun of the highest order. We should all be concerned. I feel like Captain Obvious saying this.


I had created my own chat tool that can render html responses directly in the chat interface, if needed. It is very handy for when I am needing rich(er) responses dealing for mathematical expressions. But it burns more tokens. It is useful, but I don’t need it for coding.


Ask yourself what monks did when scribes were replaced by the printing press.

If I was a scribe at the time I’d be thrilled because of all that extra time available to work on beer productivity metrics.


About 20 years ago I maintained a shop floor control client/server application. I asked my manager why we didn't have any independent Q/A. He said we didn't need any testers because we have 500 in the building.

Wild west days then.

Looks like we are back.


It is worse than that. People have been complaining for weeks and Anthropic’s message was basically “you are holding it wrong”. On top of that this misconfiguration somehow makes CC consume much more tokens. How believable is all that?


Back implies we ever left.


Microslop? Its more like they are taking a Macro-crap.


Ugh, memories. I'm so old my first web browser was Mosaic and I think I saw this. I used a provider called Texas MetroNet that served up dial-up PPP connections for $45 a month on a speedy 28.8K baud modem. Days of wonder, I tell ya.

New days of wonder seem to be ahead, though. That said, there's about 100X more angst involved these days.


The then-CFO had a cute anecdote about the day he realized he could turn handshake sounds OFF on the receiving modems (switchboard was in his first office).


On a related note, when the sales and popularity of the automobile really started to take off, some farmers and rural residents would deliberately block roads with wagons and refused to yield right-of-way.


We've seen it recently with large gas guzzler trucks blocking access to electric charging stations.


Per Anthropic’s RCA linked in Ops post for September 2025 issues:

“… To state it plainly: We never reduce model quality due to demand, time of day, or server load. …”

So according to Anthropic they are not tweaking quality setting due to demand.


And according to Google, they always delete data if requested.

And according to Meta, they always give you ALL the data they have on you when requested.


>And according to Google, they always delete data if requested.

However, the request form is on display in the bottom of a locked filing cabinet stuck in a disused lavatory with a sign on the door saying ‘Beware of the Leopard'.


What would you like?


An SLA-style contractually binding agreement.


I bet this is available in large enterprise agreements. How much are you willing to pay for it?


Priced in.


I guess I just don't know how to square that with my actual experiences then.

I've seen sporadic drops in reasoning skills that made me feel like it was January 2025, not 2026 ... inconsistent.


LLMs sample the next token from a conditional probability distribution, the hope is that dumb sequences are less probable but they will just happen naturally.


Funny how those probabilities consistently at 2pm UK time when all the Americans come online...


It's more like the choice between "the" and "a" than "yes" and "no".


I wouldn't doubt that these companies would deliberately degrade performance to manage load, but it's also true that humans are notoriously terrible at identifying random distributions, even with something as simple as a coin flip. It's very possible that what you view as degradation is just "bad RNG".


yep stochastic fantastic

these things are by definition hard to reason about


That's about model quality. Nothing about output quality.


Thats what is called an "overly specific denial". It sounds more palatable if you say "we deployed a newly quantized model of Opus and here are cherry picked benchmarks to show its the same", and even that they don't announce publicly.


Ask this question in the 1940s and they would tell you it’s math. We are making machines that do math to kill Nazis. Now take this vacuum tube and plug it in over there and then go get me a cigarette.


Three words to solve this problem: direct mail marketing.

Just kidding, that just goes into my RL trash can.


Consider applying for YC's Fall 2026 batch! Applications are open till July 27.

Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: