Hacker Newsnew | past | comments | ask | show | jobs | submit | aliljet's commentslogin

Honestly, I have a 2080ti that I use to play and I can tell you the math isn't there to upgrade it. It's much easier to just find a 3090/4090/5090 and keep pace with the software and hardware simultaneously.

I'd be curious about an.option that would allow glm use with a low end GPU like a 2080 ti...


The benchmarks here are confusing at best. Am I reading correctly that this model is essentially as good or better than all frontier models right now?


I believe the benchmark listed is about simulating the environment for the various tasks, rather than doing them. It seems that the point of this model is to generate sim data to improve other models with


Benchmarks in general are a little iffy, the whole industry is going off of vibes anyways. Can't decide before trying it out


I was just using infinity parser 2 (flash, to be fair) for pennies self-hosted to run through thousands of pages of documents with remarkable confidence. I decided to use https://huggingface.co/datasets/allenai/olmOCR-bench to determine what was the best OCR tool, yesterday, but I've got no idea what the best is now. What is the dominant OCR eval right now? Between Baidu and Mistral this morning, I wonder if there's a new tool to switch to..


I'm curious about this. What models/tools have you been using?


How does this compare with infinty parser 2 which seemed to be running the table on every other OCR tool (https://huggingface.co/datasets/allenai/olmOCR-bench). To be fair, there's no single winning OCR benchmark and this isn't showing up anywhere yet..


This sounds incredible. Have these models effectively solved the problem of trying to use a fast-processing network to predict the world's state ahead? For example, to catch a ball?


The problem here is always the cost-benefit. For $200/mo, you're receiving subsidized best of breed access. There's no model competing for that price anywhere. If a 27B param model is what you choose, show me your hardware! I would love to be wrong...


But for how long? The subsidized phase is probably short, and then what? I run Qwen 3.5 27 Dense om my old AMD RX7900XTX at about 45 t/s and barely use my Claude Code subscription anymore.


Is this just one giant marketing plot?


There's a lot of speculation that it is indeed a marketing plot and the model is just a step improvement over current capabilities... and the real reason they aren't releasing the model is they are compute constrained and cannot serve the model. To my knowledge there's no proof of this however, but given the fact that literally 60 days ago they made Mythos out to be the end of the world and last Friday they announced that they will release the model in a few weeks, I feel like it was indeed something along those lines (marketing ploy).


Their IPO is coming up soon. It would be interesting if Mythos remained mythical right up until then, wouldn't it?


Or just control of supply and demand. If they can charge twice as much serving half as many customers, that leaves a lot of potential future customers leftover.


yes


The week before they released Mythos to governments they had all their source code stolen. It's all about improving their image and creating propoganda.


It wasn't "all their source code", it was the source code to Claude Code: not really any of their internal secret sauce, at least directly.


it wasn't stolen either. an employee accidentally included a source map file with the release.


Where can a user reasonably host this in an affordable way to access the local LLM revolution?


I think their Max models are far bigger than fits on consumer hardware. People are typically using Apple, AMD Halo, or dGPUs if/when they do smaller versions. Those are all varying degrees of "affordable."


Unsloth Studio with its MTP support: https://unsloth.ai/docs/models/qwen3.6#mtp-guide


Try llama.cpp and Qwen3.6-35B-A3B

Good balance of intelligence and speed.


This one is not local


Consider applying for YC's Fall 2026 batch! Applications are open till July 27.

Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: