Hacker Newsnew | past | comments | ask | show | jobs | submit | nnx's commentslogin

> Some days, I would seriously rather work at Wendy's.

Narrator: He would not.


If such a job paid even half as much, I would quit today. I can continue to program computers as a hobby. Occupational programming has essentially eroded my passion over the years.

Money is not important to me insofar as I have enough to live an average life. I don't need anymore than that.


is flickering fully fixed yet?


Not saying there's no negative impacts, but what are the _huge_ negative impacts that have materialized so far?


Coincidentally this is how pretraining works :)


Looks really interesting but the Go binding sadly uses cgo. Could the binding be done in pure Go? Or at least purego (the cgo alternative using Go assembly for FFI) ?


I ended up just using the web version, which is actually better than the native app (multiple tabs works!).


> My M5 Pro can generate 130 tok/s (4 streams) on Gemma 4 26B.

This seems high. At which quantization? Using LM Studio or something else?

Note: Darkbloom seems to run everything on Q8 MLX.


Ah good point, this is using Q4, benchmarked total throughout serving with Llama.cpp.


The only limit is yourself!


*was


Are you `Ionstream` on OpenRouter?

If so, it would be great to provide more models through OpenRouter. This looks interesting but not enough to make me go through the trouble of setting up a separate account, funding it, etc.


second that.

for smaller start ups, it's easier to go through one provider (OpenRouter) instead of having the hassle of managing different endpoints and accounts. you might get access to many more users that way.

mid to large companies might want to go directly to the source (you) if they want to really optimize the last mile but even that is debatable for many.


Hey @nnx & @hazelnut, good question, but no, we're not IonStream on OpenRouter.

The purpose of IonRouter is to let people publicly see the speed of our engine firsthand. It makes the sales pipeline a lot easier when a prospect can just go try it themselves before committing. Signup is low friction ($10 minimum to load, and we preload $0.10) so you can test right away.

That said, we do plan to offer this as a usage-based service within our own cloud. We own every layer of the stack— inference engine, GPU orchestration, scheduling, routing, billing, all of it. No third-party inference runtime, no off-the-shelf serving framework. So there's no reason for us to go through a middleman.

No plans to be an OpenRouter provider right now.


in JS, signals and AbortController can replicate some of the functionality but it's far less ergonomic than Go.

https://github.com/ggoodman/context provides nice helpers that brings the DX a bit closer to Go.


Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: