Hacker Newsnew | past | comments | ask | show | jobs | submit | ArvidSu's commentslogin

You only need to come up with a catchy "SomethingBench" name, post it on reddit/x and now you're an ai sage. Not to disparage the launch/after comparison though, I'd genuinely enjoy a data point like that

A U-turn would make me throw up lmao

An "AI" server can do traditional server stuff but a traditional server can't do AI stuff (inference)


Why do you think they want less people subscribing?


> Why do you think they want less people subscribing?

Losing money on each subscriber?


very easy to lose money on subscription, very easy to make money on api pricing


A ThinkingCap variant of Qwen 3.8 27b would be extremely interesting.

https://huggingface.co/bottlecapai/ThinkingCap-Qwen3.6-27B

And then a Bonsai ternary on top of that model.


Re: bonsai - unsloth's quants have Q2 (UD-IQ2) variants, which are more or less same in size.

...or did Prism do something special with their "bonsai" releases? I didn't notice anything like QAT being mentioned.


There is some special sauce that they have. It’s not just a simple quant of another release. Or so they imply. I don’t have any insight into how it works or what the Bonsai special sauce is.


I suggest nobody repeats claims of "special sauce" if such "special sauce" is not understood.


That's super impressive given that it doesn't have vision! Intelligence overcomes blindness.


This explanation does so much without leaning into conspiracy that the labs are already sitting on the secret sauce but diluting it for the public or being left mystified when a lab drops out of the race for SoTA


My absolute favorite thing about Warp is the function that suggest a new command intelligently when the one I just ran failed or changes to the code when execution failed. Haven't been able to replicate that as smoothly in Ghostty which I really want to enjoy but honestly haven't really felt the same delight I got from Warp.


I'm not contradicting this but offering a contrast, I like Hermes because it simultaneously lowers barrier of entry and shows you what possibilities are unlocked by agents. I don't think I would have the time, interest or creativity to jump into the deep end by either extending an existing harness or rolling my own from the start. This also isn't an argument for doing just that, I might do so in the future, but critically only after Hermes has shown me what's possible and my preferences are developed.


I completely agree. I started with aichat[0] before the current agent trend and hit all kinds of bumps implementing agentic loops on my own. Then goose[1] showed me what a whole team working toward the same idea could do right before the official Claude Code harness which had all the bells and whistles. Now I know better what I want and its mostly less ram usage and a small set of primitives.

I still use Claude code (and codex and other big contenders) because they know what they are doing and innovate in ways I don't want to miss. And sometimes they are better at tasks.

[0]: 2023, https://github.com/sigoden/aichat [1]: 2024, https://github.com/aaif-goose/goose


Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: