Hacker Newsnew | past | comments | ask | show | jobs | submit | notfromhere's commentslogin

This really just exists so cognition can stop spending API tokens with Anthropic or OpenAI.

Basically any successful AI based service will do this because at scale the frontier models are expensive and you’ll have enough data to fine tune your own.

Same reason Harvey is doing models now and basically every other provider


Couldn't they just grab and run an open weight model to save on API tokens?

You get better performance if you also finetune it for your task

Well the models did get better but yeah their early product was godawful

its nowhere near as good as it was in Dark Sky. the apple weather app is pretty bad where I'm at.

Aren't they used doing this so they can generate training data for using Macs via the UI with computer use?

Sol Light on computer use is fantastic. I use it whenever I need to dive deep into whatever shitty web saas app menu if the API is unavailable

If your standard for free speech is “my side gets to question propaganda, but the other side’s worst posts prove the platform is biased,” that’s not a free speech principle.


I stand for no specific propaganda to be shut down.

That was not the case in the pre-Musk twitter era.


LLMs seem to be trained to work very well against a goal, especially one it can verify against. I guess because it can easily know if it passed or failed, va other tasks where good/bad output is subjective


It’s hard to manage because it’s not intelligent in a predictable way. More like a genius toddler


Na, they seem to constantly set up scenarios to create headlines. Stuff like “it hacked out of its container and tried to self replicate!” Where in reality it used provided skills and permissions while doing the thing they prompted it to do.


I have seen a lot of companies start with this, then when they hit 150 users on their team plan and start having to pay API rates they immediately start introducing other models.


You should be running your agent in a box so that’s not really a risk


Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: