Hacker Newsnew | past | comments | ask | show | jobs | submit | oh_no's commentslogin

look at token use, 3.8 flash is a huge token hog compared to openai models

Why do you think this uncracked code was so simple to solve?

And it's very possible Terra or GLM could crack it, turn off their web access and try yourself.


> Why do you think this uncracked code was so simple to solve?

I never said this. All I said is we don't have the conversation and therefore we can't determine how easy or hard of a problem it was.

> And it's very possible Terra or GLM could crack it, turn off their web access and try yourself.

I'm questioning why this should be labeled "Astra" breaking anything implying it required "the best" model to do it when in fact any other half-decent model might have been able to do this as well.

EDIT: Okay seems like the actual prompt is published, just not on the same article that was linked. Maybe I'll give it a try.


Nice try at content-free debunking though Encyclopedia Brown

shutoff for existing models, new models stopped as of that announcement, astra will never be on cursor.

which is crazy because this was grok's competitive advantage, worse than OpenAI models but better than everything else, now it's less efficient than Opus or Fable 5.1

the AA numbers are generationally bad. double token use (the one thing Grok was good at was low reasoning usage!) to gain 5% in the benchmark score. with reportedly a larger model. maybe it shows gains IRL but wow, I've never seen a new generation model look so underwhelming compared to the last.

It still works, bots can solve it but it probably increases the cost of that web call by 10x or 100x for that bot, so it won't bother. Had a recent bad experience with removing recaptcha.


I have bad experiences with recaptcha. It also increases user effort, and they might not bother either. AI doesn't get annoyed as easily.


I get this pretty frequently on windows Firefox after switching to it, note this is my work computer, Firefox works fine at home on more open network


I get reCaptchaed-to-death all the time in iOS and MacOS using both Firefox and Brave. I use a big name VPN which probably makes it worse, since I'm routinely blocked outright by Cloudfare services, assuming cluelessly that I'm a bot.


Very nice to see that this is even more token efficient than Sol, when Fable 5.1 is less so than the already bloated token budget of Fable 5.


that's all openai models but i'm very happy openai continues to focus on efficiency rather than reasoningtokenmaxxing


My assumption is in the long term that efficiency will break interpretability, which will lead to questionable alignment. As efficiency drives capitalism and evolution we'll run headlong at it and try to deal with the risk as a side effect.


either way we cannot see those thoughts anyway


the last thing nvidia would want is to make ai look more legally risky

to say nothing of their massive investments direct and indirect in openai


Yep. In fact, eliminating that risk on OpenAI's behalf was probably a small factor in the acquisition


Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: