Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

As good as Claude has gotten recently in reasoning, they are likely using RL behind the scenes too. Supposedly, o1/strawberry was initially created as an engine for high-quality synthetic reasoning data for the new model generation. I wonder if Anthropic could release their generator as a usable model too.


while i was initially excited now im having second thoughts after seeing the experiments run by people in the comments here

on X I see a totally different energy more about hyping it

on HN I see reserved and collected take which I trust more.

I do wonder why they chose gpt4o which I never bother to use for coding.

Claude is still king and looks like I won't have to subscribe to ChatGPT Plus seeing it fail on some of the important experiments run by folks on HN

If anything these type of releases that air more on the side of hype given OpenAI's track record


I think people are wrong just about as often here as anywhere else on the internet, but with more confidence. Averaging HN comments would just produce outputs similar to rudimentary LLMs with a bit snobbier of a tone, I imagine.




Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: