Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

maybe gp's use of the word "lots" is unwarranted

https://artificialanalysis.ai indicates that sonnect 4.6 beats opus 4.6 on GDPval-AA, Terminal-Bench Hard, AA Long context Reasoning, IFBench.

see: https://artificialanalysis.ai/?models=claude-sonnet-4-6%2Ccl...

 help



I was basing it off my recollection of this:

https://www.anthropic.com/_next/image?url=https%3A%2F%2Fwww-...

basically 9/13 are very close




Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: