Hacker Newsnew | past | comments | ask | show | jobs | submit | firethunder7's commentslogin

From what I understand, the goal is to train an LLM that is better at training LLMs than humans, so that it can continuously train smarter models and, once smart enough, design the successor to LLMs.


AA coding index has been updated to use DeepSWE, Terminal-Bench v2, and SWE-Atlas-QnA.


When? It literally says on the page for Gemini 3.6 Flash "Artificial Analysis Coding Index represents the weighted average of coding benchmarks in the Artificial Analysis Intelligence Index (Terminal-Bench v2.1, SciCode)"


My bad it was the coding agent index


Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: