Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

Results from my Extended NYT Connections benchmark:

GPT-5.4 extra high scores 94.0 (GPT-5.2 extra high scored 88.6).

GPT-5.4 medium scores 92.0 (GPT-5.2 medium scored 71.4).

GPT-5.4 no reasoning scores 32.8 (GPT-5.2 no reasoning scored 28.1).



How do you score this? Losing/winning the game with 4 lives?



Impressive! Do you include puzzles released before the training data cutoff date?




Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: