Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

But they score it on their own benchmark, on which coincidentally Gemini models always were the only good ones. In Nolima or Babilong we see that Gemini models still cant do long context.

Excited to see if it works this time.



Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: