Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

In my experience, Gemini 3.7 is excellent for general non-coding tasks. But for coding, especially backend development, I still find models like Opus 5 and GPT-5.6 more reliable.


If I strip the dependencies out, I'm using it on a 5MLoC C++ codebase, and I found it performs really really well. I am using the Opus 5/Fable in parallel and I couldn't tell the difference. Both models make mistakes here and there.


I’ve been using Gemini for coding daily since 3.5, mostly on some pretty complex ML engineering projects. It’s great.

What’s an example of not being “reliable”?




Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: