Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

Yes it’s very codebase dependent, and changes over model versions. In my case codex was so bad, until it suddenly became better than Opus at 5.5, which surprised me. I am sure it can be the inverse for some codebases.

The way I keep an eye on model performance is alternating between models for code review. It shows how much the model understands the codebase without disrupting my workflow.



Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: