[HN Gopher] A real-world benchmark for AI code review
___________________________________________________________________
A real-world benchmark for AI code review
Author : benocodes
Score : 23 points
Date : 2026-02-04 21:13 UTC (1 hours ago)
(HTM) web link (www.qodo.ai)
(TXT) w3m dump (www.qodo.ai)
| CuriouslyC wrote:
| I don't think LLMs are the right tool for pattern enforcement in
| general, better to get them to create custom lint rules.
|
| Agents are pretty good at suggesting ways to improve a piece of
| code though, if you get a bunch of agents to wear different hats
| and debate improvements to a piece of software it can produce
| some very useful insights.
| mbesto wrote:
| Cmd+F - "Overfitting"...nothing.
|
| Nope, no mention of how they do anything to alleviate
| overfitting. These benchmarks are getting tiresome.
| aetherspawn wrote:
| Your pricing page has a bug on it, the annual price is higher
| than the monthly price.
| zamadatix wrote:
| I'm seeing $30/m at annual and $38/m at monthly? (maybe already
| fixed, hard to tell)
| mdeeks wrote:
| I feel like pricing needs to be included here. I kind of don't
| care about 10 percentage points if the cost is dramatically
| higher. Cursor Bugbot is about the same cost but gives 10x the
| monthly quota of Qodo.
|
| I know this is focused solely on performance, but cost is a major
| factor here.
| falloutx wrote:
| Company creates a benchmark. Same company is best in that
| benchmark.
|
| Story as old as time.
| kachapopopow wrote:
| coderabbit being the worst while (presumeably) advertising the
| most seems to be check out at least, wouldn't believe the recall
| % seems bogus.
___________________________________________________________________
(page generated 2026-02-04 23:00 UTC)