eyalitki

Comparison was done in the scope of coderabbit AI code review tool, which sadly makes it practically irrelevant.

My personal experience as a software engineer, and a former security researcher who did manual code audit, is that this code review tool has such poor results that it isn't worth the "noise" and friction it causes developers during C/I code review

show comments
sdeframond

How do you guys review AI-generated code ?

In our team, frontend work is vibe-coded by the PO and merged as-is without review. Backend is coded by developers, using AI but in a slower, more controlled way.

Recently, our PO has been trying his hand at vibe-coding the backend. I must say he is a smart guy, almost technical but not quite a developer. We've just been handed a burst of stacked PRs amounting for ~15k LOC backend. We do not quite know what do to about it.

I know we are not the only ones in the situation. What's your experience and context ? What do you do ? What works for you what doesn't ?

show comments
ramon156

Both OAI and Anthropic seem to have released a model that is slightly better but cost ~2x the previous iteration. Interesting play

show comments
SneakyZero

Astra seems to be really slow. Maybe it intends to read more context. But from my experience it is definitely slower than 5.6 sol when handling same tasks.

show comments
dude250711

Given that Fable is a Sol-class model, should Astra not be compared to Mythos in those tests?

sscaryterry

From what I can tell, its shit.