Hacker News·4 min read·hard

Big Pickle on SWE Atlas – Codebase QnA

P
phillipchaffee
Big Pickle on SWE Atlas – Codebase QnA
AI Summary

A report on the performance of 'big-pickle', a stealth AI model, on the SWE Atlas Codebase QnA benchmark. The model demonstrates competitive results compared to existing leaderboard entries when using specific scaffolding.

Task Resolve Rate: 50.8% (63/124) — big-pickle , the free stealth model on OpenCode Zen, evaluated on Scale AI's SWE Atlas Codebase QnA benchmark using the mini-swe-agent scaffold.

Continue reading on Headlinne

Create a free account to read the full article.

Read full article →
technologybusiness

Get the full story

Sign up for Headlinne to unlock AI insights, political bias analysis, and your personalized news feed.

Create free account

Already have an account? Sign in