getsentry/sentry

Developer-first error tracking and performance monitoring

Run by mupt-ai · Released 2026-10-01 · 46 tasks · 2 settings

Which coding agent works best on getsentry/sentry? Most accurate: GPT-6 Luna, 45.7% at $0.024 per task. 2 model settings scored on 46 tasks from its merged pull requests.

All Settings

ModelHarnessReasoningAccessAccuracyCost / TaskFrontier
GPT-6 LunaCodexHighAPI Key45.7%$0.024Yes
GPT-6 LunaCodexLowAPI Key32.6%$0.031No

A setting is on the frontier when no other is both cheaper and more accurate.

All Repositories on SelfBench