Show HN: Self-bench – build SWE-bench style evals from private repos
By byhong03 · 2026-08-14 · 3 points · 0 comments
https://github.com/mupt-ai/self-bench
Hey HN! We all know that most public evals are saturated and are hard to trust. What matters is whether a model reliably works in your codebase in your actual day-to-day-work. To solve this, we built self-bench ( https://github.com/mupt-ai/self-bench ), an op…
Open the full discussion on BetterNews