OpenCode Zen
big-pickle, a free stealth coding model served via OpenCode Zen, scored a 50.8% Task Resolve Rate (63/124) on Scale AI's SWE Atlas Codebase QnA benchmark in August 2026, topping the mini-swe-agent class. Using the same Harbor v0.18.0 scaffold as Scale AI, it surpassed GLM 5.2 at 48.1% and GPT-5.6-Sol at 46% on the same harness. The result came from a single trial rather than the official three-run protocol and ran on reduced sandbox resources, so it should be read with caution. The model reached 58.1% accuracy on TypeScript and 60.7% on Code Onboarding tasks, while Claude Opus 5 and Opus 4.8 still lead the overall leaderboard using their native Claude Code scaffold. API signatures suggest DeepSeek infrastructure, and the underlying identity remains unconfirmed.