Supabase Releases Evals: an Open Source Benchmark That Scores Claude Code, Codex and OpenCode on Real Supabase Tasks
Supabase has open sourced Supabase Evals, its benchmark and framework for testing how nicely AI brokers construct utilizing Supabase. It runs coding brokers together with Claude Code, Codex, and OpenCode in opposition to actual duties, comparable to constructing a schema, debugging a failed Edge Function, or fixing a damaged RLS coverage, then scores the end…
