# Introducing Supabase Evals

DevFeed: [Introducing Supabase Evals](<https://devfeed.tech/articles/introducing-supabase-evals-410.md>)

Original publisher: [Read original article](<https://supabase.com/blog/introducing-supabase-evals>)

Author: Matt Rossman

Published: 2026-07-31T07:00:00Z

Content type: article

Language: en

Sources: [Supabase Blog](<https://devfeed.tech/sources/supabase-blog.md>)

Topics: [Supabase](<https://devfeed.tech/topics/supabase.md>), [AI-assisted coding](<https://devfeed.tech/topics/ai-assisted-coding.md>), [Benchmark](<https://devfeed.tech/topics/benchmark.md>), [Open Source](<https://devfeed.tech/topics/open-source.md>), [coding](<https://devfeed.tech/topics/coding.md>), [Command-line interface](<https://devfeed.tech/topics/cli.md>), [MCP Server](<https://devfeed.tech/topics/mcp-server.md>), [Agent Skills](<https://devfeed.tech/topics/agent-skills.md>), [Database](<https://devfeed.tech/topics/database.md>), [debugging](<https://devfeed.tech/topics/debugging.md>), [observability](<https://devfeed.tech/topics/observability.md>), [GitHub Issues](<https://devfeed.tech/topics/github-issues.md>)

Tags: [agent-skills](<https://devfeed.tech/tags/agent-skills.md>), [ai](<https://devfeed.tech/tags/ai.md>), [ai-coding](<https://devfeed.tech/tags/ai-coding.md>), [benchmark](<https://devfeed.tech/tags/benchmark.md>), [cli](<https://devfeed.tech/tags/cli.md>), [coding](<https://devfeed.tech/tags/coding.md>), [github-issues](<https://devfeed.tech/tags/github-issues.md>), [mcp](<https://devfeed.tech/tags/mcp.md>), [mcp-server](<https://devfeed.tech/tags/mcp-server.md>), [open-source](<https://devfeed.tech/tags/open-source.md>), [testing](<https://devfeed.tech/tags/testing.md>)

## AI overview

Supabase introduces an open-source benchmark and framework for evaluating how well AI coding agents build with Supabase. It runs agents such as Claude Code, Codex, and OpenCode through realistic tasks, including schema construction, Edge Function debugging, and RLS policy fixes, then measures their performance in benchmark and regression suites.

## Source excerpt

Our open-source benchmark for how well AI coding agents build with Supabase.