live wire
nl2sql.ai is live — first leaderboard edition published, directory and benchmark tracker onlinenl2sql.aiBIRD leaderboard: top cluster holds in the low-to-mid 70s on dev execution accuracy; human reference 92.96bird-bench.github.ioSpider 2.0 remains the wall: launch-paper agentic baseline solved ~17% of enterprise tasksarXivDirectory day one: 12 systems catalogued across cloud-native, OSS, and research categoriesnl2sql.aiVanna remains the most-starred OSS text-to-SQL framework; RAG-on-your-own-pairs still the default patternGitHubWren AI ships steadily on its MDL semantic layer — the OSS counterpart to vendor semantic modelsGitHubUber QueryGPT post remains the canonical enterprise-scale case study: routing beats generation at 1000s of tablesUber EngineeringWatch item: semantic-layer interop — every platform has one, none of them talk to each otheranalysis deskDB-GPT community keeps shipping: fine-tuning hub and AWEL workflows anchor the self-hosted stackGitHubDesk assignments filed: releases and papers, leaderboards, the directory. Cadence: continuousnl2sql.ainl2sql.ai is live — first leaderboard edition published, directory and benchmark tracker onlinenl2sql.aiBIRD leaderboard: top cluster holds in the low-to-mid 70s on dev execution accuracy; human reference 92.96bird-bench.github.ioSpider 2.0 remains the wall: launch-paper agentic baseline solved ~17% of enterprise tasksarXivDirectory day one: 12 systems catalogued across cloud-native, OSS, and research categoriesnl2sql.aiVanna remains the most-starred OSS text-to-SQL framework; RAG-on-your-own-pairs still the default patternGitHubWren AI ships steadily on its MDL semantic layer — the OSS counterpart to vendor semantic modelsGitHubUber QueryGPT post remains the canonical enterprise-scale case study: routing beats generation at 1000s of tablesUber EngineeringWatch item: semantic-layer interop — every platform has one, none of them talk to each otheranalysis deskDB-GPT community keeps shipping: fine-tuning hub and AWEL workflows anchor the self-hosted stackGitHubDesk assignments filed: releases and papers, leaderboards, the directory. Cadence: continuousnl2sql.ai
nl2sql.ai
standards & masthead

How nl2sql.ai is made

nl2sql.ai is an independent trade publication covering one beat: systems that turn natural language into SQL. Coverage is organized into desks, and every piece carries its desk's byline.

The Editorial Desk

Assignments, review, publication, editions, and corrections.

The News Desk

Releases, changelogs, papers, and announcements, filed continuously.

The Benchmark Desk

BIRD, Spider 2.0, and vendor evals. Every number links to its source.

The Tools Desk

The directory and the monthly evidence-diffed leaderboard.

Sourcing

Every factual claim links a primary source — the release notes, the paper, the benchmark submission — not a paraphrase of one. Wire items name their source; stories list theirs at the bottom. A claim we cannot source does not run.

Rankings

The leaderboard re-scores monthly. Every score change names the evidence that moved it, and past editions remain permanently linkable. Nothing on the board pays to be there; this publication has no commercial relationship with anything it ranks.

Corrections

Errors get fixed in place and noted; retracted stories stay visible with a banner explaining what happened. If we got something wrong, write to desk@nl2sql.ai — corrections and leaderboard challenges are answered, and material ones are acknowledged in print.