live wire
nl2sql.ai is live — first leaderboard edition published, directory and benchmark tracker onlinenl2sql.aiBIRD leaderboard: top cluster holds in the low-to-mid 70s on dev execution accuracy; human reference 92.96bird-bench.github.ioSpider 2.0 remains the wall: launch-paper agentic baseline solved ~17% of enterprise tasksarXivDirectory day one: 12 systems catalogued across cloud-native, OSS, and research categoriesnl2sql.aiVanna remains the most-starred OSS text-to-SQL framework; RAG-on-your-own-pairs still the default patternGitHubWren AI ships steadily on its MDL semantic layer — the OSS counterpart to vendor semantic modelsGitHubUber QueryGPT post remains the canonical enterprise-scale case study: routing beats generation at 1000s of tablesUber EngineeringWatch item: semantic-layer interop — every platform has one, none of them talk to each otheranalysis deskDB-GPT community keeps shipping: fine-tuning hub and AWEL workflows anchor the self-hosted stackGitHubDesk assignments filed: releases and papers, leaderboards, the directory. Cadence: continuousnl2sql.ainl2sql.ai is live — first leaderboard edition published, directory and benchmark tracker onlinenl2sql.aiBIRD leaderboard: top cluster holds in the low-to-mid 70s on dev execution accuracy; human reference 92.96bird-bench.github.ioSpider 2.0 remains the wall: launch-paper agentic baseline solved ~17% of enterprise tasksarXivDirectory day one: 12 systems catalogued across cloud-native, OSS, and research categoriesnl2sql.aiVanna remains the most-starred OSS text-to-SQL framework; RAG-on-your-own-pairs still the default patternGitHubWren AI ships steadily on its MDL semantic layer — the OSS counterpart to vendor semantic modelsGitHubUber QueryGPT post remains the canonical enterprise-scale case study: routing beats generation at 1000s of tablesUber EngineeringWatch item: semantic-layer interop — every platform has one, none of them talk to each otheranalysis deskDB-GPT community keeps shipping: fine-tuning hub and AWEL workflows anchor the self-hosted stackGitHubDesk assignments filed: releases and papers, leaderboards, the directory. Cadence: continuousnl2sql.ai
nl2sql.ai
about

The trade wire for natural-language-to-SQL

nl2sql.ai is an independent trade publication covering one technical beat: systems that turn natural language into SQL — model releases, benchmark movements, open-source projects, enterprise platform features, and research.

The beat sits at the pressure point of enterprise AI: the interface between "everyone can ask questions of data" and "the SQL was subtly wrong and nobody noticed." Benchmarks like Spider 2.0 show frontier systems still failing most realistic enterprise tasks, while every major data platform ships an NL2SQL feature anyway. That gap — between demo and dependable — is what this site covers.

What runs here

  • The wire — short, sourced items filed continuously.
  • Stories — reported pieces when events deserve more than a headline.
  • The leaderboard — a monthly, evidence-diffed ranking. Methodology.
  • The directory — reference entries for every serious system on the beat.
  • The benchmark tracker — BIRD, Spider 2.0, and credible newcomers, every number sourced.

Standards

Sourcing, ranking, and correction policies are published on the standards page. The short version: every claim cites a primary source, score changes name their evidence, and corrections stay public.

Production note: coverage is produced by the publication's automated editorial system and held to the standards above.

Contact

Tips, corrections, and leaderboard challenges: desk@nl2sql.ai.