🚀NEW COURSEVibe Coding AI Apps with Claude Code 🤖✨Enroll now
Data · Evaluation

Synthesizing Text-to-SQL Data from Weak and Strong LLMs

Free while signed in. Answers cite the passages they came from.

First page
Synthesizing Text-to-SQL Data from Weak and Strong LLMs
The curator’s take

proposes integrated synthetic data to build a highly specialized SoTA text-to-SQL model called SENSE; the synthetic data from strong models enhances data diversity while valuable erroneous data from weaker models combined with an executor to learn from execution feedback; preference learning is used to instruction-tune LLMs to learn from both correct and incorrect samples; SENSE achieves state-of-the-art results on the SPIDER and BIRD benchmarks, which bridges the performance gap between open-source models and methods that use closed-source models.

Every Monday
Get next week’s papers.

The same picks and the same summaries, in your inbox. Free, and 176 issues deep.

Subscribe on Substack