Introducing world's largest synthetic open-source Text-to-SQL dataset
Blog post from Gretel.ai
Gretel has introduced the world's largest synthetic open-source Text-to-SQL dataset, available on Hugging Face under Apache 2.0 license. The gretelai/synthetic_text_to_sql dataset is designed and generated using Gretel Navigator and includes over 105,851 records with diverse SQL tasks and complexity levels. This synthetic data accelerates the transition to data-centric AI by allowing teams to produce high-quality data while preserving privacy and security. The release of this dataset marks a significant milestone in the world of synthetic data and encourages developers, researchers, and data enthusiasts to leverage it for their projects.
| Trend | Post Mentions | Total Month Mentions | Posts | Companies | MoM |
|---|---|---|---|---|---|
| LLM | 14 | 3,669 | 412 | 154 | +40% |
| RAG | 2 | 1,867 | 232 | 78 | +54% |
| AI Model Fine-tuning | 1 | 787 | 151 | 83 | +58% |
Use this post, company, and trend context to find content marketing opportunities, perform competitive analysis, or address product feature gaps via the Plushcap MCP server or the Plushcap API.