Home / Companies / Gretel.ai / Blog / Post Details
Content Deep Dive

An Awesome Synthetic Multilingual Prompts Dataset

Blog post from Gretel.ai

Post Details
Company
Date Published
Author
Maarten Van Segbroeck
Word Count
652
Company Posts That Month
3
Language
English
Hacker News Points
-
Post removed?
No
Summary

Gretel has released a comprehensive "Synthetic Multilingual LLM Prompts" dataset, featuring 1,250 synthetic prompts in seven languages. The dataset is designed for use with conversational LLMs like ChatGPT and is available on GitHub and Hugging Face. Translation quality was assessed using the LLM-as-a-Judge method, ensuring accuracy, fluency, and consistency across languages. This dataset is released under the Apache 2.0 license and can be used with proper attribution.

Trends Found in this Post
Trend Post Mentions Total Month Mentions Posts Companies MoM
LLM 7 4,537 421 147 +51%
Use This Data

Use this post, company, and trend context to find content marketing opportunities, perform competitive analysis, or address product feature gaps via the Plushcap MCP server or the Plushcap API.