verified
synthetic datasets.
Each dataset includes documentation on origin, legal status, quality, ethical notes and permitted use. Filter by sector, format, language and intended use — same dossier structure on every listing.
5 datasets
Synthetic Complaints and Rights Requests Dataset
Synthetic complaints, rights requests, escalations and edge cases for testing AI customer support and compliance flows.
Records
5.4k
Format
JSONL
Version
v1.2
AI Customer Support Testing Dataset
Multi-turn synthetic conversations between users and support agents across sectors — for chatbot evaluation, escalation testing and response quality review.
Records
12k
Format
JSONL
Version
v1.0
Bias and Edge-Case Evaluation Set
A curated evaluation set for probing AI consistency, fairness behaviour and edge cases. Designed for testing, not training.
Records
2.2k
Format
CSV
Version
v0.9
Synthetic Insurance Claims Dataset
Synthetic insurance claim filings with labelled fraud signals — for testing claim-triage models, anomaly detection and underwriting AI.
Records
8.2k
Format
Parquet
Version
v1.0
Multilingual Public Services Inquiries
Synthetic citizen inquiries to public services across common topics — for testing multilingual AI assistants, intent detection and routing logic.
Records
6.5k
Format
JSONL
Version
v1.0
can't find the right shape?
Scope a custom dataset — same dossier standard, built for a use case the public catalog doesn't cover.