Free to use.
Create an account and generate up to 5,000 rows a day — no card, no contracts.
Generate synthetic people and staged T2DM patients via the API or MCP — up to 5,000 rows per day, resetting at 00:00 UTC.
- → 5,000 rows / day
- → 69 attributes that co-vary (joint realism)
- → Calibrated to NHANES, ACS, CDC, KFF, BLS
- → Async bulk jobs → CSV, JSONL & FHIR R4
- → Deterministic — same seed, same data
- → Usage & full call history in your profile
Hitting the daily limit, or want an MCP endpoint or a higher base quota? Tell us your use case and we'll get you set up — solo-built, usually a reply within one business day.
- → Higher base quota — 2×, 4×, or 8×
- → API or MCP access
- → One-off datasets beyond 5,000 rows — up to 1M
- → Custom populations & use-case questions
The engine is calibrated, deterministic, and validated on every release. No paid tier, no sales funnel: quota raises and custom datasets are free, by request.
69 attributes across 9 domains: identity (name, email, phone, SSN placeholder, DOB), geography (state, city, ZIP — correlated), social (marital status, education, employment), financial (income bracket, insurance type, EIN), behavioral (smoking, alcohol, exercise, sleep), health basics (height, weight, BMI, waist, A1c, blood pressure, eGFR), health conditions (diabetes, hypertension, CKD, cancer, depression — age-conditioned), utilization (visits, ER, hospitalizations), and medications.
Every distribution — and every dependency between attributes — is benchmarked against published US references: the American Community Survey (ACS 2022), the National Health and Nutrition Examination Survey (NHANES 2017–2020), the CDC's National Diabetes Surveillance System (NDSS 2022), the Medical Expenditure Panel Survey (MEPS 2022), KFF (formerly the Kaiser Family Foundation, 2023), the Bureau of Labor Statistics (BLS 2023), and the 2020 US Census. See the validation vs real NHANES →
No. Every record is synthetically generated from published reference distributions. No real person's data enters the system at any point.
The data is synthetic by construction — no protected health information exists. Consult your compliance officer for your specific use case.
CSV, JSONL, or FHIR R4. Generation is asynchronous — you submit a job, poll its status, then download the finished file (not a synchronous JSON response).
Yes — download the free 1,000-row sample (CSV / JSONL) with no signup, and see a full example record and every field on the product page; the OpenAPI spec documents the whole API. To generate your own data, create a free account — it takes about 10 seconds.
Yes — email hello@simpleidgen.com. Higher quotas, larger one-off datasets (up to 1M rows), and custom populations are all free, by request.