§ · Pricing

Free to use.

Create an account and generate up to 5,000 rows a day — no card, no contracts.

Free
Free tier
$0 / forever

Generate synthetic people and staged T2DM patients via the API or MCP — up to 5,000 rows per day, resetting at 00:00 UTC.

  • 5,000 rows / day
  • 69 attributes that co-vary (joint realism)
  • Calibrated to NHANES, ACS, CDC, KFF, BLS
  • Async bulk jobs → CSV, JSONL & FHIR R4
  • Deterministic — same seed, same data
  • Usage & full call history in your profile
Create a free account
Ask
Need more?
Talk to us

Hitting the daily limit, or want an MCP endpoint or a higher base quota? Tell us your use case and we'll get you set up — solo-built, usually a reply within one business day.

  • Higher base quota — 2×, 4×, or 8×
  • API or MCP access
  • One-off datasets beyond 5,000 rows — up to 1M
  • Custom populations & use-case questions
Talk to us →

The engine is calibrated, deterministic, and validated on every release. No paid tier, no sales funnel: quota raises and custom datasets are free, by request.

§ · What's in every record

69 attributes across 9 domains: identity (name, email, phone, SSN placeholder, DOB), geography (state, city, ZIP — correlated), social (marital status, education, employment), financial (income bracket, insurance type, EIN), behavioral (smoking, alcohol, exercise, sleep), health basics (height, weight, BMI, waist, A1c, blood pressure, eGFR), health conditions (diabetes, hypertension, CKD, cancer, depression — age-conditioned), utilization (visits, ER, hospitalizations), and medications.

Every distribution — and every dependency between attributes — is benchmarked against published US references: the American Community Survey (ACS 2022), the National Health and Nutrition Examination Survey (NHANES 2017–2020), the CDC's National Diabetes Surveillance System (NDSS 2022), the Medical Expenditure Panel Survey (MEPS 2022), KFF (formerly the Kaiser Family Foundation, 2023), the Bureau of Labor Statistics (BLS 2023), and the 2020 US Census. See the validation vs real NHANES →

§ · Frequently asked
Q1
Is this real patient data?

No. Every record is synthetically generated from published reference distributions. No real person's data enters the system at any point.

Q2
Can I use this for HIPAA-regulated work?

The data is synthetic by construction — no protected health information exists. Consult your compliance officer for your specific use case.

Q3
What format is the data in?

CSV, JSONL, or FHIR R4. Generation is asynchronous — you submit a job, poll its status, then download the finished file (not a synchronous JSON response).

Q4
Can I see the data first?

Yes — download the free 1,000-row sample (CSV / JSONL) with no signup, and see a full example record and every field on the product page; the OpenAPI spec documents the whole API. To generate your own data, create a free account — it takes about 10 seconds.

Q5
Can I get more than 5,000 rows a day?

Yes — email hello@simpleidgen.com. Higher quotas, larger one-off datasets (up to 1M rows), and custom populations are all free, by request.