Synthetic medical training data for Australian healthcare AI.
We build PHI-free synthetic document libraries that look and behave like real clinical PDFs - so teams can train OCR, layout-aware extraction, and clinical NLP models without waiting 18 months for ethics approval.
The bottleneck for medical document AI in Australia is training data. Real hospital PDFs are locked behind the Privacy Act. Generic synthetic medical text has no layout, no scans, no labels - useless for vision-language models like LayoutLMv3, Donut, or DocFormer. Public datasets like MIMIC are US-centric and increasingly restricted.
We sit in that gap: visually realistic, jurisdiction-specific, fully labelled, zero PHI risk.
Founded by Jack Webb - Sydney-based AI engineer, faculty advisor at the Australian Institute of Health Executives, background in clinical document AI and large-scale synthetic data generation.