Loading BCONZ...
Immunology datasets
Inflammatory bowel disease datasets: Crohn's disease and ulcerative colitis
IBD treatment has moved through several generations of therapy, and patients often switch between them. Understanding sequencing, persistence and outcomes needs longitudinal real-world data. BCONZ provides governed, de-identified US EHR datasets for inflammatory bowel disease, including Crohn's disease and ulcerative colitis.
Not a free download. These are governed, research-grade datasets licensed for a defined purpose after a feasibility review — built for studies and products that public datasets cannot support.
Research uses
What immunology datasets are used for
Treatment sequencing and persistence
Switching between conventional, biologic and small-molecule therapies.
Real-world effectiveness and safety
Outcomes, flares, hospitalisation and surgery across therapy classes in routine care.
Disease course
Progression, complications and healthcare use in Crohn's disease and ulcerative colitis.
Patient stratification
Identifying subgroups for trial design and targeted development.
How access works
- Tell us your research question through the research data request form.
- We confirm feasibility: whether a suitable cohort exists, its size, follow-up and available data types.
- Governance review and a data use agreement define the permitted purpose.
- The de-identified dataset is prepared, quality-reviewed and delivered securely.
Hold immunology data yourself? A healthcare data readiness assessment shows what it can support, and DIA shows who has published a need for it.
Frequently asked questions
Immunology datasets: common questions
What are the inflammatory bowel disease patient numbers based on?
Figures are approximate and rounded down. They count patients in de-identified US EMR/EHR data with a recorded diagnosis in the area; a patient can appear in more than one indication. The cohort for a specific study is confirmed during feasibility.
How are these inflammatory bowel disease datasets different from public datasets on Kaggle or GitHub?
Public datasets are valuable for learning and benchmarking, but they are usually small, single-source, fixed snapshots with limited clinical context and no route to more data. BCONZ datasets are research-grade, de-identified and licensed through governed partnerships, with longitudinal records and a feasibility step to confirm the cohort fits your study before any agreement.
Is this a free dataset download?
No. Access is licensed for a defined research or development purpose under a data use agreement, after a feasibility review. Tell us what you need through the research data request form.
Where does the data come from?
From US healthcare data partners, and from partner networks in other regions where a study needs them. Data is de-identified before research use and stays governed by the partner's agreement.
