For research informatics
Clinical research informatics without the warehouse queue.
Clinical Extract gives investigators study-scoped extracts from your EHR's certified bulk FHIR API. Their requests no longer wait behind warehouse builds.
Each extract holds only the study cohort and the variables the protocol approved, in one CSV with a data dictionary.
1 From request to file
Four steps from a research data request to a study export.
The investigator's approved variable list; the cohort, built from coded criteria or matched from a list; the honest broker's review of the data dictionary; the export. The Group export narrows the export to the study cohort and its resource types, and the minimum-necessary cut keeps only the approved variables, so identifiers the protocol did not approve never reach study.csv.
EHR data for research arrives coded (ICD-10-CM, LOINC, RxNorm, SNOMED CT) in study.csv, and data-dictionary.json defines every column.
2 The queue
Study pulls and the data warehouse.
A warehouse request waits in a shared queue for an analyst, a query, a chart check and a de-identification pass. A study pull goes from the approved variable list to study.csv in days, most of them the honest broker's review, and the export itself runs in hours. The warehouse still handles enterprise reporting.
Figure 2 as a table
| Stage | From request | Duration |
|---|---|---|
| Warehouse: waits in the intake queue | day 0 to 21 | 21 days, waiting |
| Warehouse: scoping meeting with a warehouse analyst | day 21 to 24 | 3 days |
| Warehouse: SQL written and run against the warehouse | day 24 to 36 | 12 days |
| Warehouse: results checked against charts | day 36 to 41 | 5 days |
| Warehouse: waits for the honest broker | day 41 to 45 | 4 days, waiting |
| Warehouse: de-identified and delivered | day 45 to 49 | 4 days |
| Warehouse: request to file | day 0 to 49 | 49 days |
| Study pull: cohort built, approved variables picked | hour 0 to 1 | 1 hour |
| Study pull: Group $export, kickoff to last NDJSON file | hour 1 to 2.8 | about 2 hours, illustrative for a study cohort |
| Study pull: study.csv and data-dictionary.json written | hour 2.8 to 3 | minutes |
| Study pull: waits for the honest broker | hour 3 to 64 | 61 hours, waiting |
| Study pull: honest broker reviews data-dictionary.json | hour 64 to 72 | 8 hours |
| Study pull: request to file | hour 0 to 72 | 3 days |
3 The honest broker
How the honest broker reviews and releases an extract.
Honest broker research centers on one document: the list of variables to be released. Clinical Extract writes that list before the export runs. data-dictionary.json names every column, its description, the FHIR element it comes from and an example value. The broker approves those variables, keeps the re-identification key, and releases the coded extract, and the analyst receives the same dictionary. One dictionary entry, explained shows what the broker reads.
Research informatics teams keep their governance as written: the export runs inside your environment, the token carries read scopes only, and no patient data reaches us. Minimum necessary in the export covers the scopes and the data path.
4 The IRB
What the IRB application quotes.
Which patients: the cohort criteria and the count the cohort builder returned. Which variables: the dictionary. Where each one comes from: the fhirSource field. Preparatory-to-research counts come from building the cohort from coded criteria before any record moves, so the application carries a number and its definition instead of an estimate.
5 Questions
Questions from research informatics
What is an honest broker in research?
The person or office that stands between the investigator and the identified record: it holds the re-identification key, approves what leaves, and releases only the coded extract. Clinical Extract gives the broker a document to approve before anything moves: data-dictionary.json, one entry per column with its FHIR source.
What is the honest broker protocol?
The written procedure for that role: who requests, who approves the variable list, how identifiers are handled, where the key is kept, how the release is logged. Clinical Extract fits it as written: the approved list is the export’s variable list, the extract is study-scoped, and the export itself runs inside your environment.
What is clinical research informatics?
The discipline that turns clinical data into research data: the systems, the governance and the people between the EHR and the investigator. Research informatics teams run the data warehouse, the honest broker service and the study data requests, and Clinical Extract adds study pulls that answer a request in days.
How does a study pull differ from a warehouse request?
A warehouse request waits for an analyst, a query, a chart check and a de-identification pass. A study pull starts from one protocol: the approved variable list becomes the export, the cohort’s Group scopes it, and the CSV lands in days, most of them the broker’s review.
Where do the files land?
In a study folder inside your environment, on the server or in the cloud tenant you run Clinical Extract on: the NDJSON, study.csv and data-dictionary.json. Patient data moves only between your EHR and that folder.
Next
See Clinical Extract run on one of your studies.
Tell us which EHR you run and what the study or registry needs. We reply within one business day to set a meeting time.