Concepts, guides and API reference.
Start with the vocabulary. The same handful of ideas explain most of the platform.
The vocabulary of the platform.
No concept matches that search.
Task-based guides.
The outline below is what we are writing for early deployments.
Getting started
Create a collection, attach a data structure, upload your first CSV.
Reproducible analysis
Pin a cohort, write a notebook, publish a finding, and re-run it.
Releasing data
Approvals, releases, retractions and the consumer view.
Administering a tenant
Members, permission groups, approval policy and quotas.
Pipelines
Launch a curated pipeline, register your own workflow, and read the outputs.
Security review pack
Architecture overview and answers to common questionnaire items.
Read any snapshot with plain SQL.
Tables are open Apache Iceberg, partitioned by collection. Passing a snapshot id reads the data as it was.
SELECT site, COUNT(DISTINCT subject_id) AS n
FROM iceberg_scan('s3:///warehouse/nda/cde_phq901/',
version = '')
WHERE collection_id = 12
GROUP BY site; Illustrative. The workbench builds this for you and always filters to the collections you may read.
Everything the app does is an API call.
Missing something?
Tell us what you need to see documented first.