About

Research data should outlive the people who collected it.

ScientHouse exists so that a result published today can still be checked, re-run and built on years from now, whoever is still on the team.

What we believe

Four principles.

A result is data plus method.If the data underneath it can change without a trace, the result cannot be trusted. So we pin everything to a snapshot.
Access is a decision, not a default.Who may see data is decided by the application on every request, and every decision can be reviewed.
Your data stays yours.Open formats in your own cloud account, so leaving is possible and staying is a choice.
Small teams should not build a platform.Labs and programs should spend their time on science, not on stitching five tools together.
Where we are

Early, and working with a few partners.

Built and runningThe core platform: collections, data structures, cohorts, releases, findings, notebooks, permissions and audit.
Optional modulesAnalysis pipelines, credits and spend limits, and access from users' own AWS accounts, enabled per deployment.
NextCard purchases for credits and inviting people by email. Both need more work before we would call them ready.
Open by design

Built on open standards.

Apache IcebergOpen table format for the data, readable by other tools.
DuckDBFast SQL in the browser and on the server.
marimoReproducible notebooks that run in the browser.
nf-core and AWS HealthOmicsCurated, peer-reviewed pipelines for optional analysis.
Team

The people behind it.

Placeholder: add founder and team bios, photos and relevant research or engineering background here. This is usually the most-read part of an About page for a research audience.

Talk to the team.

We would like to hear what you are trying to reproduce.

Contact us