Twenty-four thousand fecal samples. Twelve thousand blood samples. Continuous glucose traces, wearable streams, facial scans and deep microbial sequencing from more than 6,000 people: the GUT-FEELINGS study is as much a data-engineering project as a medical one, according to its designers.
Each participant contributes across 16 weeks of intervention, and every stream must be aligned — what a volunteer ate, how their microbes shifted, how their mood scores moved. The analysis challenge dwarfs the collection challenge, researchers acknowledge in the announcement.
The payoff for scale is pattern. Gut-brain relationships are famously individual; only datasets of this size can separate signals that hold across populations from those true for a single gut. The study’s 19-formula randomized design gives the data something rarer still: causal leverage.
The teams intend the dataset to train Mercator, a foundation model of the human gut microbiome, according to the announcement — an attempt to convert one giant trial into a reusable scientific instrument.
Whether the model delivers, the dataset itself becomes infrastructure: a reference map of American gut biology that smaller studies will be measured against for years.
The samples-to-insight ratio is the honest challenge. Twenty-four thousand fecal samples will generate data faster than any team’s hypotheses; without the randomized structure, the project would be a very expensive act of description. The 19-formula design is what converts the mountain into evidence — each formula a probe, each volunteer’s streams a before-and-after record against which a formula’s fingerprint can be isolated from the noise of ordinary life.
The foundation-model ambition deserves both the attention and the scepticism it will receive. A model trained on one trial, however large, learns that trial’s population, diets and geographies; whether Mercator generalises to the next cohort is an empirical question the field will grade harshly and fairly. But the dataset’s existence changes the baseline for everyone: future gut-brain claims will be asked, with justification, how they compare against 6,000 Americans measured properly.
Related reading:

