Introducing TRIBE v2: A Predictive Foundation Model Trained to Understand How the Human Brain Processes Complex Stimuli
Meta AI Blog
TRIBE v2 is an AI model trained to predict how the human brain responds to visual, auditory, and language stimuli by learning from fMRI scans of over 700 volunteers. The model was trained on more than 700 healthy volunteers presented with diverse media including images, podcasts, videos, and text, and can make predictions for new subjects, languages, and tasks without additional brain imaging data. Researchers can now test hypotheses about brain function computationally, reducing the need for human subjects in experimental studies and potentially accelerating neuroscience discovery.
Why it matters
Today, we're announcing TRIBE v2: our first AI model of human brain responses to sights, sounds, and language. Building on our Algonauts 2025 award-winning model, which was trained on the low-resolution fMRI recordings of four individuals, we leverage a massive dataset of more than 700 healthy volunteers who were presented with a wide variety of media, including images, podcasts, videos, and text. TRIBE v2 reliably predicts high-resolution fMRI brain activity — enabling zero-shot predictions for new subjects, languages, and tasks — and consistently outperforms standard modeling approaches. By creating a digital model of the human brain, researchers can rapidly test hypotheses about its underlying functions without the need for human subjects in every experiment.