Back

High-Fidelity Tuning of Olfactory Mixture Distances in the Perceptual Space of Smell Through a Community Effort

Satarifard, V.; Sisson, L.; Han, Y.; Ilidio, P.; Hladis, M.; Lalis, M.; Song, X.; Yin, W.; Ravia, A.; Zheng, C. X.; Andreoletti, G.; Albrecht, J.; Pellegrino, R.; Wang, Z.; Yang, S.; D'hondt, R.; Ghinis, A.; de Boer, J.; Nakano, F. K.; Gharahighehi, A.; DREAM Olfactory Mixtures Prediction Consortium, ; Sanchez-Lengeling, B.; Keller, A.; Vosshall, L. B.; Fiorucci, S.; Tewari, A.; Topin, J.; Vens, C.; Bjorkman, M.; Kragic, D.; Sobel, N.; Christakis, N. A.; Mainland, J. D.; Meyer, P.

2025-12-16 neuroscience
10.64898/2025.12.13.694160 bioRxiv
Show abstract

A central goal in sensory science is to establish quantitative mappings between physical stimuli and perceptual responses. While such mappings are well characterized in vision and audition, they remain poorly defined in olfaction, limiting progress toward understanding the representations of smell. Predicting perceptual similarity between odor mixtures offers a promising route to formalize these relationships. To advance this effort, the DREAM (Dialogue for Reverse Engineering Assessment and Methods) Olfactory Mixtures Prediction Challenge assembled a curated, cross-study dataset describing the similarity of 507 mixture pairs and an unpublished test set of 46 mixture pairs. Teams competed to predict the perceptual similarity of mixture pairs, and then collaborated post-challenge to create an ensemble combining top-performing models that notably improves predictions over the existing state-of-the-art models. Moreover, ensemble model maintains high predictive accuracy in novel validation set. Our model provides a reproducible framework for neuroscientists, chemists, and engineers to compare odor mixtures and provides a foundation for future efforts towards better understanding the olfactory properties of mixtures.

Published in Proceedings of the National Academy of Sciences (predicted rank #3) · training set

Matching journals

The top 5 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.