datasets
Public dataset·Cross-modal learning

ObjectFolder 2.0

Multisensory neural objects with visual, acoustic, and tactile signals.

About this dataset

A dataset of 1,000 objects encoded as implicit neural representations that render vision, sound, and touch on demand. Ideal for pretraining cross-modal models and generating tactile readings at arbitrary contact points.

Citation

Gao et al. “ObjectFolder 2.0.” CVPR 2022.