Loading...
Loading...
Available on 1 platform
Sign in to view source links and access this dataset
SANQI is a collection of molecular databases and samples for computational chemistry and drug discovery, created by luskyqi. It includes multiple subsets such as ChemDiv, Enamine, FDA-approved compounds, and natural products, along with algorithmically generated samples. The dataset was last updated in March 2026.
Data is stored in Python pickle (.pkl) format, requiring compatible libraries and caution regarding security. The full description and specific file contents are only available on the Hugging Face dataset page.