Loading...
Loading...
3D models, rendered datasets, physics simulation, digital twins, synthetic data generation, game engine data
1,417 datasets
Khorshed Alam published a dataset on figshare in April 2026. The dataset contains performance metrics for different synthetic data methods. The data is stored in an Excel file and is 5.5 KB in size.
The Albufera de València in Spain provides the geographic scope for this long-term dataset on its fish community. It compiles species presence-absence records, functional traits, and environmental niche descriptors for fish species from 1865 to 2025. The dataset is structured across three complementary Excel sheets.
A 2.1 MB PDF file containing supplementary material for a study, authored by Jasper Marcum and last updated on April 16, 2026. The material includes detailed derivation, numerical implementation details, and a mesh convergence study. It is licensed under CC-BY-4.0 and hosted on the figshare platform.
Roblox.Mesh likely contains 3D mesh assets for use on the Roblox platform. The dataset was published by author Hwiiiiiiii on HuggingFace and was last updated on June 14, 2026. The specific content, scale, and format of the data require verification after download.
Experimental data from a study investigating the effects of psilocybin on social aggression and activity in the mangrove rivulus fish, Kryptolebias marmoratus. The dataset includes measurements of nine distinct behaviors and analytical LC-MS results for psilocybin and psilocin concentrations within the fish. The data was authored by Dayna Forsyth and last updated on 2026-04-25.
University of Washington and Ohio University researchers conducted a collaborative project to study ice fabric and texture evolution. The work involved computer simulations and analysis of existing ice core data to model micro-structure interactions with deformation. No new observational data were collected under the University of Washington's contribution to NSF-OPP0136047.
Digital Twin Goat Health Dataset is a dataset published on Kaggle. The dataset likely contains health-related metrics for goats, potentially for modeling or simulation purposes. Metadata is minimal; actual content requires verification after download.
AI Safety Dataset is a synthetic collection intended for research in AI safety and large language models. It is published on the Kaggle platform. The specific size, structure, and creation details are not provided in the available metadata.
SynWTS is a high-fidelity synthetic dataset built as a Digital Twin of the Woven Traffic Safety dataset. It was developed by mlcglab for the 2026 AI City Challenge Track 2 to advance Sim2Real research in transportation safety understanding. The dataset provides a geometric match to real-world video for model training and evaluation.
FinOps Command Center Synthetic Data is a dataset published on Kaggle. The title suggests it contains generated data related to cloud financial operations and cost management. The dataset's specific content, scale, and origin require verification after download.
A synthetically generated image dataset for training Convolutional Neural Network models. The data is described as being in a B3D format, though specific details on volume, creation method, and authorship are not provided. The dataset is hosted on the Kaggle platform.
A synthetic dataset designed to mimic comprehensive medical check-up results, including medical history, physical exams, lab results, vaccinations, and occupational risk factors. The dataset was created by author 'udeezz' and was last updated on April 27, 2026. Its specific scale in terms of rows and columns is not detailed in the provided metadata.
VBVR-Dataset provides 1,000,000 video clips organized into 100 curated reasoning task generators, released by Video-Reason in February 2026. Each record pairs a video segment with start/end frame indices and a textual reasoning prompt to facilitate spatiotemporal understanding.
LIBERO GT Depth Aligned Hide Sites Fast contains ground-truth depth sidecar files for LIBERO robot demonstrations. Each episode is stored as a compressed .npz file, with depth aligned to original frames by re-rendering from simulator states. The dataset was authored by SeonghuJeon and last updated on 2026-04-22.
Vista4D is an evaluation dataset for video reshooting with 4D point clouds, associated with a CVPR 2026 Highlight paper. The dataset was created by a consortium including Eyeline Labs, Netflix, Columbia University, UCLA, Stony Brook University, and the University of Oxford. It was last updated on HuggingFace on April 24, 2026.
Airmesh is a dataset uploaded by ayyappanallamothu4 to Hugging Face. The dataset's specific content is not described, but its title suggests a focus on 3D mesh or simulation data. It was last updated on June 10, 2026.
122,199 high-quality, synthetic Question and Answer pairs are designed to instruction-tune Large Language Models into expert coding and architectural assistants for Unreal Engine 5.7. The dataset was created by TunstallTensor and last updated on Hugging Face in April 2026. It specifically addresses the frequent API deprecations and paradigm shifts introduced in the game engine.
Southern California coastal waters contain data from 426 fisheries-independent benthic trawls conducted by the Southern California Coastal Water Research Project (SCCWRP). The dataset originally included 288 invertebrate species, but analysis focused on 41 species present in at least 5% of 401 trawls, collected at depths from 2 to 215 meters during June-August. Site clusters were calculated using the Bray-Curtis dissimilarity coefficient.
G-LiHT airborne LiDAR point cloud data provides high-density individual return information for terrestrial ecosystems. The data includes 3D coordinates, ground classifications, height, and reflectance, with a density exceeding 10 points per square meter. NASA produced this data, last updated in March 2026.
MODIS/Terra+Aqua Burned Area data provides monthly, global gridded maps of fire-affected land at a 500-meter spatial resolution. The product is generated by NASA's LPCLOUD using a combined algorithm from MODIS Surface Reflectance and active fire observations. It identifies the specific burn date for each pixel within a calendar year.