Loading...
Loading...
Crop yield, soil data, pest surveillance, livestock, food composition, precision farming
19,247 datasets
The City of New York's Department of Transportation provides a weekly list of streets scheduled for milling or resurfacing work. The schedule is updated weekly and subject to change due to weather or equipment issues. The dataset includes tags for work types like Milling, Asphalt, Paving, and Resurfacing.
Greenstreets is a New York City program converting unused road areas into green spaces for beautification, improved air quality, and stormwater management. The dataset includes geospatial data using the NAD_1983_StatePlane_New_York_Long_Island_FIPS_3104_Feet projection, with lengths in feet and areas in square feet. It is provided by the City of New York and was last updated in March 2026.
Geospatial data from the City of New York details beach zones and sections for inspection and maintenance. The dataset uses the NAD_1983_StatePlane_New_York_Long_Island_FIPS_3104_Feet projection with measurements in feet. It was last updated in March 2026.
Active retail tobacco and vapor product vendors operating in New York State, sourced from a DOH database. The data includes vendor names, types, and location details such as street address, city, county, and zip code. The dataset was last updated on 2026-01-30 16:35:35.
Historical weather data from 2004–2023 for 77 sites across 13 Middle Eastern countries was used to simulate grain aeration effectiveness and model rice weevil (S. oryzae) populations. The dataset includes spatially interpolated calculations of accumulated daily hours below specific temperature thresholds (15, 18, and 21°C) for key months. It was generated by the Department of Agriculture using inverse distance weighted interpolation in QGIS.
Nightly updated records of food establishment licenses for the City of Hartford, sourced from the Accela Permit system. The dataset is published by the City of Hartford and was last updated on March 24, 2026. Available formats include CSV and GeoJSON, indicating it likely contains both tabular and geographic information.
Wigan Metropolitan Borough Council provides streetlighting information sourced from the Mayrise system. The dataset includes multiple file formats such as GeoJSON, CSV, and KML, suggesting geospatial asset data. It was last updated on 2026-03-17.
Approximately 6,682 tons of debris are collected annually by Norfolk's street sweeping program. The City of Norfolk maintains this schedule dataset, updated at the start of each calendar year, showing planned sweeping dates for addresses. It reflects a temporary shift to a quarterly rotation for most areas due to equipment and staffing challenges.
449 interview records and 265 three-day dietary records form this collection focused on basic education students. The data appears to originate from Magway, a region in Myanmar. The specific collection methodology and author are not detailed.
Kaggle hosts a preprocessed dataset related to food delivery. The dataset's specific content, size, and origin are not detailed in the provided metadata. Its columns and rows are unknown, requiring verification after download.
Kaggle hosts a dataset titled 'Food_Delivery_Times'. The dataset likely contains records related to food delivery logistics, such as order times or delivery durations. Its specific contents, scale, and origin require verification after download.
2022 field campaign data includes 12 measurement points per tree for 25 hydraulic and wood anatomy traits. The dataset comprises measurements from 38 individual trees across five dipterocarp species, with tree heights ranging from 7.7 to 71 meters.
British Geological Survey produced predictive raster maps for 56 chemical elements, pH, and organic matter content across western Kenya. The maps are based on 452 soil samples collected between 2015 and 2020 and were generated using random forest machine learning algorithms at a 500m spatial resolution.
A dataset used to identify the best eucalyptus seedlots for soil conservation in seasonally dry hill country. It is based on a twelfth-year assessment published in the New Zealand Journal of Forestry Science in 1991 and was later used as a machine learning benchmark in a 1996 research report from the University of Waikato. The specific features and size of the dataset are not detailed in the provided metadata.
OpenStreetMap data provides a global, community-edited map of the world. This dataset is hosted on Kaggle, though its specific contents and scope are not detailed in the provided metadata. The exact source, size, and attributes of the data require verification after download.
A study by Priyanka Rai on the effect of urban air pollution on the epidermal characteristics of roadside tree species Pongamia pinnata. Light microscopic analysis revealed marked alterations in traits, including increased stomata and epidermal cell counts per unit area in leaves from polluted sites compared to a control. The dataset likely contains quantitative measurements of these cellular features to serve as an indicator for environmental stress.
biotools is a software package by A. R. Da Silva providing statistical tools for agricultural science. It includes algorithms for cluster, discriminant, and path analysis, as well as tools for sample size calculation and spatial prediction. The package also implements tests for seed heterogeneity, genetic covariance, and Mantel's permutation test.
Over 40,600 tokens of annotated discourse relations are included from Version 2, with an additional 13,000 tokens annotated in Version 3. The Penn Discourse Treebank annotates discourse relations in the Wall Street Journal section of Treebank-2. Rashmi Prasad led the project, which includes tools for annotation, adjudication, and conversion.
An R package providing infrastructure for representing, summarizing, and visualizing tree-structured regression and classification models. It offers a unified framework for reading and coercing models from sources like 'rpart' and 'RWeka', and includes reimplementations of conditional inference trees and model-based recursive partitioning. The package was described in a 2015 paper by Hothorn and Zeileis.
crnn-synth-plates-pub is a collection of 20,000 synthetic images of license plate crops. The data was generated synthetically, likely for training optical character recognition models. The author, organization, and specific creation date are unknown.