Loading...
Loading...
Cell biology, microbiology, ecology, biodiversity, species data, evolutionary biology
27,496 datasets
Two audio datasets compiled for training neural networks to automatically identify insect species from their sounds. InsectSet47 contains 1,006 recordings from 47 species totaling 22 hours, while InsectSet66 expands to 1,554 recordings from 66 species totaling over 24 hours. The datasets were created by Marius Faiß at Leiden University, sourced from BioAcoustica, xeno-canto, iNaturalist, and private collections.
InsectSet32 contains 335 audio files totaling 57 minutes from 32 sound-producing insect species. The dataset is split between nine Orthoptera species (147 recordings) and 23 Cicadidae species (188 recordings), curated from the unpublished work of Baudewijn Odé and the Global Cicada Sound Collection on Bioacoustica. It was compiled to train neural networks for automatic insect identification, comparing adaptive waveform-based and mel-spectrogram audio frontends.
Southeast Australian marine ecosystem data from 83 CTD deployments during the RV Investigator voyage IN2024_V03 in May 2024. The data were collected using a Sea-Bird SBE911 CTD unit and processed by CSIRO's National Collections and Marine Infrastructure Information and Data Centre. Measurements include conductivity, temperature, depth, dissolved oxygen, chlorophyll-a, and other auxiliary parameters.
Synonymic checklists of the world's vascular plants, including ferns and lycophytes, compiled over 20 years. The dataset contains roughly 1 million plant names cross-checked against local floras and treatments. It was created by Michael Hassler at the Karlsruhe Institute of Technology.
Fixed cameras at the Alice Mulga SuperSite in Australia's Northern Territory provide a half-hourly time series of images during daylight hours. The collection includes processed data products like the Green Chromatic Coordinate (Gcc) for specific vegetation regions-of-interest. This long-term record, established in 2010, supports analysis of vegetation structure and condition.
Sediment grain size data and summary statistics for the greater Darwin Harbour region, derived from seabed samples. A total of 499 samples from 489 stations were collected during multiple surveys between 2011 and 2017, funded by the INPEX-led Ichthys LNG Project and co-investors. The data are published in the Marine Sediments Database (MARS) with the permission of Geoscience Australia.
28 geographically paired sites in Ontario, Canada were surveyed in 2016-2017 to compare insect communities on native and introduced subspecies of Phragmites australis. The dataset includes site characteristics, stem attack rates for fourteen insect taxa, and alpha- and beta-diversity indices. Data were produced by R. B. deJonge of the University of Toronto for a published ecological study.
Observations from May 2010 to July 2012 at the IAP tower in Beijing include radiation fluxes, turbulent fluxes (CO2, latent and sensible heat) for 2016, and meteorological data. This dataset supports the manuscript 'Simulating heat and CO2 fluxes in Beijing using SUEWS V2020b' and contains processed observations, model runs, and Python code for quality control and gap-filling. The data is intended for validating and applying the Surface Urban Energy and Water balance Scheme (SUEWS) model to urban environments.
Metagenome-assembled genomes (MAGs) from the WildR murine gut microbiome community, which was originally sourced from wild mice and reconstituted in germ-free laboratory mice in 2017. The dataset includes MAG sequences in FASTA and Prokka-annotated GenBank formats, along with assembly metrics and taxonomic annotations. It was generated by researchers at the University of Washington using long and short-read metagenomic sequencing on the F7 generation of the community.
11.7 GB of raw multibeam echosounder data collected aboard the RV Investigator during the 55-day COOKIES voyage to the Antarctic region in early 2026. The EM2040MK2 system recorded bathymetry, backscatter, and watercolumn data at 400 kHz nominal frequency, with motion, position, and tidal corrections applied. Data are managed by the Australian Ocean Data Network and stored in .kmall and .kmwcd formats at CSIRO.
Darwin Harbour seabed mapping and habitat classification surveys completed in 2011 and 2013. The dataset includes high-resolution multibeam sonar bathymetry, acoustic backscatter, video observations, and physical sediment samples. It was produced by Geoscience Australia, the Australian Institute of Marine Science, the Department of Land Resource Management, and the Darwin Port Corporation.
Mesozooplankton community composition and structure data collected from the D’Entrecasteaux Channel, Huon Estuary, and North West Bay in Tasmania on 13 October 2005. The data was aggregated by the Australian Ocean Data Network and shows seasonal abundance patterns and spatial differences in species representation. Copepods were the largest contributors to total abundance across all seasons and stations.
Tasmanian coastal waters, including the Huon Estuary, D'Entrecasteaux Channel, and North West Bay, were sampled for mesozooplankton on 05/04/2005. The data, provided by the Australian Ocean Data Network, captures community composition and structure typical of inshore temperate habitats. Copepods were the largest contributors to total abundance, with spatial variations in species representation between marine and estuarine zones.
Digital Earth Australia Intertidal provides annual continental-scale elevation and exposure products for Australia's intertidal zone. The data is mapped at a 10-meter resolution from Digital Earth Australia's archive of open-source Landsat and Sentinel-2 satellite data. It is hosted by the Australian Ocean Data Network and was last updated in June 2026.
Annual monitoring of stream benthic invertebrates in PEI National Park assesses aquatic health through biodiversity and pollution tolerance. Samples are sorted and classified to the lowest possible taxonomic level, with community health evaluated using Simpson’s reciprocal index and the Hilsenhoff Biotic Index. The dataset is produced by Parks Canada using Environment Canada's CABIN stream monitoring network methods.
Australian Ocean Data Network provides dissolved oxygen and temperature data collected from Macquarie Harbour, Tasmania, between 2017 and 2018. The data was gathered using HOBO Dissolved Oxygen loggers deployed at two locations under FRDC project 2016-067. It documents environmental conditions relevant to finfish aquaculture and benthic ecosystem health.
432 species of moths, mostly macro-moths, have their national population trends modeled from 1968 to 2016. Trends are calculated using a Generalised Abundance Index (GAI) model and presented as year coefficients, Annual Growth Rates (AGR), and total percentage changes with confidence intervals. A related subset of trends was produced specifically for the Atlas of Britain & Ireland’s Larger Moths, using data from Great Britain between 1970 and 2016.
The Swan-Canning Estuary (WA) Seagrass 2011 dataset contains polygon data showing areas of seagrass habitat derived from aerial imagery. It was collected by the WA Department of Water and Geoscience Australia as part of a 2013 Australia-wide risk assessment of seagrass. The data serves as baseline information on seagrass composition and distribution in key estuaries of southern and south-western Western Australia.
Peter T. Harris et al. published a study in 2013 analyzing seabed data from the Great Barrier Reef. The dataset indicates that only about 39% (16,110 km²) of available seabed on submerged banks is capped by near-sea-surface coral reefs, while the other 61% (25,600 km²) is submerged at a mean depth of around 27 m. Predictive habitat modelling suggests more than half of this submerged area (around 14,000 km²) is suitable for coral communities.
Global Affairs Canada developed a Privacy Impact Assessment for its Blackberry cellular phone and email services. The assessment focuses on personal information collected from departmental employees to provide them access to the approved service. It was created to meet the Department's commitment to personal information protection and Management of Information Technology Security requirements.