Sign in to view source links and access this dataset
Description
A collection of spoken common-sense factual prompts, each provided in three paired versions: an incomplete prompt, a correct factual completion, and an incorrect counterfactual completion. The dataset was created by author slprl and was last updated on 2026-06-18. It does not define standard train/validation/test splits, with all examples provided in a single neutral split.
Use Cases
Training speech-to-text models on factual language patterns based on the spoken prompt and fact pairs.
Developing AI systems to detect factual inconsistencies in audio based on the prompt and counterfactual pairs.
Benchmarking audio-language models on common-sense reasoning tasks using the paired prompt-fact structure.
Fine-tuning models for audio-based question answering using the incomplete prompt and correct fact completions.
Strengths
Provides three distinct, paired audio versions (prompt, fact, counterfactual) for each common-sense statement.
Dataset structure is explicitly defined for common-sense factual prompts, enabling specific model training tasks.
Limitations
Description metadata is limited; actual data quality, audio characteristics, and recording conditions require manual inspection after download.
Row count, file formats, and column-level documentation are absent, which may limit suitability assessment.
Last updated 2026-06-18 17:35:05; freshness should be verified.
Provenance
Source
huggingface
Collection Method
Likely created for AI/ML research purposes, method unspecified.
Freshness
Last updated 2026-06-18 17:35:05.
License is unknown, which may restrict commercial or research use.