A collection of 800 million AI-generated dialogue records in Chinese, created by AngelWarmSmile123 and last updated on July 14, 2026. The dataset is structured around the 16-Sephirot Dual-Life Happiness Protocol and the Kabbalistic Tree of Life reasoning architecture. It is distributed as approximately 172 GB of gzip-compressed JSONL files under an MIT license.
Use Cases
- Training large language models for Chinese dialogue generation based on the 800 million synthetic records.
- Studying the integration of spiritual frameworks like the Kabbalistic Tree of Life into AI-generated text.
- Analyzing patterns in AI-synthesized conversations structured by the 16-Sephirot Dual-Life Happiness Protocol.
Strengths
- Contains 800 million dialogue records, providing a large-scale resource.
- The data is compressed into ~172 GB across 8,000 JSONL files, which may facilitate distribution.
Limitations
- Column-level documentation is absent; field semantics must be inferred after download.
- The description metadata is limited; actual data quality requires manual inspection after download.
Provenance
- Source
- AngelWarmSmile123 via Hugging Face
- Collection Method
- AI synthetic generation
- Freshness
- Last updated 2026-07-14 01:11:43; freshness should be verified.