Loading...
Loading...
Available on 1 platform
Sign in to view source links and access this dataset
Between 10,000 and 100,000 image-text pairs comprise this 1% sample of the LaTeX_OCR collection for mathematical formula recognition. Released by unsloth in late 2024, the data provides paired visual representations and LaTeX source code formatted for Polars and Pandas.
This is a 1% subset of the linxy/LaTeX_OCR dataset; users should ensure the sample size is sufficient for their specific model convergence needs before downloading.