Skip to content
JournalsWorldThe Global Research Discovery Platform
Featured Dataset

The Rodrigo corpus

The Rodrigo corpus was obtained from the digitisation of the book “Historia de España del arçobispo Don Rodrigo”, written in ancient Spanish in 1545. It is a single writer book where most pages consist of a single block of well-separated lines of calligraphical

👤
CreatorEmilio Granell
📅
Published2018-11-16
🔗
DOI10.5281/zenodo.1490009
📊
Downloads939
⚖️
Licensecc-by-4.0
File Size382.3 MB
Data TypeDataset
Published2018
Licensecc-by-4.0
Total Views2,294
Total Downloads939

The Rodrigo corpus was obtained from the digitisation of the book “Historia de España del arçobispo Don Rodrigo”, written in ancient Spanish in 1545. It is a single writer book where most pages consist of a single block of well-separated lines of calligraphical text.

This dataset is free available for research purposes. It contains 15,010 images of text lines with their paleographic transcription. It is divided into three partitions: 9000 text lines for training, 1000 for validation and 5010 for testing.

📤 Share this page

Found this useful? Share it with your network.

✓ Link copied! Paste it on ResearchGate / Academia.edu
📦
The Rodrigo corpus (Full Dataset)382.3 MB
⬇
📄
ReadmeVia DOI record
↗

Files are hosted on the source repository. Click download to access the full dataset.

Emilio Granell (2018). The Rodrigo corpus. https://doi.org/10.5281/zenodo.1490009