Dataset Viewer
Auto-converted to Parquet Duplicate
Image Patch
string
image
image
3200797029-1.tif
3200797029-2.tif
3200797029-3.tif
3200797029-4.tif
3200797029-5.tif
3200797029-6.tif
3200797029-7.tif
3200797029-8.tif
3200797029-9.tif
3200797029-10.tif
3200797029-11.tif
3200797029-12.tif
3200797029-13.tif
3200797029-14.tif
3200797029-15.tif
3200797029-16.tif
3200797029-17.tif
3200797029-18.tif
3200797037-1.tif
3200797037-2.tif
3200797037-3.tif
3200797037-4.tif
3200797037-5.tif
3200797037-6.tif
3200797037-7.tif
3200797037-8.tif
3200797037-9.tif
3200797037-10.tif
3200797037-11.tif
3200797037-12.tif
3200797037-13.tif
3200797037-14.tif
3200797037-15.tif
3200797037-16.tif
3200797037-17.tif
3200797037-18.tif
3200797037-19.tif
3200797037-20.tif
3200797037-21.tif
3200801612-1.tif
3200801612-2.tif
3200801612-3.tif
3200801612-4.tif
3200801612-5.tif
3200801612-6.tif
3200801612-7.tif
3200801612-8.tif
3200801612-9.tif
3200801612-10.tif
3200801612-11.tif
3200801612-12.tif
3200801612-13.tif
3200801612-14.tif
3200801612-15.tif
3200801612-16.tif
3200801612-17.tif
3200801612-18.tif
3200801612-19.tif
3200801612-20.tif
3200801612-21.tif
3200801612-22.tif
3200801612-23.tif
3200801612-24.tif
3200801612-25.tif
3200801612-26.tif
3200801612-27.tif
3200801612-28.tif
3200801612-29.tif
3200801612-30.tif
3200801612-31.tif
3200801612-32.tif
3200801612-33.tif
3200801612-34.tif
3200801612-35.tif
3200801612-36.tif
3200801612-37.tif
3200801612-38.tif
3200801612-39.tif
3200801612-40.tif
3200801612-41.tif
3200801612-42.tif
3200801612-43.tif
3200801612-44.tif
3200801612-45.tif
3200801612-46.tif
3200801612-47.tif
3200801612-48.tif
3200801613-1.tif
3200801613-2.tif
3200801613-3.tif
3200801613-4.tif
3200801613-5.tif
3200801613-6.tif
3200801613-7.tif
3200801613-8.tif
3200801613-9.tif
3200801613-10.tif
3200801613-11.tif
3200801613-12.tif
3200801613-13.tif
End of preview. Expand in Data Studio

BLN600 Image Patches

This dataset provides BLN600's image patches for fine-tuning vision-language models on post-OCR correction, introduced in "Image-Informed Post-OCR Correction with Vision-Language Models" (EMNLP 2026 Findings). Each patch corresponds to a text sequence in BLN600 and is cropped from the Gale British Library Newspapers collection, using word-level bounding boxes from the collection's ALTO XML OCR layout data. It is intended to be used alongside the code and CSVs in Shef-AIRE/vlms_post-ocr_correction, which provide the corresponding OCR text, ground truth, metadata, and CER/WER for each patch.

Data Fields

Field Type Description
Image Patch string Filename, joins to the Image Patch column in the code repo's CSVs
image image The cropped newspaper-page region (TIFF)

License

Released under CC BY-NC-ND 4.0 (non-commercial, no derivatives). Permission was granted by Gale on behalf of the company and the British Library partners for non-commercial release, publicly accessible with no additional access stipulations.

Citation

Citation to follow on publication.

@inproceedings{thomas-etal-2026-image,
    title = "Image-Informed Post-OCR Correction with Vision-Language Models",
    author = "Thomas, Alan  and
      Liu, Xianyuan  and
      Lu, Haiping  and
      Gaizauskas, Robert",
    booktitle = "Findings of the Association for Computational Linguistics: EMNLP 2026",
    year = "2026",
    publisher = "Association for Computational Linguistics",
}
Downloads last month
51