InscriptionOCR:古婆罗米铭文数据集与方法
Get the story
2026年10月8日,arXiv 计算机视觉(一手来源)报道了 InscriptionOCR 项目,提出一个端到端 AI 框架,用于处理石刻低质量图像,完成图像修复、Brahmi 字符识别、映射为 Roman 字符,并把 Prakrit 文本翻译成英文。项目同时发布 InscriptionOCR Dataset,含 200,000+ 字符图像、约 600 个类别,报道称为目前最大可公开使用的 Brahmi OCR dataset;另发布 2,000+ 句对的 Prakrit–English 平行语料,用于神经机器翻译(NMT)。目前报道仅涉及数据集与方法发布,未见后续验证或应用进展。
Generated from reports · updated 3 hr ago
Timeline
Follow the coverage from different angles.
- arXiv · Computer VisionInscriptionOCR:用于理解铭刻的 dataset 与方法
提出一个端到端 AI 框架,处理石刻低质量图像,完成图像修复、Brahmi 字符识别、映射为 Roman 字符并把 Prakrit 文本翻译成英文。发布 InscriptionOCR Dataset,含 200,000+ 字符图像、约 600 个类别,为目前最大可公开使用的 Brahmi OCR dataset;另发布 2,000+ 句对的 Prakrit–English 平行语料用于 NMT。
Heat trend
Current heat 9·Comparable peak 10(Oct 8)·Comparable change over 24 hours –
The trend compares only the same participants observed continuously; its range may be smaller than the current heat count. Move or click on the chart to inspect hourly heat; use the left and right arrow keys to switch.