课题基金 / 基金详情

面向手写藏文古籍的跨字体文字识别方法研究

批准号:
62066042
项目类别:
地区科学基金项目
资助金额:
36.0 万元
负责人:
格桑多吉
依托单位:
学科分类:
模式识别与数据挖掘
结题年份:
2024
批准年份:
2020
项目状态:
已结题
项目参与者:
格桑多吉

项目摘要

结项摘要

格桑多吉的其他基金

相似基金

相关文献

中文摘要
藏文古籍文献数字化是解决古籍“藏”与“用”的矛盾的有效手段,而藏文古籍文本识别技术是数字化的关键技术之一。尽管雕版印刷藏文古籍文本识别已经得到较好解决,但是手写藏文古籍因书写风格导致的不规则和多样化字体,因书写条件限制导致的复杂版面样式,以及因保存条件影响导致的图像退化等问题,准确识别其中的文字依然极具挑战。为此,本项目提出藏文古籍版面内容的多粒度表征模型,并据此基于多任务深度学习技术实现鲁棒的手写藏文古籍版面分析;提出基于图像生成和对抗学习的手写藏文古籍文本行图像修复方法,克服古籍文档图像退化、文字破损的问题;提出藏文字字形与字体特征的解耦模型和字丁级语义特征提取模型,据此实现跨字体的手写藏文字识别方法和针对开放和长尾类别的藏文字识别方法,提升处理相似字、生僻字和错别字的能力。本项目研究成果有望突破手写藏文古籍文本识别的技术瓶颈,为作为中华文化重要部分的藏文化的传承与发展提供技术支撑。
英文摘要
The digitization of Tibetan ancient books is an effective means to solve the contradiction between “preservation” and “utilization” of ancient books, and the text recognition technology of Tibetan ancient books is one of the key technologies of digitization. Although the text recognition of Tibetan ancient books in engraving printing has been well resolved, it is still very challenging to accurately recognize the text of handwritten Tibetan ancient books because of the irregular and diversified fonts from different writing styles, the complex layout styles caused by the limitation of writing conditions, the degrade images caused by the preservation conditions. To this end, the project proposes a multi-granular representation model for Tibetan ancient books, based on which we propose a robust layout analysis method for handwritten Tibetan ancient books using multi-task deep learning technology; to overcome the challenges from degraded image and incomplete text in ancient documents, we propose a restoration method for handwritten Tibetan ancient book images based on image generation and adversarial learning; we finally propose a decoupling model to extract font and glyph features separately, together with a character-level semantic feature extraction model for cross-font, open and long-tailed category based handwritten Tibetan character recognition for similar and unfamiliar characters and typos. This project is expected to provide solutions and technical support for handwritten Tibetan ancient book text recognition, which can further help the inheritance and development of Tibetan culture and the Chinese culture.
藏文识别技术在过去二十年间不断发展,印刷体藏文识别和联机手写藏文识别已经走向实用,但 是针对藏文古籍文献中的文字识别主要集中在基于常用字符集的木刻典籍文本识别,而对手写藏文古籍文本识别的研究涉猎甚少,距离形成系统理论和方法、开发有效技术和工具还相差甚远。针对这一现状,本项目首先构建了可用于手写藏文古籍文献文本检测与识别相关科学研究的数据集,为该领域的研究提供了基础支持。在此基础上,项目针对手写藏文古籍文献中常见的问题,如版面结构复杂、字体类别多样、书写风格不一、字符类别呈长尾分布以及编码随垂直位置变化等,提出了一系列创新性方法。这些方法包括基于多粒度表征的手写藏文古籍版面分析与检测、低质手写藏文古籍文本行修复以及面向多字体和开放长尾的手写藏文古籍文本识别等。经实验,本项目提出的方法均取得较好效果,针对性解决了手写藏文古籍文本识别技术面临的多项技术困难,为领域研究提供了一定的基础和借鉴。.截至2024年12月31日,本项目已发表学术论文20篇(其中,JRC一区论文1篇,二区论文1篇,CCF-A类学术会议论文1篇,CCF-C类以上学术会议论文5篇,其他12篇),申请发明专利4项,培养4名博士和8名硕士。
基于群体智能涌现的藏文网络舆情分析及突发事件预警机制研究
  • 批准号:
    61165013
  • 项目类别:
    地区科学基金项目
  • 资助金额:
    50.0万元
  • 批准年份:
    2011
  • 负责人:
    格桑多吉
  • 依托单位:
国内基金
海外基金