Camera-based Text Recognition from Complex Backgrounds for the Blind or Visually
Camera-based Text Recognition from Complex Backgrounds for the Blind or Visually
批准号:
7977496
负责人:
YingLi Tian
金额:
$19.0万
依托单位:
依托单位国家:
美国
项目类别:
财政年份:
2010
资助国家:
美国
项目状态:
已结题
起止时间:
2010-09-01 至 2012-08-31
关键词:
AddressAfrican AmericanAgeAlgorithmsAmericasAreaBlindnessBooksCellular PhoneCitiesColorCommunitiesComplexComputer Systems DevelopmentComputer Vision SystemsComputer softwareComputersDatabasesDevelopmentDevicesEffectivenessEnvironmentEvaluationEventFacial Expression RecognitionFoundationsGoalsGrantHeadHispanicsImageImpairmentIndividualInstitutionInstructionInternationalLabelLettersLifeMailsMarketingMedicineMethodsMinorityNeighborhoodsNew YorkNew York CityNonprofit OrganizationsOutputPersonal Digital AssistantPricePrintingReadingRehabilitation therapyResearchResearch Project GrantsResolutionRunningScientistShapesSolutionsSpeechSurfaceSystemTechniquesTechnologyTextThickTimeUnited States National Institutes of HealthVertebral columnVisionVisual impairmentVisually Impaired PersonsWorkWritingbaseblindcollegecomputer generatedcomputer human interactioncostdesigndigitalexperiencelaptopnext generationoptical character recognitionprototyperehabilitation sciencerehabilitation serviceresearch and developmentsunglassestechnology developmentvisual information
中文摘要
今天,美国有1000多万盲人和视障人士。计算机视觉、数码相机和便携式计算机方面的最新技术发展,使得开发基于相机的辅助医疗系统成为可能
英文摘要
There are more than 10 million blind and visually impaired people living in America today. Recent technology developments in computer vision, digital cameras, and portable computers make it possible to assist these individuals by developing camera-based
products that combine computer vision technology with other existing products. Although a number of reading assistants have been designed specifically for people who are blind or visually impaired, reading text from complex backgrounds or non-flat
surfaces is very challenging and has not yet been successfully addressed. Many everyday tasks involve these challenging conditions, such as reading instructions on vending machines, titles of books aligned on a shelf, instructions on medicine bottles or labels on soup cans.
This proposal focuses on the development of new computer vision algorithms to recognize text from complex backgrounds: 1) from backgrounds with multiple different colors (e.g .. the titles of books lined up on a shelf) and 2) from non-flat surfaces (e.g .. labels on medicine bottles or soup cans). The newly developed computer vision techniques will be integrated with off-the-shelf optical character recognition (OCR) and speech-synthesis software products. Visual information will be captured via a head-mounted
camera (on sunglasses or hat) and analyzed by a portable computer (PDA or cell phone), while the speech display will be outputted via mini speakers, earphones, or Bluetooth device. A practical reading system prototype will be produced to read text from complex backgrounds and non-flat surfaces. The system will be cost-effective since it requires only a head mounted camera (<US$100 for 1M resolution), a wearable computer (<US$300), and two mini-speakers or earphones. The price of "ReadIRlS" [74] OCR software is under $150 and the "TextAloud" speech synthesis software is about $30 [75].
This project will be executed over two years at the City College of New York (CCNY) and Lighthouse International, New York. CCNY, located in the Harlem neighborhood of New York City, is designated as both a Minority Institution and a Hispanic-serving Institution (37% Hispanic and 27% African American). Lighthouse International is a leading non-profit organization dedicated to preserving vision and to providing critically needed vision and rehabilitation services to help people of all ages overcome the challenges of vision loss. During the two years, we will 1) develop new algorithms to recognize text from backgrounds with multiple different colors; 2) develop new algorithms to recognize text from non-flat surfaces; and 3) develop a cost-effective prototype reading system for blind users by integrating with off-the-shelf optical character recognition (OCR) and speech-synthesis software products. The effectiveness of the prototype and algorithms will be evaluated by people with normal vision and people with vision impairment. A database of text on complex backgrounds (multiple colors and non-flat surfaces) will be created for algorithm and system evaluation. The database will be made available to research communities in the areas of computer vision and vision rehabilitation science. In summary, this effort will provide a research-based foundation to inform the design of next generation reading assistants for blind persons, as well as produce a practical prototype to help the blind user read text from complex backgrounds in real-world environments.
期刊论文(12)
专著(0)
科研奖励(0)
会议论文
登录
查看更多内容
DOI:
10.1109/wocc.2011.5872294
发表时间:
2011
期刊:
WOCC ... : Wireless & Optical Communications Conference : the ... Annual Wireless & Optical Communications Conference. Annual Wireless & Optical Communications Conference
影响因子:
--
作者:
[Hasanuzzaman FM, Yang X, Tian Y]
通讯作者:
Tian Y
DOI:
10.1007/s13721-013-0026-x
发表时间:
2013-07-01
期刊:
NETWORK MODELING AND ANALYSIS IN HEALTH INFORMATICS AND BIOINFORMATICS
影响因子:
2.3
作者:
[Yi, Chucai, Flores, Roberto W, Chincha, Ricardo, Tian, Yingli]
通讯作者:
Tian, Yingli
DOI:
10.1007/s00138-012-0431-7
发表时间:
2013-04-01
期刊:
MACHINE VISION AND APPLICATIONS
影响因子:
3.3
作者:
[Tian, YingLi, Yang, Xiaodong, Yi, Chucai, Arditi, Aries]
通讯作者:
Arditi, Aries
Detecting Signage and Doors for Blind Navigation and Wayfinding.
检测标牌和门以进行盲导航和寻路。
DOI:
10.1007/s13721-013-0027-9
发表时间:
2013
期刊:
Network modeling and analysis in health informatics and bioinformatics
影响因子:
2.3
作者:
[Wang,Shuihua, Yang,Xiaodong, Tian,Yingli]
通讯作者:
Tian,Yingli
DOI:
10.1016/j.cviu.2012.11.002
发表时间:
2013-02-01
期刊:
COMPUTER VISION AND IMAGE UNDERSTANDING
影响因子:
4.5
作者:
[Yi, Chucai, Tian, Yingli]
通讯作者:
Tian, Yingli
共 11 条
海外基金