TH105604B - A process for semi-automatic character image recognition. - Google Patents
A process for semi-automatic character image recognition.Info
- Publication number
- TH105604B TH105604B TH901003430A TH0901003430A TH105604B TH 105604 B TH105604 B TH 105604B TH 901003430 A TH901003430 A TH 901003430A TH 0901003430 A TH0901003430 A TH 0901003430A TH 105604 B TH105604 B TH 105604B
- Authority
- TH
- Thailand
- Prior art keywords
- character
- image
- database
- character image
- string
- Prior art date
Links
Abstract
------21/02/2563------(OCR) หน้าที่ 1 ของจำนวน 1 หน้า บทสรุปการประดิษฐ์ การประดิษฐ์นี้นำเสนอระเบียบวิธีการรู้จำภาพตัวอักษรกึ่งอัตโนมัติ ซึ่งเป็นการรู้จำตัวอักษรโดยที่ผู้ใช้ต้องป้อนข้อมูลตัวอักษรก่อนการใช้งาน จากนั้นระบบจะทำการรู้จำภาพตัวอักษร โดยใช้การคำนวณค่าความใกล้เคียงของภาพตัวอักษรที่อินพุตกับตัวอักษรที่ผู้ใช้กำหนดให้ในตอนแรก และถ้าตัวอักษรมีการเชื่อมติดกัน 2 ตัวขึ้นไปหรือ เป็นตัวอักษรที่ไม่สมบูรณ์เนื่องจากการขาดของเส้นตัวอักษร ก็จะใช้ลักษณะเฉพาะของตัวอักษร เช่น ความกว้างของตัวอักษรมาช่วยกำหนดขนาดของตัวอักษรและอาศัยค่าความน่าจะเป็นของคู่ตัวอักษรช่วยคาดเดาตัวอักษรที่ติดกันว่าควรเป็นตัวอักษรใด เพื่อช่วยเพิ่มความถูกต้องในการรู้จำมากขึ้น นอกจากนี้ระเบียบวิธีที่นำเสนอยังสามารถนำไปประยุกต์ใช้กับภาษาใดก็ได้ ------------ DC60 การประดิษฐ์นี้นำเสนอระเบียบวิธีการรู้จำภาพตัวอักษรกึ่งอัตโนมัติ ซึ่งเป็นการรู้จำ ตัวอักษร โดยที่ผู้ใช้ต้องป้อนข้อมูลตัวอักษรก่อนการใช้งาน จากนั้นระบบจะทำการรู้จำภาพ ตัวอักษร โดยใช้การคำนวณค่าความใกล้เคียงของภาพตัวอักษรที่อินพุตกับตัวอักษรที่ผู้ใช้ กำหนดให้ในตอนแรก และถ้าตัวอักษรมีการเชื่อมติดกัน 2 ตัวขึ้นไปหรือ เป็นตัวอักษรที่ไม่สมบูรณ์ เนื่องจากการขาดของเส้นตัวอักษร ก็จะใช้ลักษณะเฉพาะของตัวอักษร เช่น ความกว้างของ ตัวอักษรมาช่วยกำหนดขนาดของตัวอักษรและอาศัยค่าความน่าจะเป็นของคู่ตัวอักษรช่วยคาดเดา ตัวอักษรที่ติดกันว่าควรเป็นตัวอักษรใด เพื่อช่วยเพิ่มความถูกต้องในการรู้จำมากขึ้น นอกจากนี้ ระเบียบวิธีที่นำเสนอยังสามารถนำไปประยุกต์ใช้กับภาษาใดก็ได้------February 21, 2020------(OCR) Page 1 of 1 Abstract of the Invention This invention presents a semi-automatic character image recognition method. This method recognizes characters where the user must input character data before use. The system then recognizes the character image by calculating the proximity of the input character image to the character initially specified by the user. If two or more characters are connected or are incomplete due to broken character strokes, it uses character characteristics such as character width to determine the character size and relies on the probability of character pairs to predict which adjacent character should be identified, thereby improving recognition accuracy. Furthermore, the proposed method can be applied to any language. ------------ DC60 This invention presents a semi-automatic character image recognition method. This method recognizes characters where the user must input character data before use. The system then recognizes the character image by calculating the proximity of the input character image to the character initially specified by the user. If two or more characters are connected or are incomplete due to broken character strokes, it uses character characteristics such as character width to determine the character size and relies on the probability of character pairs to predict which adjacent character should be identified, thereby improving recognition accuracy. In addition, the proposed method can be applied to any language. This method uses unique characteristics of characters, such as character width, to determine character size and relies on the probability of character pairs to predict which adjacent characters should be, thereby improving recognition accuracy. Furthermore, the proposed method can be applied to any language.
Claims (1)
Publications (3)
| Publication Number | Publication Date |
|---|---|
| TH103248A TH103248A (en) | 2010-08-11 |
| TH105604A TH105604A (en) | 2010-12-30 |
| TH105604B true TH105604B (en) | 2024-12-18 |
Family
ID=
Similar Documents
| Publication | Publication Date | Title |
|---|---|---|
| CN110114776B (en) | System and method for character recognition using fully convolutional neural networks | |
| Firmani et al. | Towards Knowledge Discovery from the Vatican Secret Archives. In Codice Ratio-Episode 1: Machine Transcription of the Manuscripts. | |
| Bai et al. | Keyword spotting in document images through word shape coding | |
| US20090304282A1 (en) | Recognition of tabular structures | |
| JP2014106961A (en) | Method executed by computer for automatically recognizing text in arabic, and computer program | |
| CN111630521A (en) | Image processing method and image processing system | |
| US20050182760A1 (en) | Apparatus and method for searching for digital ink query | |
| Reffle et al. | Unsupervised profiling of OCRed historical documents | |
| TWI567569B (en) | Natural language processing systems, natural language processing methods, and natural language processing programs | |
| CN109815452A (en) | Text comparative approach, device, storage medium and electronic equipment | |
| Okamoto et al. | Performance evaluation of a robust method for mathematical expression recognition | |
| US9934429B2 (en) | Storage medium, recognition method, and recognition apparatus | |
| Ramakrishnan et al. | Global and local features for recognition of online handwritten numerals and Tamil characters | |
| Springmann et al. | Automatic quality evaluation and (semi-) automatic improvement of OCR models for historical printings | |
| US9384304B2 (en) | Document search apparatus, document search method, and program product | |
| JP2012043385A (en) | Character recognition device and character recognition method | |
| Li et al. | A fast keyword-spotting technique | |
| US20150063698A1 (en) | Assisted OCR | |
| TH103248A (en) | Process for semi-automatic character image recognition | |
| Tariq et al. | Softconverter: A novel approach to construct OCR for printed Urdu isolated characters | |
| Madhavaraj et al. | Improved recognition of aged Kannada documents by effective segmentation of merged characters | |
| KR20160053544A (en) | Method for extracting candidate character | |
| KR20160053587A (en) | Method for minimizing database size of n-gram language model | |
| JP5853488B2 (en) | Information processing apparatus and program | |
| US11270153B2 (en) | System and method for whole word conversion of text in image |