Techniques & Architectures

Optical Character Recognition (OCR)

Optical Character Recognition (OCR) is a technology that converts different types of documents, such as scanned paper documents, PDFs, or images, into editable and searchable data.

OCR works by analysing the patterns of light and dark that make up individual printed or handwritten characters. The system processes these patterns through complex algorithms that can recognise the distinctive shapes and features of letters, numbers, and symbols, converting them into machine-readable text.

The technology typically follows a multi-step process, beginning with image pre-processing to enhance quality, followed by character recognition, and finally post-processing to improve accuracy. Modern OCR systems often employ machine learning and neural networks to achieve higher accuracy rates and handle various fonts, styles, and languages.

Examples

  • Document digitisation in offices
  • Passport scanning at airports
  • Number plate recognition systems
  • Converting printed books to digital text
  • Processing bank cheques and receipts
Related Terms
×