NEURAL NETWORK FOR UNICODE OPTICAL CHARACTER RECOGNITION

📄 Item Type: Project Material| 📋 45 pages| 📚 1–5 chapters| Amount: ₦5,000

NEURAL NETWORK FOR UNICODE OPTICAL CHARACTER RECOGNITION

📄 Project Material 📋 45 pages 📚 Chapters 1–5 💾 MS-Word & PDF

Chapters 1–5  |  ₦5,000

Download Complete Project Now

NEURAL NETWORK FOR UNICODE OPTICAL CHARACTER RECOGNITION  (CASE STUDY OF DHL, ENUGU)

ABSTRACT

Optical character Recognition (OCR) refers to the process of converting printed tamil text documents into software translated Unicode tamil text. The printed documents available in the form of books, projects, magazines etc are scanned using standard scanners which produce an image of the scanned documents. As part of the preprocessing phase the image like is checked for skewing. If the image is skewed, it is corrected by a simple rotation technique in the appropriate direction. Then the image is passed through a noise elimination phase and is binarized. The preprocessed image is segmented using an algorithm which decomposes the scanned text into paragraphs using special space detection technique and then the paragraphs into lines using vertical histograms, and lines into words using horizontal histograms, and words into character image glyphs using horizontal histograms. Each image glyph is comprised of 32 x 32 pixels, thus a data base of character image glyphs is created out of the segmentation phase. Then all the image glyphs are considered for recognition using Unicode mapping. Each image glyph is passed through various routines which extract  the features of the glyph. The various features that are considered for classification are the character height, character width, then number of horizontal lines (Long and short), the number of vertical lines (long and short), the horizontally oriented curves, the vertically oriented curves, the number of circles, number of slope lines, image centroid and special dots. The glyphs are now set ready for classification based on these features. The extracted features are passed to a support vector machine (SVM) where the characters are classified by supervised learning algorithm. These classes are mapped into Unicode for recognition. Then the text is reconstructed using Unicode fonts. 

TABLE OF CONTENTS

Title page         -       -       -       -       -       -       -      -       ii

Certification    -       -       -       -       -       -       -      -       iii

Approval page         -       -       -       -       -       -      -       iv

Dedication       -       -       -       -       -       -       -      -       v

Acknowledgement   -       -       -       -       -       -       -      vi

Abstract -       -       -       -       -       -       -       -      -       vii

Table of contents     -       -       -       -       -       -      -       ix

CHAPTER ONE

 1.0 INTRODUCTION    -      -      -      -      -      -       1

1.1      Statement of the problem       -       -       -       -       5

1.2      Purpose of the study       -       -       -       -       -      6

1.3      Aims and objectives        -       -       -       -       -      6

1.4      Scope of study         -       -       -       -       -       -      8

1.5      Limitations of the study -       -       -       -       -       8

1.6      Definition of terms.-       -       -       -       -       -       9

CHAPTER TWO

 2.0 LITERATURE REVIEW -      -      -      -      -      11

CHAPTER THREE

3.0      Methods for fact finding and details discussions on the subject matter.        -       -       -       -       -      -       15

3.1      Methodologies for fact finding         -       -       -      15

3.2      Discussions     -       -       -       -       -       -       -      16

CHAPTER FOUR

4.0      Futures, Implications and challenges of the subject matter for the society             -       -       -       -      20

4.1      Futures   -       -       -       -       -       -       -       -      20

4.2      Implications    -       -       -       -       -       -       -      21

4.3      Challenges      -       -       -       -       -       -       -      22

CHAPTER FIVE

5.0      SUMMARY, RECOMMENDATION AND CONCLUSION 24

5.1      Summary        -       -       -       -       -       -       -      24

5.2      Recommendation    -       -       -       -       -       -      25

5.3      Conclusion      -       -       -       -       -       -       -      28

References       -       -       -       -       -       -      -       30

CHAPTER ONE

1.0 INTRODUCTION

Character is the basic building block of any language that is used to build different structures of a language. Characters are the alphabets and the structures are the words, strings and sentences.

Optical character Recognition (OCR) is the process of converting an image of text, such as a scanned project character, document or electronic fax file, into computer-editable text. The text in an image is not editable. The letters/characters are made of tiny dots (pixels) that together form a picture of text. During OCR, the software analyzes an image and converts the pictures of the characters to editable text based on the patterns of the pixels in the image. After OCR, you can expert the converted text and use it with a variety of word-processing, page layout and spreadsheet applications. OCR also enables screen readers and refreshable bralle displays to read the text contained in images.

Optical character Recognition (OCR) deals with machine recognition of characters present in an input image obtained using scanning operation. It refers to the process by which scanned images are electronically processed and converted to an editable text. The need for OCR arises in the context of digitizing tamil documents from the ancient and old era to the latest, which helps in sharing the data through the internet.

A properly printed document is chosen for scanning. It is placed over the scanner, A scanner software is invoked which scans the document. The document is sent to a program that saves it in preferably TIF, JPG or GIF format, so that the image of the document can be obtained when needed. This is the first step in OCR (Vijaya Kumar, 2001), the size of the input image is as specific by the user and can be of any length but is inherently restricted by the scope of the vision and by the scanner software length.

This is the first step in the processing of scanned image. The scanned image is checked for skewing, there are possibilities of image getting skewed with either left or right orientation.

Here, the image is first brightened and binarized the function for skew detection checks for an angle of orientation between +15 degrees and if detected than a simple image rotation is carried out till the lines match with the true horizontal axis, which produce a skew corrected image.

After pre-processing, the noise free image is passed to the segmentation phase, where the image is decomposed into individual characters.

Algorithm for Segmentation:

This Project is for You If:

  • ✅You are writing your final year project for the first time and don't want to make costly mistakes
  • ✅This is your exact topic, or very close to what your supervisor gave you
  • ✅Your supervisor has rejected one or more chapters and you don't know what to fix
  • ✅You are stuck on Chapter 3 or data analysis and nothing is making sense
  • ✅Your deadline is close and you are still far behind
  • ✅You can't find enough materials anywhere for your topic
  • ✅My project supervisor is very strict and I don't want to have any problem with my project work — I need a well structured and referenced project
  • ✅I saw my school project requirement and I am confused on where or what to do
  • ✅Everyone in my class has submitted their project and I haven't even started yet
  • ✅My supervisor rejected my research methodology and I don't know what to do
  • ✅I thought my topic was simple and easy but on Chapter 2 I cannot find any material and my work is due for submission
  • ✅I have the content somehow but my references and bibliography are all over the place and I don't know how to arrange them properly
  • ✅I am combining school with work and I genuinely do not have the time to write everything from scratch — I just need something to work with
  • ✅My supervisor changed the topic I wanted and approved another one — I am confused on where to start from
  • ✅I am scared of submitting something that will be flagged for plagiarism — I need something original that I can use as a proper guide
  • ✅I have been browsing the internet for days looking for materials on this topic and I cannot find anything useful anywhere
  • ✅This is not my first time doing this project — I have carried it over before and I cannot afford to do it again
  • ✅I have written some chapters already but I am stuck midway and need to see how others handled the same topic
Yes, This Is My Situation — Send Me the Complete Material

Here's everything you get with the complete project package:

✔Chapters 1–5 (Word & PDF)
✔Abstract & Table of Contents
✔References, Citation & Bibliography
✔Fully Editable Microsoft Word Format
✔Original, Plagiarism-Free Content
✔Instant Email Delivery
Complete package —₦10,000₦5,000
Price may return to ₦10,000 at any time.

See Our Conversations With Students on WhatsApp

Students sharing their experience after using our service

I've Seen Enough — Get Me This Project Material

What Other Students Are Saying

↩ View More Reviews

Can't find what you're looking for?

Search thousands of project topics across every department.

Frequently Asked Questions

Yes. All project materials on iProject are original, well-researched, and crafted to be plagiarism-free. They are intended as academic guides and reference materials.
You will receive the complete project in both Microsoft Word (.docx) and PDF formats, making it easy to edit and submit.
Delivery is instant. Once payment is confirmed, the material is sent directly to your email address. You can also download it immediately from the checkout page.
Absolutely. The Word format is fully editable, allowing you to modify the content, swap out case studies, and adapt the material to your specific research context.
All My Questions Are Answered — Download the Complete Project