Machine Learning OCR Processor

Urgent

Job Description

Job Description (JD) – Machine Learning OCR Processor | G2T Solutions
Job Title

Machine Learning OCR Processor

Employment Type
Full-Time (not explicitly stated in the posting)
Experience Required
5+ Years in OCR (Optical Character Recognition), Machine Learning, or Document AI projects.
Job Summary

The Machine Learning OCR Processor will be responsible for designing, training, optimizing, and deploying OCR solutions for intelligent document processing. The role requires expertise in OCR frameworks, machine learning, computer vision, cloud OCR services, and automation technologies to deliver highly accurate document extraction systems.

Key Responsibilities
OCR Model Development
Train and fine-tune OCR models.
Improve document recognition accuracy.
Build intelligent document processing pipelines.
Optimize OCR workflows for production environments.
OCR Framework Implementation

Work with industry-leading OCR technologies including:

Nanonets
Super.AI
Tesseract OCR
EasyOCR
ABBYY OCR
Machine Learning & Computer Vision
Develop machine learning models for OCR applications.
Improve text recognition using deep learning techniques.
Apply image processing techniques for better OCR accuracy.
Build document classification and extraction solutions.
Cloud OCR Integration

Integrate OCR solutions with cloud platforms such as:

AWS Textract
Google Vision API
Azure OCR
Database & API Integration
Design and integrate REST APIs.
Store and retrieve OCR data using relational and NoSQL databases.
Build scalable OCR services.
Performance Optimization

Improve OCR accuracy through:

Image preprocessing
Noise reduction (Denoising)
Thresholding
Parallel processing
Performance tuning
Deployment & Automation
Deploy OCR applications.
Automate OCR workflows.
Support production environments.
Containerize applications.
Required Technical Skills
OCR Technologies
Nanonets
Super.AI
Tesseract OCR
EasyOCR
ABBYY OCR
Programming Languages
Python
Machine Learning Frameworks
TensorFlow
PyTorch
Scikit-learn
Computer Vision
OpenCV
Cloud OCR Services
AWS Textract
Google Vision API
Azure OCR
Databases
SQL
NoSQL
MongoDB
PostgreSQL
API Development
RESTful APIs
Performance Optimization

Knowledge of:

Image preprocessing
Image denoising
Thresholding
Parallel processing
Automation & DevOps (Preferred)

Experience with:

Selenium
UiPath
Docker
Kubernetes
Required Experience
Minimum 5+ years working on OCR or Document AI projects.
Hands-on experience developing production-grade OCR systems.
Experience integrating cloud-based OCR APIs.
Strong background in Machine Learning and Computer Vision.
Preferred Skills
Intelligent Document Processing (IDP)
AI-powered document automation
Cloud-native deployments
OCR performance optimization
Workflow automation
Scalable API development
Soft Skills
Strong analytical and problem-solving skills
Attention to detail
Excellent debugging abilities
Good communication skills
Ability to work independently and in teams
Ideal Candidate Profile

The ideal candidate should have expertise in:

OCR technologies (Nanonets, Super.AI, Tesseract, EasyOCR, ABBYY)
Python
TensorFlow / PyTorch
OpenCV
Cloud OCR services (AWS Textract, Google Vision API, Azure OCR)
SQL / MongoDB / PostgreSQL
REST APIs
Docker & Kubernetes
Image processing and OCR optimization techniques

Candidates with experience in Document AI, Intelligent Document Processing (IDP), Computer Vision, and enterprise OCR deployments will be highly preferred.

How to Apply

Interested candidates can share their updated CV to:

Email: priyadharshini.r@g2tsolutions.com

Contact Numbers:

8754128555
8069299816

Location