Machine Learning OCR Processor
Job Description
Job Description (JD) – Machine Learning OCR Processor | G2T Solutions
Job Title
Machine Learning OCR Processor
Employment Type
Full-Time (not explicitly stated in the posting)
Experience Required
5+ Years in OCR (Optical Character Recognition), Machine Learning, or Document AI projects.
Job Summary
The Machine Learning OCR Processor will be responsible for designing, training, optimizing, and deploying OCR solutions for intelligent document processing. The role requires expertise in OCR frameworks, machine learning, computer vision, cloud OCR services, and automation technologies to deliver highly accurate document extraction systems.
Key Responsibilities
OCR Model Development
Train and fine-tune OCR models.
Improve document recognition accuracy.
Build intelligent document processing pipelines.
Optimize OCR workflows for production environments.
OCR Framework Implementation
Work with industry-leading OCR technologies including:
Nanonets
Super.AI
Tesseract OCR
EasyOCR
ABBYY OCR
Machine Learning & Computer Vision
Develop machine learning models for OCR applications.
Improve text recognition using deep learning techniques.
Apply image processing techniques for better OCR accuracy.
Build document classification and extraction solutions.
Cloud OCR Integration
Integrate OCR solutions with cloud platforms such as:
AWS Textract
Google Vision API
Azure OCR
Database & API Integration
Design and integrate REST APIs.
Store and retrieve OCR data using relational and NoSQL databases.
Build scalable OCR services.
Performance Optimization
Improve OCR accuracy through:
Image preprocessing
Noise reduction (Denoising)
Thresholding
Parallel processing
Performance tuning
Deployment & Automation
Deploy OCR applications.
Automate OCR workflows.
Support production environments.
Containerize applications.
Required Technical Skills
OCR Technologies
Nanonets
Super.AI
Tesseract OCR
EasyOCR
ABBYY OCR
Programming Languages
Python
Machine Learning Frameworks
TensorFlow
PyTorch
Scikit-learn
Computer Vision
OpenCV
Cloud OCR Services
AWS Textract
Google Vision API
Azure OCR
Databases
SQL
NoSQL
MongoDB
PostgreSQL
API Development
RESTful APIs
Performance Optimization
Knowledge of:
Image preprocessing
Image denoising
Thresholding
Parallel processing
Automation & DevOps (Preferred)
Experience with:
Selenium
UiPath
Docker
Kubernetes
Required Experience
Minimum 5+ years working on OCR or Document AI projects.
Hands-on experience developing production-grade OCR systems.
Experience integrating cloud-based OCR APIs.
Strong background in Machine Learning and Computer Vision.
Preferred Skills
Intelligent Document Processing (IDP)
AI-powered document automation
Cloud-native deployments
OCR performance optimization
Workflow automation
Scalable API development
Soft Skills
Strong analytical and problem-solving skills
Attention to detail
Excellent debugging abilities
Good communication skills
Ability to work independently and in teams
Ideal Candidate Profile
The ideal candidate should have expertise in:
OCR technologies (Nanonets, Super.AI, Tesseract, EasyOCR, ABBYY)
Python
TensorFlow / PyTorch
OpenCV
Cloud OCR services (AWS Textract, Google Vision API, Azure OCR)
SQL / MongoDB / PostgreSQL
REST APIs
Docker & Kubernetes
Image processing and OCR optimization techniques
Candidates with experience in Document AI, Intelligent Document Processing (IDP), Computer Vision, and enterprise OCR deployments will be highly preferred.
How to Apply
Interested candidates can share their updated CV to:
Email: priyadharshini.r@g2tsolutions.com
Contact Numbers:
8754128555
8069299816