Program Manager Ocr
Current👉Developed URDU OCR End to End Pipeline Using Tesseract and LSTM - WER 20%👉Developed URDU Data Scrapping Pipeline and Data Augmentation using 512 different Fonts👉Developed Urdu Speech Recognition App using Transformer based approaches in Huggingface -9 speakers 👉Speech Data Augmentation using Librosa and Essentia for Noise removal and Noise AdditionSpeech Cloning for Speech data Augmentation Machine Translation using Transformer based approaches like seq2seq modelling -… Show more 👉Developed URDU OCR End to End Pipeline Using Tesseract and LSTM - WER 20%👉Developed URDU Data Scrapping Pipeline and Data Augmentation using 512 different Fonts👉Developed Urdu Speech Recognition App using Transformer based approaches in Huggingface -9 speakers 👉Speech Data Augmentation using Librosa and Essentia for Noise removal and Noise AdditionSpeech Cloning for Speech data Augmentation Machine Translation using Transformer based approaches like seq2seq modelling - WER and Bleu 15%👉Docker Speech Data Preparation Pipeline Developed Flask API for URDU OCRWord Segmentation of URDU Ligatures👉Developed Real-time Data acquisition pipeline for machine translation using Kafka Using Fuzzy Based approached like levenshtein distance n-gram and cosine similarity for machine translation data preparation. 👉Deployed Machine Learning Models and Developed Sophisticated system design architecture to support million of users using Load balancing consistent hashing and consensus algorithms. 👉Organization of Team Task by maintaining the product backlog- Running Sprints for Data Science Activities Show less