Ml Engineer
CurrentML Engineer at Resana, specializing in computer vision algorithms, integrating NLP and multimodal techniques for enhanced performance. My responsibilities are:- Development and implementation of cutting-edge computer vision algorithms and solutions for a wide range of applications such as image processing, object detection, video analysis, and action detection.- Developing cutting-edge natural language processing models for text classification, sentiment analysis, and machine translation using techniques like BERT and GPT-3 to enhance model performance and accuracy.- Apply advanced image segmentation techniques to partition images into meaningful regions, enabling robust applications like semantic segmentation and instance segmentation.- Implementing and utilizing video analysis algorithms, including tracking, activity recognition, and action detection, leveraging motion analysis and temporal modeling to extract valuable insights from video data.- Implementing innovative fusion techniques to effectively combine text and image modalities for multimodal sentiment analysis, exploring attention mechanisms and strategies to improve model performance.- Apply generative adversarial networks (GANs) for various computer vision tasks, including image generation, data augmentation, and domain adaptation.- Implement diffusion algorithms for image and video enhancement, denoising, and inpainting, resulting in improved quality and visual fidelity.- Explore and implement zero-shot learning techniques, expanding the capabilities of computer vision systems to recognize objects or concepts not seen during training.- Stay up-to-date with the latest advancements in computer vision, machine learning, and deep learning research, actively contributing to the company's knowledge base and driving innovation.