We are seeking a skilled Computer Vision Engineer with 3 to 4 years of proven experience in developing and deploying practical Computer Vision and AI solutions. The ideal candidate will have strong expertise in Python, PyTorch, OpenCV, and deep learning frameworks, with hands-on experience in image processing, segmentation, and real-time AI systems. Experience working on projects involving human parsing, virtual try-on, 3D human reconstruction, and generative AI will be highly valued.
Key Responsibilities:
- Develop and deploy advanced Computer Vision and Deep Learning solutions tailored to real-world applications.
- Work extensively on image and video processing tasks including object detection, segmentation, classification, and pose estimation.
- Design and implement human parsing and body segmentation algorithms.
- Build, optimize, and maintain real-time Computer Vision applications ensuring high performance and low latency.
- Train, fine-tune, test, and evaluate AI and Computer Vision models for accuracy and efficiency.
- Create, clean, annotate, and manage datasets specific to Computer Vision projects.
- Optimize models for GPU inference using tools like ONNX, TensorRT, and CUDA.
- Deploy AI and Computer Vision models in production environments, including cloud platforms.
- Integrate Computer Vision models seamlessly into software products and applications.
- Stay updated with the latest research and implement cutting-edge Computer Vision and AI techniques.
Required Qualifications:
- 3 to 4 years of professional experience in Computer Vision and AI development.
- Strong programming skills in Python, with proficiency in PyTorch, OpenCV, and NumPy.
- Solid understanding of Computer Vision and Deep Learning fundamentals.
- Practical experience in image segmentation, object detection, classification, and pose estimation.
- Proven track record of working on real-world Computer Vision projects and deploying models in production.
- Familiarity with real-time and video-based Computer Vision applications.
- Knowledge of GPU inference optimization tools such as ONNX, TensorRT, and CUDA.
- Experience with Linux operating systems, version control using Git, containerization with Docker, and RESTful APIs.
- Competence in dataset creation, annotation, and preprocessing workflows.
Preferred Qualifications and Additional Skills:
- Expertise in human parsing and body segmentation techniques.
- Experience with Virtual Try-On technologies (VTON/VITON).
- Familiarity with generative AI models including diffusion models, stable diffusion, inpainting, and ControlNet.
- Knowledge of 3D human body reconstruction using SMPL/SMPL-X models.
- Understanding of depth maps, point clouds, meshes, and 3D vision concepts.
- Experience with 3D libraries and tools such as Open3D, PyTorch3D, Trimesh, or Blender.
- Programming skills in C++ and familiarity with FastAPI for backend development.
- Experience working with cloud GPU infrastructure on AWS or GCP.
This position requires working on-site to collaborate closely with the team and contribute effectively to ongoing projects. If you are passionate about pushing the boundaries of Computer Vision and AI technology and meet the above qualifications, we encourage you to apply.