Senior CV Engineer Job
Employer: John Henry
SpiderID: 14239578
Location: New York., New York
Posted: 9/15/2026
Wage: 60
Priority Review Date: 10/15/2026
Job Code / NOC / SOC:
Category: Information Technology
Job Description:
We are looking for Senior CV Engineer.
Your main tasks will be:
Train and fine-tune generative models (text2image, image2image, video2image, IP-adapters) to produce photorealistic and stylized visuals
Work with image classifiers, rankers, and image/video captioning models that understand visual context
Track cutting-edge CV research and open-source developments; translate them into a CV roadmap for the ML team
Build and own the iterative model improvement process, including quality validation systems for CV outputs
Optimize inference performance: diffusion optimization, quantization, kernel and framework-level work
Optionally: train multimodal AI companions that combine CV and NLP for deeper visual understanding
We expect from you:
5+ Years of experience training diffusion models and modifying their architectures
Proficiency with diffusers and transformers libraries
Hands-on experience with IP-adapters
Backend engineering experience (Python, Go, C#) and knowledge of scalable deployment systems is an advantage
Nice to have:
Flow matching training
Diffusion model distillation
DPO and other text2image fine-tuning approaches
Text2video and CLIP fine-tuning
Experience with multimodal LLMs
NLP / LLM training and chatbot development background
Your main tasks will be:
Train and fine-tune generative models (text2image, image2image, video2image, IP-adapters) to produce photorealistic and stylized visuals
Work with image classifiers, rankers, and image/video captioning models that understand visual context
Track cutting-edge CV research and open-source developments; translate them into a CV roadmap for the ML team
Build and own the iterative model improvement process, including quality validation systems for CV outputs
Optimize inference performance: diffusion optimization, quantization, kernel and framework-level work
Optionally: train multimodal AI companions that combine CV and NLP for deeper visual understanding
We expect from you:
5+ Years of experience training diffusion models and modifying their architectures
Proficiency with diffusers and transformers libraries
Hands-on experience with IP-adapters
Backend engineering experience (Python, Go, C#) and knowledge of scalable deployment systems is an advantage
Nice to have:
Flow matching training
Diffusion model distillation
DPO and other text2image fine-tuning approaches
Text2video and CLIP fine-tuning
Experience with multimodal LLMs
NLP / LLM training and chatbot development background
Contact Information:
| Contact Name: John Henry | Type: |
| Company: Jobodds | |
| Web Site: https://www.oddsjob.org | |