Logo-of-Abaka-Ai-hiring-for-jobs-in-US-on-GrabJobs

Research Intern (Vision / VLM)

Job Description - Research Intern (Vision / VLM)

















Our recent related work: EditReward (submitted to ICLR '26)


Responsibilities




  • Design workflows to curate high-quality image editing and generation datasets for controllable diffusion and instruction tuning.




  • Conduct evaluations of vision-language models, including image understanding, caption alignment, and editing precision.




  • Assist in the training and evaluation of diffusion models or reward models.




  • Explore visual reasoning datasets that bridge images and text prompts.




Qualifications




  • Strong background in computer science, data engineering, artificial intelligence, or related fields, with hands-on experience in large-scale vision data systems.




  • 1+ years of experience in computer vision or multimodal machine learning (e.g., PyTorch, Diffusers, CLIP, BLIP, etc.).




  • Solid understanding of image-text alignment and latent-space editing.




  • (Preferred) Familiarity with aesthetic models, diffusion-based editing, vision-language modeling (VLM), or visual question answering (VQA) tasks.




  • (Preferred) Relevant publications in top conferences.








 


 











 





 
Original job Research Intern (Vision / VLM) posted on GrabJobs ©. To flag any issues with this job please use the Report Job button on GrabJobs.
Share Job
Share Job

Similar Research Intern Jobs in the US

GrabJobs is the no1 job portal in the US, connecting you to thousands of jobs fast! Find the best jobs in the US, apply in 1 click and get a job today!

Mobile Apps

Copyright © 2026 Grabjobs Pte.Ltd. All Rights Reserved.