Наши требования
Knowledge of large-scale model training and optimization
Experience in duplex-conversational model
Broader understanding of generative AI across modalities
Exposure to software development best practices
A flexible, experimental mindset i.e. comfortable working across research and engineering
(Bonus) Publications at EMNLP, COLING, NeurIPS, ICLR, CVPR, ICCV
Location
Preferred: San Francisco (hybrid) or London
Remote within the U.S. or Europe available for exceptional candidates
A PhD (or near completion) in a relevant field, or equivalent research experience
Hands-on experience with Large Multimodal Models and a strong foundation in generative (language) models. This could be in the context of tasks such as VQA, Audio/Video understanding tasks, captioning behavioral analysis, Translation tasks, Speech to Speech systems
Experience in fine-tuning/adapting VLMs for control, conditioning, or downstream tasks
Solid background in deep learning and foundation modes
Strong PyTorch skills and comfort building deep learning pipelines