Jobs and Careers
ME

Research Engineer, Computer Vision and Multimodal - GenAI

Meta
Menlo Park, United Statesfull_timeVerifiedPosted 30 Jul 2024
💰 $208,000/yr($141,340/yr$208,000/yr)

About the role

Meta is seeking a Research Engineer to join our Llama Multimodal team building really large scale Llama models on some of the largest training clusters with extremely large volumes of data leading to unforeseen multimodal capabilities. We are looking for research engineer specialized in generative AI, multimodal reasoning, computer vision, with experience in areas like multimodal model training; data processing for pretraining and fine-tuning; LLM alignment; reinforcement learning for model tuning; efficient training and inference; image and video generation. The ideal candidate will have an interest in producing and applying new science to help us develop and deploy large multimodal datasets and models.Research Engineer, Computer Vision and Multimodal - GenAI Responsibilities
  • Leading, collaborating, and executing on research that pushes forward the state of the art in multimodal reasoning and generation research
  • Contributing to experimental design, implementation, evaluation, and reporting results
  • Working with and creating large datasets and benchmarks
Minimum Qualifications
  • Bachelor's degree in Computer Science, Computer Engineering, relevant technical field, or equivalent practical experience.
  • Masters in computer science, computer engineering, or relevant technical field
  • 3+ years of industry experience in computer vision and/or machine learning
  • Research experience in one or more of these areas: computer vision, multimodal perception, NLP, vision-language
  • Experience writing software and executing complex experiments
  • Must obtain work authorization in the country of employment at the time of hire, and maintain ongoing work authorization during employment
Preferred Qualifications
  • PhD in computer vision and/or machine learning or related areas
  • 2+ years of industry, academic, or government lab experience in computer vision and/or machine learning
  • Publication record at AI conferences (e.g., CVPR, ICCV, ECCV, NeurIPS, ICML, ICLR, EMNLP, and ACL)
  • Hands on experience working with really large scale models and datasets
  • Background in multimodal LLM or computer vision
LocationsAbout Meta Meta builds technologies that help people connect, find communities, and grow businesses. When Facebook launched in 2004, it changed the way people connect. Apps like Messenger, Instagram and WhatsApp further empowered billions around the world. Now, Meta is moving beyond 2D screens toward immersive experiences like augmented and virtual reality to help build the next evolution in social technology. People who choose to build their careers by building with us at Meta help shape a future that will take us beyond what digital connection makes possible today—beyond the constraints of screens, the limits of distance, and even the rules of physics. Meta is committed to providing reasonable support (called accommodations) in our recruiting processes for candidates with disabilities, long term conditions, mental health conditions or sincerely held religious beliefs, or who are neurodivergent or require pregnancy-related support. If you need support, please reach out to accommodations-ext@fb.com. $70.67/hour to $208,000/year + bonus + equity + benefits

Individual compensation is determined by skills, qualifications, experience, and location. Compensation details listed in this posting reflect the base hourly rate, monthly rate, or annual salary only, and do not include bonus, equity or sales incentives, if applicable. In addition to base compensation, Meta offers benefits. Learn more about benefits at Meta.

Apply for this role

Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.

Apply Now →Generate Application Kit

Free account required — sign up in 30s

Company

Meta

View company profile →