ActiveJobs

Research Scientist, Multi-Modal Human Understanding

Meta · Pittsburgh, PA

Full-timeOn-sitePosted 10 September 2026
Apply on Company Site →

Job description

Meta is seeking a Research Scientist to advance multi-modal AI technologies for human understanding and synthesis. In this role, you will develop Vision-Language Models (VLMs) and video foundation models that enable machines to perceive, interpret, and generate rich representations of human behavior, expression, and interaction. Your research will span multi-modal reasoning, video understanding, and generative synthesis, enabling more natural and intuitive human-computer interaction at scale.

Verified and listed by ActiveJobs. Applications are made directly on Meta's own career page — we never sit in the middle.