Research Scientist, Multi-Modal Human Understanding
Meta
Pittsburgh, PA, Burlingame, CAMid
Sign in to applyVerified 2h ago
- Location
- Pittsburgh, PA, Burlingame, CA
- Work model
- On-Site
- Level
- Mid
About this role
Meta is seeking a Research Scientist to advance multi-modal AI technologies for human understanding and synthesis. In this role, you will develop Vision-Language Models (VLMs) and video foundation models that enable machines to perceive, interpret, and generate rich representations of human behavior, expression, and interaction. Your research will span multi-modal reasoning, video understanding, and generative synthesis, enabling more natural and intuitive human-computer interaction at scale.