The incredible potential of multimodal foundation models and large language models has unlocked machine learning applications that were previously thought infeasible. The Video Computer Vision (VCV) group is looking for a highly motivated and skilled Machine Learning Systems Engineer to help us ship cutting-edge computer vision technology on Apple devices.
The VCV organization has pioneered groundbreaking features like FaceID/FaceKit, Gaze/Hand Gesture Control, Body Tracking, and 2D/3D Scene Understanding fundamentally changing how millions of users interact with technology. We seamlessly balance research and product requirements to deliver pioneering, Apple-quality experiences. By innovating across the full stack and partnering closely with hardware, software, and AI teams, we shape future products and bring our architectural vision to life.
As a member of the Video Computer Vision team, you will train, evaluate, and deploy purpose-built vision models on Apple hardware. You will develop innovative techniques to optimize model performance, efficiency, and scalability, ensuring a seamless user experience under strict on-device constraints.