Staff AI Inference and Acceleration Engineer
Figure is an AI robotics company developing autonomous general-purpose humanoid robots. The goal of the company is to ship humanoid robots with human level intelligence. Its robots are engineered to perform a variety of tasks in the home and commercial markets. Figure is headquartered in San Jose, CA.
We are looking for a Staff AI Inference & Acceleration Engineer to join the Platform Software team and own the on-board inference architecture for Figure’s humanoid robots. You will be the technical authority on how AI workloads are mapped, optimized, and executed across the robot’s compute hardware — driving down power consumption and cost while meeting the strict latency and reliability demands of a real-time autonomous system.
Responsibilities:
- Own the on-board inference architecture — mapping models to available accelerators (NPU, GPU, DSP, CPU) based on latency, power, and memory budgets.
- Partition inference workloads across heterogeneous compute resources, balancing real-time performance with power and thermal constraints.
- Define and maintain a system-level compute budget across all inference tasks running on the robot.
- Evaluate next-generation acceleration hardware and contribute to the definition of future compute platform requirements.
- Optimize inference toolchains end-to-en