Senior MLOps Engineer
Palo Alto, CA | Full-Time | On-site
About Nace AI:
Nace AI is an enterprise AI product and research company in Palo Alto (backed by General Catalyst, Walden Catalyst, and Intel). We build long-running AI agents powered by our own specialized SLMs — we started with financial audit and accounting workflows and are expanding from there. Real enterprise deployments, not demos.
Role Overview:
As a Senior MLOps Engineer, you will own the infrastructure that takes Nace.AI's models from research to reliable, production-grade systems. Our infrastructure generates task-specific Small Language Models (SLMs) in real time — which means our training, serving, and evaluation infrastructure isn't an afterthought; it is the product. You will design and operate the pipelines, orchestration, and serving layers that allow us to train, deploy, monitor, and continuously improve many specialized models at once, with the reliability that high-stakes audit, compliance, and finance workflows demand. This role sits at the intersection of ML engineering, LLM inference infrastructure, and platform reliability, and requires both strong systems instincts and hands-on execution.
Key Responsibilities
