Deployment Engineering Director, Systems Engineering
About Nscale
Nscale is the GPU cloud engineered for AI. We provide cost-effective, high-performance infrastructure for AI start-ups and large enterprise customers. Nscale enables AI-focused companies to achieve superior results by reducing the complexity of AI development. Our GPU cloud bolsters technical capabilities and directly supports strategic business outcomes, including cost management, rapid innovation, and environmental responsibility.
We thrive on a culture of relentless innovation, ownership, and accountability, where every team member takes pride in their work and drives it with excellence and urgency. As an Nscaler, you’ll build trust through openness and transparency, where everyone is inspired to do their best work. If you join our team, you’ll be contributing to building the technology that powers the future.
About The Role
We are seeking a Deployment Engineering Director to lead and scale a high-performing engineering team responsible for core platform systems used across the company. You will drive multi-quarter platform initiatives that improve delivery velocity, system reliability, and product quality through systems validation. Working closely with Product Management and partner engineering teams, you will turn complex, ambiguous challenges into clear execution plans and raise the bar for engineering quality and operational excellence.
What You'll Be Doing
- Lead and scale a high-performing engineering team responsible for core platform systems used across the company.
- Define and drive multi-quarter platform initiatives that improve delivery velocity, system reliability/quality based on systems validations, and evolve the product in partnership with Product Management and partner engineering teams.
- Partner closely with product and engineering leaders to evolve the platform as a product, balancing long-term architecture with near-term business needs.
- Take ambiguous, high-impact problem areas (e.g., system scalability, vendor experience, platform abstractions) and turn them into clear execution plans.
- Drive alignment across multiple teams working on complex, interdependent systems including internal platforms, infrastructure, and developer tooling.
- Raise the bar on engineering quality, operational excellence, and execution across the platform organization.
About You
- Bachelor's degree in CS or related engineering field with 10+ years of Systems Engineering Management experience.
- Experience working in a large cloud provider or hyperscale data center environment. Experience with distributed systems or virtualization platforms is a plus.
- Experience working in a systems or compute operations leadership role.
- Must be able to travel up to 50% of the time.
- Experience with virtualization, containerization (Kubernetes, Docker), and distributed systems design.
- Strong knowledge of Linux/Unix systems administration, OS-level tuning, and server/GPU hardware architecture.
- Experience with infrastructure-as-code and configuration management (Ansible, Terraform, Puppet/Chef).
- Experience with scripting or automation and data center design.
- Ability to use professional concepts and company objectives to resolve complex issues in creative and effective ways.
- Excellent organizational, verbal, and written communication skills.
- Excellent judgment in influencing product roadmap direction, features, and priorities.
- Familiarity with system-level architecture, data synchronization, fault tolerance, and state management.
- General enterprise storage, networking, or computing experience.
- Experience with systems monitoring, observability, and telemetry solutions.
- Experience with GPU burn-in and validation testing at scale (thermal, power, and stress qualification prior to production acceptance).
What We Can Offer You
You’ll have the opportunity to help shape the operating standards behind a next-generation AI cloud platform, working on complex infrastructure challenges with real ownership and impact. This is a chance to play a meaningful role in scaling high-performance, sustainable data centre operations in a fast-moving environment.
Equal Opportunities Statement
We strongly encourage applications from people of colour, the LGBTQ+ community, people with disabilities, neurodivergent people, parents, carers, and people from lower socio-economic backgrounds.
If there’s anything we can do to accommodate your specific situation, please let us know.
The responsibilities outlined in this job description are not exhaustive and are intended to provide a general overview of the position. The employee may be required to perform additional duties, tasks, and responsibilities as assigned by management, consistent with the skills and qualifications required for the role.
For information on how Nscale handles candidate personal data, please see our Employee & Candidate Privacy Notice: Here.