
Firmus Technologies is a global leader pioneering the development and operation of efficient AI infrastructure from model to grid. Founded in Australia in 2019, our mission is to create the most efficient AI infrastructure by combining cutting-edge technology with a steadfast commitment to sustainability.
At Firmus, we are unique in our approach. We design, build, and operate a new class of digital infrastructure – the AI Factory. Through our model-to-grid technology approach, we have pushed the boundaries of multi-generational liquid cooling systems, energy management, AI software orchestration, and construction. This co-designed approach from model to grid allows us to make every watt count and deliver low-cost AI tokens globally.
Our large-scale GPU cloud platform, Firmus AI Cloud, is purpose-built to deliver energy-efficient AI compute at scale. It empowers developers, enterprises, educational institutions, and government users to train and deploy AI models with unmatched efficiency and cost savings. With an ever-growing suite of services and applications, we are committed to delivering a cloud experience that is market-leading, proprietary, and built to scale.
At Firmus, you’ll work at the intersection of sustainability and artificial intelligence in a fast-paced environment powered by next-generation technology. You’ll be helping to transform an entire industry — and you’ll feel it every day.
Our team is made up of true innovators and leaders in their fields, and as an emerging company, you won’t be lost in a crowd. You’ll work closely with the founders, build a strong network, and see the impact of your work first-hand as we democratise AI tools for everyone — more sustainably and more affordably.
We believe great things happen when people from diverse backgrounds come together to do their best work and be their authentic selves. We are proud to be an equal opportunity employer.
Firmus Technologies is seeking an Engineering Manager to lead the Platform Engineering and Observability team. You are accountable for both the people and the delivery: the engineers you grow and the platform software they ship. Your team builds and operates the platform beneath the Firmus AI Cloud, from bare-metal GPU compute and high-performance networking to the internal platform services, self-service tooling, and the observability platform that our engineering teams and customers depend on. You run these as products for the engineering teams that build on them, investing in self-service, reliability, and developer experience. You are the escalation point above first-line operations and own the cross-cutting decisions your team shares. This is a build-and-grow leadership role: you build the team that scales our AI platform toward gigawatt-scale AI factories across the regions and grow your scope as you earn it.
Be the Single-Threaded Owner (STO) for multiple squads: their engineers and tech leads report to you, and you work alongside product, architecture, and delivery peers. Hire strong engineers, develop them through coaching and clear feedback, and manage performance directly. Own the org design, headcount planning, and engineering culture for your group, building a team that raises its own bar rather than depending on you.
Own end-to-end delivery for your team. Turn the product roadmap into sequenced, predictable delivery, shaping prioritisation with product and working with team leads on sprint planning and release gates. Own the cross-squads' dependencies that determine whether the platform ships on time, keep delivery health visible through metrics, and balance velocity against reliability and technical debt.
Stay technically credible and set the engineering quality bar across your squads. You are accountable for making sure the right cross-cutting decisions get made and driven to closure, such as build, buy, or open source, and the product SLAs your teams commit to. Keep design and code review rigorous, champion AI-assisted development so your teams ship faster without lowering the design, review, or security bar, and unblock the hard problems that stall delivery.
Own the reliability and security of the services your squads run in production. Act as the escalation point above the operations centre, accountable for L3 incident resolution, SLA-breach response, and post-mortems that convert into runbook and prevention work. Set the standards for SLOs, on-call, and change management, backed by the observability platform your team owns. Govern how AI-generated code and agentic workloads reach production, extend Firmus' existing SOC 2 Type 2 and ISO 27001 controls as the platform scales into new data centres, and own the cost efficiency of what your team operate, balancing reliability, performance, and speed against infrastructure spend.
Represent your team to engineering leadership and the CTO, clear about progress, risk, and the decisions you need. Own the engineering side of the customer relationship: lead customer technical briefings, and represent engineering directly when escalations turn on delivery, reliability, or architecture. Align with product, who own the roadmap and customer outcomes, with the architects on technical direction, and with delivery and operations so the platform ships and runs as one system, not a set of parts.

Firmus Technologies creates the most efficient AI infrastructure: AI Factories designed to operate at the max-q of AI token generation and profitability. With advanced liquid-everywhere data center technology at its core, Firmus innovates across all layers of the AI Factory, from energy grid management and thermal innovations to GPU and networking telemetry and control. This holistic approach creates AI Factories that operate at peak efficiency, delivering high uptime, throughput, and profitability with future-proofed designs.