
Positron AI specializes in developing custom hardware systems to accelerate AI inference. These inference systems offer significant performance and efficiency gains over traditional GPU-based systems, delivering advantages in both performance per dollar and performance per watt. Positron exists to create the world's best AI inference systems.
Positron AI is hiring platform software architects to define the hardware/software contract for Asimov and the generations that follow. You will work across silicon architecture, firmware, operating systems, runtime, simulation, security, performance, and developer tooling.
This is a direct technical-leadership role reporting to the Director of Accelerator Platform Software. It is intentionally not buried inside a delivery team: you will maintain a cross-generation view, make durable interface decisions, and still contribute reference implementations where code is the clearest way to resolve ambiguity.
While this role is currently posted at a specific level, we are a growth-oriented organization and are open to hiring at a more senior level for the right candidate. Please note that this job description serves as a focused but generalized overview of the role; specific responsibilities and impact expectations will be tailored to the experience and seniority of the final hire.
In the first 12 months, the end-to-end Asimov software architecture and ownership map are clear, reviewable, and tied to implementation milestones. High-risk hardware/software contracts are specified and exercised in Terminus or Titan before silicon. A forward-looking architecture track is active for a non-evolutionary future device, while current-silicon delivery remains unblocked.
The base salary range for this role is $200,000 – $350,000.
Please note that the figures provided represent the base salary range only and do not include other elements of our total compensation package, equity, or comprehensive benefits.
At Positron AI, we value the unique expertise each candidate brings. While the range above reflects our typical expectation for the position, we reserve the flexibility to exceed this range for candidates whose specialized skills, significant experience, or unique qualifications fall outside the standard scope of the role. Final offers are determined based on a variety of factors, including internal equity, and individual impact.
We want you to do your best work and feel confident that you and your family are taken care of. That means comprehensive coverage, real time to rest, and support for your future.
This position is open to candidates currently authorized to work in the U.S. We cannot provide new visa sponsorship for this role but are open to facilitating H-1B visa transfers for eligible candidates.
Equal Opportunity Employer. If you're excited about the role but don't meet every bullet, we'd still love to hear from you.

Positron delivers vendor freedom and faster inference for both enterprises and research teams, by allowing them to use hardware and software explicitly designed from the ground up for generative and large language models (LLMs).
Through lower power usage and drastically lower total cost of ownership (TCO), Positron enables you to run popular open source LLMs to serve multiple users at high token rates and long context lengths. Positron is also designing its own ASIC to expand from inference and fine tuning to also support training and other parallel compute workloads.