Location: San Francisco, CA
Work Model: Onsite, 5 days per week. Relocation support provided.
Industry: AI infrastructure / model inference
Compensation: $180,000 – $240,000 base, plus equity
Our partner is a venture-backed AI infrastructure company delivering the fastest inference available on open models, serving both enterprise and serverless customers. They run their own compute, and customer demand currently outpaces the capacity they can bring online. The team is small, flat, and deliberately staying that way as they scale.
This is a forward-deployed engineering role that owns the customer, not just the code. Sales takes the first meeting. From there you are both the customer's engineer and their point of contact: you decide what to prove, you build it, you keep it running in production, and you carry the relationship.
Because you sit closer to real production load than anyone else on the team, what you learn goes straight into the roadmap. You will report directly into engineering leadership, and the longer-term plan is for engineers in this function to own their own pods as the team grows. These are among the first hires into the function, so its shape is still yours to influence.

talentpluto is the AI-native platform connecting elite GTM professionals with high-growth startups. Our mission is simple: streamline the hiring journey for exceptional GTM talent and innovative companies, ensuring the perfect match every time.
Talentpluto empowers professionals, even those not actively job hunting, to discreetly explore exciting career paths and opportunities.
Join talentpluto today and transform your hiring experience, connecting the best tech sales talent with visionary startups, faster and smarter than ever.