Ema

AI Resident

Ema  •  California (Hybrid)  •  3 hours ago
Apply
AI can make mistakes so check important info. Chat history is never stored.

Job Description

About Ema

Ema is building the world’s leading Agentic AI platform to transform enterprise productivity. We enable organizations to delegate repetitive tasks to Ema, the Universal AI Employee, delivering 10x gains in workforce efficiency, across functions. Founded by former executives from Google, Coinbase, Flipkart, and Okta, our team includes engineers from premier tech companies and graduates of Stanford, MIT, UC Berkeley, CMU, and IITs.

We are backed by industry leading investors including Accel, Naspers/Prosus, Section32, and angels like Sheryl Sandberg and Dustin Moskovitz. Headquartered in Silicon Valley and with offices in London, Bangalore and Vancouver and Bangalore, Ema is at the frontier of what Agentic AI can do in production — we ship real systems that run real business processes at scale.

The residency

You own one hard problem end to end. You write the proposal, build the system, design the evaluation, ship behind a gate, and finish with a write-up of what turned out to be true, including the parts that didn't work. You'll sit in the production codebase with a senior mentor and real production data. Recent residents have shipped self-improving harnesses, inference-cost work, agent memory, and eval infrastructure. Your project gets scoped with you, not handed to you.

The problem space

The loop we care about: production traces become data, data becomes training and evaluation, and better agents produce better traces. Projects live somewhere on that loop.

  • Harness and inference-time work. Context engineering, tool and skill design, orchestration, and deciding where extra inference compute actually pays. Self-improvement loops run behind hard fences.

  • Post-training for agents. SFT on curated trajectories, preference optimization, RL on real agent tasks. Reward design where outcomes are verifiable, process vs. outcome supervision, distilling frontier behavior into cheaper models.

  • Environments and rewards. Turning enterprise workflows into training and eval environments: fixture tenants, simulated users (some of whom get impatient and leave), verifiable rewards, and defenses against reward hacking. Agents will exploit a lazy grader.

  • Data engines. Mining production agent-steps into training and eval corpora: failure mining, labeling with calibrated judges, synthetic augmentation that stays useful.

  • Evaluation. Behavior-level benchmarks from real workflows, LLM judges calibrated against human labels, reliability statistics for stochastic agents.

  • Efficiency. Routing, ensembles, caching, small-model specialization. Quality per dollar is a research metric here.

What we're looking for

  • No specific degree required. Strong undergrads, grad students, and self-taught builders are all welcome; what matters is demonstrated depth in ML or agent systems.

  • Solid ML fundamentals and strong engineering: Python, PyTorch, and the discipline to ship in a large production codebase.

  • Real depth in at least one of: post-training (SFT/DPO/GRPO-family RL), reward modeling or LLM judges, agent and tool-use systems, retrieval and memory, eval design. One area you can teach us beats five you've touched.

  • Statistical literacy. You can size an experiment, and you know 25 samples at one seed is a datapoint, not a result.

  • Honest measurement as a habit. You'd rather kill your own feature with a clean experiment than ship it on a hunch.

Nice to have

  • Hands-on post-training with open models (TRL, veRL, OpenRLHF, or your own loop). Bonus points if you've debugged a reward-hacked run.

  • Built or trained in interactive agent environments (SWE, web, or tool-use gyms).

  • Large-scale trace analysis, data curation, or synthetic data work.

  • Serving and efficiency experience: vLLM/SGLang, distillation, quantization.

  • Multi-node GPU training, or the infra fluency to get there fast.

  • Publications, open-source work, or writing that shows how you think.

  • Security instincts: prompt injection, data governance, why a self-improving agent needs a fence.

Logistics

  • SF Bay Area, on-site/hybrid, half/full-time for the term. Flexible start.

  • Salary: $4,000 month

Compensation offered will be determined by factors such as location, level, job-related knowledge, skills, and experience. Certain roles may be eligible for variable compensation, equity, and benefits.

Ema Unlimited is an equal opportunity employer and is committed to providing equal employment opportunities to all employees and applicants for employment without regard to race, color, religion, sex, national origin, age, disability, sexual orientation, gender identity, or genetics.

Ema

About Ema

Ema is an enterprise partner of choice in building and deploying Agentic AI solutions. We are:

›› Simple - Conversationally build AI employees that learn and excel beyond human limits. Multiply your workforce in minutes, not months.

›› Trusted - Take control back by replacing 100s of vulnerable co-pilots with Ema’s secure and compliant AI Employees. Ema can be deployed both on-cloud and on-prem, and is compliant with industry-leading standards, including SOC 2 Type I & II, HIPAA, GDPR, ISO 27001, NIST CSF, NIST SP 800-171, and NIST AI RMF. We are also one of the world's only companies to be certified ISO 42001, the world's first AI management system standard.

›› Accurate - Achieve the highest accuracy at the lowest costs and latency with Ema’s proprietary 2T+ EmaFusion (TM) model, built for wide-ranging enterprise use-cases. Get the power of 100+ LLMs at your fingertips.

Experience the future of work with Ema, every enterprise’s best performing employee. Hire Ema today.

Industry
IT & Software
Company Size
51-200 employees
Headquarters
San Francisco
Year Founded
2023
Website
ema.co
Social Media