Tailored answers, filled into supported job forms.Your tailored resume and answers, filled into supported job forms for you to review.

Download the Chrome Extension
AI AdoptionFunded CompaniesJob SimulationCertificationsRoadmapsJobsPricing
Sign In
OneRoadmap

OneRoadmap is a career platform built around ORI, its AI career agent. ORI finds overlooked job opportunities, matches them to your profile and shows the skill gaps to close, with roadmaps, challenges, job simulations and certifications to close them. When you are ready, it prepares a tailored resume, application answers and an application strategy, with a cover letter where the application asks for one. The OneRoadmap Chrome extension fills supported application forms for you to review and submit, and you keep track of every application in one place.

gaurav.ghai@oneroadmap.in
Delhi NCR, India

Platform

  • AI Roadmaps
  • Free Certifications
  • Learning Resources
  • Pricing

Training

  • AI Adoption Workshops
  • Expert Sessions
  • Upcoming Events
  • Workshop Gallery

Company

  • About
  • Blog
  • Contact

Legal

  • Privacy
  • Terms
  • Refunds
  • Delete your data

© 2026 OneRoadmap

Operated by Ghai Technologies, India · International operations through One Roadmap Marketing Management, Dubai, UAE

Built for your next chapter.

Open roles

AI / ML · Engineering

Senior Backend / ML Ops Engineer

Drafted

SeniorOn-site · San FranciscoListed 1d ago
Apply now

Backed by

Y Combinator

HQ

🇺🇸 San Francisco

Open roles

6

How to stand out for Senior Backend / ML Ops Engineer at Drafted

Auto Match agent

Let ORI find you the best jobs.

Set up your Auto Match agent once - your target role, level and where you want to work. It searches every day, scores each opening against your profile and resume, and delivers the ones worth applying to, with a prepared application a click away.

Searches every day Scored against your profile Applications prepared for you
Sign in & set up Auto Match agent

Resume & career call

Get your resume reviewed for this role - 30-minute 1:1 call

Line-by-line resume feedback for this application, how to position your Role Readiness, and a clear plan for what to do next - with a OneRoadmap career coach.

Experience Senior · 6+ yrs

About the role

structured by ORI

CareersSenior Backend / ML Ops EngineerOverviewOverviewapplicationDrafted's ValuesWe're a small team working fully in-person in San Francisco. We value high-ownership builders who want to be a part of a talented, highly motivated team.

What you will do

  • Build the backend systems, model infrastructure, and production pipelines that power Drafted's generative home design workflows.

What they are looking for

  • Building and scaling GPU-based inference services, optimizing for both low latency and high resource utilization.
  • Job orchestration and load balancing with parallel generations, heterogeneous resource constraints (GPU, CPU, I/O), and multi-tiered queues.
  • Implementing observability for latency attribution and failure diagnosis for multi-stage, asynchronous, and cross-platform pipelines.
  • Designing fan-out architectures where upstream job completion triggers multiple independent downstream consumers that have mixed criticality, with some consumers blocking and others best-effort.
  • Familiarity with modern cloud infrastructure: managed databases, job queues, edge compute/CDN, and PaaS deployment platforms.
  • Knowledge of training infrastructure, especially distributed GPU training across multiple nodes.
GPU-based inferenceJob orchestrationLoad balancingObservabilityFan-out architecturesDistributed GPU trainingPaaS deploymentGPUCPU
Full posting text

CareersSenior Backend / ML Ops EngineerOverviewOverviewapplicationDrafted's ValuesWe're a small team working fully in-person in San Francisco. We value high-ownership builders who want to be a part of a talented, highly motivated team. We're guided by the following values:Own the mission. We take agency, act like owners, and see problems through to real outcomes.Build in the open. We value direct feedback, fast learning, and growth through honest collaboration.Move with care and speed. We iterate quickly while staying deeply respectful of our teammates.Seek the why. We challenge assumptions, think from first principles, and never stop asking questions.Design for everyone. We believe anyone should be able to design and build a home they love.Solve what matters. We embrace hard problems and create new paths forward when none exist.The RoleBuild the backend systems, model infrastructure, and production pipelines that power Drafted's generative home design workflows.Example ProjectsBuilding parallel generation pipelines where multiple workers race to fill output slots, with dynamic filtering based on post-processing results. Implementing claim coordination to prevent duplicate work, fallback logic to use best-available generations when hitting retry limits, and caching mechanisms to reuse generations across jobs (same user regenerating with the same prompt).Developing coordination mechanisms for capacity-constrained pipelines where maximum concurrency is fixed (reserved GPU instances, instance quotas, API rate limits) and peak demand exceeds available capacity -- implementing backpressure, admission control, and retry logic to prevent overwhelming downstream consumers.Implementing timeout and cleanup policies that account for high variance of computational complexity (p99 is 10x p50) and variable parallelism where completion time depends on concurrent worker count, which fluctuates dynamically based on queue dynamics and capacity constraints, without being overly conservative or prematurely terminating legitimately slow work.Ideal ExperienceBuilding and scaling GPU-based inference services, optimizing for both low latency and high resource utilization.Job orchestration and load balancing with parallel generations, heterogeneous resource constraints (GPU, CPU, I/O), and multi-tiered queues.Implementing observability for latency attribution and failure diagnosis for multi-stage, asynchronous, and cross-platform pipelines.Designing fan-out architectures where upstream job completion triggers multiple independent downstream consumers that have mixed criticality, with some consumers blocking and others best-effort.Familiarity with modern cloud infrastructure: managed databases, job queues, edge compute/CDN, and PaaS deployment platforms.Desired SkillsetKnowledge of training infrastructure, especially distributed GPU training across multiple nodes.Apply NowRole DetailsLocationSan FranciscoEmployment TypeFull timeLocation TypeIn personSalary$150k-$300k + 0.5%-1.5%

Apply on company site

Meet Ori - your career agent on WhatsApp

Find jobs, get your roadmap, check if you're ready for a role and prepare applications - in chat, any language.

Ask Ori about this role
Checking your fit…
Drafted

Consumer

Design your home instantly with AI

Backed by Y Combinator

Company pageWebsite

More at Drafted

DraftedEngineeringFunded

On-site · San Francisco

Senior · 1d ago

Senior Full Stack Engineer
Vibe codingReactTanstack+10 more
View role
DraftedEngineeringFunded

On-site · San Francisco

Senior · 1d ago

Senior ML Engineer
diffusion models+6 more
View role
DraftedFunded

On-site · San Francisco

Senior · 1d ago

Senior ML Researcher
diffusion models+1 more
View role
Apply