Tailored answers, filled into supported job forms.Your tailored resume and answers, filled into supported job forms for you to review.

Download the Chrome Extension
AI AdoptionFunded CompaniesJob SimulationCertificationsRoadmapsJobsPricing
Sign In
OneRoadmap

OneRoadmap is a career platform built around ORI, its AI career agent. ORI finds overlooked job opportunities, matches them to your profile and shows the skill gaps to close, with roadmaps, challenges, job simulations and certifications to close them. When you are ready, it prepares a tailored resume, application answers and an application strategy, with a cover letter where the application asks for one. The OneRoadmap Chrome extension fills supported application forms for you to review and submit, and you keep track of every application in one place.

gaurav.ghai@oneroadmap.in
Delhi NCR, India

Platform

  • AI Roadmaps
  • Free Certifications
  • Learning Resources
  • Pricing

Training

  • AI Adoption Workshops
  • Expert Sessions
  • Upcoming Events
  • Workshop Gallery

Company

  • About
  • Blog
  • Contact

Legal

  • Privacy
  • Terms
  • Refunds
  • Delete your data

© 2026 OneRoadmap

Operated by Ghai Technologies, India · International operations through One Roadmap Marketing Management, Dubai, UAE

Built for your next chapter.

Open roles

SOFTWARE · Platform

Staff / Principal Software Engineer - USA

Inworld AI

SeniorOn-site · Mountain View, California, USAFullTimeListed 1y ago
Apply now

Backed by

Lightspeed India

HQ

🇮🇳 India

Open roles

19

How to stand out for Staff / Principal Software Engineer - USA at Inworld AI

Auto Match agent

Let ORI find you the best jobs.

Set up your Auto Match agent once - your target role, level and where you want to work. It searches every day, scores each opening against your profile and resume, and delivers the ones worth applying to, with a prepared application a click away.

Searches every day Scored against your profile Applications prepared for you
Sign in & set up Auto Match agent

Resume & career call

Get your resume reviewed for this role - 30-minute 1:1 call

Line-by-line resume feedback for this application, how to position your Role Readiness, and a clear plan for what to do next - with a OneRoadmap career coach.

Experience Senior · 6+ yrs

About the role

structured by ORI

About Inworld Inworld is a research lab and inference provider focused on realtime AI for consumer-facing applications. We build first-party speech models, serve LLMs, and run the inference behind modular APIs designed for high-volume, realtime workloads.

What you will do

  • Establish significant scope: Collaborate with the PMs, engineers and leads to determine the biggest product needs to focus on now.
  • Operate with technical autonomy: You have considerable leeway to suggest how to address a given focus area, including bringing in new technical dependencies or standards where it's the best choice.
  • Collaborate, execute, deliver: This is the core of the building loop. We aim to optimize for both speed and quality, despite it being decidedly non-obvious how to manage that tradeoff exceptionally well.
  • Reflect and drive improvements: Especially as a Staff Engineer, advocate for and realize system improvements, both related to and independent of key features.

What they are looking for

  • Candidates must be based in the SF Bay Area or willing to relocate (you will be working on-site in our South Bay office a few days a week)
  • Excellent programming skills and experience in a statically typed backend programming language, preferably Go, Python, C++ or Rust
  • Experience developing and deploying cloud-based services to at least hundreds of qps (preferably more)
  • Experience with relational databases (PostgreSQL or MySQL)
  • Hands-on experience with caching (Redis or Memcached), pubsub/queues, data pipelines (Flink, Beam), and Cloud storage
  • Excellent verbal and written communication skills, can collaborate and coordinate with other roles and engineer with ease, trusted and well-regarded teammate

Nice to have

  • Experience building API gateways, routing/proxy layers, or multi-provider orchestration systems
  • Experience with analytics or timeseries databases (ClickHouse, Timescale, InfluxDB)
  • Experience with OpenTelemetry
  • Experience with C++

Benefits

  • Equity
GoPythonC++RustPostgreSQLMySQLRedisMemcachedFlinkBeamClickHouseTimescaleInfluxDBOpenTelemetry
Full posting text

About Inworld Inworld is a research lab and inference provider focused on realtime AI for consumer-facing applications. We build first-party speech models, serve LLMs, and run the inference behind modular APIs designed for high-volume, realtime workloads. Hundreds of millions of users interact with Inworld powered apps every day and we serve over 10 trillion LLM tokens per month. Our models and infrastructure support consumer applications across companions, healthcare, fitness, education, media, and more. Our work spans model research, realtime inference, large-scale serving infrastructure, and the APIs developers use to bring these capabilities into production. We’ve raised more than $125M from Lightspeed Venture Partners, Section 32, Kleiner Perkins, Microsoft’s M12 venture fund, Founders Fund, Meta, Stanford, and others. Our technology has powered experiences from companies including NVIDIA, Microsoft Xbox, Niantic, Logitech Streamlabs, Wishroll, Little Umbrella, and Bible Chat. Inworld has also been recognized by CB Insights as one of the 100 most promising AI companies globally and named one of LinkedIn’s Top 10 Startups in the USA. ABOUT THE ROLE: Inworld recently launched a few exciting new products (Inworld TTS https://inworld.ai/tts, Inworld STT https://inworld.ai/speech-to-text, Speech-to-Speech / Realtime API https://inworld.ai/realtime-api and Inworld Router https://inworld.ai/router) for consumer AI applications, and we're looking for an ambitious and capable Staff/Principal Backend Engineer to join us and help take the Inworld AI platform even farther. Here is what you are going to work on: - Inworld Router: an intelligent routing layer that gives developers a single API to access 200+ LLMs. You'll own core systems for multi-provider failover, cost/latency-based routing, live A/B experimentation, and real-time observability at massive scale. - Realtime API - API-based model services: Our custom TTS/STT models and API includes free instant voice cloning. Learn more and hear examples at https://inworld.ai/ttsinworld.ai/tts http://inworld.ai/tts. Better yet, sign up yourself at https://platform.inworld.ai/platform.inworld.ai http://platform.inworld.ai, try out the premade voices, clone your own voice in just a few seconds, and let us know what you think! Beyond TTS, there is also LLM, Knowledge/RAG, STT, and more. - New exciting products, ambitious and large-scale, in the lineup for the launch later this year. - Services for control and optimization. We're just getting started on these deeper capabilities. - Finally complicated and exciting Infrastructural projects: platformization of new product upcoming offering, development and integration of best development tools, projects like system-wide billing and so on. As a Staff/Principal Software Engineer, you would be a significant part of one or more of these areas. The key challenges are: - Shipping quickly. AI is evolving weekly, so there's a ton of opportunity to be had. We want to move fast to capture those opportunities while they are still fresh and full of potential. - Zero to one. The platform is not a simple copycat. We have a vision for a deep platform/suite of capabilities that make it dramatically simpler for developers to scale and evolve their AI. - Realtime, online. As consumer applications become more capable of listening and talking, performance will matter, and AI has to adapt in realtime as well. These are bold but exciting challenges. - Multi-provider complexity at scale. Inworld Router must intelligently route across hundreds of models and providers while handling failover, sticky sessions, cost optimization, and conditional logic, all with minimal latency overhead. You'll design systems where every millisecond and every routing decision matters. Finally, almost everything here is a collaboration with our sibling ML teams, since ML and AI are critical to providing the learning and adaptability central to this vision. Please note: This is an IC-focused role. We are looking for someone who loves direct technical contribution alongside very capable peers. WHAT YOU’LL DO: - Establish significant scope: Collaborate with the PMs, engineers and leads to determine the biggest product needs to focus on now. - Operate with technical autonomy: You have considerable leeway to suggest how to address a given focus area, including bringing in new technical dependencies or standards where it's the best choice. - Collaborate, execute, deliver: This is the core of the building loop. We aim to optimize for both speed and quality, despite it being decidedly non-obvious how to manage that tradeoff exceptionally well. - Reflect and drive improvements: Especially as a Staff Engineer, advocate for and realize system improvements, both related to and independent of key features. EXPECTED EXPERIENCE: Must Haves - Excellent programming skills and experience in a statically typed backend programming language, preferably Go, Python, C++ or Rust - Experience developing and deploying cloud-based services to at least hundreds of qps (preferably more) - Experience with relational databases (PostgreSQL or MySQL) - Hands-on experience with caching (Redis or Memcached), pubsub/queues, data pipelines (Flink, Beam), and Cloud storage - Excellent verbal and written communication skills, can collaborate and coordinate with other roles and engineer with ease, trusted and well-regarded teammate Bonus Qualifications - Experience building API gateways, routing/proxy layers, or multi-provider orchestration systems - Experience with analytics or timeseries databases (ClickHouse, Timescale, InfluxDB) - Experience with OpenTelemetry - Experience with C++ Candidates must be based in the SF Bay Area or willing to relocate (you will be working on-site in our South Bay office a few days a week). The US base salary range for this full-time position is $280,000 - $350,000. In addition to base pay, total compensation includes equity and benefits. Within the range, individual pay is determined by work location, level, and additional factors, including competencies, experience, and business needs. The base pay range is subject to change and may be modified in the future. Inworld Jobs Privacy https://inworld.ai/jobs-privacy

Apply on company site

Meet Ori - your career agent on WhatsApp

Find jobs, get your roadmap, check if you're ready for a role and prepare applications - in chat, any language.

Ask Ori about this role
Checking your fit…
Inworld AI

India

Backed by Lightspeed India

Company pageWebsite

More at Inworld AI

Inworld AIGTMFunded

On-site · Mountain View, California, USA

5+ yrs · 4mo ago

Founding AI Solutions Engineer - USA
PythonJavaScriptTypeScript+9 more
View role
Inworld AIML EngineeringFunded

On-site · Switzerland

Senior · 5mo ago

Staff / Principal Machine Learning Engineer, Serving - Switzerland
Model AccelerationQuantization+14 more
View role
Inworld AIML EngineeringFunded

On-site · UK

Senior · 5mo ago

Staff / Principal Machine Learning Engineer, Serving - UK
Model AccelerationQuantization+14 more
View role
Inworld AIML EngineeringFunded

On-site · Serbia

Senior · 5mo ago

Senior / Lead Machine Learning Engineer, Serving - Serbia
Model AccelerationQuantization+14 more
View role
Apply