RefolkCandidates
Open nowEngineeringPlatform

Staff / Principal Software Engineer - USA

Inworld AI · Mountain View, California

Location
Mountain View, California
Workplace
Hybrid
Employment
Full time
Level
Principal
Posted
16 months ago

About this role

About Inworld

Inworld is a research lab of top researchers and engineers, building the world’s top-ranked realtime voice models.

Today our models are the #1 ranked realtime voice models in the world. They are used to power the largest consumer-facing AI applications available, across categories like health, fitness, learning, therapy, companions, customer experience and media; representing 100s of millions of end users. Our work spans areas like research and development of state-of-the-art models, optimizing realtime inference, and creating best-in-class APIs and products that allow developers to engage their users.

We’ve raised more than $125M from Lightspeed, Section 32, Kleiner Perkins, Microsoft’s M12 venture fund, Founders Fund, Meta and Stanford, among others. Our technology has powered experiences from companies such as NVIDIA, Microsoft Xbox, Niantic, Logitech Streamlabs, Wishroll, Little Umbrella and Bible Chat. We’ve also been recognized by CB Insights as one of the 100 most promising AI companies globally and have been named one of LinkedIn’s Top 10 Startups in the USA.

About the role:

Inworld recently launched a few exciting new products (Inworld TTS, Inworld STT, Speech-to-Speech / Realtime API and Inworld Router) for consumer AI applications, and we're looking for an ambitious and capable Staff/Principal Backend Engineer to join us and help take the Inworld AI platform even farther. Here is what you are going to work on:

  • Inworld Router: an intelligent routing layer that gives developers a single API to access 200+ LLMs. You'll own core systems for multi-provider failover, cost/latency-based routing, live A/B experimentation, and real-time observability at massive scale.

  • Realtime API

  • API-based model services: Our custom TTS/STT models and API includes free instant voice cloning. Learn more and hear examples at inworld.ai/tts. Better yet, sign up yourself at platform.inworld.ai, try out the premade voices, clone your own voice in just a few seconds, and let us know what you think! Beyond TTS, there is also LLM, Knowledge/RAG, STT, and more.

  • New exciting products, ambitious and large-scale, in the lineup for the launch later this year.

  • Services for control and optimization. We're just getting started on these deeper capabilities.

  • Finally complicated and exciting Infrastructural projects: platformization of new product upcoming offering, development and integration of best development tools, projects like system-wide billing and so on.

As a Staff/Principal Software Engineer, you would be a significant part of one or more of these areas. The key challenges are:

  • Shipping quickly. AI is evolving weekly, so there's a ton of opportunity to be had. We want to move fast to capture those opportunities while they are still fresh and full of potential.

  • Zero to one. The platform is not a simple copycat. We have a vision for a deep platform/suite of capabilities that make it dramatically simpler for developers to scale and evolve their AI.

  • Realtime, online. As consumer applications become more capable of listening and talking, performance will matter, and AI has to adapt in realtime as well. These are bold but exciting challenges.

  • Multi-provider complexity at scale. Inworld Router must intelligently route across hundreds of models and providers while handling failover, sticky sessions, cost optimization, and conditional logic, all with minimal latency overhead. You'll design systems where every millisecond and every routing decision matters.

Finally, almost everything here is a collaboration with our sibling ML teams, since ML and AI are critical to providing the learning and adaptability central to this vision.

Please note: This is an IC-focused role. We are looking for someone who loves direct technical contribution alongside very capable peers.

What you’ll do:

  • Establish significant scope: Collaborate with the PMs, engineers and leads to determine the biggest product needs to focus on now.

  • Operate with technical autonomy: You have considerable leeway to suggest how to address a given focus area, including bringing in new technical dependencies or standards where it's the best choice.

  • Collaborate, execute, deliver: This is the core of the building loop. We aim to optimize for both speed and quality, despite it being decidedly non-obvious how to manage that tradeoff exceptionally well.

  • Reflect and drive improvements: Especially as a Staff Engineer, advocate for and realize system improvements, both related to and independent of key features.

Expected experience:

Must Haves

  • Excellent programming skills and experience in a statically typed backend programming language, preferably Go, Python, C++ or Rust

  • Experience developing and deploying cloud-based services to at least hundreds of qps (preferably more)

  • Experience with relational databases (PostgreSQL or MySQL)

  • Hands-on experience with caching (Redis or Memcached), pubsub/queues, data pipelines (Flink, Beam), and Cloud storage

  • Excellent verbal and written communication skills, can collaborate and coordinate with other roles and engineer with ease, trusted and well-regarded teammate

Bonus Qualifications

  • Experience building API gateways, routing/proxy layers, or multi-provider orchestration systems

  • Experience with analytics or timeseries databases (ClickHouse, Timescale, InfluxDB)

  • Experience with OpenTelemetry

  • Experience with C++

Candidates must be based in the SF Bay Area or willing to relocate (you will be working on-site in our South Bay office a few days a week).

The US base salary range for this full-time position is $280,000 - $350,000. In addition to base pay, total compensation includes equity and benefits. Within the range, individual pay is determined by work location, level, and additional factors, including competencies, experience, and business needs. The base pay range is subject to change and may be modified in the future.

As published by Inworld AI. Applications are handled on their site.

Skills this posting mentions

Apache FlinkData EngineeringApache Beam

About Inworld AI

Inworld AI is a realtime AI model and infrastructure company, and the leading consumer AI infrastructure platform. Inworld provides industry-leading realtime generative models, including the world’s #1-ranked voice AI models, intelligent model routing and optimization, and an Agent Runtime, enabling developers to build and deploy interactive AI applications to millions of concurrent users. Inworld primarily serves use cases where realtime interaction and sophisticated agent capabilities are critical, such as companion apps, developer assistants, and agents for learning & education, health & wellness, interactive media and enterprise. At its core, Inworld is a product-oriented research lab of top AI researchers and engineers. The founding team led product for LLMs at DeepMind and built Dialogflow, the conversational AI platform acquired by Google. Inworld has raised $125M+ from Lightspeed, Kleiner Perkins, Founders Fund, CRV, Stanford, Microsoft M12, Meta, Intel Capital, Samsung NEXT, LG Tech Ventures, and Bitkraft among others.

All 18 openings at Inworld AI

One click, then it is written

Apply to Inworld AI with a resume written for this role.

Queue Staff / Principal Software Engineer - USA and I read the posting, rewrite your resume against it, draft the cover letter, and score the fit. Then you press send, or press one button and I fill in Inworld AI’s form for you.

  1. 01Drop your resume

    A PDF or a LinkedIn URL. About a minute, once.

  2. 02I rank the openings

    Every weekday morning, the live catalog scored against your history. Up to 20 worth your time, not two hundred links.

  3. 03Each one is written up

    Resume rewritten for the posting, a cover letter, a fit score. Press send, or let me fill in the form.

  • 25 sent a week, free
  • No card
  • Nothing sent until you say so

More roles at Inworld AI

See all

Similar roles elsewhere

See more

Put this to work

Paste your career in once. Every application after that is written for you.

Drop a resume or a LinkedIn URL. I rank the live openings against it, rewrite the resume and write a cover letter for the best of them, and fill in the employer's form when you press the button. You read, you decide what goes out.

  1. 01Drop your resume

    A PDF or a LinkedIn URL. About a minute, once.

  2. 02I rank the openings

    Every weekday morning, the live catalog scored against your history. Up to 20 worth your time, not two hundred links.

  3. 03Each one is written up

    Resume rewritten for the posting, a cover letter, a fit score. Press send, or let me fill in the form.

  • New matches ranked and written before you are up.
  • Every bullet stays inside what your history supports. Nothing invented.
  • Queued, submitted, interviewing, offer: one screen, not a spreadsheet.

500 free credits on sign-up. No card. Nothing is sent until you say so.

Listed from the job board Inworld AI publishes. Refolk is not the employer and does not handle their hiring. Applications go to Inworld AI directly.