Inworld AI Logo

Inworld AI

Staff / Principal Software Engineer - USA

Reposted One Month Ago
Hybrid
Mountain View, CA, USA
180K-280K Annually
Senior level
Hybrid
Mountain View, CA, USA
180K-280K Annually
Senior level
As a Staff Backend Engineer, you will drive technical contributions, collaborate with peers, optimize product needs, and improve system capabilities in AI software development.
The summary above was generated by AI

About Inworld

Inworld is a research lab and inference provider focused on realtime AI for consumer-facing applications. We build first-party speech models, serve LLMs, and run the inference behind modular APIs designed for high-volume, realtime workloads.

Hundreds of millions of users interact with Inworld powered apps every day and we serve over 10 trillion LLM tokens per month. Our models and infrastructure support consumer applications across companions, healthcare, fitness, education, media, and more. Our work spans model research, realtime inference, large-scale serving infrastructure, and the APIs developers use to bring these capabilities into production.

We’ve raised more than $125M from Lightspeed Venture Partners, Section 32, Kleiner Perkins, Microsoft’s M12 venture fund, Founders Fund, Meta, Stanford, and others. Our technology has powered experiences from companies including NVIDIA, Microsoft Xbox, Niantic, Logitech Streamlabs, Wishroll, Little Umbrella, and Bible Chat. Inworld has also been recognized by CB Insights as one of the 100 most promising AI companies globally and named one of LinkedIn’s Top 10 Startups in the USA.

About the role:

Inworld recently launched a few exciting new products (Inworld TTS, Inworld STT, Speech-to-Speech / Realtime API and Inworld Router) for consumer AI applications, and we're looking for an ambitious and capable Staff/Principal Backend Engineer to join us and help take the Inworld AI platform even farther. Here is what you are going to work on:

  • Inworld Router: an intelligent routing layer that gives developers a single API to access 200+ LLMs. You'll own core systems for multi-provider failover, cost/latency-based routing, live A/B experimentation, and real-time observability at massive scale.

  • Realtime API

  • API-based model services: Our custom TTS/STT models and API includes free instant voice cloning. Learn more and hear examples at inworld.ai/tts. Better yet, sign up yourself at platform.inworld.ai, try out the premade voices, clone your own voice in just a few seconds, and let us know what you think! Beyond TTS, there is also LLM, Knowledge/RAG, STT, and more.

  • New exciting products, ambitious and large-scale, in the lineup for the launch later this year.

  • Services for control and optimization. We're just getting started on these deeper capabilities.

  • Finally complicated and exciting Infrastructural projects: platformization of new product upcoming offering, development and integration of best development tools, projects like system-wide billing and so on.

As a Staff/Principal Software Engineer, you would be a significant part of one or more of these areas. The key challenges are:

  • Shipping quickly. AI is evolving weekly, so there's a ton of opportunity to be had. We want to move fast to capture those opportunities while they are still fresh and full of potential.

  • Zero to one. The platform is not a simple copycat. We have a vision for a deep platform/suite of capabilities that make it dramatically simpler for developers to scale and evolve their AI.

  • Realtime, online. As consumer applications become more capable of listening and talking, performance will matter, and AI has to adapt in realtime as well. These are bold but exciting challenges.

  • Multi-provider complexity at scale. Inworld Router must intelligently route across hundreds of models and providers while handling failover, sticky sessions, cost optimization, and conditional logic, all with minimal latency overhead. You'll design systems where every millisecond and every routing decision matters.

Finally, almost everything here is a collaboration with our sibling ML teams, since ML and AI are critical to providing the learning and adaptability central to this vision.

Please note: This is an IC-focused role. We are looking for someone who loves direct technical contribution alongside very capable peers.


What you’ll do:
  • Establish significant scope: Collaborate with the PMs, engineers and leads to determine the biggest product needs to focus on now.

  • Operate with technical autonomy: You have considerable leeway to suggest how to address a given focus area, including bringing in new technical dependencies or standards where it's the best choice.

  • Collaborate, execute, deliver: This is the core of the building loop. We aim to optimize for both speed and quality, despite it being decidedly non-obvious how to manage that tradeoff exceptionally well.

  • Reflect and drive improvements: Especially as a Staff Engineer, advocate for and realize system improvements, both related to and independent of key features.


Expected experience:

Must Haves

  • Excellent programming skills and experience in a statically typed backend programming language, preferably Go, Python, C++ or Rust

  • Experience developing and deploying cloud-based services to at least hundreds of qps (preferably more)

  • Experience with relational databases (PostgreSQL or MySQL)

  • Hands-on experience with caching (Redis or Memcached), pubsub/queues, data pipelines (Flink, Beam), and Cloud storage

  • Excellent verbal and written communication skills, can collaborate and coordinate with other roles and engineer with ease, trusted and well-regarded teammate

Bonus Qualifications

  • Experience building API gateways, routing/proxy layers, or multi-provider orchestration systems

  • Experience with analytics or timeseries databases (ClickHouse, Timescale, InfluxDB)

  • Experience with OpenTelemetry

  • Experience with C++

Candidates must be based in the SF Bay Area or willing to relocate (you will be working on-site in our South Bay office a few days a week).  

The US base salary range for this full-time position is $280,000 - $350,000. In addition to base pay, total compensation includes equity and benefits. Within the range, individual pay is determined by work location, level, and additional factors, including competencies, experience, and business needs. The base pay range is subject to change and may be modified in the future.

Inworld Jobs Privacy

HQ

Inworld AI Mountain View, California, USA Office

1975 W El Camino Real, Mountain View, CA, United States, 94040

Similar Jobs

21 Minutes Ago
Easy Apply
Remote or Hybrid
San Jose, CA, USA
Easy Apply
119K-170K Annually
Senior level
119K-170K Annually
Senior level
Cloud • Information Technology • Security • Software • Cybersecurity
As a Staff Site Reliability Engineer, you'll oversee Zscaler production data center services, optimize code, and ensure cloud service availability and performance. Collaborate with cross-functional teams to improve processes and resolve escalated issues.
Top Skills: BashDnsFirewallsGrafanaHTTPIcmpLoad BalancingNagiosOsi ModelPrometheusPythonTcp/Ip
27 Minutes Ago
Hybrid
50K-90K Annually
Senior level
50K-90K Annually
Senior level
Information Technology
Drives strategic growth and revenue by selling technology products and services to federal Navy and public-sector customers. Develops account strategies, leads consultative sales engagements, negotiates with stakeholders, identifies expansion opportunities, and serves as a trusted executive advisor. Coordinates cross-functional teams to design AI-enabled technology solutions, achieve product gross profit and services targets, and support long-term client success.
Top Skills: AICiscoData ModelingDellForecastingHpeIbmMicrosoft
27 Minutes Ago
Hybrid
80K-100K Annually
Entry level
80K-100K Annually
Entry level
Information Technology
Provides Tier 1 and Tier 2 service desk support in a secure environment. Responsibilities include ticket intake and management, desktop and endpoint troubleshooting, Microsoft Office and application support, account and access assistance, basic network triage, user communications, VIP support, documentation, and escalation of complex issues to Tier 3 teams. Requires an active TS/SCI clearance and enterprise end-user support experience.
Top Skills: Basic Network ConnectivityCollaboration PlatformsEmailEndpoint SupportMS OfficeOperating SystemsShared DrivesTicketing Systems

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account