Inworld AI Logo

Inworld AI

Staff / Principal Software Engineer - USA

Reposted One Month Ago
Hybrid
Mountain View, CA, USA
180K-280K Annually
Senior level
Hybrid
Mountain View, CA, USA
180K-280K Annually
Senior level
As a Staff Backend Engineer, you will drive technical contributions, collaborate with peers, optimize product needs, and improve system capabilities in AI software development.
The summary above was generated by AI

About Inworld

Inworld is a research lab and inference provider focused on realtime AI for consumer-facing applications. We build first-party speech models, serve LLMs, and run the inference behind modular APIs designed for high-volume, realtime workloads.

Hundreds of millions of users interact with Inworld powered apps every day and we serve over 10 trillion LLM tokens per month. Our models and infrastructure support consumer applications across companions, healthcare, fitness, education, media, and more. Our work spans model research, realtime inference, large-scale serving infrastructure, and the APIs developers use to bring these capabilities into production.

We’ve raised more than $125M from Lightspeed Venture Partners, Section 32, Kleiner Perkins, Microsoft’s M12 venture fund, Founders Fund, Meta, Stanford, and others. Our technology has powered experiences from companies including NVIDIA, Microsoft Xbox, Niantic, Logitech Streamlabs, Wishroll, Little Umbrella, and Bible Chat. Inworld has also been recognized by CB Insights as one of the 100 most promising AI companies globally and named one of LinkedIn’s Top 10 Startups in the USA.

About the role:

Inworld recently launched a few exciting new products (Inworld TTS, Inworld STT, Speech-to-Speech / Realtime API and Inworld Router) for consumer AI applications, and we're looking for an ambitious and capable Staff/Principal Backend Engineer to join us and help take the Inworld AI platform even farther. Here is what you are going to work on:

  • Inworld Router: an intelligent routing layer that gives developers a single API to access 200+ LLMs. You'll own core systems for multi-provider failover, cost/latency-based routing, live A/B experimentation, and real-time observability at massive scale.

  • Realtime API

  • API-based model services: Our custom TTS/STT models and API includes free instant voice cloning. Learn more and hear examples at inworld.ai/tts. Better yet, sign up yourself at platform.inworld.ai, try out the premade voices, clone your own voice in just a few seconds, and let us know what you think! Beyond TTS, there is also LLM, Knowledge/RAG, STT, and more.

  • New exciting products, ambitious and large-scale, in the lineup for the launch later this year.

  • Services for control and optimization. We're just getting started on these deeper capabilities.

  • Finally complicated and exciting Infrastructural projects: platformization of new product upcoming offering, development and integration of best development tools, projects like system-wide billing and so on.

As a Staff/Principal Software Engineer, you would be a significant part of one or more of these areas. The key challenges are:

  • Shipping quickly. AI is evolving weekly, so there's a ton of opportunity to be had. We want to move fast to capture those opportunities while they are still fresh and full of potential.

  • Zero to one. The platform is not a simple copycat. We have a vision for a deep platform/suite of capabilities that make it dramatically simpler for developers to scale and evolve their AI.

  • Realtime, online. As consumer applications become more capable of listening and talking, performance will matter, and AI has to adapt in realtime as well. These are bold but exciting challenges.

  • Multi-provider complexity at scale. Inworld Router must intelligently route across hundreds of models and providers while handling failover, sticky sessions, cost optimization, and conditional logic, all with minimal latency overhead. You'll design systems where every millisecond and every routing decision matters.

Finally, almost everything here is a collaboration with our sibling ML teams, since ML and AI are critical to providing the learning and adaptability central to this vision.

Please note: This is an IC-focused role. We are looking for someone who loves direct technical contribution alongside very capable peers.


What you’ll do:
  • Establish significant scope: Collaborate with the PMs, engineers and leads to determine the biggest product needs to focus on now.

  • Operate with technical autonomy: You have considerable leeway to suggest how to address a given focus area, including bringing in new technical dependencies or standards where it's the best choice.

  • Collaborate, execute, deliver: This is the core of the building loop. We aim to optimize for both speed and quality, despite it being decidedly non-obvious how to manage that tradeoff exceptionally well.

  • Reflect and drive improvements: Especially as a Staff Engineer, advocate for and realize system improvements, both related to and independent of key features.


Expected experience:

Must Haves

  • Excellent programming skills and experience in a statically typed backend programming language, preferably Go, Python, C++ or Rust

  • Experience developing and deploying cloud-based services to at least hundreds of qps (preferably more)

  • Experience with relational databases (PostgreSQL or MySQL)

  • Hands-on experience with caching (Redis or Memcached), pubsub/queues, data pipelines (Flink, Beam), and Cloud storage

  • Excellent verbal and written communication skills, can collaborate and coordinate with other roles and engineer with ease, trusted and well-regarded teammate

Bonus Qualifications

  • Experience building API gateways, routing/proxy layers, or multi-provider orchestration systems

  • Experience with analytics or timeseries databases (ClickHouse, Timescale, InfluxDB)

  • Experience with OpenTelemetry

  • Experience with C++

Candidates must be based in the SF Bay Area or willing to relocate (you will be working on-site in our South Bay office a few days a week).  

The US base salary range for this full-time position is $280,000 - $350,000. In addition to base pay, total compensation includes equity and benefits. Within the range, individual pay is determined by work location, level, and additional factors, including competencies, experience, and business needs. The base pay range is subject to change and may be modified in the future.

Inworld Jobs Privacy

HQ

Inworld AI Mountain View, California, USA Office

1975 W El Camino Real, Mountain View, CA, United States, 94040

Similar Jobs

3 Minutes Ago
Hybrid
San Francisco, CA, USA
Senior level
Senior level
Financial Services
Provides executive administrative support through complex calendar management, call screening, meeting and event coordination, domestic and international travel arrangements, invoice and expense processing, onboarding and offboarding support, document maintenance, and preparation of professional communications, spreadsheets, and presentations. The role requires discretion, strong organization, communication skills, Microsoft Office proficiency, and five days of on-site work.
Top Skills: MS Office
7 Minutes Ago
Hybrid
San Francisco, CA, USA
Senior level
Senior level
Artificial Intelligence • HR Tech • Information Technology • Machine Learning • Software • App development • Industrial
Own the full order-to-cash function, including invoicing and payment policy, credit decisions, collections, AR aging, customer holds, and outsourced collections management. Partner with Sales, Finance, Product, Engineering, and Operations to resolve aged receivables and automate credit, invoicing, PO matching, customer-portal delivery, cash application, and VMS billing workflows. Lead roadmap requirements, rollout, KPIs, and process improvements in a high-growth environment.
Top Skills: Ai-Native Finance Automation ToolsAr DashboardsCustomer PortalsNetSuiteQuickbooks Online (Qbo)Vms/Msp Billing Systems
Junior
Artificial Intelligence • Healthtech • Logistics • Social Impact • Software • Telehealth
Engages patients through high-volume inbound and outbound calls, educates them about ordered at-home healthcare services, schedules appointments, answers questions, and escalates concerns. The role requires approximately 150–200 outbound calls and 80–100 patient conversations daily, collaboration with healthcare professionals, and proficiency with EHR and healthcare software. Zendesk and Five9 experience are preferred.
Top Skills: Auto-Dialer SystemsElectronic Health Records (Ehr)Five9Zendesk

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account