Velaura AI, Inc. Logo

Velaura AI, Inc.

Principal AI SoC Runtime Software Architect

Posted 3 Hours Ago
Be an Early Applicant
In-Office
Santa Clara, CA, USA
200K-300K Annually
Expert/Leader
In-Office
Santa Clara, CA, USA
200K-300K Annually
Expert/Leader
Own the architecture and technical direction of a heterogeneous AI SoC runtime spanning sensor ingest, preprocessing, inference, postprocessing, and application delivery. Define execution, memory, synchronization, driver, firmware, compiler-runtime, observability, resilience, and validation architectures. Lead cross-functional engineers while contributing hands-on to production systems software, performance optimization, APIs, and low-level interfaces for robotics and edge AI workloads.
The summary above was generated by AI
Role Overview

We are looking for a Principal AI SoC Runtime Software Architect to own the software architecture that turns Velaura’s heterogeneous AI SoC into a coherent, high-performance execution platform.

This role will define and lead development of the end-to-end runtime spanning sensor ingest, preprocessing, AI inference, postprocessing, and delivery of results to robotics and other physical and edge AI applications. The runtime will coordinate execution and data movement across the SoC’s CPU cores, AI accelerator, vision and multimedia engines, and other embedded processors. The complete runtime must coordinate heterogeneous workloads, manage ownership, synchronization, and safe reuse of shared data buffers, minimize data movement, provide predictable low-latency execution, recover from failures, and expose cohesive APIs and observability to applications and SDK components.

The ideal candidate combines deep runtime and systems-software expertise with a strong understanding of heterogeneous compute, Linux kernel and driver interfaces, DMA and shared-memory architectures, and performance-sensitive AI or multimedia pipelines. This will be a hands-on principal architect and technical lead who directs engineers across runtime, kernel, driver, firmware, multimedia, and SDK development.


    Responsibilities
  • Define the SoC-wide execution model for coordinating workloads across heterogeneous compute and media engines, including dependency management, resource arbitration, priority and QoS, and concurrent pipeline behavior.

  • Set the architecture and technical direction for the AI inference runtime, including compiled-model execution, integration with industry-standard AI execution frameworks, and interfaces to applications and the broader Velaura SDK.

  • Own the end-to-end dataflow architecture for sensor-to-application pipelines, ensuring that camera, media, preprocessing, inference, and postprocessing components operate as an efficient and coherent system.

  • Define the SoC-wide memory and buffer-sharing architecture across user space, the kernel, and heterogeneous hardware engines, establishing clear ownership, coherency, isolation, synchronization, and lifecycle semantics.

  • Define the division of responsibility and interface contracts among the runtime, kernel drivers, firmware, and hardware engines, including execution, completion, telemetry, fault management, and recovery semantics.

  • Partner with the compiler team to define the compiler–runtime contract, ensuring compiled artifacts contain the metadata and execution information needed for the runtime to load, validate, execute, profile, and maintain compatibility across releases.

  • Establish a system-wide observability and performance architecture that correlates behavior across software and hardware layers and enables optimization against latency, throughput, bandwidth, power, utilization, and predictability goals.

  • Define the runtime resilience and validation architecture, including fault-containment and recovery policies, architecture-level acceptance criteria, and qualification across correctness, concurrency, compatibility, performance, and sustained workloads.

    Required Qualifications
  • Extensive experience designing and building production runtime systems, embedded middleware, multimedia frameworks, or other performance-critical systems software.

  • Strong C/C++ programming skills and demonstrated ability to architect and contribute hands-on to production runtime software spanning application-facing APIs, user-space libraries, and low-level driver, firmware, and hardware interfaces.

  • Strong understanding of heterogeneous and asynchronous execution, including command submission, queues, events, dependencies, synchronization, concurrency, scheduling, and resource management.

  • Strong understanding of device memory, DMA, IOMMU/SMMU, cache coherency, memory mapping, shared buffers, buffer lifetimes, and kernel/user-space memory interfaces.

  • Experience optimizing end-to-end data movement and execution across multiple hardware engines rather than focusing solely on individual kernels or accelerator performance.

  • Experience designing stable runtime APIs with well-defined compatibility, versioning, error handling, diagnostics, and recovery behavior.

  • Demonstrated ability to debug complex cross-layer correctness and performance problems using disciplined, data-driven methods, profiling, and tracing.

  • Demonstrated technical leadership across component and organizational boundaries, including translating system requirements into clear architectures, interfaces, implementation guidance, and validation strategies.

    Preferred Qualifications
  • Direct experience developing or extending AI inference runtimes or execution providers, such as ONNX Runtime, TensorRT-like runtimes, OpenVINO, TensorFlow Lite delegates, Qualcomm QNN/SNPE, TVM runtimes, or comparable systems for NPUs, GPUs, DSPs, or other accelerators.

  • Linux kernel development or upstream contribution experience involving device, accelerator, media, or shared-memory subsystems.

  • Experience with robotics, autonomous systems, edge AI, ROS 2, camera pipelines, ISP integration, V4L2/media, GStreamer, or other sensor-driven workloads.

  • Familiarity with compiled-model artifacts, quantized execution, tensor layouts, graph partitioning, memory planning, and compiler/runtime integration.

  • Experience with embedded Linux SDKs, production deployment, long-term runtime/API compatibility, or the security and isolation requirements of multi-process accelerator systems.

Why Velaura?

Velaura is building next-generation compute technology for cloud, edge, and Physical AI. Our solutions will enable robots, autonomous systems, drones, and other intelligent machines to operate efficiently in the physical world.
This is an opportunity to help build foundational technology at a time when the industry is undergoing fundamental change. You will work alongside experienced leaders, architects, engineers, and operators who have delivered industry-defining products across mobile, cloud, semiconductor, and AI platforms. If you enjoy solving difficult problems, working across disciplines, and helping shape the future of Physical AI, we would love to hear from you.

Equal Employment Opportunity and Accommodations

Velaura is an Equal Opportunity Employer that is committed to inclusion and diversity. Qualified applicants will receive consideration for employment without regard to race, color, religion, national origin, gender, sexual orientation, gender identity, disability or protected veteran status. We also take affirmative action to offer employment
opportunities to minorities, women, individuals with disabilities, and protected veterans.

Velaura is committed to working with qualified individuals with physical or mental disabilities. Applicants who would like to contact us regarding the accessibility of our website or who need special assistance or a reasonable accommodation for any part of the application or hiring process may contact us at: [email protected]. This contact
information is for accommodation requests only. Evaluation of requests for reasonable accommodation will be determined on a case-by-case basis.

Similar Jobs

11 Minutes Ago
Hybrid
Sunnyvale, CA, USA
86K-135K Annually
Mid level
86K-135K Annually
Mid level
Cloud • Computer Vision • Information Technology • Sales • Security • Cybersecurity
Build and manage AI technology alliances and go-to-market partnerships in cybersecurity. Responsibilities include researching and sourcing partners, managing relationships and integration roadmaps, coordinating onboarding, supporting partnership agreements, tracking pipeline and KPIs, enabling co-selling, and coordinating joint marketing activities. The role collaborates with Product, Engineering, Legal, Marketing, and Sales to grow partnerships and generate measurable pipeline and revenue impact.
Top Skills: Ai/MlAPIsCrm SystemsCrossbeamCybersecuritySalesforceSdks
Senior level
Machine Learning • Payments • Security • Software • Financial Services
The Security Specialist manages the workforce identity access management product and backlog, collaborates with engineering teams, supports delivery planning, and ensures compliance with security protocols while acting as a liaison to business stakeholders.
Top Skills: Oracle OimSailpointSaviynt
An Hour Ago
In-Office
San Francisco, CA, USA
95K-154K Annually
Mid level
95K-154K Annually
Mid level
Fintech • Information Technology • Payments • Sharing Economy • Financial Services • Cryptocurrency
Develop and maintain enterprise AI/ML solutions using Generative AI, RAG, and agentic frameworks on AWS. Design scalable architectures with Terraform, build CI/CD pipelines using GitLab, automate deployments with Python and Bash, and develop React-based internal tools and APIs. The role also involves troubleshooting, reliability, governance, security, compliance, project collaboration, technical mentorship, and occasional production support or travel.
Top Skills: Agentic AiAWSAws BedrockBashCloudwatchDockerEcsEksGenerative AiGitGitlab Ci/CdGrafanaLangchainLanggraphNoSQLPythonRagReactSQLTerraform

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account