Cogniify Logo

Cogniify

Software Engineer AI/ML Systems - USA

Sorry, this job was removed at 09:07 a.m. (PST) on Saturday, Sep 19, 2026
Be an Early Applicant
In-Office
Santa Clara, CA, USA
150K-170K Annually
Senior level
In-Office
Santa Clara, CA, USA
150K-170K Annually
Senior level

Similar Jobs

An Hour Ago
Remote or Hybrid
United States
71K-88K Annually
Junior
71K-88K Annually
Junior
Digital Media • Gaming • Information Technology • Software • Sports • Esports • Big Data Analytics
Administer and enhance Salesforce for VIP teams by configuring flows, reports, dashboards, access, and data management. Gather stakeholder requirements, troubleshoot user issues, support adoption, and collaborate with development, analytics, and data engineering teams on integrations and technical solutions. Query and validate data using SQL and Snowflake, test system enhancements, and identify process improvements while maintaining Salesforce data accuracy and integrity.
Top Skills: ApexExcelGoogle SheetsSalesforceSalesforce Flow BuilderSnowflakeSQL
3 Hours Ago
In-Office
126K-165K Annually
Expert/Leader
126K-165K Annually
Expert/Leader
Cloud • Information Technology • Internet of Things • Machine Learning • Software • Cybersecurity • Infrastructure as a Service (IaaS)
Leads safety, strategy, financial performance, governance, operations, customer relationships, and service delivery across a defined geographic or business area. Oversees project teams, budgets, cost-efficiency initiatives, compliance, execution quality, performance reviews, coaching, customer escalations, and adjacent sales opportunities. Requires extensive project management and leadership experience with financial responsibility, strong stakeholder management skills, and regional travel as needed.
3 Hours Ago
Hybrid
San Jose, CA, USA
249K-399K Annually
Expert/Leader
249K-399K Annually
Expert/Leader
AdTech • eCommerce • Information Technology • Software • Travel • Generative AI
Lead AI-powered reporting and insights strategy, design reporting data architecture and pipelines, and implement LLM-powered workflows (RAG, prompt engineering, agentic pipelines). Integrate BI tools and enterprise AI platforms, establish governance for AI outputs, and provide technical leadership to deliver trusted executive insights and scalable self-service reporting.
Top Skills: AnthropicAws BedrockAzure OpenaiDatabricksDelta LakeFlinkGoogle Vertex AiIcebergLangchainLlamaindexLlmLookerPythonRagSnowflakeSparkSQLTableau
Develop and operate production AI/ML systems, including distributed model infrastructure, automated evaluation and testing, GPU optimization, secure CI/CD, and performance analytics. The role owns features end to end from architecture through production, performs model benchmarking and error analysis, and manages reliable deployment across multi-node clusters.
The summary above was generated by AI

The Role

We're seeking a Software Engineer specialized in AI/ML applications to independently drive the development, evaluation, deployment, and end-to-end lifecycle management of AI-powered systems. This role sits at the intersection of advanced AI application development, robust software engineering, and continuous automation: you'll build the systems, and you'll build the machinery that keeps them tested, secure, and running at scale.

This position goes deep on infrastructure and reliability: deploying and scaling open-source models across distributed GPU infrastructure, designing automated testing frameworks that cover both deterministic code and stochastic AI outputs, and implementing secure CI/CD pipelines that enforce quality on every merge.

This role is built for an engineer who takes high ownership: comfortable carrying a feature from ideation through architecture, build, evaluation, and production, and rigorous enough to prove with benchmarks and dashboards that the system actually works.

What You'll Do

  • Architect and scale AI infrastructure: deploy and scale open-source models using distributed orchestration frameworks (Kubernetes, Ray, or Slurm) to run highly available, fault-tolerant AI workloads across multi-node clusters.

  • Design and evaluate AI systems: run experiments, prompt-tune, evaluate, and deploy production-grade models and AI agents, with flexible mechanisms to benchmark performance and swap models quickly as use cases evolve.

  • Drive error and gap analysis: run comprehensive model benchmarks, perform deep error and gap analysis on model outputs, and build analytics dashboards that communicate system performance clearly to stakeholders.

  • Build automated testing at depth: develop extensive automated test suites that validate end-to-end application code, aggressively improving coverage across both standard software and stochastic AI outputs.

  • Automate DevSecOps: keep systems clean and secure by automating vulnerability scanning on GitLab Merge Requests and implementing fast remediation pipelines for identified issues.

  • Execute independently: own features from ideation to production, including architectural decisions, public/private repository synchronization, and open-source community interactions.

What We're Looking For

  • 6+ years of professional experience writing production-grade, asynchronous Python, with a strong focus on decoupled, clean system architecture and design patterns.

  • Deployment & orchestration: hands-on experience deploying, monitoring, and scaling models in production using Kubernetes, Ray, or Slurm, including multi-node cluster configurations.

  • Hardware & scaling optimization: strong understanding of GPU memory management and infrastructure-level tuning for high-throughput, low-latency AI inference workloads.

  • AI evaluation & frameworks: deep experience building with LangChain, Hugging Face libraries, MLOps, vLLM, and SGLang, with proven expertise in prompt engineering, automated model benchmarking, and systematic LLM evaluations.

  • Data analysis: proficient in Python-based analysis (pandas, NumPy, or similar), able to extract insights from evaluation results and communicate findings clearly to technical and non-technical audiences.

  • CI/CD & security automation: advanced knowledge of GitLab pipelines, including automated test jobs and vulnerability scanners integrated directly into the MR workflow.

  • Testing toolchains: expert familiarity with Python testing frameworks (PyTest), mocking libraries, and automated test generation approaches for AI workloads.

  • Advanced version control: high proficiency in advanced Git workflows, including rebase strategies, cryptographic commit signing, and complex public/private repository mirroring.

  • Bachelor's or Master's degree in Computer Science, Engineering, or a related field (or equivalent experience).

Salary Range: US East/West Coast: $150000 - $170000

Disclaimer: The base salary range is a guideline and may vary based on factors such as candidate experience, specialized skills, and geographical location. Actual compensation may include additional benefits and bonuses.

Perks and Benefits of Working With Us

  • Unlimited PTO.

  • Please ask us about our very generous parental leave, much above industry standards!

  • Entrepreneurial culture where pushing limits and taking risks is everyday business.

  • Open communication with management and company leadership.

  • Small, dynamic teams = massive impact.

  • Medical, Dental and Vision coverage for employees.

  • Access to Disability & Life insurance.

  • Mental health and wellbeing support.

  • Annual bonus program.

  • Employer Stock Purchase Program (ESPP).

  • Yearly team building experiences.

  • Mentorship and sponsorship opportunities.

  • Manager resources and support.

Cogniify is an equal opportunity employer. We celebrate diversity and are committed to creating an inclusive environment for all employees. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, disability, or any other protected characteristic

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account