Lattice Semiconductor Logo

Lattice Semiconductor

Sr Staff Engineer

Posted 7 Days Ago
Be an Early Applicant
In-Office
San Jose, CA, USA
221K-221K Annually
Senior level
In-Office
San Jose, CA, USA
221K-221K Annually
Senior level
Develop, fine-tune, evaluate, optimize, and deploy large language and multimodal models in local, on-premises, edge, and restricted environments. Build scalable training and inference pipelines, integrate models into production APIs, monitor performance and resource usage, and ensure security and compliance. Collaborate with engineering and business teams, define KPIs, document workflows, evaluate emerging AI techniques, and mentor other engineers.
The summary above was generated by AI
Lattice Overview

There is energy here…energy you can feel crackling at any of our international locations. It’s an energy generated by enthusiasm for our work, for our teams, for our results, and for our customers. Lattice is a worldwide community of engineers, designers, and manufacturing operations specialists in partnership with world-class sales, marketing, and support teams, who are developing programmable logic solutions that are changing the industry. Our focus is on R&D, product innovation, and customer service, and to that focus, we bring total commitment and a keenly sharp competitive personality.

Energy feeds on energy. If you flourish in a fast paced, results-oriented environment, if you want to achieve individual success within a “team first” organization, and if you believe you can contribute and succeed in a demanding yet collegial atmosphere, then Lattice may well be just what you’re looking for.

Job Description:

We are looking for a Sr. Staff Engineer to join our team.

Key Responsibilities

  • Train, fine-tune, and evaluate machine learning models, including LLMs, using techniques such as supervised fine-tuning (SFT), LoRA/QLoRA, and RLHF.

  • Optimize models for local inference through quantization, pruning, and distillation.

  • Deploy models on-prem or at the edge using frameworks such as PyTorch, TensorRT, ONNX, vLLM, or llama.cpp.

  • Build and maintain training and inference pipelines for reproducibility and scalability.

  • Integrate locally deployed models into production systems via APIs and internal services.

  • Monitor model performance, drift, latency, and resource utilization in production.

  • Collaborate with software engineers, infrastructure teams, and domain experts to deliver end-to-end AI solutions.

  • Ensure models meet security, privacy, and compliance requirements, especially in restricted or offline environments.

  • Document model architectures, training procedures, and deployment workflows.

Required Qualifications

  • Master’s or Ph.D. in computer science, Engineering, or a related field, or equivalent practical experience.

  • 8+ years of experience in AI and machine learning, with at least 3 years of experience working on LLMs, code generation, large-scale neural networks, RAG, or AI-powered automation.

  • Hands-on experience with LLMs (e.g., GPT-OSS, Nemotron, Kimi, Qwen, LLaMA, Mistral, Falcon, or similar open-weight models).

  • Proficiency in Python and ML frameworks such as PyTorch or TensorFlow.

  • Expertise in vector databases (FAISS, Weaviate, Chroma, Pinecone, Milvus) and retrieval models

  • Experience deploying models in local, on-prem, or resource constrained environments.

  • Experience with multi-agent AI systems (LangGraph, CrewAI, AutgoGen, OpenAI Assistants API) for autonomous coding tasks

  • Hands-on model development, working with business stakeholders to define KPIs and develop and deliver multi-modal (Text and Images) and ensemble models.

  • Solid understanding of model optimization techniques (quantization, batching, memory optimization).

  • Familiarity with Linux, containers (Docker), and basic cloud/on-prem infrastructure concepts.

  • Research & Innovation:

    • Continuous Learning: Commitment to stay updated with the latest research and advancements in AI and machine learning.

    • Innovation: Ability to think creatively and propose innovative solutions to complex problems.

Soft Skills:

  • Mentoring: Act as mentor/guide to a small set of engineers

  • Communication: Excellent verbal and written communication skills.

  • Adaptability: Ability to adapt to changing technologies and project requirements.

  • Team Player: Strong interpersonal skills and the ability to work well in a team environment.

Pay & Benefits

Consistent with Lattice Semiconductor values and applicable law, we provide the following information to promote pay transparency and equity. We have a market-based pay structure which varies by location.  Please note that the base pay range is a guideline, and our compensation range reflects the cost of labor in the U.S. geographic market based on the location of the role. Pay within these ranges varies and depends on job-related knowledge, skills, and relevant work experience. 

For candidates who receive and offer, the starting salary will vary based on various factors including, but not limited to, such qualifications as, skill level, competencies, and work location.  The range provided may represent a candidate range and may not reflect the full range for an individual tenured employee.

Base Pay Range

220800

In addition to base pay, this role may be eligible for variable/ incentive compensation and/ or equity.  In addition, this role is eligible for a comprehensive, competitive benefits package which may include healthcare and retirement plans, paid time off, and more! 

Additional Information:

This position requires a successful background and reference checks and satisfactory proof of your right to work in the United States.

Lattice recognizes that employees are its greatest asset and the driving force behind success in a highly competitive, global industry.  Lattice continually strives to provide a comprehensive compensation and benefits program to attract, retain, motivate, reward and celebrate the highest caliber employees in the industry. 


Lattice is an international, service-driven developer of innovative low cost, low power programmable design solutions.  Our global workforce, some 1,000 strong, shares a total commitment to customer success and an unbending will to win.  For more information about how our FPGA, CPLD and programmable power management  devices help our customers unlock their innovation, visit 
www.latticesemi.com.  You can also follow us via Twitter, Facebook, or RSS. At Lattice, we value the diversity of individuals, ideas, perspectives, insights and values, and what they bring to the workplace.  Applications are welcome from all qualified candidates.

 

As an E-Verify employer, we use this system to confirm the employment eligibility of all new hires in accordance with federal law. All applicants will be required to complete a Form I-9, Employment Eligibility Verification, upon hire. We do not use E-Verify to pre-screen job candidates and will comply with all E-Verify regulations.

Lattice Semiconductor San Jose, California, USA Office

2115 Onel Dr, San Jose, CA, United States, 95131

Similar Jobs

Yesterday
Hybrid
2 Locations
286K-392K Annually
Senior level
286K-392K Annually
Senior level
Fintech • Machine Learning • Payments • Software • Financial Services
Lead the architecture, development, deployment, optimization, governance, and observability of scalable AI platforms and systems. Build foundation model training, LLM inference, agentic workflows, similarity search, guardrails, and evaluation capabilities. Establish enterprise AI safety and transparency standards, drive long-term infrastructure strategy, mentor technical leaders, and guide organization-wide adoption of responsible AI solutions.
Top Skills: Ai AgentsAWSAws UltraclustersAzureC#C++CudaGoGCPHugging FaceJavaLarge Language ModelsMulti-Agent WorkflowsPythonPyTorchScalaSimilarity SearchVectordbs
Yesterday
Hybrid
Santa Clara, CA, USA
191K-334K Annually
Senior level
191K-334K Annually
Senior level
Artificial Intelligence • Cloud • HR Tech • Information Technology • Productivity • Software • Automation
Designs, builds, and operates production telephony and CCaaS connectivity for ServiceNow Voice. Integrates SIP, signaling, media, WebSocket, and real-time audio systems across carriers and platforms such as Genesys, NICE, Five9, Amazon Connect, and Twilio. Owns call reliability, observability, security, failover, customer-embedded deployments, technical specifications, and incident resolution. Provides architectural leadership, mentors engineers, and collaborates with product and voice AI teams without building agent logic.
Top Skills: Amazon ConnectCtiFive9GenesysMedia GatewaysNative Voice ControlsNicePbxPstnReal-Time Audio InterfacesRtpServicenow Voice/CtiSipSip TrunkingSrtpStreaming Media ApisTracing And Monitoring ToolsTwilioVoice CodecsWebrtcWebsockets
2 Days Ago
Hybrid
Santa Clara, CA, USA
191K-334K Annually
Senior level
191K-334K Annually
Senior level
Artificial Intelligence • Cloud • HR Tech • Information Technology • Productivity • Software • Automation
Provides technical vision and architecture for AI-native and cloud-native systems at organizational scale. Leads development of AI platforms, evaluation and monitoring frameworks, retrieval pipelines, security guardrails, and distributed systems. Drives ServiceNow platform strategy, technical roadmaps, architectural standards, production reliability, and responsible AI practices. Mentors senior engineers and leaders, influences executives and cross-functional teams, supports strategic customers, leads complex technical initiatives, and contributes to critical software development and innovation.
Top Skills: Agent DesignAWSAzureCloud-Native SystemsData PipelinesDistributed SystemsDockerEmbedding SystemsEvaluation FrameworksGCPGlidescriptInfrastructure As CodeJavaJavaScriptKubernetesLlm Orchestration FrameworksLlmsManaged Ai ServicesMicroservicesObservabilityProduction MonitoringPrompt EngineeringPythonRetrieval-Augmented Generation (Rag)ServicenowServicenow ApisVector Databases

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account