Weekday, Inc. Logo

Weekday, Inc.

MLOps Engineer (JAX, PyTorch, Pallas/Triton)

Reposted 22 Days Ago
Remote
Hiring Remotely in United States
70-110 Hourly
Junior
Remote
Hiring Remotely in United States
70-110 Hourly
Junior
Remote full-time MLOps Engineer role to design, evaluate, and improve ML infrastructure and large-scale training systems. Develop and review challenging MLOps tasks, optimize GPU kernels (Pallas/Triton), create evaluation rubrics, and collaborate with researchers to improve model reasoning and training pipelines. Requires production JAX/PyTorch experience and strong distributed training knowledge.
The summary above was generated by AI

This role is for one of our clients

Compensation: $70-$110 per hour

Join a cutting-edge AI research initiative at the forefront of Generative AI and contribute to the development of next-generation Large Language Models. We are seeking experienced MLOps Engineers with deep expertise in modern machine learning frameworks, large-scale training infrastructure, and kernel-level optimization.

In this role, you'll leverage your knowledge of JAX, PyTorch, and custom GPU kernel programming (Pallas/Triton) to create, evaluate, and refine high-quality technical tasks that help train frontier AI systems. You'll collaborate with AI researchers and engineering teams to improve model reasoning across MLOps, distributed training, and ML infrastructure topics.

This is a full-time, 40-hour-per-week remote engagement requiring full weekday availability.


RequirementsKey Responsibilities
  • Partner with research and engineering teams to strengthen AI model capabilities in MLOps, ML infrastructure, and large-scale training systems.
  • Design challenging, real-world MLOps and machine learning systems tasks that reflect production engineering scenarios.
  • Develop accurate, well-documented solutions to complex ML infrastructure and training pipeline problems.
  • Review and evaluate technical tasks and AI-generated solutions, providing clear and actionable written feedback.
  • Create detailed evaluation rubrics and scoring frameworks for topics including:
    • Distributed training architectures
    • ML pipeline design
    • Infrastructure optimization
    • Kernel-level programming
    • Performance tuning
  • Collaborate with fellow subject matter experts to maintain consistency, quality, and technical accuracy across training datasets.
  • Contribute domain expertise to improve the reasoning capabilities of advanced AI systems.
Required Qualifications
  • Minimum 2 years of professional experience in MLOps, Machine Learning Infrastructure, or ML Systems Engineering within a recognized technology organization.
  • Hands-on production experience with JAX and/or PyTorch in large-scale machine learning environments.
  • Practical experience developing or optimizing custom GPU kernels using Pallas (JAX) or Triton.
  • Strong understanding of distributed training systems, model optimization, and scalable ML infrastructure.
  • Demonstrated career growth and increasing technical responsibility.
  • Availability to work 40 hours per week during standard weekday business hours.
  • Excellent written communication skills with the ability to clearly explain technical concepts and architectural decisions.
Preferred Skills
  • Experience designing and optimizing large-scale ML training pipelines.
  • Knowledge of distributed computing and GPU performance optimization.
  • Familiarity with evaluation methodologies for AI models and ML systems.
  • Experience collaborating with research teams on advanced machine learning projects.
  • Passion for advancing AI infrastructure and frontier model development.
Why Join
  • Help build and improve next-generation Large Language Models.
  • Work alongside leading AI researchers and experienced machine learning engineers.
  • Apply your expertise to high-impact projects involving large-scale ML systems and infrastructure.
  • Contribute directly to the development of cutting-edge AI technologies.
  • Enjoy a fully remote engagement with meaningful technical challenges.
Equal Opportunity

We welcome applications from qualified professionals regardless of legally protected characteristics and are committed to providing reasonable accommodations throughout the application and engagement process upon request.

Contract & Payment Terms
  • Engagement is offered on an independent contractor basis.
  • This is a fully remote opportunity that can be completed according to your own schedule.
  • Project duration may be extended, shortened, or concluded based on project requirements and individual performance.
  • The engagement does not require access to confidential or proprietary information belonging to any current employer, client, or institution.
  • Payments are processed weekly through Stripe or Wise based on approved work completed.
  • Please note: Applicants requiring H-1B sponsorship or participating in the STEM OPT program are not eligible for this opportunity.

Similar Jobs

11 Minutes Ago
Easy Apply
Remote or Hybrid
United States
Easy Apply
134K-178K Annually
Entry level
134K-178K Annually
Entry level
eCommerce • Healthtech • Kids + Family • Retail • Social Media
Leads payer onboarding, transitions, ongoing payer operations, licensing, credentialing, audits, site inspections, and regulatory compliance for Babylist Health. Builds scalable processes, monitors payer performance, manages commercial and Medicaid operations, program-manages new business launches, and partners with billing, legal, product, contracting, and customer support. Manages and develops two direct reports while establishing the operational foundation for expanding clinical services.
Top Skills: Artificial IntelligenceElectronic Funds Transfer (Eft)MedicaidRevenue Cycle Management (Rcm)
12 Minutes Ago
Easy Apply
Remote
United States
Easy Apply
Entry level
Entry level
Cloud • Security • Software • Cybersecurity • Automation
Sources and engages passive candidates for sales roles across the AMER region, primarily Account Executives. Builds talent pipelines, maps competitive markets, screens candidates, partners with recruiters and hiring managers, maintains applicant tracking data, and improves outbound sourcing strategies and candidate experience in a fully remote environment.
Top Skills: AIGitlabGreenhouseLinkedInLinkedin Recruiter
13 Minutes Ago
Remote or Hybrid
136K-245K Annually
Expert/Leader
136K-245K Annually
Expert/Leader
Blockchain • Fintech • Mobile • Payments • Software • Financial Services
Manage enterprise KYC due diligence policy and strategy across Block’s business lines. Develop and maintain KYC standards, assess regulatory changes, guide implementation, define AI/ML risk-segmentation and identity-verification controls, monitor program performance, advise stakeholders, support audits and examinations, and remediate compliance risks including model bias, drift, false positives, false negatives, and explainability gaps.
Top Skills: Ai/Ml

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account