Design, train, evaluate, and deploy reinforcement learning systems for complex decision-making problems. Develop reward functions, simulation environments, and neural network policies; scale training on GPU clusters; and move RL solutions from research into reliable production systems. The role requires strong theoretical and engineering expertise, with preferred experience in RLHF, multi-agent or hierarchical RL, robotics, autonomous driving, and open-source or published RL work.
RL Engineer – Remote
Bright Vision Technologies is a technology consulting and software development company delivering cloud, AI, data, and enterprise solutions across the United States.
This is a fantastic opportunity to join an established and well-respected organization offering tremendous career growth potential.
Job Title: RL Engineer
Location: 100% Remote (U.S.)
Position Type: Full-time, Direct W2
Salary Range: $80,000–$100,000 Annually
Experience Required: 6+ years
Sponsorship: U.S. Citizens, Green Card Holders, EAD Holders, and H-1B transfer candidates are encouraged to apply. We are unable to sponsor new H-1B visa petitions for this position.
Job Summary
We are looking for a RL Engineer to design, train, and deploy RL-based systems for high-impact decision-making problems where supervised learning alone is insufficient. The role requires deep familiarity with modern reinforcement learning algorithms, simulation environments, reward modeling, and the engineering complexity of training and evaluating policies at scale. The ideal candidate has both research depth and engineering pragmatism, with experience taking RL solutions out of the lab and into production where stability, safety, and ongoing improvement are critical.
Required Qualifications
Would you like to know more about this opportunity? For immediate consideration, please send your resume to [email protected]. Learn more about Bright Vision Technologies at www.bvteck.com.
Bright Vision Technologies is an Equal Opportunity Employer.
Bright Vision Technologies is a technology consulting and software development company delivering cloud, AI, data, and enterprise solutions across the United States.
This is a fantastic opportunity to join an established and well-respected organization offering tremendous career growth potential.
Job Title: RL Engineer
Location: 100% Remote (U.S.)
Position Type: Full-time, Direct W2
Salary Range: $80,000–$100,000 Annually
Experience Required: 6+ years
Sponsorship: U.S. Citizens, Green Card Holders, EAD Holders, and H-1B transfer candidates are encouraged to apply. We are unable to sponsor new H-1B visa petitions for this position.
Job Summary
We are looking for a RL Engineer to design, train, and deploy RL-based systems for high-impact decision-making problems where supervised learning alone is insufficient. The role requires deep familiarity with modern reinforcement learning algorithms, simulation environments, reward modeling, and the engineering complexity of training and evaluating policies at scale. The ideal candidate has both research depth and engineering pragmatism, with experience taking RL solutions out of the lab and into production where stability, safety, and ongoing improvement are critical.
Required Qualifications
- Master’s or PhD in Computer Science, Machine Learning, or a related field; or equivalent applied experience.
- Six or more years of combined RL research and engineering experience.
- Strong proficiency in Python and modern deep learning frameworks.
- Hands-on experience with at least one major RL library or in-house RL stack.
- Solid understanding of probability, optimization, and the theoretical foundations of RL.
- Experience designing and tuning reward functions in non-trivial environments.
- Familiarity with simulation environments and large-scale experience collection.
- Experience training neural network policies on GPU clusters.
- Strong written and verbal communication skills.
- Track record of shipping or publishing impactful RL work.
- Experience with RLHF for large language models.
- Familiarity with multi-agent RL or hierarchical RL.
- Exposure to robotics, control systems, or autonomous driving.
- Publications in RL or related research venues.
- Open-source contributions to RL libraries or environments.
Would you like to know more about this opportunity? For immediate consideration, please send your resume to [email protected]. Learn more about Bright Vision Technologies at www.bvteck.com.
Bright Vision Technologies is an Equal Opportunity Employer.
Similar Jobs
Artificial Intelligence • Generative AI
Build scalable platforms and pipelines that transform real-world, vendor, tutor, acquired, and synthetic data into realistic reinforcement-learning environments and tasks. Develop self-service APIs, tooling, automated quality checks, and data validation layers for internal and external contributors. Partner with research and ML platform teams to translate model capabilities into effective training environments and ensure trustworthy data production at scale.
Top Skills:
APIsData PipelinesData PlatformsDistributed SystemsMachine Learning InfrastructureReinforcement LearningSynthetic Data
Artificial Intelligence • Big Data • Machine Learning
Own the technical foundation for building, packaging, executing, verifying, and scaling reinforcement learning environments. Design sandboxed execution, rollout orchestration, trajectory capture, verifier frameworks, environment versioning, and authoring tools. Build high-throughput infrastructure and reliable reward signals, instrument real applications, develop task suites, and defend against reward hacking. Provide staff-level technical leadership across engineering and research teams while shipping complex production systems.
Top Skills:
AWSAzureCi/CdDockerFirecrackerGCPGoGrpoGvisorInfrastructure As CodeKubernetesMcpPpoPythonRayReactRlaifRlhfRlvrRustSglangTrlTypescriptVerlVirtual MachinesVllm
Artificial Intelligence • Generative AI
Design and iterate task sets, rewards, and environments for coding agents. Analyze agent traces and evaluations to identify failure modes, build systems to surface recurring behaviors, and convert one-off solutions into reusable training infrastructure and high-quality datasets. Partner with research teams to validate that datasets teach intended capabilities. The role requires strong software engineering fundamentals and comfort with infrastructure, data, or distributed systems.
Top Skills:
Data InfrastructureDistributed SystemsReinforcement Learning
What you need to know about the San Francisco Tech Scene
San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.
Key Facts About San Francisco Tech
- Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
- Major Tech Employers: Google, Apple, Salesforce, Meta
- Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
- Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
- Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
- Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine


.png)