Courses and programmes from universities and institutions worldwide.
Reinforcement Learning Engineer
Get alerts for roles like this
More Reinforcement Learning Engineer roles in United States — straight to your inbox. No account needed.
Applying to this role? Tailor your résumé to this job description in one click, then download it clean — no watermark, no subscription.
Job Description
Reinforcement Learning Engineer - Remote Bright Vision Technologies is a technology consulting and software development company delivering cloud, AI, data, and enterprise solutions across the United States.
This is a fantastic opportunity to join an established and well-respected organization offering tremendous career growth potential. Job Title: Reinforcement Learning Engineer Location: 100% Remote (U. S.)
Position Type: Full-time, Direct W2 Salary Range: $100,000–$150,000 Annually Experience Required: 6+ years Sponsorship: U. S. Citizens, Green Card Holders, EAD Holders, and H-1B transfer candidates are encouraged to apply. We are unable to sponsor new H-1B visa petitions for this position.
Job Summary
We are looking for a Reinforcement Learning Engineer to design, train, and deploy RL-based systems for high-impact decision-making problems where supervised learning alone is insufficient.
The role requires deep familiarity with modern reinforcement learning algorithms, simulation environments, reward modeling, and the engineering complexity of training and evaluating policies at scale.
The ideal candidate has both research depth and engineering pragmatism, with experience taking RL solutions out of the lab and into production where stability, safety, and ongoing improvement are critical.
Key Responsibilities
Design and implement reinforcement learning solutions for sequential decision-making problems in real and simulated environments. Develop, calibrate, and maintain simulation environments suitable for large-scale agent training.
Implement and evaluate modern RL algorithms including policy gradient, actor-critic, off-policy, and offline RL methods. Engineer reward functions and shaping strategies that align agent behavior with desired outcomes and safety constraints.
Apply offline RL and imitation learning techniques where exploration is costly or unsafe. Use RLHF, DPO, and related techniques for fine-tuning large language models when relevant. Build scalable tra
Required Skills
Partner picks for Reinforcement Learning Engineer in United States
Matched to the skills this page calls for and the candidate's location.
- edXVerified partnerPartner course provider
- UdemyVerified partnerPartner course provider
A marketplace of instructor-created courses across technology, business and creative skills.
- Partner competition
Hackathon platform used by student and community hackathons across India.
BengaluruVisit partner →
Partners are ResumeKart affiliates or institutes it works with; ResumeKart may earn a commission when a candidate enrols. Placement is decided by relevance, not payment. How ResumeKart earns
Prepare to Win This Role
Everything you need to ace the interview and negotiate top-of-band compensation.