Bright Vision Technologies

Reinforcement Learning Engineer

Ann Arbor, Michigan · Posted 3 days ago

Opens brightvisiontechnologies.applytojob.com

Get a version of your resume written for this job.

Salary
Not listed
Job type
Full-time
Work mode
Not specified
Source
Jazzhr (employer's hiring system)

Skills mentioned

Python, Machine Learning, Deep Learning

About the role

Reinforcement Learning Engineer – Remote

Bright Vision Technologies is a technology consulting and software development company delivering cloud, AI, data, and enterprise solutions across the United States.
This is a fantastic opportunity to join an established and well-respected organization offering tremendous career growth potential.

Job Title: Reinforcement Learning Engineer
Location: 100% Remote (U.S.)
Position Type: Full-time, Direct W2
Salary Range: $96,000–$120,000 Annually
Experience Required: 8+ years

Sponsorship: U.S. Citizens, Green Card Holders, EAD Holders, and H-1B transfer candidates are encouraged to apply. We are unable to sponsor new H-1B visa petitions for this position.

Job Summary
We are looking for a Reinforcement Learning Engineer to design, train, and deploy RL-based systems for high-impact decision-making problems where supervised learning alone is insufficient. The role requires deep familiarity with modern reinforcement learning algorithms, simulation environments, reward modeling, and the engineering complexity of training and evaluating policies at scale. The ideal candidate has both research depth and engineering pragmatism, with experience taking RL solutions out of the lab and into production where stability, safety, and ongoing improvement are critical.

Required Qualifications
  • Master’s or PhD in Computer Science, Machine Learning, or a related field; or equivalent applied experience.
  • Six or more years of combined RL research and engineering experience.
  • Strong proficiency in Python and modern deep learning frameworks.
  • Hands-on experience with at least one major RL library or in-house RL stack.
  • Solid understanding of probability, optimization, and the theoretical foundations of RL.
  • Experience designing and tuning reward functions in non-trivial environments.
  • Familiarity with simulation environments and large-scale experience collection.
  • Experience training neural network policies on GPU clusters.
  • Strong written and verbal communication skills.
  • Track record of shipping or publishing impactful RL work.
Preferred Qualifications
  • Experience with RLHF for large language models.
  • Familiarity with multi-agent RL or hierarchical RL.
  • Exposure to robotics, control systems, or autonomous driving.
  • Publications in RL or related research venues.
  • Open-source contributions to RL libraries or environments.
How to Apply
Would you like to know more about this opportunity? For immediate consideration, please send your resume to hilda@bvteck.com. Learn more about Bright Vision Technologies at www.bvteck.com.
Bright Vision Technologies is an Equal Opportunity Employer.

 

Job ID jz-brightvisiontechnologies-20260921203201_a1iylasdrsbrcvhu · Original posting ↗