Senior RL Engineer - Optimization & LLM Alignment (Winnipeg)

Senior RL Engineer - Optimization & LLM Alignment (Winnipeg)

19 Sep
|
Appit
|
Winnipeg

19 Sep

Appit

Winnipeg

APPIT Software Solutions in Montreal is seeking a Reinforcement Learning Engineer to design RL systems for enterprise optimization, building adaptive agents and RLHF alignment of large language models.

You will implement RL algorithms (PPO, SAC, DQN, MCTS), build simulation environments for training and evaluation, and collaborate with research teams to translate advances into production applications in a rapid-paced AI product company.

📌 Senior RL Engineer - Optimization & LLM Alignment (Winnipeg)
🏢 Appit
📍 Winnipeg

Reply to this offer

Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.

Subscribe to this job alert:

Get the latest job offers by email for: senior rl engineer - optimization & llm alignment (winnipeg) / winnipeg