Summary
Pokee AI is a company focused on reinforcement learning for AI agents, and they are seeking an RL AI Research Intern to join their research team. The role involves designing experiments, implementing algorithms, and contributing to core technology with mentorship from experienced researchers.
Responsibilities
- Investigate and prototype novel RL approaches for context selection, reward shaping, or policy optimization in agent workflows
- Design and execute experiments, producing clear analyses and actionable insights
- Implement and benchmark RL algorithms against internal and public baselines
- Collaborate with senior researchers and engineers in a fast-paced startup environment
- Document findings and present results to the broader team; strong work may lead to publication opportunities
Skills
- Currently pursuing a PhD or Master's degree in Reinforcement Learning, Machine Learning, Computer Science, or a related field
- Solid understanding of RL fundamentals (MDPs, policy gradients, Q-learning, actor-critic methods)
- Strong programming skills in Python with experience in PyTorch or similar frameworks
- Ability to independently design experiments and iterate quickly on ideas
- Excellent written and verbal communication skills
- Publications or preprints at top ML venues
- Experience with LLM fine-tuning, RLHF, or AI agent frameworks
- Familiarity with distributed training or large-scale experiment infrastructure
- Prior industry research internship experience
Qualifications
Must Haves
- Currently pursuing a PhD or Master's degree in Reinforcement Learning, Machine Learning, Computer Science, or a related field
- Solid understanding of RL fundamentals (MDPs, policy gradients, Q-learning, actor-critic methods)
- Strong programming skills in Python with experience in PyTorch or similar frameworks
- Ability to independently design experiments and iterate quickly on ideas
- Excellent written and verbal communication skills
Nice to Haves
- Publications or preprints at top ML venues
- Experience with LLM fine-tuning, RLHF, or AI agent frameworks
- Familiarity with distributed training or large-scale experiment infrastructure
- Prior industry research internship experience