> Markdown version of [/jobs/ext/3017213-humanoid-robot-reinforcement-learning-researcher](https://www.wearedevelopers.com/jobs/ext/3017213-humanoid-robot-reinforcement-learning-researcher). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Humanoid Robot Reinforcement Learning Researcher - **Company:** REK PARTNERS, INC. - **Location:** San Francisco, CA, United States - **Experience:** Experienced - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Algorithm Design, C++ (Programming Language), Cloud Computing, Computer Clusters, Communications Protocols, Python (Programming Language), Kinematics, Microsoft Dynamics, Motion Capture, Software Engineering, Reinforcement Learning, Data Logging, Scripting, Graphics Processing Unit (GPU), Pytorch, Low Latency - **Published:** September 20, 2026 - **Apply:** https://www.careerbuilder.com/job-details/reinforcement-learning-researcher-humanoid-san-francisco-ca--bdd6f050-9ef0-47b6-8504-b607ca51af8e ## About the Role * 2+ years of hands-on experience in reinforcement learning for humanoid robots at a university lab, research institute, or corporation. * Deep understanding of deep RL algorithms (PPO, SAC, TD3, DDPG, etc.). * Experience using Isaac Gym, MuJoCo, PyBullet, or Gazebo for simulation training. * Strong software engineering skills in Python, PyTorch, and C++. * Understanding of robot kinematics, dynamics, and control systems. * Ability to run large-scale experiments efficiently and interpret quantitative results. * Strong communication skills and comfort working in an experimental, cross-disciplinary team. Preferred / Bonus Qualifications * Experience with Unitree G1 EDU * Background in teleoperation, imitation learning, or motion retargeting. * Familiarity with low-latency communication protocols and embedded robot control. * Understanding of VR systems, motion capture, or human-in-the-loop RL. * Mandarin proficiency a major plus for collaboration with Chinese robotics partners and manufacturers., Algorithms, Artificial Intelligence (AI), Benchmarking, C++ Programming Language, Cloud Computing, Communication Skills, Communications Protocols, Control Systems, Embedded Systems, Engineering, GPU (Graphics Processing Unit), Human Interaction, Mandarin Chinese Language, Performance Metrics, Physics, Policy Evaluation, Preferred Provider Organization (PPO), Publications, Python Programming/Scripting Language, Quality Management, Reinforcement Learning, Research Laboratory, Research Skills, Robotics, Simulation, Software Engineering, Startup, Strategic Planning, Testing, White Papers ## Description Were seeking a Humanoid Robot Reinforcement Learning Researcher to develop and train advanced control policies for our REK humanoid robots. Youll work closely with REKs robotics, simulation, and teleoperation teams to design reinforcement learning pipelines that improve movement quality, stability, responsiveness, and adaptability all optimized for real-time fighting performance. This role sits at the intersection of research and applied engineering: youll design algorithms, implement simulation environments, train agents, and test them on full-scale humanoid robots. Responsibilities * RL Algorithm Development Design and implement state-of-the-art reinforcement learning algorithms for humanoid locomotion, balance, and reactive movement. * Simulation Environments Build and customize physics-based simulation environments (Isaac Gym, MuJoCo, PyBullet, etc.) for efficient training and domain randomization. * Sim-to-Real Transfer Develop robust transfer strategies that ensure trained policies perform reliably on physical humanoid robots. * Policy Evaluation Define performance metrics, run experiments, and benchmark results for both simulated and real-world tests. * Integration & Collaboration Work closely with teleoperation and control system teams to blend RL policies with operator input in hybrid control architectures. * Research & Publication Stay current with cutting-edge humanoid and robotics RL research and contribute to internal whitepapers or external publications as appropriate. * Data & Infrastructure Maintain scalable training pipelines and data logging systems using GPU clusters or cloud resources. ## Related Videos - [Robots Among us: Advances in AI for Everyday Androids](https://www.wearedevelopers.com/videos/100230-robots-among-us-advances-in-ai-for-everyday-androids) - [How I built my own intelligent Robot Arm from Scratch](https://www.wearedevelopers.com/videos/100097-how-i-built-my-own-intelligent-robot-arm-from-scratch) - [JavaScript? No. Java Scripts! - Scripting with Java](https://www.wearedevelopers.com/videos/2094-javascript-no-java-scripts-scripting-with-java) - [Photonic Computing: Programming a New Class of AI Accelerators (incl. Live Coding)](https://www.wearedevelopers.com/videos/100196-photonic-computing-programming-a-new-class-of-ai-accelerators-incl-live-coding) - [How Robots Learn to be Robots](https://www.wearedevelopers.com/videos/1632-how-robots-learn-to-be-robots) - [Robots are coming into the wild! Full-Stack Robotics Engineers, be ready!](https://www.wearedevelopers.com/videos/479-robots-are-coming-into-the-wild-full-stack-robotics-engineers-be-ready) ## Related Articles - [MLOps And AI Driven Development](https://www.wearedevelopers.com/magazine/82-mlops-and-ai-driven-development) - [DeepSeek R1 vs ChatGPT o1: How Do They Compare?](https://www.wearedevelopers.com/magazine/542-deepseek-r1-vs-chatgpt-o1-how-do-they-compare) - [Got AI ideas but no money? Here are 10 free ways to level up your AI skills with Google Cloud](https://www.wearedevelopers.com/magazine/600-got-ai-ideas-but-no-money-here-are-10-free-ways-to-level-up-your-ai-skills-with-google-cloud) - [Résumé-Driven Development: How IT trends affect the job market for software developers](https://www.wearedevelopers.com/magazine/59-resume-driven-development-how-it-trends-affect-the-job-market-for-software-developers) - [Dev Digest 137 - AI'm not sure about this](https://www.wearedevelopers.com/magazine/485-dev-digest-137-ai-m-not-sure-about-this) - [Coffee with Developers - Maria Apazoglou - Making AI understandable for all in production](https://www.wearedevelopers.com/magazine/475-coffee-with-developers-maria-apazoglou-making-ai-understandable-for-all-in-production)