> Markdown version of [/jobs/ext/2971109-python-engineer-remote-hired](https://www.wearedevelopers.com/jobs/ext/2971109-python-engineer-remote-hired). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Python Engineer (Remote) - Hired - **Company:** Hired - **Location:** Gijón, Spain (Remote available) - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Python (Programming Language), Reinforcement Learning, Large Language Models, Data Generation - **Published:** September 18, 2026 - **Apply:** https://www.buscojobs.com.es/python-engineer-remote-hired-en-gijon-ID-372178858 ## About the Role * Proficiency in Python and related frameworks/libraries for AI/ML tasks. * Experience with supervised fine-tuning (SFT) and reinforcement learning with human feedback (RLHF). * Strong understanding of evaluation strategies and benchmarking processes for AI models. * Ability to design and implement Python code for data generation and model optimization. * Familiarity with AI model response evaluation and ranking methodologies. More About the Opportunity: This role offers a unique opportunity to work with a global leader in artificial intelligence, contributing to the advancement of large language models. Candidates will collaborate with top researchers and engineers in the field. Equal Opportunity Employer: We hire based on skills and expertise. All qualified candidates are welcome regardless of background, experience, or prior employment history. Applications are reviewed solely on demonstrated technical ability and qualifications. ## Description Role:Python Engineer (Remote)Location:Remote (Work from Anywhere)Role Overview:We are hiring for one of our clients, seeking a Senior Python Developer to assist a foundational LLM company in enhancing their large language models.The goal is to provide high-quality proprietary data for fine-tuning and benchmarking model performance.Key Responsibilities:* Design, develop, and maintain efficient, high-quality Python code to train and optimize AI models.* Conduct evaluations to benchmark model performance and analyze results for continuous improvement.* Evaluate and rank AI model responses to user queries across diverse domains, ensuring alignment with predefined criteria.* Lead efforts in supervised fine-tuning, including creating and maintaining high-quality, task-specific datasets.* Collaborate with researchers and annotators to execute reinforcement learning with human feedback and refine reward models.Required Skills & Qualifications:* Proficiency in Python and related frameworks/libraries for AI/ML tasks.* Experience with supervised fine-tuning (SFT) and reinforcement learning with human feedback (RLHF).* Strong understanding of evaluation strategies and benchmarking processes for AI models.* Ability to design and implement Python code for data generation and model optimization.* Familiarity with AI model response evaluation and ranking methodologies.More About the Opportunity:This role offers a unique opportunity to work with a global leader in artificial intelligence, contributing to the advancement of large language models.Candidates will collaborate with top researchers and engineers in the field.Equal Opportunity Employer:We hire based on skills and expertise.All qualified candidates are welcome regardless of background, experience, or prior employment history.Applications are reviewed solely on demonstrated technical ability and qualifications.Apply Now! ## Related Videos - [Enhancing AI-based Robotics with Simulation Workflows](https://www.wearedevelopers.com/videos/472-enhancing-ai-based-robotics-with-simulation-workflows) - [AI in the Open and in Browsers - Tarek Ziadé](https://www.wearedevelopers.com/videos/1787-ai-in-the-open-and-in-browsers-tarek-ziade) - [On the straight and narrow path - How to get cars to drive themselves using reinforcement learning and trajectory optimization](https://www.wearedevelopers.com/videos/205-on-the-straight-and-narrow-path-how-to-get-cars-to-drive-themselves-using-reinforcement-learning-and-trajectory-optimization) - [Creating Industry ready solutions with LLM Models](https://www.wearedevelopers.com/videos/899-creating-industry-ready-solutions-with-llm-models) - [Bringing the power of AI to your application.](https://www.wearedevelopers.com/videos/1010-bringing-the-power-of-ai-to-your-application) - [Your next 10x engineer isn't in your city. Refactor accordingly.](https://www.wearedevelopers.com/videos/100103-your-next-10x-engineer-isn-t-in-your-city-refactor-accordingly) ## Related Articles - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [What Are Large Language Models?](https://www.wearedevelopers.com/magazine/304-what-are-large-language-models) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Prompt Engineering is a Job of the Past](https://www.wearedevelopers.com/magazine/342-prompt-engineering-is-a-job-of-the-past) - [MLOps And AI Driven Development](https://www.wearedevelopers.com/magazine/82-mlops-and-ai-driven-development) - [Dev Digest 137 - AI'm not sure about this](https://www.wearedevelopers.com/magazine/485-dev-digest-137-ai-m-not-sure-about-this)