> Markdown version of [/jobs/ext/2259306-machine-learning-speech-contract](https://www.wearedevelopers.com/jobs/ext/2259306-machine-learning-speech-contract). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Machine Learning - (Speech) - Contract - **Company:** microTECH Global Limited - **Location:** London, UK - **Experience:** Expert - **Salary:** £150,000.0 - £200,000.0 - **Contract:** Contract - **Skills:** Java (Programming Language), Agile Methodology, Artificial Intelligence, Amazon Web Services, Android Software Development, Microsoft Azure, C++ (Programming Language), Distributed Computing Environment, Python (Programming Language), Machine Learning, Natural Language Processing, Open Source Technology, Tensorflow, Signal Processing, Software Engineering, Systems Modeling Language, Speech Recognition, Google Cloud, Cloud Platform System, Pytorch, Generative AI, Git, Kotlin, Information Technology, HuggingFace, Speech Synthesis, Software Version Control - **Published:** August 26, 2026 - **Apply:** https://www.collegerecruiter.com/job/2815157532-machine-learning-speech-contract ## About the Role * MSc/PhD in Artificial Intelligence, Computer Science/Engineering, Electrical Engineering, Mathematics, or a related discipline. * Professional software development experience with Python (experience with C++, Java, or Kotlin is a plus). * Deep understanding of machine learning and deep learning fundamentals, including various architectures, training techniques, and evaluation metrics. * Strong experience in audio/speech processing, including speech recognition, speech enhancement, audio analysis, text-to-speech synthesis, and natural language processing. * Proficiency with machine learning frameworks such as TensorFlow or PyTorch. * Solid understanding of software engineering principles, including version control (Git), CI/CD pipelines, and agile development methodologies. * Excellent communication, collaboration, and problem-solving skills. * Demonstrated ability to translate research ideas into practical, production-ready solutions. * Experience with generative AI, particularly in the context of audio/speech technologies. * A strong publication record in top-tier machine learning, artificial intelligence, or signal processing conferences and journals (e.g., ICML, NeurIPS, ICLR, CVPR, SysML, INTERSPEECH, ICASSP, IEEE/ACM TASLP, IEEE TPAMI, JMLR). * Experience with open-source speech processing toolkits (e.g., Hugging Face Transformers, SpeechBrain, ESPnet, Kaldi, NeMo). * Experience developing and deploying AI models on Android mobile platforms. * Proven experience in building and maintaining large-scale, distributed training pipelines. * Experience with cloud computing platforms (e.g., AWS, Azure, GCP). ## Description Senior Machine Learning Research Engineer - Speech, Audio, and Generative AI This role is available on a permanent basis or as a 6-month contractor via an agency. The role is inside IR35. Role and Responsibilities * Drive the research, design, development, and evaluation of innovative AI algorithms and models, with a primary focus on audio and speech processing. * Lead the development of robust and scalable software solutions for deployment on flagship mobile devices. * Independently own and deliver significant components of complex research projects, from initial concept to production readiness. * Design, implement, and maintain high-quality, well-documented code, adhering to best software development practices. * Collaborate closely with a multi-disciplinary team of researchers and engineers, providing technical guidance and mentorship. * Proactively identify and address technical challenges, proposing creative solutions and ensuring the successful delivery of projects. * Contribute to the development of internal tools and infrastructure to support research and development efforts. Skills and Qualifications Required Skills * MSc/PhD in Artificial Intelligence, Computer Science/Engineering, Electrical Engineering, Mathematics, or a related discipline. * Professional software development experience with Python (experience with C++, Java, or Kotlin is a plus). * Deep understanding of machine learning and deep learning fundamentals, including various architectures, training techniques, and evaluation metrics. * Strong experience in audio/speech processing, including speech recognition, speech enhancement, audio analysis, text-to-speech synthesis, and natural language processing. * Proficiency with machine learning frameworks such as TensorFlow or PyTorch. * Solid understanding of software engineering principles, including version control (Git), CI/CD pipelines, and agile development methodologies. * Excellent communication, collaboration, and problem-solving skills. * Demonstrated ability to translate research ideas into practical, production-ready solutions. * Experience with generative AI, particularly in the context of audio/speech technologies. * A strong publication record in top-tier machine learning, artificial intelligence, or signal processing conferences and journals (e.g., ICML, NeurIPS, ICLR, CVPR, SysML, INTERSPEECH, ICASSP, IEEE/ACM TASLP, IEEE TPAMI, JMLR). * Experience with open-source speech processing toolkits (e.g., Hugging Face Transformers, SpeechBrain, ESPnet, Kaldi, NeMo). * Experience developing and deploying AI models on Android mobile platforms. * Proven experience in building and maintaining large-scale, distributed training pipelines. * Experience with cloud computing platforms (e.g., AWS, Azure, GCP). Seniority Level * Associate Employment Type * Contract Job Function * Software Development Location: London, England, United Kingdom Salary: $150,000 - $200,000 per annum ## Related Videos - [How a Small Team Shrank a Microsoft Monorepo by 94%](https://www.wearedevelopers.com/videos/1236-how-a-small-team-shrank-a-microsoft-monorepo-by-94) - [Kotlin Multiplatform - True power of native code reuse](https://www.wearedevelopers.com/videos/4-kotlin-multiplatform-true-power-of-native-code-reuse) - [Photonic Computing: Programming a New Class of AI Accelerators (incl. Live Coding)](https://www.wearedevelopers.com/videos/100196-photonic-computing-programming-a-new-class-of-ai-accelerators-incl-live-coding) - [Raise your voice!](https://www.wearedevelopers.com/videos/10-raise-your-voice) - [Serverless deployment of (large) NLP models ](https://www.wearedevelopers.com/videos/158-serverless-deployment-of-large-nlp-models) - [Git for Code Reviews](https://www.wearedevelopers.com/videos/429-git-for-code-reviews) ## Related Articles - [What Are Large Language Models?](https://www.wearedevelopers.com/magazine/304-what-are-large-language-models) - [MLOps And AI Driven Development](https://www.wearedevelopers.com/magazine/82-mlops-and-ai-driven-development) - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [Got AI ideas but no money? Here are 10 free ways to level up your AI skills with Google Cloud](https://www.wearedevelopers.com/magazine/600-got-ai-ideas-but-no-money-here-are-10-free-ways-to-level-up-your-ai-skills-with-google-cloud) - [MLOps – What’s the deal behind it?](https://www.wearedevelopers.com/magazine/125-mlops-what-s-the-deal-behind-it) - [Navigating the AI Shift](https://www.wearedevelopers.com/magazine/629-navigating-the-ai-shift)