Data Scientist

Semrush
Municipality of Madrid, Spain
4 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Working hours
Regular working hours
Languages
English
Experience level
Senior

Job location

Municipality of Madrid, Spain

Tech stack

Artificial Intelligence
Algorithm Design
Big Data
Software Quality
Python
PostgreSQL
Machine Learning
Natural Language Processing
Named Entity Recognition
Performance Tuning
Standard Sql
Search Technologies
Google Cloud Platform
Large Language Models
Gitlab-ci
Production Code
Vertica
Docker
Web Api

Job description

  • Build and improve LLM- and NLP-based systems used in AI Search optimization products

  • Design evaluation frameworks and benchmarks for LLM outputs, prompts, and model behavior

  • Work on brand extraction, domain mapping, entity extraction, and related algorithmic tasks

  • Optimize LLM pipelines for quality, cost, throughput, and reliability

  • Write clean, production-quality Python code

  • Collaborate with Product, Engineering, and other Data Scientists to turn product problems into working technical solutions Examples of our projects:

  • Brand Extraction and Domain Mapping - extracting brand names, mapping corresponding domains, and other relevant entities from LLM-generated responses

  • LLM quality evaluation and optimization - building benchmarks to evaluate prompt and model quality, while improving throughput, latency, and cost

  • Algorithm development - designing and improving algorithms for product aliases, brand hierarchies, and other domain-specific problems, We are a Data Science team focused on LLM- and NLP-powered products for AI Search optimization. We work on AIO, AI Summarization, Brand Extraction and Domain Mapping, building AI-driven solutions for SEO analysis, entity extraction, and domain understanding. Our stack:

  • Python

  • Google Cloud Platform

  • ClickHouse, PostgreSQL

  • ClearML, Docker, LangFuse, DVC

  • LLM APIs: OpenAI, Google, Anthropic, Perplexity

About the perks

  • Unlimited PTO
  • Hobby & team building budget allowance
  • Employee Support Program
  • Loss of family member financial aid
  • Employee Resource Groups

Requirements

Move together. Raise the bar. Learn fast-grow faster. That's the default. And here's what else is needed to succeed in this role: Hard Skills:

  • 5+ years of experience as a Data Scientist, Machine Learning Engineer, or in a similar role

  • Strong knowledge of machine learning, statistics, and classical NLP techniques

  • Strong Python and SQL skills

  • Practical experience working with LLMs or LLM-based systems

  • Understanding of how to evaluate model quality using datasets, benchmarks, or experiments

  • Ability to write clean, maintainable, and reliable code

  • Ownership mindset and ability to work with ambiguous product problems Soft Skills:

  • You care about delivering working solutions, not just experiments

  • You balance execution speed with code quality, reliability, and maintainability

  • You communicate clearly and collaborate well across Data Science, Product, and Engineering, + Hands-on experience with LLM PEFT (Parameter-Efficient Fine-Tuning) methods

  • Experience with RAG, AI agents, or agent-based systems

  • Experience deploying ML or LLM-based solutions in production

  • Experience with GCP, Vertex AI, or GitLab CI

  • Experience with asynchronous Python, high-throughput API calls, or large-scale data processing

  • Experience with monitoring, experiment tracking, or reproducible research practices

Apply for this position