Project Perseus | Data Labeling Analyst - Italian Speakers (Speech & Voice AI)

Welocalize, Inc.
Redmond, WA, United States
5 days ago
Apply on www.careerbuilder.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Compensation
$54,080.0 - $58,240.0
Working hours
Regular working hours
Languages
Italian

Tech stack

Training Data Artificial Intelligence Data Analysis Data Files Data Integrity Machine Learning Natural Language Processing Programming Languages

Job description

  • Execute high-volume data labeling and annotation tasks across speech and voice datasets
  • Follow detailed guidelines to ensure consistency, accuracy, and data integrity at scale
  • Work with audio and language data, including transcription, categorization, and tagging
  • Maintain strong throughput while meeting quality expectations
  • Escalate unclear or ambiguous cases appropriately
  • Adapt to evolving guidelines and workflows as systems and requirements change
  • Support baseline data production needs for AI training pipelines
  • Contribute to team calibrations and quality alignment sessions, In compliance with federal law, all persons hired will be required to verify identity and eligibility to work in the United States and to complete the required employment eligibility verification form upon hire. In addition, we employ anti-fraud checks to ensure all candidates meet the requirements of the program.

Requirements

Welo Data is looking for detail-oriented and reliable individuals to join our team as Data Labeling Analysts, supporting speech and voice AI systems.

This is a high-impact production role focused on building the datasets that power real-world AI systems. Youll be working with audio, speech, and language data helping ensure models are trained on accurate, well-structured, and representative inputs.

While this role is more execution-focused than evaluation-heavy roles, it still requires strong judgment, attention to detail, and consistency. The work sits at the intersection of language, data, and AI systems where precision and discipline matter at scale.

Were looking for people who are dependable, focused, and take pride in producing high-quality work, even across repetitive workflows., * Native-level fluency in Italian

  • Strong written communication skills and language fundamentals
  • 12 years of experience in data labeling, annotation, or content-focused work
  • Ability to follow detailed instructions and apply guidelines consistently
  • High attention to detail and ability to maintain accuracy in repetitive tasks
  • Comfort working in structured, process-driven environments
  • Ability to manage time effectively and maintain steady output
  • Willingness to ask questions and escalate when needed
  • Basic familiarity with AI, speech technology, or language data is a plus, To perform this job successfully, an individual must be able to perform each essential duty satisfactorily. The requirements listed below are representative of the knowledge, skill, and/or ability required. Reasonable accommodations may be made to enable individuals with disabilities to perform essential functions., Artificial Intelligence (AI), Artificial Intelligence (AI) Programming Languages, Communication Skills, Content Delivery/Distribution, Corporate Policies, Data Analysis, Data Modeling Language, Data Quality, Data Sets, Dental Insurance, Detail Oriented, Employee Assistance Plan, Human Intelligence (HUMINT), Intelligence Gathering, International Business, Italian Language, Large-Scale Systems, Localization, Machine Learning, Multilingual, Natural Language Processing (NLP), Production Support, Speech Technology, State Laws and Regulations, Time Management, Training Data Sets, Translation Services, Vision Plan, Voice Applications, Writing Skills

Benefits & conditions

  • Paid Vacation: 6 days
  • Paid Company Holidays: 2 days (Memorial Day and Labor Day)
  • Paid Sick Leave: accrued per applicable state law and company policy
  • Medical, Dental, and Vision Insurance (eligibility applies)
  • Health Savings Account (HSA)
  • 401(k) Retirement Plan
  • Employee Assistance Program
  • Additional voluntary benefits (life, accident, critical illness, etc.)

Onsite Perks (where applicable):

Free breakfast, lunch, and dinner

Stocked micro-kitchens with snacks and beverages

Commuter benefits, including shuttles and bike-to-work options

Unique campus features depending on location

$26 - $28 an hour

Why This Role

This role is a critical part of how modern AI systems are built. The data you produce directly impacts how speech and voice models understand and interact with real users.

Its a great entry point into AI operations, offering exposure to large-scale systems and the opportunity to build foundational experience in data, language, and AI workflows.

Please note that in order to verify work authorization as is required by Federal law (I-9 process), all new employees must complete a live video verification with their selected IDs and provide photos of these selected IDs within their first 3 days of employment.

About the company

As a trusted global transformation partner, Welocalize accelerates the global business journey by enabling brands and companies to reach, engage, and grow international audiences. Welocalize delivers multilingual content transformation services in translation, localization, and adaptation for over 250 languages with a growing network of over 400,000 in-country linguistic resources. Driving innovation in language services, Welocalize delivers high-quality training data transformation solutions for NLP-enabled machine learning by blending technology and human intelligence to collect, annotate, and evaluate all content types. Our team works across locations in North America, Europe, and Asia serving our global clients in the markets that matter to them. www.welocalize.com

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.careerbuilder.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:38 min

Reusing email software standards for HTTP file uploads

Imran Nazar · World Congress 2023

2:20 min

Architecting language translation with focused training data

Jaroslaw Kutylowski Jaroslaw Kutylowski +1 · World Congress 2023

1:48 min

Automating exploratory data analysis within training pipelines

Dora Petrella · World Congress 2023

3:13 min

Addressing language barriers and the data science talent deficit

Markus Hacker Markus Hacker +3 · World Congress 2024

5:21 min

Extracting and preprocessing HTML data into markdown files

Rainer Stropek Rainer Stropek · LIVE

4:20 min

Combating human workforce shortages with specialized language models

Markus Hacker Markus Hacker +3 · World Congress 2024

Videos

See all

Related articles

See all