Architect Organization

LARMAUR, INC.
San Francisco, CA, United States
6 days ago
Apply on www.careerbuilder.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
8 years minimum
Compensation
$156,000.0 - $234,000.0
Working hours
Regular working hours

Tech stack

Agile Methodology Artificial Intelligence Amazon Web Services JIRA Microsoft Azure Unix Cloud Computing Configuration Management Information Systems Computer Programming Continuous Delivery Continuous Integration
+32 more
Data Architecture Data Validation Information Engineering Data Systems Data Warehousing Linux Distributed Computing Environment Python (Programming Language) Unix Shell Project Management Software Operational Databases Performance Tuning Query Optimization Standard Sql Requirements Management Shell Script SQL Databases Structured Text Unstructured Data Management of Software Versions Extensible Markup Language (XML) Data Processing Scripting Data Storage Technologies Large Language Models Git Build Management Pyspark Information Technology Data Management Software Version Control Jenkins

Job description

We are looking for an experienced Senior Data Architect to own data product architecture end to end, working closely together with Trinity domain experts to design and build data products that provide biopharma leaders the accurate, integrated evidence they need to inform high-stakes commercial decisions. This is a hands-on role. We are looking for an experienced architect who wants to both design and build. Much of our raw material is unstructured or loosely structured text data; the core responsibility of the data architect in this role is to lead the teams that convert this unstructured data to enterprise-quality data products.

What You’ll Do

  • Work closely with domain experts, data scientists, and AI engineers to translate business requirements into data product architecture.
  • Own the architecture for Trinity data product(s): tooling/stack, schemas, attribute inclusion/exclusion, grain and key design, entity resolution, data provenance, schema evolution strategy, etc.
  • Design and build robust normalization pipelines for heterogeneous and loosely-structured text-based sources
  • Build multi-stage pipelines that combine deterministic parsing, LLM-assisted extraction and inference, and human review and adjudication
  • If desired, there is significant opportunity to take on AI-assisted full stack development: for example, to design and build internal interfaces analysts use to review, approve, and modify pipeline outputs (optional, if of interest)

Requirements

  • 8+ years of hands-on data engineering and data architecture experience, including owning architecture and technical direction for complex data systems end-to-end, from scoping and design through delivery and production support
  • Bachelor’s degree in Computer Science, Data Engineering, Information Systems, or a related technical field.
  • Equivalent practical experience will be considered for strong candidates
  • Depth in data modeling, with practical judgment on grain, keys, entity resolution, data/attribute provenance, and schema evolution.
  • Strong SQL and relational and data warehouse expertise, including query and performance optimization
  • Strong Python for production data pipelines, with orchestration, versioning, testing, and observability practices you have applied at scale
  • Demonstrated experience normalizing messy semi-structured and unstructured sources - XML, PDFs, etc. - into stable, documented schemas (core requirement)
  • Ability to achieve significant productivity increases without sacrificing quality while leveraging AI coding assistants as part of the day-to-day engineering process
  • Excellent written and verbal communication skills; ability to present technical concepts clearly to non-technical stakeholders and engage directly with business teams

Additional Skills

  • Familiarity with PySpark and distributed data processing frameworks.
  • Hands-on experience with Git, Jenkins, and Code Pipeline for version control and CI/CD processes.
  • Knowledge of Unix/Linux environments and shell scripting.
  • Exposure to cloud platforms such as AWS, Azure, or GCP for data storage and processing.
  • Experience with data validation and data quality checks.
  • Understanding of Agile methodologies and project management tools like JIRA., Adjudication, Agile Programming Methodologies, Amazon Web Services (AWS), Artificial Intelligence (AI), Atlassian JIRA, Biology, Biotech and Pharmaceutical, Cloud Computing, Communication Skills, Computer Science, Continuous Deployment/Delivery, Continuous Integration, Data Management, Data Processing, Data Quality, Data Science, Data Storage, Data Warehousing, DataArchitect Data Modeling Tool, GCP (Good Clinical Practices), Git, Human Intelligence (HUMINT), Industry Standards, Information Technology & Information Systems, Jenkins, Linux Operating System, Machine Tool, Microsoft Windows Azure, Performance Tuning/Optimization, Presentation/Verbal Skills, Product Development, Product/Service Launch, Production Support, Productivity Management, Project Management Software, Python Programming/Scripting Language, Quality Assurance Methodology, Query Optimization, Requirements Management, SQL (Structured Query Language), Source Code/Configuration Management (SCM), Structured Data, Team Lead/Manager, Technical Presentation, Unix Operating Systems, Unix Shell Programming, Unstructured Data, Writing Skills, XML (EXtensible Markup Language)

Benefits & conditions

Trinity’s salary bands account for a wide range of factors that are considered in making compensation decisions including but not limited to skill sets and market demand for skills, level of experience and training, specific qualifications, geographic location, internal equity, and other business/organizational needs. The base salary range for this role is $156,000 - $234,000 per year. In addition to your base salary, you will also be eligible for an annual discretionary performance bonus.

About the company

Trinity powers the future of life sciences commercialization through the fusion of human and artificial intelligence. By blending deep therapeutic expertise and trusted human ingenuity with a purpose-built technology platform, Trinity accelerates clarity and confidence at every step of the commercialization journey-from pre-launch to scale to loss of exclusivity. For more than 30 years, the world’s leading pharmaceutical, biotech, and medtech companies have relied on Trinity’s foresight, execution, and partnership to deliver confident product launches, decisive market advantage, and measurable patient impact. During that time, Trinity expanded from its first office in Waltham, MA to 1,300 professionals across 14 offices and five continents, setting new industry standards in quality, responsiveness, and client partnership. For more information, visit Trinity at www.trinitylifesciences.com.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.careerbuilder.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

6:21 min

Investigating push inefficiencies with upstream Git experts

Jonathan Creamer · Coffee With Developers

3:05 min

Integrating an assistant application with Jira software

Felix Augenstein · LIVE

2:03 min

Microsoft integrating native Unix coreutils into Windows environments

Chris Heilmann Chris Heilmann +2 · LIVE

3:14 min

Structuring career paths and localized data architectures

Ulrich Wurstbauer +1 · LIVE

56 sec

Favorite git commands and the importance of patch commits

Eileen Uchitelle Eileen Uchitelle +1 · Coffee With Developers

3:09 min

Balancing data science skillings alongside systems engineering rigor

Nico Schmidt · LIVE

Videos

See all

Related articles

See all