> Markdown version of [/jobs/ext/2021507-principal-data-scientist-immunology](https://www.wearedevelopers.com/jobs/ext/2021507-principal-data-scientist-immunology). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Principal Data Scientist - Immunology - **Company:** Johnson & Johnson - **Location:** Titusville, FL, United States (Remote available) - **Experience:** Expert - **Salary:** $117,000.0 - $201,250.0 - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Data Analysis, Microsoft Azure, Bioinformatics, Health Informatics, Clinical Data Repository, Computer Programming, Computer Literacy, Continuous Delivery, Continuous Integration, Data Dictionary, Data Files, Data Governance, Data Visualization, DevOps, Document Management Systems, EHealth, Graph Database, Health Information Technology, Interoperability, Linked Data, Natural Language Processing, Neo4j, Resource Description Framework (RDF), Requirements Management, Semantic Web, SPARQL, SQL Databases, Data Processing, Data Storage Technologies, System Availability, Gitlab, Git, Containerization, Information Technology, Data Lineage, Graphql, Restful APIs, Docker, Jenkins - **Published:** August 11, 2026 - **Apply:** https://www.careerbuilder.com/job-details/principal-data-scientist-immunology-2-positions-titusville-nj--4ee88958-460f-49db-a64b-32fe31be372b ## About the Role * Desired Ph.D. or master's degree in bioengineering, computer science, IT, bioinformatics, physics, mathematics, or related fields, emphasis on semantic technologies for biomedical application. * 5+ years professional experience in health informatics. * Demonstrated experience in large-scale knowledge graphs construction, ontology development, pharmaceutical or healthcare domains integration. * Programming background in parser combinators, natural language processing, and linked data (RDF Triple Stores and property graphs). * Proficiency in semantic web technologies (e.g. SPARQL, RDF, OWL), familiarity with graph databases (Neo4j, Amazon Neptune). * Proven work with complex biomedical datasets (e.g. clinical, genomics, proteomics) * Proficiency in various data storage solutions (SQL, key-value, column, document, graph stores) and data modeling techniques (semantic data, ontologies, taxonomies). * Experience in CI/CD implementations, git usage, CI/CD stacks (Jenkins, GitLab, Azure DevOps), DevOps tools, metrics/monitoring, and containerization technologies (Docker, Singularity). * Demonstrated stakeholder management capabilities- including requirements gathering, business analysis and planning. Must have the capacity to translate discussions into user requirements and project plans. * Ability to manage a numerous projects simultaneously, prioritize work, exhibit organizational skills and flexibility to deliver maximum business value. * Willingness to conduct periodic travel (<15% of time) to conferences and internal meetings. This position will be located on site at one of our campuses in either Spring House PA, Cambridge MA, Titusville NJ, Raritan NJ, or San Diego, CA (NO fully remote option available). Occasional travel for cross-functional workshops, design sessions, and team meetings may be required., Advanced Analytics, Coaching, Critical Thinking, Data Analysis, Data Privacy Standards, Data Quality, Data Reporting, Data Savvy, Data Science, Data Visualization, Digital Fluency, Econometric Models, Organizing, Process Improvements, Strategic Thinking, Technical Credibility, Workflow Analysis, Artificial Intelligence (AI), Bioengineering, Bioinformatics, Biology, Biomedical Software, Biomedicine, Borland ObjectWindows Library (OWL) Programming Libraries, Business Analysis, Business Plan, Clinical Data, Clinical Research, Coaching, Compensation and Benefits, Computer Science, Conferences, Construction, Continuous Deployment/Delivery, Continuous Integration, Cross-Functional, Data Analysis, Data Modeling, Data Processing, Data Quality, Data Science, Data Sets, Data Storage, Data Visualization, Database Administration, DevOps, Disease, Disease Prevention and Control, Diversity, Docker, Document Management, Drug Development, Econometric Modeling, Establish Priorities, Genomics, Git, Graph Database Data Format, GraphQL, Health Information Technology, Healthcare, High Availability, Immunology, Internet Technology, Interoperability, Jenkins, Mathematics, Medicine, Metrics, Microsoft Windows Azure, Multitasking, Natural Language Processing (NLP), Neo4j, Ontology, Organizational Skills, Physics, Predictive Modeling, Process Improvement, Product Lifecycle, Project Planning, Proteomics, RDF (Resource Description Framework), REST (Representational State Transfer), Requirements Management, Research & Development (R&D), SPARQL, SQL (Structured Query Language), Taxonomies, Technical Strategy, Technical/Engineering Design, Willing to Travel, Workflow Analysis ## Description * Be a key contributor to the design and implementation of a scalable knowledge graph infrastructure focused on data standardization and interoperability, focusing on Immunology R&D data. * Apply graph-based data modeling for efficient Immunology R&D organization, integration and retrieval to ensure system flexibility and long-term maintainability. * Work with a larger community of Data Scientists, Clinical Scientists, and Discovery Scientists to standardize, curate and create AI-Ready data sets. * Curate and extend ontologies for clear mapping into established biomedical ontologies and controlled terminologies using resource description framework (RDF) standards. * Work with SPARQL/GraphQL/REST services; develop ingestion and curation pipelines to ingest, normalize and map concepts across data sources. * Extend and curate Immunology R&D-relevant ontologies (e.g., diseases, drugs, targets, pathways, etc.) and maintain synonyms, cross-references, and provenance. * Partner with cross-functional teams to enable NLP/RAG over graphs, features for predictive modeling and terminology services for search and study design tools. * Work with Data Science & Digital Health colleagues, IT and DevOps teams to deploy and manage the graph database infrastructure, focusing on high availability, scalability, and recovery operations specifically geared toward Immunology R&D needs and applications. * Draft and manage documentation, such as data dictionaries, data lineage, and data flow diagrams, to facilitate understanding of the knowledge graph., The anticipated base pay range for this position is $117,000 to $201,250. The Company maintains highly competitive, performance-based compensation programs. Under current guidelines, this position is eligible for an annual performance bonus in accordance with the terms of the applicable plan. The annual performance bonus is a cash bonus intended to provide an incentive to achieve annual targeted results by rewarding for individual and the corporation's performance over a calendar/performance year. Bonuses are awarded at the Company's discretion on an individual basis. Employees and/or eligible dependents may be eligible to participate in the following Company sponsored employee benefit programs: medical, dental, vision, life insurance, short- and long-term disability, business accident insurance, and group legal insurance. Employees may be eligible to participate in the Company's consolidated retirement plan (pension) and savings plan (401(k)). Employees are eligible for the following time off benefits: Vacation - up to 120 hours per calendar year Sick time - up to 40 hours per calendar year Holiday pay, including Floating Holidays - up to 13 days per calendar year of Work, Personal and Family Time - up to 40 hours per calendar year Additional information can be found through the link below. https://www.careers.jnj.com/employee-benefits The compensation and benefits information set forth in this posting applies to candidates hired in the United States. Candidates hired outside the United States will be eligible for compensation and benefits in accordance with their local market. ## Related Videos - [Geometric deep learning for drug discovery](https://www.wearedevelopers.com/videos/264-geometric-deep-learning-for-drug-discovery) - [How a Small Team Shrank a Microsoft Monorepo by 94%](https://www.wearedevelopers.com/videos/1236-how-a-small-team-shrank-a-microsoft-monorepo-by-94) - [Docker Compose: Rediscovered](https://www.wearedevelopers.com/videos/1978-docker-compose-rediscovered) - [Putting the Graph In GraphQL With The Neo4j GraphQL Library](https://www.wearedevelopers.com/videos/257-putting-the-graph-in-graphql-with-the-neo4j-graphql-library) - [Cyber Sleuth: Finding Hidden Connections in Cyber Data](https://www.wearedevelopers.com/videos/893-cyber-sleuth-finding-hidden-connections-in-cyber-data) - [Blueprints for Success: Steering a Global Data & AI Architecture](https://www.wearedevelopers.com/videos/1577-blueprints-for-success-steering-a-global-data-ai-architecture) ## Related Articles - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Best Coding Boot Camps in Germany](https://www.wearedevelopers.com/magazine/237-best-coding-boot-camps-in-germany) - [The Most Popular IT Jobs on the Market](https://www.wearedevelopers.com/magazine/376-the-most-popular-it-jobs-on-the-market) - [The Fastest-Growing Tech Sectors to Look Out for in 2025](https://www.wearedevelopers.com/magazine/373-the-fastest-growing-tech-sectors-to-look-out-for-in-2025) - [Data Analyst Salary in the UK](https://www.wearedevelopers.com/magazine/278-data-analyst-salary-in-the-uk) - [Graph and AI Trends 2026: Why Is AI Running but Not Yet Delivering?](https://www.wearedevelopers.com/magazine/680-graph-and-ai-trends-2026-why-is-ai-running-but-not-yet-delivering)