Principal Developer, Oncology Data & AI Systems (2 available)
Johnson & Johnson is hiring a Principal Developer, Oncology Data & AI Systems (2 available). It pays $117k-$201k a year and Level rates it ; you can apply on Level.
AI in this role
Lead efforts to modernize data capture and build scalable data pipelines to power AI-ready data foundations in Oncology R&D.
At Johnson & Johnson, we believe health is everything. Our strength in healthcare innovation empowers us to build a world where complex diseases are prevented, treated, and cured, where treatments are smarter and less invasive, and solutions are personal. Through our expertise in Innovative Medicine and MedTech, we are uniquely positioned to innovate across the full spectrum of healthcare solutions today to deliver the breakthroughs of tomorrow, and profoundly impact health for humanity. Learn more at jnj.com.
As guided by Our Credo, Johnson & Johnson is responsible to our employees who work with us throughout the world. We provide an inclusive work environment where each person is considered as an individual. At Johnson & Johnson, we respect the diversity and dignity of our employees and recognize their merit.
Job Function:
Data Analytics & Computational SciencesJob Sub Function:
Data ScienceJob Category:
Scientific/TechnologyAll Job Posting Locations:
Cambridge, Massachusetts, United States of America, Raritan, New Jersey, United States of America, San Diego, California, United States of America, Spring House, Pennsylvania, United States of America, Titusville, New Jersey, United States of AmericaJob Description:
Our expertise in Innovative Medicine is informed and inspired by patients, whose insights fuel our science-based advancements. Visionaries like you work on teams that save lives by developing the medicines of tomorrow.
Join us in developing treatments, finding cures, and pioneering the path from lab to life while championing patients every step of the way.
Learn more at https://www.jnj.com/innovative-medicine
Johnson and Johnson Innovative Medicine is recruiting a Principal Developer, Oncology Data & AI Systems (2 available positions) to strengthen the AI-ready data foundation that powers Oncology R&D. This role will lead efforts to modernize and standardize data capture, translate scientific and business needs into engineering requirements, and design, build, and optimize scalable data pipelines and workflows that ensure data is high-quality, well-governed, interoperable, and fit for advanced analytics and AI/ML.
This role focuses on data science initiatives across Oncology R&D— including Clinical, Pre-Clinical, RWD and ‘omics platforms by enabling trusted, reusable datasets and foundational data products that accelerate downstream insights and model development. This role will be a leading data science contributor and creative problem solver with developing AI-ready data and other routinely used data applications that improve speed, reliability and impact of data-driven decision making for Oncology R&D.
This position will be located in either Cambridge, MA; Spring House, PA; Titusville, NJ; Raritan, NJ; or San Diego, CA (no remote option).
Key Responsibilities:
- Partner with Oncology R&D and Data Science stakeholders to identify, prioritize, and deliver high-impact data, AI/ML, and GenAI use cases that accelerate scientific discovery, study execution, and evidence generation.
- Own the end-to-end design, build, and lifecycle management of Oncology R&D data products, including requirements, architecture, ETL/ELT development, documentation, and operational support.
- Integrate and harmonize multi-modal R&D data across biomarker labs, translational platforms, clinical trials, real-world data/evidence (), genomics/other ‘omics, and pre-clinical research systems to create trusted, reusable, AI-ready datasets.
- Co-develop an AI-ready data ecosystem in collaboration with Data Product Engineers, Data Scientists, Knowledge Graph Engineers, Enterprise Architecture, and IT—enabling advanced analytics, ML, and GenAI applications.
- Translate complex scientific, clinical, and operational needs into scalable engineering solutions, ensuring alignment with Oncology R&D priorities, enterprise data standards, and interoperability requirements.
- Design and optimize data pipelines for structured and unstructured data, leveraging Python, R, SQL, AWS services, and other relevant technologies to improve throughput, reliability, and maintainability.
- Implement robust data quality and reliability controls, including validation frameworks, automated monitoring, and KPI-driven measurement of data product performance, adoption, and business impact.
- Establish strong data governance foundations by maintaining data lineage, metadata, and versioning practices that support transparency, traceability, reproducibility, and compliance with internal and regulatory expectations. Embed FAIR data principles and compliance requirements into workflows
Required Qualifications:
- Bachelor’s Degree in Computer Science, Engineering, Life Sciences, or other relevant field. Advanced degree preferred
- 5+ years of experience in data engineering, including data modeling, schema design and database architecture, preferably in the healthcare industry
- Demonstrated proficiency in data engineering tools such as Python, R and SQL for data processing, transformation, and automation across large-scale datasets
- Hands-on experience designing and operating cloud-based data platforms, preferably on AWS (e.g., Redshift, FSx, Glue, Lambda) or equivalent services
- Experience working with multiple data storage patterns, including relational, NoSQL/unstructured, and graph data technologies (e.g., knowledge graph–enabling platforms). Strong analytical and problem-solving skills, with the ability to troubleshoot complex data pipeline, quality, and performance issues in production environments.
- Proven capability to lead cross-functional delivery and continuous improvement initiatives with multidisciplinary, distributed teams; experience coordinating with external vendors/partners.
- Demonstrated strength in stakeholder management, including requirements discovery, business analysis, and planning; ability to translate conversations into clear user stories, engineering requirements, and executable delivery plans.
- Ability to manage multiple concurrent projects, prioritize work, exhibit organizational skills and flexibility to deliver maximum business value.
- Willingness to conduct periodic travel (<15% of time) to conferences and internal meetings.
Preferred Qualifications:
- Experience with healthcare data standards (e.g. CDISC, HL7, FHIR, SNOMED CT, OMOP, DICOM).
- Exposure to high dimensional data technologies and handling, including imaging.
- Familiarity with machine learning operations (MLOps) and GenAI model deployment.
This position will be located in either Spring House, PA; Titusville, NJ; Raritan, NJ; San Diego, CA; or Cambridge, MA, (no remote option), and may require up to approximately 10% travel.
Johnson & Johnson is an Equal Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, age, national origin, disability, protected veteran status or other characteristics protected by federal, state or local law. We actively seek qualified candidates who are protected veterans and individuals with disabilities as defined under VEVRAA and Section 503 of the Rehabilitation Act.
If you are under 18 years of age, you (the candidate) may need to obtain the necessary working papers or other documentation required by state law to start the assignment, as well as get a parent’s consent for the background check.
Johnson & Johnson is committed to providing an interview process that is inclusive of our applicants’ needs. If you are an individual with a disability and would like to request an accommodation, external applicants please contact us via https://www.jnj.com/contact-us/careers , internal employees contact AskGS to be directed to your accommodation resource.
The anticipated base pay range for this position is $117,000 to $201,250. The Company maintains highly competitive, performance-based compensation programs. Under current guidelines, this position is eligible for an annual performance bonus in accordance with the terms of the applicable plan. The annual performance bonus is a cash bonus intended to provide an incentive to achieve annual targeted results by rewarding for individual and the corporation’s performance over a calendar/performance year. Bonuses are awarded at the Company’s discretion on an individual basis. Employees and/or eligible dependents may be eligible to participate in the following Company sponsored employee benefit programs: medical, dental, vision, life insurance, short- and long-term disability, business accident insurance, and group legal insurance.
Employees may be eligible to participate in the Company’s consolidated retirement plan (pension) and savings plan (401(k)).
Employees are eligible for the following time off benefits:
Vacation – up to 120 hours per calendar year
Sick time - up to 40 hours per calendar year
Holiday pay, including Floating Holidays – up to 13 days per calendar year of Work, Personal and Family Time - up to 40 hours per calendar year
Additional information can be found through the link below. https://www.careers.jnj.com/employee-benefits
The compensation and benefits information set forth in this posting applies to candidates hired in the United States. Candidates hired outside the United States will be eligible for compensation and benefits in accordance with their local market.
#LI-SL
#JNJDataScience
#JNJIMRND-DS
#JRDDS
#LI-Hyrbid
Required Skills:
Preferred Skills:
Advanced Analytics, Coaching, Critical Thinking, Data Analysis, Data Privacy Standards, Data Quality, Data Reporting, Data Savvy, Data Science, Data Visualization, Digital Fluency, Econometric Models, Organizing, Process Improvements, Strategic Thinking, Technical Credibility, Workflow AnalysisHow we rate this
Principal Developer, Oncology Data & AI Systems (2 available) at Johnson & Johnson rates 65 out of 100 for how much of the daily work is AI. That makes it Works on AI (AI Level 3 of 4). The level is about AI in the job, not seniority.
Works on AI. The daily work is on AI products, without building the model.
- ●●●● Builds AI80 to 100
- ●●●○ Works on AI60 to 79
- ●●○○ Uses AI40 to 59
- ●○○○ Little AI0 to 39
Levels come from how often the tools, models and workflows of the role are named in the posting itself. Open the description and count.
Prepare for this job
A free preview built only from this posting: what it asks for, what you could be asked in an interview, and how to adjust your resume.
Skills and AI tools this role asks for
Questions you could be asked
- How do you monitor a model once it's live, and how do you know it needs retraining?
- Tell me about a project where data engineering was part of your work. What did you do?
- Tell me about a project where machine learning was part of your work. What did you do?
- Tell me about a project where data pipelines was part of your work. What did you do?
- Tell me about a project where data science was part of your work. What did you do?
Adapt your resume
- List these exact terms on your resume: ML Ops, Data Engineering, Machine learning, Data Pipelines, and Data Science. An applicant tracking system matches the wording, not the idea.
- Attach one line of real, concrete experience to at least one of them — a tool named with nothing behind it rarely survives a human read.
- Show where AI is part of your daily process, not a one-off project — this role expects it to be a running habit.
Want an expert to read your CV for this job?
Free. Send your CV and the role you want next. We reply by email within 2 to 4 business days.
Get new software engineering jobs (Works on AI ●●●○ or higher) by email
One email a week with the new software engineering jobs (Works on AI ●●●○ or higher), each rated for how much AI is in the work. No recruiter spam, unsubscribe in one click.
Free. One email a week. Unsubscribe in one click.
Similar roles
Software Engineering roles that work on AI, at other companies.
What kind of AI work fits you?
Answer 12 practical questions in about three minutes. Get a simple profile, the work it points to, and live roles to explore next.
Find my next step