Skip to main content
Tallo logoTallo logo

Find Jobs

Find Jobs Near You – Available Work in Your Location

Skip to job details

Back to Results

Apply for this opportunity

To apply for this job, you'll continue to an external website or email application.

Medasource

Data Scientist

Career Insights for Data Scientist

See where this job fits in the broader career landscape. Knowing your career path helps you see what's possible from here.

Scorecard

Based on Pennsylvania data

Review key factors to help you decide if this role fits your goals. How is this calculated?

Were these scores useful?

What they do

A Data Scientist utilizes skills and experience to systematically answer questions using data to provide actionable recommendations. Commonly utilizes advanced statistical analysis and machine learning techniques. Common responsibilities also include data cleaning and data management.

$105,420 / year median in Pennsylvania

+18% projected growth

Explore Career

Job Description

Data Scientist / Data Engineer - AI/LLM & Metadata Harmonization Spring House, PA (Hybrid - 3 days/week onsite) As the Data Scientist / Data Engineer, you will work across business and technology teams at a Fortune 50 Life Sciences company to support metadata harmonization and AI/LLM-driven initiatives across its data science organization. This role blends data engineering, data curation, and applied AI work, and offers the opportunity to build automated pipelines and explore emerging LLM tooling as part of a high-visibility, cross-functional data integration effort.
MINIMUM QUALIFICATIONS
Master's or PhD in Data Science, Computer Science, or a related field Strong Python scripting skills, with prior experience building ETL pipelines Proficiency in Python and SQL for designing and developing automated ETL pipelines Experience with Information Retrieval, Search technologies, and NLP Hands-on experience with AI/LLM frameworks such as the OpenAI API, LangChain, and GPT models Prompt engineering experience
RESPONSIBILITIES
Design and develop automated ETL pipelines to extract, transform, and harmonize metadata across the data science team Define and maintain metadata standards and formatting conventions used across the broader team Map and reconcile data across multiple resource types, each containing different data structures, into a consistent format Build automation to match and validate large sets of fields across data sources Support a second initiative applying AI tools and prompt engineering to surface and understand initiatives from internal and public data sources, including BI data points and API-pulled data Transform raw internal and public data into standardized, usable formats Balance technical build work with data curation responsibilities across two concurrent projects
Pay:
$60.00 - $80.00 per hour
Benefits:
Dental insurance Health insurance Vision insurance Application Question(s): Do you have advance Python scripting skills? Do you have proficiency in Python and SQL for designing and developing automated ETL pipelines? Do you have hands-on experience with AI/LLM frameworks such as the OpenAI API, LangChain, and GPT models?
Education:
Master's (Preferred)
Work Location:
Hybrid remote in Spring House, PA 19477