Data Engineer Databricks SME

Job

PLANIT GROUP

Remote

Full-Time

Posted 3 days ago (Updated 9 hours ago) • Actively hiring

Expires 6/11/2026

Apply for this opportunity

This job application is on an outside website. Be sure to review the job posting there to verify it's the same.

See Job Scorecard

Review key factors to help you decide if the role fits your goals.

How is this calculated?

Pay Growth

out of 5

Not enough data

Not enough info to score pay or growth

Job Security

out of 5

Not enough data

Calculating job security score...

Total Score

out of 100

Average of individual scores

Were these scores useful?

Skill Insights

Compare your current skills to what this opportunity needs—we'll show you what you already have and what could strengthen your application.

Job Description

Back to Jobs Data Engineer

Databricks SME #28155184

Raleigh, NC Contract On-Site Flexibility/Remote:

100% Apply now Posted on May 9, 2026

Job Title:

Data Engineer

Databricks SME.

Location:

Raleigh, NC preferred

Remote candidate might be considered.

6-month assignment, with high probability of extension and longer term. This position is currently unfunded with award expected soon.

Description:

We are seeking a Data Engineer to support our client with data ingestion, data deduplication and data tagging for migration of a large-scale data environment into Databricks. The ideal candidate will also bring hands-on expertise in end-to-end data pipeline management , including data ingestion from diverse sources, de-duplication of large-scale datasets, and data tagging to support downstream analytics, governance, and machine learning workflows. Roles and Responsibilities (including but not limited to)

Design, develop, and maintain scalable data ingestion pipelines to onboard structured, semi-structured, and unstructured data from batch and streaming sources (e.g., APIs, databases, flat files, message queues) into the Azure/Databricks environment.
Implement de-duplication strategies across large-scale datasets using deterministic and probabilistic matching techniques to ensure data integrity and reduce redundancy within the Data Lake.
Develop and enforce data tagging frameworks to classify, label, and annotate datasets with appropriate metadata (e.g., sensitivity, source, domain, lineage) to support data governance, discoverability, and compliance requirements.
Assist with Operationalizing deployments and support of Cloud services for ETL Operations. This will include standardizing and automating processes and workflows, creating documentation/knowledge articles, and overall assisting Operations staff who have limited experience in Cloud.
Written and oral presentations to high-level CIO management on status of current efforts.
Possesses skills and experience related to business management, systems engineering, operations research, and management engineering. Typically has specialization in a particular technology or business application. Keeps abreast of technological developments and industry trends.
Assist with deployment, configuration, and management of Azure Cloud environment.
Assist with migration efforts of existing ETL jobs into Azure/Databricks cloud environment.
Ability to share optimization and efficiencies with the larger team and management.
Ability to automate solutions to repetitive problems/tasks. Basic Qualifications
Must be eligible for a Position of Public Trust, including U.S. citizenship or permanent residency, five years of U.S. residency, and no more than six months of international travel in the past five years (excluding travel for U.S
based work).
Bachelor's degree and 13 years of experience. A degree from an accredited College/University in the applicable field of services is preferred. Four additional years of relevant experience in lieu of a college degree is required. If Degree is not in the applicable field, then four additional years of related experience is required.
3+ years demonstrated experience designing and implementing data ingestion pipelines using tools such as Azure Data Factory, Apache Kafka, Apache NiFi, Spark Structured Streaming, or equivalent technologies.
3+ years of experience applying de-duplication techniques at scale, including record linkage, fuzzy matching, and entity resolution across structured and unstructured datasets.
3+ Hands-on experience with data tagging and metadata management , including the use of tagging schemas, data catalogs (e.g., Azure Purview, Apache Atlas), and automated classification tools to support data governance and lineage tracking.
3 + Demonstrated experience working with unstructured data.
2 + years of experience in using Databricks or other Spark-based platforms.
Fluency in at least one scripting language (Python, Perl, Ruby, or equivalent).

Desired Skills:

Integration of Git in continuous deployment and experience with DevOps monitoring tools.
Experience with one or more of the following products and technologies: SAS, Python, C++, Hadoop, SQL Database/Coding, Teradata, Oracle, Amazon S3, Apache Spark, Machine Learning, Natural Language Processing, and visualization tools such as Tableau, Strategy and QLIK.
Strong skills and experience in Cloud Operations support in Azure.

Apply now

Similar remote jobs

Job
Psychologist- Top Market Pay - New Hyde Park, NY
LH
LifeStance Health
New Hyde Park, NY
Posted2 days ago
Updated9 hours ago
Job
Sr. Azure Architect
T
TEKsystems
Atlanta, GA
Posted2 days ago
Updated9 hours ago
Job
Watershed Stewardship Manager (Temp)
AC
Albemarle County Public Schools
Charlottesville, VA
Posted2 days ago
Updated9 hours ago
Job
Data Analyst - Technical - Senior
IH
Intermountain Health
Frankfort, KY
Posted2 days ago
Updated9 hours ago
Job
Senior Amazon Channel Manager
S
STERRY
Posted2 days ago
Updated9 hours ago

Similar jobs in Raleigh, NC

Job
Floor Staff
TB
THE BLUE ROOM
Raleigh, NC
Posted2 days ago
Updated9 hours ago
Job
Part Time Educator | North Hills
LA
lululemon athletica
Raleigh, NC
Posted2 days ago
Updated9 hours ago
Job
Executive Meeting Manager-The Westin Raleigh-Durham Airport
TW
The Westin Raleigh-Durham Airport
Raleigh, NC
Posted2 days ago
Updated9 hours ago
Job
Dual Sales and Catering Manager
HH
Hyatt House RDU / Brier Creek
Raleigh, NC
Posted2 days ago
Updated9 hours ago
Job
Financial Coordinator
C
Confidential
Raleigh, NC
Posted2 days ago
Updated9 hours ago

Similar jobs in North Carolina

Job
Occupational Therapist (OT)
PR
Powerback Rehabilitation
Pinehurst, NC
Posted2 days ago
Updated9 hours ago
Job
Floor Staff
TB
THE BLUE ROOM
Raleigh, NC
Posted2 days ago
Updated9 hours ago
Job
MEAT MANAGER
HF
HOUCHENS FOOD GROUP INC
Murphy, NC
Posted2 days ago
Updated9 hours ago
Job
Housekeeper Part Time-101020
EM
ESA Management, LLC
Wilmington, NC
Posted2 days ago
Updated9 hours ago
Job
Part Time Educator | North Hills
LA
lululemon athletica
Raleigh, NC
Posted2 days ago
Updated9 hours ago