We are looking for a Data Engineer to support data-focused initiatives with a strong emphasis on controls, process clarity, and technical documentation. This is a Long-term Contract position expected to begin as a 3-4 month engagement at 40 hours per week, with remote work flexibility. The ideal candidate will help strengthen data workflows, improve reliability across engineering processes, and create well-organized documentation that supports ongoing delivery and compliance.
Responsibilities:
- Design, build, and maintain data pipelines that support reliable movement and transformation of information across platforms.
- Develop clear technical documentation, data process records, and control-related artifacts to improve transparency and audit readiness.
- Use Python, Apache Spark, and ETL frameworks to prepare, cleanse, and transform large datasets for downstream consumption.
- Work with Hadoop- and Kafka-based environments to support scalable data processing and streaming or batch integration needs.
- Review existing data workflows to identify gaps in controls, consistency, and documentation quality, then recommend practical improvements.
- Collaborate with cross-functional stakeholders to clarify data requirements, align engineering deliverables, and support operational continuity.
- Monitor pipeline performance and troubleshoot data issues to maintain dependable processing and accurate outputs.