Required Education
- Bachelor's degree in Computer Science Information Technology Engineering Data Science or a related fieldRequired Experience
- 7+ years of experience in data engineering software development or a related technology field
- Strong hands-on experience with Python
- Strong experience with SQL and relational databases
- Strong experience with data modeling
- Hands-on experience with Apache Spark or PySpark
- Experience developing scalable data pipelines and data processing solutions
- Experience working with large and complex datasets
- Experience with data integration and transformation
- Experience developing enterprise data solutions
- Experience working in Agile development environmentsTechnical Skills
- Python
- SQL•Data Modeling•Apache Spark•PySpark•Data Engineering•Data Pipelines•ETL and Data Transformation•Data Integration•Relational Databases•Database Design•Git• REST APIsPreferred Skills•Generative AI•Machine Learning•Google ADK•Large Language Models•AI Agent Development•Python-based AI and ML development•Experience integrating AI capabilities into data solutions•Cloud-based data platformsSoft Skills•Strong analytical and problem-solving skills•Excellent communication and collaboration skills•Strong attention to detail•Ability to work with technical and business stakeholders•Ability to translate business requirements into scalable data solutions•Strong ownership and accountability•Ability to manage multiple priorities•Ability to learn and adopt emerging technologiesJob SummaryThe Senior Python Data Engineer will design develop and support scalable data solutions and data-intensive applications within an enterprise environment.
The role will focus on Python data engineering, data modeling, SQL, and Apache Spark while supporting data processing, integration, and analytics initiatives.
The ideal candidate will have strong experience building data pipelines, developing data models, working with large datasets, and creating scalable data processing solutions using Python and Spark. Experience with Generative AI, Machine Learning, and Google ADK is highly desirable for supporting emerging AI and data capabilities.
Responsibilities
- Design develop and maintain scalable data pipelines using Python and Apache Spark
- Develop data processing and transformation solutions for large and complex datasets
- Design and maintain data models to support business and analytical requirements
- Develop optimize and maintain complex SQL queries
- Build data integration processes across multiple enterprise data sources
- Perform data transformation validation and quality checks
- Develop reusable Python-based data engineering components
- Optimize data pipelines and processing jobs for performance and scalability
- Collaborate with application developers data architects analysts and business stakeholders
- Translate business and technical requirements into scalable data engineering solutions
- Troubleshoot data pipeline and application issues and perform root cause analysis
- Participate in code reviews and maintain development standards
- Develop unit and integration tests for data processing solutions
- Support production deployments and resolve data-related production issues
- Maintain technical documentation for data pipelines models and processes
- Identify opportunities to improve data quality performance automation and operational efficiency
- Support initiatives involving Generative AI and Machine Learning where applicable
- Explore and implement AI-enabled data solutions using Python and Google ADK
- Follow enterprise security data governance and development standards
- Participate in Agile ceremonies and development activities
- Only those lawfully authorized to work in the designated country associated with the position will be considered.
- Please note that all Position start dates and duration are estimates and may be reduced or lengthened based upon a client's business needs and requirements.
•