Austin, TX Job Details Full-time | Contract | Internship $30 - $35 an hour 1 day ago Benefits Relocation assistance 401(k) Tuition reimbursement Qualifications Teamwork ETL pipeline development Spark AWS Hadoop Full Job Description Overview Join our innovative team as a Jr Data Engineer and become a vital contributor to our data-driven initiatives! In this dynamic role, you will support the development and maintenance of scalable data pipelines, manage large datasets, and collaborate across teams to enable insightful business intelligence. This position offers an exciting opportunity to grow your expertise in big data systems, cloud platforms, and data management technologies while working on impactful projects that shape strategic decision-making. If you are passionate about data engineering, eager to learn cutting-edge tools, and thrive in a fast-paced environment, this role is perfect for you! Duties Assist in designing, developing, and maintaining ETL (Extract, Transform, Load) pipelines to ensure seamless data flow across systems Support the integration of data from diverse sources such as AWS cloud services, Hadoop clusters, and SQL databases including Microsoft SQL Server and Oracle Collaborate with senior engineers to optimize data models and implement dimensional modeling techniques for efficient data warehousing Contribute to the development of data management solutions using tools like Informatica, Talend, Apache Hive, Spark, and Azure Data Lake Help monitor and troubleshoot big data systems ensuring high availability and performance of data pipelines Participate in the creation of business intelligence dashboards using Looker and other analytics tools to visualize key metrics Support agile development practices by documenting processes, writing scripts in Python or Shell Scripting (Bash), and assisting with API integrations such as RESTful APIs Skills Proficiency in SQL programming with experience in SQL databases like Microsoft SQL Server or Oracle; strong understanding of query management and database design Familiarity with big data technologies including Hadoop ecosystem components (Hadoop, Spark, Hive) and cloud-based platforms such as AWS or Azure Data Lake Knowledge of ETL tools like Informatica or Talend for data pipeline development and data warehousing design Experience with programming languages such as Python and Java for software development and automation tasks Understanding of data modeling principles including dimensional modeling for data warehouse architecture Ability to work with cloud databases and Big Data systems within a Public Cloud environment using Azure or AWS services Strong analysis skills coupled with experience in analytics tools like Looker or similar business intelligence platforms Basic scripting skills using Bash or Shell Scripting for automation tasks within Unix/Linux environments Familiarity with RESTful API integration for model training or data exchange purposes Knowledge of Agile methodologies to support iterative development cycles in a collaborative team setting Embark on this exciting journey where your contributions directly impact our ability to harness the power of data! We're committed to fostering growth, innovation, and success—join us today!