What are the responsibilities and job description for the Mid-Level Data Engineer, Surface Transportation position at Jobright.ai?
Jobright is an AI-powered career platform that helps job seekers discover the top opportunities in the US. We are NOT a staffing agency. Jobright does not hire directly for these positions. We connect you with verified openings from employers you can trust.
Job Summary:
Amazon is seeking a Data Engineer to join their Surface Transportation team, which is focused on building the largest transportation network. The role involves designing and maintaining data pipelines, managing data infrastructure, and collaborating with various teams to create scalable data solutions.
Responsibilities:
• Design, build, and maintain batch and streaming data pipelines using modern Big Data technologies such as AWS Redshift, S3, Glue, Lake Formation, Athena, Kinesis, Lambda, etc.
• Manage data infrastructure as code using AWS CloudFormation and the AWS Cloud Development Kit (CDK)
• Automate processes using scripting languages such as Python, Scala, etc.
• Optimize, tune, and govern data warehouses and data lake environments to ensure performance for large user bases
• Work collaboratively with software engineers, data scientists, business analysts, and other internal partners to identify opportunities and build scalable solutions
• Possess strong verbal and written communication skills, be self-driven, and deliver high quality results in a fast-paced environment
Qualifications:
Required:
• Experience as a data engineer or related specialty (e.g., software engineer, business intelligence engineer, data scientist) with a track record of manipulating, processing, and extracting value from large datasets
• Knowledge of cloud services such as AWS or equivalent
• Experience with data modeling, warehousing and building ETL pipelines
• Experience with one or more query language (e.g., SQL, PL/SQL, DDL, MDX, HiveQL, SparkSQL, Scala)
• Experience with one or more scripting language (e.g., Python, KornShell, Scala)
• Design, build, and maintain batch and streaming data pipelines using modern Big Data technologies such as AWS Redshift, S3, Glue, Lake Formation, Athena, Kinesis, Lambda, etc.
• Manage data infrastructure as code using AWS CloudFormation and the AWS Cloud Development Kit (CDK)
• Automate processes using scripting languages such as Python, Scala, etc.
• Optimize, tune, and govern data warehouses and data lake environments to ensure performance for large user bases
• Work collaboratively with software engineers, data scientists, business analysts, and other internal partners to identify opportunities and build scalable solutions
• Possess strong verbal and written communication skills, be self-driven, and deliver high quality results in a fast-paced environment
Preferred:
• Experience with big data technologies such as: Hadoop, Hive, Spark, EMR
Company:
Amazon is a tech firm with a focus on e-commerce, cloud computing, digital streaming, and artificial intelligence. Founded in 1994, the company is headquartered in Seattle, Washington, USA, with a team of 10001 employees. The company is currently Public Company. Amazon has a track record of offering H1B sponsorships.