logo
AI Resume Tailoring Sign in to use this AI Cover Letter Sign in to use this

Job Description

AWS Data Center Capacity Delivery (DCCD) is looking for a Data Engineer to support data center construction globally. We work on the most challenging problems, with thousands of variables impacting the data center delivery — and we’re looking for talented people who want to help.

You’ll join a diverse team of software, hardware, and network engineers, construction specialists, security experts, operations managers, and other vital roles. You’ll collaborate with people across AWS to help us deliver the highest standards for safety and security while providing seemingly infinite capacity at the lowest possible cost for our customers. You’ll experience an inclusive culture that welcomes bold ideas and empowers you to own them to completion.

We’re looking for Data Engineer to help us grow our Data Lake and Data Warehouse Systems, which is being built using a serverless architecture, with 100% native AWS components including Redshift Spectrum, Athena, S3, Lambda, Glue, EMR, Kinesis, SNS, CloudWatch and more! We own a world-class data lake that is used to drive multi-billion dollar decisions on a regular cadence and we're looking to improve on filling the lake quickly, with as little human intervention needed and democratize the data in the lake.

Our Data Engineers build the ETL and analytics solutions for our internal customers to answer questions with data and drive critical improvements for the business. Our Data Engineers use best practices in software engineering, data management, data storage, data compute, and distributed systems. We are passionate about solving business problems with data!


Key job responsibilities
- Design and implement scalable, fault-tolerant data pipelines using AWS technologies and internal Amazon tools to extract, transform, and load data from multiple sources leveraging and implementing AI solutions as required.
- Collaborate cross-functionally with BIEs, Data Scientists, PMs, and SDEs to understand data requirements and deliver customized data solutions.
- Automate infrastructure deployment with CI/CD pipelines and ensure streamlined processes for deployment and maintenance.
- Ensure data quality through robust validation, cleansing, and deduplication techniques.
- Implement data governance standards, including access control, encryption, data retention, deletion policies, and audit mechanisms to ensure compliance and security.
- Continuously improve and optimize data pipelines and infrastructure, staying up to date with emerging technologies and implementing automation and monitoring tools.
- Build a scalable and reliable data platform supporting analytics for intuitive, self-service data products.
- Write high quality code and build scalable applications that interface with critical services and APIs to extract and process unstructured data
- Work with a range of data technologies, including Python, EMR, Spark, Iceberg, Airflow, and many AWS data services like Glue, Athena, Redshift to create end-to-end pipelines that consolidate data from disparate systems.

About the team
DCCD -CAT is a central data and analytics team within the DCCD tooling org that plays a pivotal role in supporting analytics for data center construction space suporting cost,controls and commissioning domains. We own data platform, reporting, dashboards, measurement, and analytical solutions for DCCD org.

Basic qualifications

- Bachelor's degree
- 3+ years of data engineering experience
- Experience with data modeling, warehousing and building ETL pipelines
- Experience with SQL
- Knowledge of professional software engineering & best practices for full software development life cycle, including coding standards, software architectures, code reviews, source control management, continuous deployments, testing, and operational excellence
- Knowledge of distributed systems as it pertains to data storage and computing
- Knowledge of batch and streaming data architectures like Kafka, Kinesis, Flink, Storm, Beam
- Experience as a data engineer or related specialty (e.g., software engineer, business intelligence engineer, data scientist) with a track record of manipulating, processing, and extracting value from large datasets
- Experience in at least one modern scripting or programming language, such as Python, Java, Scala, or NodeJS
- Experience with Apache Spark / Elastic Map Reduce

Preferred qualifications

- Experience with AWS technologies like Redshift, S3, AWS Glue, EMR, Kinesis, FireHose, Lambda, and IAM roles and permissions
- Experience with non-relational databases / data stores (object storage, document or key-value stores, graph databases, column-family databases)
- Master's degree in computer science, engineering, analytics, mathematics, statistics, IT or equivalent
- Experience programming with at least one modern language such as C++, C#, Java, Python, Golang, PowerShell, Ruby
- Experience building/operating highly available, distributed systems of data extraction, ingestion, and processing of large data sets
- Experience in translating business needs into detailed feature requirements

Amazon is an equal opportunity employer and does not discriminate on the basis of protected veteran status, disability, or other legally protected status.

Our inclusive culture empowers Amazonians to deliver the best results for our customers. If you have a disability and need a workplace accommodation or adjustment during the application and hiring process, including support for the interview or onboarding process, please visit https://amazon.jobs/content/en/how-we-hire/accommodations for more information. If the country/region you’re applying in isn’t listed, please contact your Recruiting Partner.

The base salary range for this position is listed below. Your Amazon package will include sign-on payments and restricted stock units (RSUs). Final compensation will be determined based on factors including experience, qualifications, and location. Amazon also offers comprehensive benefits including health insurance (medical, dental, vision, prescription, Basic Life & AD&D insurance and option for Supplemental life plans, EAP, Mental Health Support, Medical Advice Line, Flexible Spending Accounts, Adoption and Surrogacy Reimbursement coverage), 401(k) matching, paid time off, and parental leave. Learn more about our benefits at https://amazon.jobs/en/benefits.



USA, WA, Seattle - 132,100.00 - 178,800.00 USD annually
Do you match this job?

Here is what this employer asked for. Sign in and we will fill in your half.

  • Role Data Engineer II, Data Engineer,Data Center Capacity Delivery
  • Experience 3-4 years
  • Education Any
  • Work type On-site
  • Location United States
Check my match (free)
Amazon
Consumer Apps & Media · 500+ Members · Seattle, WA, United States

Amazon operates a broad portfolio spanning online retail, third-party marketplaces, logistics, cloud computing, advertising, devices, entertainment, and subscription services. Its consumer businesses include Amazon stores and Prime, while Amazon Web Services provides cloud infrastructure, databases, analytics, and AI services. The company also develops products such as Alexa and Kindle, produces and distributes media, and runs fulfillment and delivery networks. Amazon serves consumers, sellers, developers, enterprises, creators, and public-sector customers through interconnected commerce and technology platforms.

All jobs at Amazon
Job Overview

Approx. salary range

152K – 321K

Our estimate — this employer did not publish a salary

Our estimate, not the employer’s. Worked out from the middle half of 35 comparable roles on Neural Jobs that did publish a salary, in the same field, country and experience band. The real figure for this job may be different.

Eligibility
United States Right to work in the United States required.
Workplace
On-site
Job Posted:
1 day ago
Job Type
Full Time
Education
Any
Experience
3-4 years

Share This Job: