AWS Data Engineer

BIG DATA INC.
Atlanta, GA, United States
5 days ago
Apply on www.dice.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Working hours
Regular working hours
Job source

Tech stack

Airflow Amazon Web Services Amazon S3 Big Data Data as a Services Extract Transform Load (ETL) Python (Programming Language) Cloud Services Standard Sql Enterprise Data Management Snowflake Git
+8 more
Data Lakes Pyspark AWS Glue AWS Data Analytics Apache Kafka Restful APIs Amazon Elastic Mapreduce (EMR) Docker

Job description

We are seeking an AWS Data Engineer with expertise in big data processing, cloud-native data engineering, and streaming architecture. The ideal candidate should possess strong experience building enterprise data platforms using AWS analytics services., * Develop scalable AWS data pipelines.

  • Build ETL jobs using AWS Glue.
  • Develop PySpark applications on EMR.
  • Design Redshift data warehouse solutions.
  • Develop streaming pipelines using Kafka.
  • Build data orchestration workflows using Airflow.
  • Optimize cloud data processing.
  • Develop REST APIs for data services.
  • Implement data quality monitoring.

Requirements

  • AWS Glue
  • Amazon EMR
  • Redshift
  • Athena
  • S3
  • Python
  • PySpark
  • SQL
  • Kafka
  • Apache Airflow
  • Docker
  • Git

Preferred Skills

  • Snowflake
  • dbt
  • Iceberg
  • Delta Lake
  • Lambda
  • Step Functions

Mandatory Skills

  • AWS Glue
  • EMR
  • Redshift
  • PySpark
  • SQL
  • Kafka
  • Airflow
  • S3
  • ETL
  • Python

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.dice.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

6:21 min

Investigating push inefficiencies with upstream Git experts

Jonathan Creamer · Coffee With Developers

2:07 min

Inspecting default bridge architectures and custom Docker networks

Oliver Seitz Oliver Seitz · World Congress 2025

4:32 min

Harnessing Spark with Python using PySpark and Py4J

Ayon Roy · LIVE

1:43 min

AWS infrastructure stack and data flow pipeline overview

Artem Volk Artem Volk +1 · World Congress 2024

56 sec

Favorite git commands and the importance of patch commits

Eileen Uchitelle Eileen Uchitelle +1 · Coffee With Developers

2:57 min

Core technical practices for robust data engineering

Sandhya Menon Sandhya Menon · World Congress 2026 Europe

Videos

See all

Related articles

See all