AWS Data Engineer (Associate)
4 г. назад
USAMiddleRemote
awsdata engineeringdata products
Full-stack AWS Data Engineer working on data platforms and modernization to build scalable data products for better decision-making.
О компании
- is the agent-native AWS modernization firm. We ship modernization to production in weeks — data platforms migrated, applications refactored, AI agents running against real data. Our delivery is built on the agent platform built by our team, and executed by forward-deployed engineers who own outcomes end-to-end. Fixed-date commitments. Real production systems. Retired legacy. We are an AWS Premier Tier Services Partner with the AWS Agentic AI Specialization, seven AWS Consulting Competencies (including Migration and Modernization, Data and Analytics, Machine Learning, and AI Services), and seventeen AWS Service Validations. Our named production customers include Synaptics, Flipboard, Poshmark, Tilia, KlearTrust, and Safaricom. Most modernization work doesn't ship. Ours does. That's
Обязанности
- Write efficient code in - PySpark, Amazon Glue
- Write SQL Queries in - Amazon Athena, Amazon Redshift
- Explore new technologies and learn new techniques to solve business problems creatively
- Collaborate with many teams - engineering and business, to build better data products and services
- Deliver the projects along with the team collaboratively and manage updates to customers on time
Будет плюсом
- Prior experience in working on AWS EMR, Apache Airflow
- Certifications AWS Certified Big Data – Specialty OR Cloudera Certified Big Data Engineer OR Hortonworks Certified Big Data Engineer
- Understanding of DataOps Engineering
Другое
- 1 to 3 years of experience in Apache Spark, PySpark, Amazon Glue
- 2+ years of experience in writing ETL jobs using pySpark, and SparkSQL
- 2+ years of experience in SQL queries and stored procedures
- Have a deep understanding of all the Dataframe API with all the transformation functions supported by Spark 2.7+