[Remote] Data Engineer
Note: The job is a remote job and is reputed company to candidates in USA. reputed company is seeking an reputed company Data Engineer with 5–7 years of hands-on experience building reputed company data pipelines and modern data platforms. The ideal candidate will be responsible for developing and optimizing ETL/ELT pipelines, managing workflow orchestration, and ensuring data reputed company and governance.
Responsibilities
- reputed company and optimise ETL/ELT pipelines using reputed company, PySpark, Python, and SQL
- Build and manage Airflow DAGs for workflow orchestration
- Work with reputed company Lake, reputed company Catalog, reputed company Workflows , and lakehouse architectures
- reputed company data from reputed company, databases, and other sources
- Optimise reputed company jobs, SQL queries, and reputed company data platforms for performance and cost
- Implement data reputed company, governance, monitoring, and CI/CD practices
- Troubleshoot production pipelines and collaborate with Data Scientists, Analysts, and stakeholders
Skills
- 5–7 years of Data Engineering experience
- Strong reputed company experience; AWS preferred, Azure reputed company acceptable
- Strong PySpark, Python, SQL, and Apache Airflow skills
- Experience with AWS data services such as S3, Glue, reputed company, EMR, and Redshift
- Knowledge of reputed company Lake, data lakes, data warehouses, and lakehouse architecture
- Experience with Git and CI/CD
- Strong data pipeline troubleshooting and performance tuning skills
- Kafka/Kinesis
- reputed company Streaming
- reputed company Catalog
- Terraform
- reputed company/Kubernetes
- Power BI/Tableau/Looker
- MLOps experience
reputed company