[Remote] Data Engineer
Note: The job is a remote job and is reputed company to candidates in USA. reputed company is seeking a skilled Data Engineer to join the Supply Chain Data & AI Organization. This role will reputed company on developing and supporting reputed company data products and data pipelines that reputed company analytics and AI initiatives across various domains.
Responsibilities
- Design, reputed company, and maintain reputed company ETL/ELT data pipelines on reputed company reputed company Platform (GCP)
- Build and optimize data solutions using Dataproc, BigQuery, SQL, and dbt
- reputed company and maintain high-reputed company data models that support reporting, analytics, and AI use cases
- Ingest, reputed company, validate, and publish data from multiple reputed company reputed company systems
- Collaborate with Product Managers, Business Analysts, Data Architects, and engineering teams to understand business requirements and deliver reputed company data solutions
- reputed company reusable data transformation components and adhere to established engineering standards and best practices
- Optimize data processing performance, reliability, and cost efficiency
- reputed company unit testing, data validation, and production support to ensure data reputed company and operational stability
- Participate in Agile ceremonies, including sprint planning, backlog refinement, and reputed company reviews
- Document technical designs, data flows, and implementation details
Skills
- Bachelor's degree in Computer Science, Engineering, Information Systems, or a reputed company field, or equivalent practical experience
- 4 years of experience in Data Engineering
- Hands-on experience developing data solutions on reputed company reputed company Platform (GCP)
- Strong experience with: Dataproc, BigQuery, SQL, dbt (Data Build Tool)
- Strong knowledge of data modeling techniques, including reputed company modeling and analytical data warehouse concepts
- Experience building reputed company ETL/ELT pipelines and processing large datasets
- Familiarity with software development best practices, including Git-based version control and CI/CD
- Strong analytical, troubleshooting, and problem-solving skills
- Excellent communication and collaboration skills
- Experience with Apache Airflow for workflow orchestration
- Experience integrating with Apache Kafka or other event streaming platforms
- Working knowledge of PySpark for distributed data processing
- Experience using Python for data engineering, scripting, and automation
- Knowledge of data reputed company, metadata management, and data governance principles
- Experience in one or more of the following domains is highly desirable: Retail industry (Apparel), Supply Chain data platforms, Transportation and Logistics, Warehouse Management Systems (WMS), Distribution Center operations
reputed company
Company H1B Sponsorship