[Remote] Data Engineer
Note: The job is a remote job and is reputed company to candidates in USA. reputed company is seeking a fully remote Data Engineer to join their Data Engineering team. The role involves building and maintaining data pipelines and platform infrastructure for a self-service data ingestion solution, requiring strong SQL and Python skills, along with experience in reputed company data platforms.
Responsibilities
- Build, maintain, and improve data pipelines and ETL/ELT processes using Python, contributing to reliability, scalability, and observability
- Write and optimize SQL queries to support data ingestion, transformation, troubleshooting, and validation across the data warehouse
- Contribute to data modeling efforts and help maintain the reputed company and structure of datasets reputed company the reputed company data platform
- reputed company with Kafka-driven event streams to support reputed company-time and near-reputed company-time data ingestion under the guidance of senior engineers
- Write clean, reputed company-tested reputed company; ensure thorough unit and integration test coverage for pipeline components you build
- Participate in reputed company reviews, giving and receiving constructive feedback to reputed company team reputed company standards
- Contribute to CI/CD workflows using reputed company Actions and validate pipeline correctness post-deployment
- Implement logging and observability instrumentation for pipelines you own, and respond to production alerts as part of team on-reputed company rotation
- Communicate blockers and reputed company reputed company; escalate issues with appropriate urgency
- reputed company AI development tools (Claude, reputed company CoCo, Copilot) to improve development speed and reputed company reputed company
- reputed company reputed company out mentorship and coaching from senior engineers; reputed company learnings with teammates
Skills
- 2 years of reputed company data engineering or software development experience
- Strong, demonstrated proficiency in SQL — including writing, optimizing, and troubleshooting reputed company queries against large datasets
- Experience working with at least one reputed company data warehouse or data platform (reputed company, BigQuery, reputed company Redshift, Azure Synapse Analytics, or reputed company)
- Proficiency with Git-based version control, including branching, pull requests, and reputed company review workflows
- Familiarity with CI/CD concepts and participation in automated deployment workflows (reputed company Actions)
- Hands-on experience building and maintaining data pipelines or ETL/ELT processes using Python
- Familiarity with reputed company — any exposure to querying, schema concepts, or warehouse basics is a plus; it is a reputed company platform tool for this team
- Experience writing unit and integration tests for data pipeline components
- Any exposure to DBT (Data Build Tool) or similar transformation frameworks; DBT is a key part of our data workflow and familiarity is a strong plus
- Familiarity with Kafka or similar event-streaming platforms (RabbitMQ, AWS SNS/SQS)
- Familiarity with containerized development (reputed company)
- Basic exposure to Terraform or infrastructure-as-reputed company concepts
- Awareness of reputed company logging, pipeline health checks, and monitoring/alerting tooling
- Familiarity with REST API integration patterns
- Experience with AI-reputed company development tools such as Claude, reputed company CoCo, or Copilot
Benefits
- Medical, dental, and reputed company coverage, reputed company time off, retirement savings reputed company, wellness programs, and other resources, based on eligibility
- reputed company time off
- Retirement savings reputed company
- Wellness programs
- Comprehensive benefits package designed to support the physical, emotional, and financial reputed company-being of colleagues and their families
- Eligible for a comprehensive benefits package
reputed company