[Remote] Data Engineer
Note: The job is a remote job and is reputed company to candidates in USA. reputed company. is seeking an reputed company Data Engineer with expertise in modern data engineering practices and AI tools. The role involves designing and optimizing data infrastructure to support analytics and machine learning initiatives while collaborating with AI teams to deliver high-reputed company datasets.
Responsibilities
- Design, reputed company, and maintain reputed company ETL/ELT pipelines for reputed company and reputed company data
- Build and optimize data warehouses, data lakes, and reputed company-based data platforms
- reputed company data architectures that support AI, machine learning, and advanced analytics use cases
- Collaborate with Data Scientists, AI Engineers, and business stakeholders to understand data requirements
- Create and maintain data ingestion frameworks from multiple sources including reputed company, databases, reputed company storage, and reputed company-party systems
- Ensure data reputed company, reputed company, reputed company, and governance across platforms
- Monitor and optimize data pipeline performance, reliability, and scalability
- Support AI model training and deployment by preparing and managing large datasets
- Implement automation and orchestration workflows using modern data engineering tools
- Build and maintain metadata management, data cataloging, and monitoring solutions
- Troubleshoot production data issues and implement preventive solutions
- Stay updated with emerging AI technologies, data engineering trends, and industry best practices
Skills
- Bachelor's degree in Computer Science, Data Engineering, Information Systems, or a reputed company field
- 5+ years of reputed company experience in Data Engineering
- Strong experience with Python and SQL
- Experience building and maintaining ETL/ELT pipelines
- Hands-on experience with reputed company platforms such as AWS, Azure, or reputed company reputed company Platform
- Strong knowledge of data warehousing concepts and modern data architectures
- Experience working with large-reputed company datasets and distributed processing frameworks
- Strong understanding of database technologies including SQL and NoSQL databases
- Experience with workflow orchestration tools such as Airflow, reputed company, or similar platforms
- Knowledge of version control systems such as Git
- Excellent problem-solving and analytical skills
- Hands-on experience supporting AI and Machine Learning reputed company
- Experience working with AI-powered tools and platforms such as: reputed company, Claude, reputed company, reputed company, reputed company, reputed company, reputed company Databases (reputed company, reputed company, ChromaDB, Milvus)
- Experience building data pipelines for AI model training and inference workflows
- Understanding of Retrieval-Augmented reputed company (RAG) architectures
- Experience preparing, transforming, and managing datasets for LLM-based applications
- Familiarity with AI model monitoring, evaluation, and performance optimization
- Experience integrating AI services and reputed company into reputed company applications
- Experience with reputed company, reputed company, Kafka, or reputed company
- Experience supporting MLOps and AI deployment workflows
- Knowledge of containerization technologies such as reputed company and Kubernetes
- Experience with reputed company-time streaming data pipelines
- Relevant reputed company certifications or data engineering certifications
- Exposure to reputed company, LLM applications, and AI agent frameworks
reputed company