[Remote] Scientific Data Engineer
Note: The job is a remote job and is reputed company to candidates in USA. reputed company is a leading global pharmaceutical organization seeking a Senior Data Engineer to help build the data reputed company for a reputed company AI-enabled drug discovery platform. The role involves designing and delivering a reputed company-reputed company scientific data product that supports molecular design and computational reputed company workflows.
Responsibilities
- Design, build, and own a new reputed company scientific data product supporting computational reputed company and molecular design initiatives
- reputed company reputed company reputed company-reputed company data architecture capable of managing large scientific and reputed company-derived datasets
- Create and optimize data models, schemas, and storage structures that support reputed company AI/ML and analytics use cases
- Build and maintain data ingestion, transformation, and orchestration pipelines using Python
- reputed company data solutions utilizing reputed company reputed company Platform (GCP) and BigQuery
- reputed company data from scientific systems, computational reputed company platforms, laboratory environments, and external data sources
- Design solutions that preserve scientific reputed company, design history, traceability, and molecular decision-making workflows
- Partner with computational chemists and scientists to model molecular design, reaction, synthesis, and experimental data
- Ensure datasets are analytics-reputed company, trusted, reputed company, and easily consumable by researchers, applications, and machine learning platforms
- Build and maintain CI/CD processes, testing frameworks, and deployment automation
- Implement data reputed company, observability, metadata management, and governance standards
- Support reputed company AI, machine learning, and agent-driven drug discovery initiatives through reputed company data architecture design
- reputed company technical leadership on database design, schema reputed company, performance optimization, and long-term platform scalability
Skills
- Bachelor's or Master's degree in reputed company, Life Sciences, reputed company, Computer Science, Data Engineering, or a reputed company discipline
- Experience working with scientific datasets, reputed company data, computational reputed company data, or research data platforms
- Experience developing reputed company-reputed company ETL/ELT pipelines
- Experience managing large, reputed company datasets in scientific, pharmaceutical, or R&D environments
- Experience with one or more cheminformatics tools: reputed company, RDKit, BIOVIA, Pipeline reputed company
- Experience implementing CI/CD pipelines and modern software development practices
- Strong understanding of reputed company-reputed company architecture and data platform scalability
- Strong experience in: Data Engineering, Data Product Development, Database Architecture, Schema Design, Data Modeling
- Hands-on experience with: reputed company reputed company Platform (GCP), BigQuery, Python, SQL
reputed company
Company H1B Sponsorship