About Us Data, analytics and cybersecurity staffing. We connect professionals and companies to deliver successful projects. Job Description We are looking for an experienced Data Engineer to design, build, and maintain cloud-based data pipelines supporting pharmaceutical manufacturing and laboratory operations. The role focuses on integrating data from enterprise and laboratory systems into the AWS ecosystem and ensuring that data is reliable, traceable, and accessible for analytics. Key Responsibilities Design, develop, and maintain scalable data pipelines within AWS. Ingest and transform data using Amazon S3, AWS Glue, and Athena . Develop complex SQL queries and optimize distributed data processing. Use Python, Scala, or PySpark for data transformation and automation. Integrate data from enterprise and laboratory platforms such as SAP, GLIMS, and Veeva . Support data quality, integrity, lineage, and validation requirements. Troubleshoot pipeline failures, data inconsistencies, and performance issues. Collaborate with data engineers, architects, business analysts, laboratory teams, and manufacturing stakeholders. Participate in Agile delivery, code reviews, testing, and release activities. Contribute to automated deployments through CI/CD pipelines . Ensure solutions comply with applicable GxP and pharmaceutical data-integrity requirements . Requirements Candidates must meet one of the following criteria: Bachelor’s degree in Computer Science, Information Technology, Engineering, or a related field, with at least 4 years of relevant technical experience ; or At least 8 years of equivalent professional experience in data engineering, cloud analytics, or a related technical area without a degree. Required Technical Skills Hands-on experience with the AWS ecosystem , particularly: Amazon S3 AWS Glue Amazon Athena Advanced proficiency in SQL . Experience with at least one distributed data or query platform: Trino Presto Databricks Strong programming or scripting skills in at least one of: Python Scala PySpark Experience building data transformations and production-grade data pipelines. Understanding of data integration involving ERP, manufacturing, or laboratory systems . Strong problem-solving, troubleshooting, and communication skills. Preferred Skills Previous experience with SAP ERP data. Knowledge of laboratory or pharmaceutical platforms such as GLIMS or Veeva . Experience working in pharmaceutical manufacturing or another regulated industry. Familiarity with GxP regulations , validation, auditability, and data-integrity principles. Experience working in an Agile environment . Familiarity with Git and CI/CD pipelines . Understanding of data governance, lineage, quality controls, and access management. Ideal Candidate Profile The strongest candidate will combine hands-on AWS data engineering experience with advanced SQL and Python/PySpark skills. Experience integrating SAP or laboratory data and working in a GxP-regulated pharmaceutical environment would be a significant advantage. Benefits Location: Prague - Smíchov Work Model: Hybrid (2-3 days/week onsite), Contract Type: Freelance / Contract Start date: Summer, 2026 Time Allocation: 40 hours/week Global Pharmaceutical Company in Prague
Data Science Team Lead
Process& GmbH
Data Engineer (Clinical, laboratory, scientific data)
Fproof
Business Development Representative
Worki | Asana Expert & Solutions Partner
S4O9 Change Readiness Lead
Huxley
Project Manager with French
Havas
Mid-Level BE Engineer
Oddin