Relomote
Remote JobsRelocation Jobs
Add companySaved
Relomote

Relomote is a job board for remote, hybrid, and relocation jobs — every listing AI-classified for the countries it actually hires from, or the visa and relocation support it offers.

LinkedInCrunchbase

Remote jobs by category

  • Remote Engineering & Development jobs
  • Remote Customer Support jobs
  • Remote Design jobs
  • Remote Marketing jobs
  • Remote Sales jobs
  • Remote Product jobs
  • Remote Data & Analytics jobs
  • Remote People & Talent jobs
  • Remote Writing & Content Creation jobs
  • Remote Finance jobs
  • Remote Legal & Compliance jobs
  • Remote Operations & Admin jobs
  • Remote Data Entry jobs
  • Remote Virtual Assistant jobs
  • Remote Education & Training jobs
  • Remote Healthcare & Nursing jobs
  • Remote Other jobs

Remote jobs by location

  • Work from anywhere jobs
  • Remote jobs in Africa
  • Remote jobs in Asia
  • Remote jobs in Europe
  • Remote jobs in Latin America
  • Remote jobs in Middle East
  • Remote jobs in North America
  • Remote jobs in Oceania
  • All remote jobs →

Relocation & visa sponsorship

  • Visa sponsorship jobs
  • Relocation package jobs
  • Relocate to Europe
  • Relocate to Germany
  • Relocate to Netherlands
  • Relocate to Spain
  • Relocate to Portugal
  • Relocate to Greece
  • Relocate to United Kingdom
  • Relocate to Canada
  • Relocate to Australia
  • Relocate to Sweden
  • Relocate to Switzerland
  • Relocate to Japan
  • Relocate to United Arab Emirates
  • All relocation jobs →

© 2026 RelomoteAboutPrivacyTermsLogos provided by Logo.dev

Contact mahmoud@relomote.com · Built by Mahmoud

Relomote
Remote JobsRelocation Jobs
Add companySaved
Egen logo

Lead Data Engineer (Databricks, PySpark & GCP)

Egen
Posted 1 hour ago
🇮🇳India🏢Hybrid📁Data & Analytics
Is this job info correct?

Job Overview: We are looking for a skilled and motivated Lead Data Engineer with strong experience in Python programming, PySpark, Databricks and Google Cloud Platform (GCP) to join our data engineering team. The ideal candidate will be responsible for requirements gathering, designing, architecting the solution, developing, and maintaining robust and scalable ETL (Extract, Transform, Load) & ELT data pipelines. The role involves working with customers directly, gathering requirements, discovery phase, designing, architecting the solution, using various GCP services, implementing data transformations, data ingestion, data quality, and consistency across systems, and post post-delivery support. Experience Level: 10 to 16 years of relevant IT experience Key Responsibilities: Design, develop, test, and maintain scalable ETL data pipelines using Python, PySpark, Databricks & GCP / Azure. Architect the enterprise solutions with various technologies like GCP, Azure, Databricks, PySpark and Spark SQL. Work extensively on Google Cloud Platform (GCP) services such as: Dataflow for real-time and batch data processing Cloud Functions for lightweight serverless compute BigQuery for data warehousing and analytics Cloud Composer for orchestration of data workflows (on Apache Airflow) Google Cloud Storage (GCS) for managing data at scale IAM for access control and security Cloud Run for containerized applications Should have experience in the following areas : Develop production-grade Databricks notebooks and workflows. Build data transformation pipelines using PySpark and Spark SQL . Implement Delta Lake architecture. Design Bronze, Silver, and Gold data layers using the Medallion Architecture. Implement Databricks Workflows/Jobs and dependency management. Tune Spark jobs for large-scale data processing. Optimize cluster configuration and compute utilization. Implement appropriate partitioning, caching, and file-size optimization strategies. Perform data ingestion from various sources and apply transformation and cleansing logic to ensure high-quality data delivery. Implement and enforce data quality checks, validation rules, and monitoring. Collaborate with data scientists, analysts, and other engineering teams to understand data needs and deliver efficient data solutions. Manage version control using GitHub and participate in CI/CD pipeline deployments for data projects. Write complex SQL queries for data extraction and validation from relational databases such as SQL Server, Oracle, or PostgreSQL. Document pipeline designs, data flow diagrams, and operational support procedures. Required Skills: 10+ years of hands-on experience in Python for backend or data engineering projects. Strong understanding and working experience with GCP cloud services (especially Dataflow, BigQuery, Cloud Functions, Cloud Composer, etc.). Working experience with Azure Data Factory (ADF), Azure Databricks, Azure Data Lake Storage Gen2 (ADLS). Solid understanding of data pipeline architecture, data integration, and transformation techniques. Experience in working with version control systems like GitHub and knowledge of CI/CD practices. Experience in Apache Spark, Kafka, Redis, Fast APIs, Airflow, GCP Composer DAGs. Strong experience in SQL with at least one enterprise database (SQL Server, Oracle, PostgreSQL, etc.). Experience with PySpark is required. Experience in data migrations from on-premise data sources to Cloud platforms. Good to Have (Optional Skills): Experience with AWS services. Additional Details: Excellent problem-solving and analytical skills. Strong communication skills and ability to collaborate in a team environment. Education : ● Bachelor's degree in Computer Science, a related field, or equivalent experience.

Similar jobs

Similar jobs

Enable Data Incorporated logo

Senior AWS Data Engineer & Lead AWS Data Engineer (Individual contributor)

Enable Data Incorporated

🇮🇳India4 days ago
Genpact logo

Lead Data Engineer - Data Engineering 4C

Genpact

🇮🇳India3 days ago
DC

Lead Test Engineer - Lead Data Tester

DTCC Candidate Experience Site

🇮🇳IndiaJul 13, 2026, 1:49 AM UTC
Enable Data Incorporated logo

Lead Data Engineer/ Data Architect

Enable Data Incorporated

🇮🇳IndiaJun 26, 2026, 9:15 PM UTC
Bms logo

Data Solutions Engineer II , Enabling Functions Data

Bms

🇮🇳India18 hours ago
ShyftLabs logo

Lead Data Engineer (Databricks)

ShyftLabs

🇮🇳India4 days ago