SS

[English Only] Data Engineer | GCP / BigQuery / Databricks / PySpark | AI & Data Platform Engineering

Hiring from
Japan
Work type
Hybrid
Posted
Sep 25, 2026
Is this job info correct?

[English Only] Data Engineer | GCP / BigQuery / Databricks / PySpark | AI & Data Platform Engineering



Why should you apply:


【Build and Improve Cloud Data Pipelines with GCP and Databricks】

Design, build, and maintain data pipelines using GCP services such as BigQuery, Dataflow, Cloud Composer, GCS, and Pub/Sub. You will work on data processing workflows with PySpark and Databricks, including pipeline optimization and data quality improvement.


【Work on Data Platforms Supporting AI and Data Solutions】

Join an AI & Data engineering organization that develops data-oriented solutions using data collected from multiple business services. You will contribute to the data foundations that support analytics, applications, and business initiatives.


【Expand Your Data Engineering Scope Across Cloud-Native Technologies】

Work beyond traditional pipeline development by contributing to cloud operations, deployment processes, APIs, internal tools, and cloud-native application development when required.



Hiring Company:

A Japanese internet services company, operating more than 70 businesses across e-commerce, fintech, digital content, and communications. The company has billions of members worldwide and operates across 30 countries and regions.



Responsibilities:

  • Design, build, and maintain scalable data pipelines and ETL/ELT workflows on Google Cloud Platform.
  • Develop and optimize data processing jobs using PySpark and Databricks.
  • Build and manage data lakes, data warehouses, and cloud data platforms.
  • Monitor, tune, and troubleshoot pipeline performance and data quality issues.
  • Collaborate with analysts, data scientists, product teams, and development teams to deliver data solutions.
  • Contribute to APIs, dashboards, and internal tools when required.
  • Participate in code reviews, technical discussions, and architectural decisions.
  • Support data security, governance, and compliance practices.



Required Skills:

  • 6+ years of hands-on experience as a Data Engineer.
  • Strong experience with GCP services including BigQuery, Dataflow, Cloud Composer, GCS, and Pub/Sub.
  • Experience with PySpark and distributed data processing.
  • Hands-on experience with Databricks.
  • Experience with Apache Airflow or similar orchestration tools.
  • Strong proficiency in Python and SQL.
  • Experience building and troubleshooting data pipelines.
  • Knowledge of data security and governance practices.
  • English: Advanced level (TOEIC 800+)
  • Japanese: Not required


Preferred Skills:

  • Experience with Terraform or Infrastructure as Code.
  • Experience with streaming data frameworks such as Kafka or Pub/Sub.
  • Experience with ML pipelines or MLOps workflows.
  • GCP or Databricks certifications.
  • Experience with Docker, Kubernetes, and CI/CD.
  • Experience building APIs using FastAPI, Flask, or Node.js.



Working Hours:

9:00–17:30 (Mon–Fri)


Working Style:

Hybrid — 4 days office / 1 day WFH


Employment Type:

Haken Contract — Long-term renewable contract


Holidays:

Saturday, Sunday, National Holidays, Year-end and New Year Holidays, Paid Holidays, and other special holidays


Services / Benefits:

Social insurance, DC pension plan, transportation allowance, and other benefits.



ID: 506052

Similar jobs

Apply on LinkedIn