[English Only] Data Engineer | GCP / BigQuery / Databricks / PySpark | AI & Data Platform Engineering
- Hiring from
- Japan
- Work type
- Hybrid
- Posted
- Sep 25, 2026
[English Only] Data Engineer | GCP / BigQuery / Databricks / PySpark | AI & Data Platform Engineering
Why should you apply:
【Build and Improve Cloud Data Pipelines with GCP and Databricks】
Design, build, and maintain data pipelines using GCP services such as BigQuery, Dataflow, Cloud Composer, GCS, and Pub/Sub. You will work on data processing workflows with PySpark and Databricks, including pipeline optimization and data quality improvement.
【Work on Data Platforms Supporting AI and Data Solutions】
Join an AI & Data engineering organization that develops data-oriented solutions using data collected from multiple business services. You will contribute to the data foundations that support analytics, applications, and business initiatives.
【Expand Your Data Engineering Scope Across Cloud-Native Technologies】
Work beyond traditional pipeline development by contributing to cloud operations, deployment processes, APIs, internal tools, and cloud-native application development when required.
Hiring Company:
A Japanese internet services company, operating more than 70 businesses across e-commerce, fintech, digital content, and communications. The company has billions of members worldwide and operates across 30 countries and regions.
Responsibilities:
- Design, build, and maintain scalable data pipelines and ETL/ELT workflows on Google Cloud Platform.
- Develop and optimize data processing jobs using PySpark and Databricks.
- Build and manage data lakes, data warehouses, and cloud data platforms.
- Monitor, tune, and troubleshoot pipeline performance and data quality issues.
- Collaborate with analysts, data scientists, product teams, and development teams to deliver data solutions.
- Contribute to APIs, dashboards, and internal tools when required.
- Participate in code reviews, technical discussions, and architectural decisions.
- Support data security, governance, and compliance practices.
Required Skills:
- 6+ years of hands-on experience as a Data Engineer.
- Strong experience with GCP services including BigQuery, Dataflow, Cloud Composer, GCS, and Pub/Sub.
- Experience with PySpark and distributed data processing.
- Hands-on experience with Databricks.
- Experience with Apache Airflow or similar orchestration tools.
- Strong proficiency in Python and SQL.
- Experience building and troubleshooting data pipelines.
- Knowledge of data security and governance practices.
- English: Advanced level (TOEIC 800+)
- Japanese: Not required
Preferred Skills:
- Experience with Terraform or Infrastructure as Code.
- Experience with streaming data frameworks such as Kafka or Pub/Sub.
- Experience with ML pipelines or MLOps workflows.
- GCP or Databricks certifications.
- Experience with Docker, Kubernetes, and CI/CD.
- Experience building APIs using FastAPI, Flask, or Node.js.
Working Hours:
9:00–17:30 (Mon–Fri)
Working Style:
Hybrid — 4 days office / 1 day WFH
Employment Type:
Haken Contract — Long-term renewable contract
Holidays:
Saturday, Sunday, National Holidays, Year-end and New Year Holidays, Paid Holidays, and other special holidays
Services / Benefits:
Social insurance, DC pension plan, transportation allowance, and other benefits.
ID: 506052