Relomote
Remote JobsRelocation Jobs
Add companySaved
Relomote

Relomote is a job board for remote, hybrid, and relocation jobs — every listing AI-classified for the countries it actually hires from, or the visa and relocation support it offers.

LinkedInCrunchbase

Remote jobs by category

  • Remote Engineering & Development jobs
  • Remote Customer Support jobs
  • Remote Design jobs
  • Remote Marketing jobs
  • Remote Sales jobs
  • Remote Product jobs
  • Remote Data & Analytics jobs
  • Remote People & Talent jobs
  • Remote Writing & Content Creation jobs
  • Remote Finance jobs
  • Remote Legal & Compliance jobs
  • Remote Operations & Admin jobs
  • Remote Data Entry jobs
  • Remote Virtual Assistant jobs
  • Remote Education/Training jobs
  • Remote Healthcare/Clinical jobs
  • Remote Other jobs

Remote jobs by location

  • Work from anywhere jobs
  • Remote jobs in Africa
  • Remote jobs in Asia
  • Remote jobs in Europe
  • Remote jobs in Latin America
  • Remote jobs in Middle East
  • Remote jobs in North America
  • Remote jobs in Oceania
  • All remote jobs →

Relocation & visa sponsorship

  • Visa sponsorship jobs
  • Relocation package jobs
  • Relocate to Europe
  • Relocate to Germany
  • Relocate to Netherlands
  • Relocate to Spain
  • Relocate to Portugal
  • Relocate to Greece
  • Relocate to United Kingdom
  • Relocate to Canada
  • Relocate to Australia
  • Relocate to Sweden
  • Relocate to Switzerland
  • Relocate to Japan
  • Relocate to United Arab Emirates
  • All relocation jobs →

© 2026 RelomoteAboutPrivacyTerms

Contact [email protected] · Built by Mahmoud

Relomote
Remote JobsRelocation Jobs
Add companySaved
3Pillar logo

Lead Data Engineer with AI experience

3Pillar
Posted Jun 12, 2026, 8:45 PM UTC
🇮🇳India🏠Remote📁Data & Analytics
Is this job info correct?

3Pillar is an AI transformation partner on a mission to help enterprises build the AI-native products and intelligent agents that will define the next era of business. With teams across North America, Europe, Latin America, and Asia, we work with the most ambitious companies in financial services, healthcare, media, and technology — helping them move faster, modernize boldly, and compete on their own terms. Our HelixAI platform and Helix Pods delivery model put our engineers at the center of real agentic transformation — doing work that is open, portable, and built to last. We are building the future of enterprise AI We are looking Lead Data Engineer to build, operate, and continuously improve the data pipelines, retrieval infrastructure, and ML/LLMOps foundations that power our AI initiatives. The resource will work on turning reference architectures and data contracts into robust, production-grade implementations that serve conversational AI assistants, dashboard copilots, autonomous agents, RAG applications, and predictive ML models. Key Responsibilities: Data Pipeline Engineering : Build, test, and maintain production pipelines (batch & real-time) on Snowflake, PySpark, Delta Lake, and Kafka. Implement data quality checks, schema validation, and alerting at every pipeline stage. Migrate legacy ETL/DWH to cloud-native AWS/Azure architectures with measurable latency and cost improvements. Maintain CI/CD pipelines: automated testing, deployment, rollback, and IaC (Terraform, GitHub Actions). RAG, Vector & Retrieval Infrastructure: Build end-to-end retrieval infrastructure: document ingestion, embedding pipelines, vector store management (Pinecone, FAISS, ChromaDB, OpenSearch), and hybrid retrieval layers. Implement chunking, metadata filtering, and re ranking — tuning for precision, recall, and latency. Maintain data freshness and index consistency; instrument with context relevance and faithfulness metrics. Semantic Layer & Knowledge Infrastructure: Implement and maintain business entity mappings, ontologies, and knowledge graphs (Neo4j) per Architect design. Build and version the feature store and semantic data contracts serving both ML models and LLM applications. Manage metadata, data lineage, and audit trail instrumentation across the platform. ML/LLMOps Pipeline Support: Build ML data infrastructure: training curation, feature engineering, MLflow experiment tracking, dataset versioning. Support LLM fine-tuning workflows — corpus curation, quality filtering, dataset formatting. Implement automated evaluation pipelines: factual accuracy, hallucination detection, regression tracking. Maintain production monitoring dashboards for pipeline health, model metrics, and alerting. Agentic Data Infrastructure: Build and maintain data APIs, tool schemas, and memory/state stores that autonomous agents depend on. Implement agent observability: capture inputs, retrieved context, tool calls, reasoning traces, and outputs. Maintain text-to-SQL layers, semantic query interfaces, and context APIs for conversational AI consumers. Governance, Security & Data Quality: Implement RBAC, attribute-based access, PII detection/masking, data classification, and audit logging. Enforce data contracts and schema governance with automated breaking-change detection and versioned migrations. Build data quality monitoring (completeness, freshness, consistency) with automated alerting and root-cause tooling. Support compliance readiness: audit trails, data provenance, and regulatory documentation. Qualifications: 7+ years data engineering using Cloud services 2+ years production AI/ML or LLM-era data infrastructure. Proven experience building production pipelines at scale — batch and streaming, Snowflake,AWS/Azure. Deep expertise: Python, PySpark, Snowflake, Delta Lake, Kafka, Spark Structured Streaming. Hands-on with vector stores, embedding pipelines, and retrieval infrastructure in production RAG environments. Working knowledge of MLOps: MLflow, CI/CD for AI, automated evaluation, and production monitoring. Strong grounding in data governance, quality frameworks, and compliance- aligned engineering. Technical Skills: Primary skills: Python, SQL, PySpark, Kafka, Snowflake/DataBricks, Delta Lake, AWS (S3, Glue, Kinesis, EKS, Redshift), Docker, Kubernetes, GitHub Actions. Secondary Skills : LangChain, LlamaIndex, LLM APIs (OpenAI, Bedrock, Claude, HuggingFace), Pinecone, FAISS, ChromaDB, OpenSearch, MLflow, FastAPI, Neo4j, LangGraph, prompt engineering, RLHF dataset prep, LLM fine-tuning workflows Connect: Regards, Kiran Dhanak Talent Acquisition Manager

Similar jobs

Similar jobs

SonicWall logo

Solutions Engineer - Chennai (Pre-Sales experience is a must)

SonicWall

🇮🇳India9 hours ago
SonicWall logo

Senior Director - Data & AI Engineering (Cybersecurity/Network Security Domain Experience Required)

SonicWall

🇮🇳India1 weeks ago
Terac logo

Experienced Software Engineers: Coding Tasks for AI Systems

Terac

🇮🇳India2 weeks ago
This is Gain Ltd logo

GAIN - Experience - Drupal Developer 1

This is Gain Ltd

🇮🇳India3 weeks ago
Sapsol Technologies Inc 7 logo

SAP Freshers Training & Simulated Project Experience-Remote - Sapsol Technologie

Sapsol Technologies Inc 7

🌍India, United States3 weeks ago
Coinbase logo

Engineering Manager - Customer Experience AI

Coinbase

🇮🇳India3 weeks ago