Gcp data engineer
Innodata Inc. · Lucknow, India
FULL TIMEpermanent
Job Description
Role: GCP data engineer
Experience: 5+years
need immediate joiners.
Responsibilities:
• Design and implement data-driven solutions on GCP including Big Query, Cloud Storage, Dataflow, Pub/Sub, and Looker/BI.
• Build ETL scripts using SQL and Python to extract, clean, and transform structured and unstructured data from ERP, procurement, logistics, and facility management systems.
• Develop and optimize data pipelines for ingestion, transformation, and loading into enterprise data lakes and warehouses.
• Build and extend end-to-end data and BI solutions, spanning extraction, storage, transformation, and visualization layers.
• Partner with supply chain, real estate, and AI/ML teams to provide pipelines for AI solutions (e.g., RAG ingestion, Copilot integration, multi-agent workflows).
• Ensure data governance, lineage, and compliance across supply chain datasets.
• Continuously optimize query performance, ETL processes, and pipeline reliability.
Skills Required:
• Strong hands-on expertise with GCP services: Big Query, Dataflow, Pub/Sub, Cloud Storage, Looker/BI (or similar).
• Advanced proficiency in SQL (complex queries, optimization) and Python (data engineering, scripting, APIs).
• Experience building ETL/ELT pipelines operating on structured and unstructured data sources.
• Knowledge of enterprise data warehouse and data lake architectures.
Experience: 5+years
need immediate joiners.
Responsibilities:
• Design and implement data-driven solutions on GCP including Big Query, Cloud Storage, Dataflow, Pub/Sub, and Looker/BI.
• Build ETL scripts using SQL and Python to extract, clean, and transform structured and unstructured data from ERP, procurement, logistics, and facility management systems.
• Develop and optimize data pipelines for ingestion, transformation, and loading into enterprise data lakes and warehouses.
• Build and extend end-to-end data and BI solutions, spanning extraction, storage, transformation, and visualization layers.
• Partner with supply chain, real estate, and AI/ML teams to provide pipelines for AI solutions (e.g., RAG ingestion, Copilot integration, multi-agent workflows).
• Ensure data governance, lineage, and compliance across supply chain datasets.
• Continuously optimize query performance, ETL processes, and pipeline reliability.
Skills Required:
• Strong hands-on expertise with GCP services: Big Query, Dataflow, Pub/Sub, Cloud Storage, Looker/BI (or similar).
• Advanced proficiency in SQL (complex queries, optimization) and Python (data engineering, scripting, APIs).
• Experience building ETL/ELT pipelines operating on structured and unstructured data sources.
• Knowledge of enterprise data warehouse and data lake architectures.
Details
| Company | Innodata Inc. |
| Location | Lucknow, India |
| Type | FULL TIME |
| Niche | tech |
| Experience | permanent |
