Senior GCP Data Engineer with 4 years at Accenture building production data platforms. I design and build BigQuery data warehouses, automated ETL/ELT pipelines, and real-time streaming with Kafka and Apache Flink - delivered as clean, documented, Terraform-managed infrastructure.
Proof over promises:
- Migrated SQL Server to BigQuery, cutting query latency by 35%
- Automated ingestion from 5+ sources, cutting manual pipeline effort by 50%
- Sustained 99.9% platform uptime on production workloads
- Certified GCP Professional Data Engineer (May 2025) + GCP Generative AI Leader
Core stack: BigQuery, Cloud Composer (Airflow), Pub/Sub, Cloud Functions, Kafka/Confluent, Apache Flink, Terraform, Azure DevOps, Python, SQL.
I work in your environment, ship production-ready code with documentation, and communicate proactively. Let us build a data stack your team can actually maintain - message me to discuss your project.
SQL
Google Cloud Platform
Azure DevOps
BigQuery
Terraform
Apache Kafka
Apache Airflow
Python
Apache Flink
ETL
Data Warehousing
Data Migration
Microsoft Power BI
Streaming Platform
Data Engineering
Shivam W.
Shahdara, India
$20/hr
5.0
8 jobs
I'm a Senior Data Engineer with 4.5+ years of experience building scalable, cloud-native data platforms that turn raw data into reliable, business-ready insights. I've delivered enterprise solutions across banking (NAB), healthcare (Molina), and CPG (PepsiCo), specializing in end-to-end pipeline architecture, data modeling, and cloud migrations.
What I bring to your project:
๐น Cloud Data Engineering โ Deep expertise in Azure (Databricks, Data Factory, Synapse) and AWS (EMR, Glue, S3, RedShift), with hands-on migration experience from on-prem and Teradata to cloud.
๐น Pipeline Architecture & ETL โ I design and build robust ingestion frameworks handling batch, incremental, and real-time data (Event Hub, Kafka) across formats like JSON, CSV, Parquet, and fixed-width files.
๐น Data Modeling & Warehousing โ Skilled in dimensional modeling, Data Vault, star/snowflake schemas, and silver/gold layer design. I've modeled 50+ tables across Oracle Fusion, SAP S/4, and healthcare domains.
๐น Transformation & Orchestration โ I translate complex business rules into DBT models, orchestrate workflows with Apache Airflow or AutoSys, and automate CI/CD via Jenkins and Azure DevOps.
๐น Performance & Governance โ I tune PostgreSQL and Spark jobs, implement data quality checks, reconciliation frameworks, and ensure compliance with data governance standards.
๐น Generative AI & MLOps โ Databricks-certified in Generative AI, with experience integrating MLflow for experiment tracking and building LLM-based automation using OpenAI and LangChain.
Tech Stack: Python | SQL | Scala | Apache Spark | DBT | PostgreSQL | Snowflake | Airflow | Databricks | Azure | AWS | Git | Jenkins | MLflow | Power BI
Certifications: Databricks Certified Data Engineer Professional | Azure Data Engineer (DP-203) | Snowflake SnowPro Core | Fabric Analytics Engineer (DP-600) | Generative AI Engineer Associate
Whether you need a production-grade pipeline, a cloud migration, or a well-modeled data warehouse, I deliver clean, documented, and scalable solutions โ on time and with clear communication. Let's discuss your project!
Data Extraction
Data Mining
Artificial Intelligence
ETL Pipeline
Machine Learning
Database Design
Database Modeling
PySpark
Databricks Platform
Snowflake
Data Warehousing
Apache Airflow
Python
Web Scraping
Data Engineering
Generative AI
Exploratory Data Analysis
Scala
Data Integration
Anil Kumar M.
Hyderabad, India
$25/hr
5.0
6 jobs
Iโm a Senior Data Engineer with 9+ years of experience building scalable data platforms, automated ETL/ELT pipelines, and end-to-end analytics systems for startups to enterprise clients.
I help organizations turn raw, messy data into clean, reliable, and analytics-ready datasets using modern cloud data stacks.
What I Can Help You With
โค ETL/ELT Development (Python, Airflow, AWS Glue, Talend)
โค Cloud Data Engineering (AWS, Snowflake, Redshift, S3, Lambda)
โค Data Warehouse Architecture & Optimization
โค API Integrations (SaaS - Databases / Warehouses)
โค SQL Development & Performance Tuning
โค Dashboard/Data Model Preparation (Power BI, QuickSight, Tableau)
โค Data Migration from CRMs/ERPs (Salesforce, HubSpot, Workday, Custom APIs)
โค Automation of repeated data processes (Python + Cloud services)
Why Clients Work With Me
๐น Clear communication and structured delivery
๐น Clean, audit-ready pipelines with proper logging & monitoring
๐น Fast turnaround and long-term maintainable code
๐น Experience across Healthcare, Finance, Retail, and SaaS domains
My Recent Projects:
โญCRM/ERP to Snowflake migration with automated Python pipelines
โญAWS Data Platform modernization using Glue, Lambda, and Redshift
โญFinance & Claims data processing using PySpark on AWS
โญCustom API integrations to centralize business data
โญDashboard backend (SQL + models) for Power BI & QuickSight
If you need
โ Faster data pipelines
โ A scalable cloud data architecture
โ Clean data models for BI
โ Help with Talend/Airflow/Snowflake
โ A partner who delivers on time, every time
I can help you.
Send me a message, and letโs discuss your project.
Talend Open Studio
Snowflake
SQL
Data Engineering
Python
AWS Glue
Amazon Redshift
ETL
Amazon Web Services
Apache Superset
Apache Airflow
Amazon QuickSight
Tableau
AWS Lambda
Streamlit
Adarsh R.
Bengaluru, India
$70/hr
5.0
38 jobs
I'm a Senior Data Engineer with 8+ years of strong technical expertise in building reliable and scalable data infrastructure, from data ingestion to transformation to warehousing, streaming, and data analytics, specializing in dbt, Snowflake, Airflow, Databricks (and more) across AWS, Azure, and GCP, with robust ELT and ETL pipelines. If your data pipelines are brittle, your data warehouse is slow, or your data was never built to scale, that is exactly what I fix, with fault tolerance, observability, and audit-ready quality engineered in from day one.
I cover the full data engineering lifecycle: batch and real-time data pipelines, Modern Data Stack builds, lakehouse architecture, cloud and warehouse data migration, governance, and the data foundations that feed modern systems.
๐ฏ Core Expertise:
โ Data Pipelines & Orchestration: End-to-end batch and real-time pipelines with Apache Airflow, Dagster, Prefect, AWS Step Functions, and Azure Data Factory. Idempotent, schema-drift tolerant, and monitored so failures surface before they reach your stakeholders.
โ Cloud Warehousing & Lakehouse: Snowflake, BigQuery, Amazon Redshift, Databricks, and Microsoft Fabric, with Delta Lake and Apache Iceberg lakehouse foundations governed through the Glue Data Catalog and Lake Formation, with Athena and Redshift Spectrum for serverless queries, Medallion Architecture, partitioning, and performance tuning.
โ Data Transformation & Modeling: dbt (Core and Cloud), SQLMesh, Spark and PySpark on EMR and AWS Glue, Star Schema and dimensional modeling, analytics engineering best practices, full test coverage, and CI/CD for data models.
โ Streaming & Real-Time Analytics: Distributed streaming with Apache Kafka, Flink, Spark Structured Streaming, Kinesis, and Pub/Sub, including exactly-once semantics, dead-letter queues, CDC, and end-to-end latency guarantees.
โ Data Ingestion & Integration: Fivetran, Airbyte, Matillion, Stitch, Hevo, Meltano, and custom CDC pipelines for near-real-time sync across structured, semi-structured, and unstructured sources.
โ Data Quality, Governance & Observability: Automated data quality frameworks, SLA monitoring, auditable lineage, data catalog and metadata management, and observability that catches bad data early.
โ Cloud Migration & Modernization: Zero-downtime migration handled end to end, from legacy warehouse assessment through cutover, with zero data loss and minimal downtime, replacing brittle ETL and ELT with a clean Modern Data Stack.
โ AI-Ready Data Infrastructure: Pipelines engineered to feed LLMs and ML systems with clean, structured, high-quality data, from ingestion through transformation to serving.
------------------------------------------------------
โ๏ธTech Stack:
โก Warehouses & Lakehouse: Snowflake | BigQuery | Redshift | Databricks | Microsoft Fabric | Athena | Delta Lake | Iceberg
โก Transformation: dbt | SQLMesh | Spark | PySpark | AWS Glue | EMR | Star Schema | Medallion Architecture
โก Orchestration: Airflow (GCP Cloud Composer and AWS MWAA) | Dagster | Prefect | Azure Data Factory | Step Functions
โก Streaming: Kafka | Flink | Kinesis | Pub/Sub | Spark Structured Streaming | ClickHouse
โก Ingestion: Fivetran | Airbyte | Matillion | Stitch | Hevo | Meltano | CDC
โก Governance & Catalog: Glue Data Catalog | Lake Formation | Unity Catalog | Microsoft Purview | Dataplex
โก Cloud: AWS | GCP | Azure
โก Languages: Python | SQL (Snowflake, BigQuery, T-SQL, PL/pgSQL) | FastAPI
โก Databases: PostgreSQL | MySQL | SQL Server | DynamoDB | MongoDB
โก BI & Reporting: Looker | Tableau | Power BI | GA4 | Metabase | Superset | Streamlit | Grafana
------------------------------------------------------
โญ What Clients Say:
๐ "Adarsh rebuilt our analytics pipeline on Snowflake, Airflow, and dbt, giving us reliable, version-ready data. Reporting accuracy improved overnight, and we can finally trust the numbers." โ Anita, Head of Product, FinTech SaaS
๐ "He designed a zero-downtime migration to a modern data warehouse that cut query latency by more than half while keeping our SLAs intact." โ Daniel, VP of Data, AdTech Firm
๐ "Clean architecture, solid dbt models, and Airflow pipelines running without issues for months. He brought a level of engineering discipline we hadn't seen from a data consultant before." โ Mark, Director of Data Engineering, E-commerce Startup
๐ "We came to him with a Spark pipeline costing us a fortune and delivering stale data. He restructured the workflow logic and cut processing time by 70%." โ Leo, Head of Analytics, HealthTech SaaS
------------------------------------------------------
๐ TOP RATED PLUS | EXPERT-VETTED | Top 1% on Upwork | 8+ Years Experience | 100% Job Success
๐ Ready to build a scalable, production-ready data infrastructure to turn your raw data into reliable, actionable business insights? Click the 'Invite to Job' button on the top right, and let's discuss your data pipeline!
Data Engineering
Snowflake
dbt
Apache Airflow
Python
SQL
Amazon Web Services
Google Cloud Platform
Microsoft Azure
Databricks Platform
PostgreSQL
ETL Pipeline
Data Warehousing
API Integration
Apache Kafka
PySpark
BigQuery
Data Modeling
Data Extraction
Big Data
Pavithravathy S.
Chennai, India
$35/hr
5.0
11 jobs
I am 14 years experienced Integration Specialist. For the last 8 years, gained experience in Boomi Integration,Mulesoft,Boomi API, Boomi Flow, EDI.
Kindly find my experience below in detail.
Experience Summary:
--------------------------
๏ 14+ yearsโ experience in integration technologies which includes Dell Boomi, Websphere Enterprise Bus (IBM Integration Designer), Mulesoft with 3 years experience,Websphere Application Server, Websphere Message Broker, IBM Datapower, Websphere Message Queue and IBM Integration Bus with 8+ experience in Dell Boomi
๏Have Built Web Applications in Boomi Flow
๏ Certification and hands-on experience using Dell Boomi as the iPaaS platform, to build solutions integrating enterprise applications across systems hosted on-premise and in the Cloud.
๏ Experience working with Integration technologies such as HTTP Web Services (SOAP/REST), XML, JSON, Flat file
๏ Experienced in working with JDBC adapter, JMS adapter, Web Service (SOAP and REST), WSDL.
๏ Understanding of enterprise integration tools, middleware technologies, and data integration patterns, including Pub/Sub Messaging, EIP, Event-Driven Architecture, SOA, and iPaaS.
๏ Maintain existing APIs, review, enhance code to make it readable and clean
๏ Building Boomi processes, managing/deploying/supporting APIs, Designing Boomi Flows and leveraging common frameworks such as message queuing, caching, complex mapping, and scripting (JavaScript, Groovy)
๏ Experienced on Boomi Atom Management, Process Reporting, Deployments, API Management, Flow Management
๏ Experienced on Oracle Natsuite, Mediusz, Slate and many more
๏ Extensive database development skills using SQL, Stored Procedures, Functions for various Relation Databases like Oracle, MS-SQL Server; MySQL
๏ Experience working with Salesforce, NetSuite, Big Commerce, Work Day and custom software applications
๏ Experienced and have good understanding of EDI X12 transaction sets such as 850, 855, 856, 810, 860,824,870 and 997
๏ Developed and maintained UNIX shell scripts for data-driven automatic processing
๏ Experience working with Realtime and Scheduled Processes.
๏ Roles and responsibilities mainly included technical designing, Interface designing, research and development of complex requirements, Client interaction, Onsite Coordination and Estimations, Application Development, Documentations and Code Review.
๏ Have worked in migration projects using Websphere Enterprise Service Bus, Websphere Message Broker, Web methods, Boomi
๏ Has undergone intensive training and experienced with quality process like CMMI5, Lean and agile technologies
๏ Good understanding of network (load balancer & firewall), security (digital certificates, encryption)
๏ Excellent communication & interpersonal skills.
I can ensure below parameter for my clients,
******Quality of the deliverables as per client's expectation and as set by the project delivery guidelines
******Ensure Quality standards are followed
******Adherence of Schedule
******Documentation of the work done
******Showing flexibility for working in different timezones
When working on a new project, I like to connect with my clients so that we can have a clear understanding of their needs and vision of the project. Thank you in advance for your time and consideration.
WordPress
PHP
Adobe Photoshop
IBM MQ
Agile Software Development
API Development
AngularJS
Microsoft Excel
IBM WebSphere
XML
IBM DataPower
XSLT
Dell Boomi
API
API Integration
Shantanu K.
Nagpur, India
$18/hr
5.0
7 jobs
I am a Data Architect with 5+ yearsโ experience in data engineering in Healthcare, Retail and Banking domains.
Worked on multiple data warehousing and reporting projects with open source tools such as Apache Spark, Apache Airflow, Apache Kafka, Apache Superset. I also have experiance in open building open source data warehouses.
I am microsoft certified data engineer and Azure administarator. With 5+ years of expirance in Microsoft azure technologies such as Azure Data Factory, Azure Databricks, MS SQL, Azure Synapse etc.
I am Microsoft, Databricks and GCP Certified Data Engineer and am currently working as a Data Architect at FulzTech.
Databricks Platform
Python
Java
SQL
Snowflake
Data Modeling
Microsoft Power BI
Data Warehousing & ETL Software
ETL
Microsoft Azure SQL Database
Machine Learning
Android
Angular
NoSQL Database
How it works
Post a job for freePost a job
Tell us what you need. Create your own job post or generate one with AI then filter talent matches.
Hire top talent fast
Consult, interview, and hire quickly, so you can meet the freelancers you're excited about.
Collaborate easily
Use Upwork to chat or video call, share files, and track project progress right from the app.
Payment simplified
Manage payments in one place with flexible billing options. Only pay for approved work, hourly or by milestone.
Don't just take our word for it
โUpwork provides an umbrella-level of security. I can see a talentโs work history and ratings. I can hold payments in escrow. I can communicate through Upwork Messages instead of working through my email address.โ
KD
Kim Darling
Emerald Tiger
โUpwork is the best platform to hire skilled professionals when we're not looking for a full-time employee. All the companies in our portfolio use Upwork to find talent across a wide range of fields.โ
DM
David Merry
Kinetic Investments
โOur very specific requirements can be a challengeโWith Upwork, weโre able to access a bigger community to ensure the success of our projects.โ
KK
Katja Krohn
Summa Linguae
How do I hire a IBM InfoSphere DataStage Specialist in India on Upwork?
You can hire a IBM InfoSphere DataStage Specialist in India on Upwork in four simple steps:
Create a job post tailored to your IBM InfoSphere DataStage Specialist project scope. We'll walk you through the process step by step.
Browse top IBM InfoSphere DataStage Specialist talent on Upwork and invite them to your project.
Once the proposals start flowing in, create a shortlist of top IBM InfoSphere DataStage Specialist profiles and interview.
Hire the right IBM InfoSphere DataStage Specialist for your project from Upwork, the world's largest work marketplace.
At Upwork, we believe talent staffing should be easy.
How much does it cost to hire a IBM InfoSphere DataStage Specialist?
Rates charged by IBM InfoSphere DataStage Specialists on Upwork can vary with a number of factors including experience, location, and market conditions. See hourly rates for in-demand skills on Upwork.
Why hire a IBM InfoSphere DataStage Specialist in India on Upwork?
As the world's work marketplace, we connect highly-skilled freelance IBM InfoSphere DataStage Specialists and businesses and help them build trusted, long-term relationships so they can achieve more together. Let us help you build the dream IBM InfoSphere DataStage Specialist team you need to succeed.
Can I hire a IBM InfoSphere DataStage Specialist in India within 24 hours on Upwork?
Depending on availability and the quality of your job post, it's entirely possible to sign up for Upwork and receive IBM InfoSphere DataStage Specialist proposals within 24 hours of posting a job description.
Find more freelancers
Top cities for IBM InfoSphere DataStage Specialists in India