Most data pipelines don’t fail because of code. They fail because they weren't built for scale.
With 8+ years of experience engineering data systems at companies like Microsoft and Coreweave, I help businesses move away from "brittle prototypes" to production-grade, scalable infrastructure.
I don’t just move data; I build the "Source of Truth" that leadership and AI systems actually trust.
💬What I Solve for You:
Productionizing AI Pipelines: Hardening Python prototypes into scalable RAG and LLM infrastructures (AWS/Azure).
➔Infrastructure-as-Code: Building automated, modular ETL/ELT pipelines that don't require daily manual fixes.
➔The "One-Source" Dashboard: Integrating messy data from APIs, SaaS (Shopify, HubSpot), and DBs into clean Snowflake/BigQuery layers.
➔Performance Recovery: Optimizing slow SQL queries and high-cost cloud warehouses to save you thousands in monthly spend.
🛠 Tech Stack:
Languages: Python (FastAPI, Pandas, PySpark), SQL
Cloud & Warehousing: AWS (Glue, Lambda, S3), Snowflake, BigQuery, Azure
Orchestration: Airflow, dbt, GitHub Actions
Data Ops: API Integrations, Vector DBs, Data Validation
✅ Why Me?
8+ Years Experience: I’ve seen what breaks at the enterprise level and how to prevent it in your startup.
Speed over Perfection: I focus on shipping high-impact systems that drive revenue, not just technical documentation.
Transparent Communication: You get regular updates and a partner who challenges requirements to find better solutions.
Ready to clean up your data debt?
📩 Message me for a FREE 15-minute technical consultation. Let’s discuss your architecture and see if I’m the right fit for your system.
Data Engineering
Python
ETL Pipeline
SQL
Apache Spark
Apache Airflow
Snowflake
Amazon Web Services
BigQuery
Data Warehousing
Data Modeling
Apache Kafka
PostgreSQL
Data Integration
Tableau
Docker
Jonas N.
Kill Devil Hills, North Carolina
$100/hr
5.0
29 jobs
I am a Google Cloud Certified Data Engineer who builds production-ready data pipelines, scalable RAG/vector infrastructure, and resilient AI workflow automations.
I help growing companies and enterprise teams modernize their cloud architecture, optimize runaway BigQuery costs, and structure complex data for real-world AI applications.
How I Can Help Your Team:
• AI & Vector Infrastructure: Building end-to-end RAG pipelines, unstructured document extraction (PDFs/docs to LLM formats), and BigQuery Vector Search integration.
• GCP Cost Optimization & FinOps: Auditing and refactoring expensive BigQuery queries and Dataflow jobs to significantly cut monthly cloud bills.
• Modern Data Engineering: Scalable ETL/ELT pipelines, dbt modeling, and robust cloud orchestration using Prefect, Cloud Run, and n8n.
Core Tech Stack: GCP (BigQuery, Cloud Run, Dataflow, Vertex AI), Python, SQL, dbt, Prefect, n8n, Snowflake, Terraform, Vector Databases.
If you need clean, enterprise-grade data infrastructure built to scale without the headache, let's talk.
Microsoft Excel
Python
Data Visualization
SQL
Looker
Visualization
ETL Pipeline
GIS
Database
Data Engineering
Data Analytics
Analytics
ETL
Kashif S.
Gudja, Malta
$25/hr
5.0
8 jobs
I build data platforms that work at scale and keep working as your
business grows.
Over the past 10 years I've served as the lead or founding data engineer
across fintech, e-commerce, ride-hail, legal tech, and cybersecurity
companies. That means I've designed systems from scratch, made architecture
decisions with no one to fall back on, and delivered platforms that
product teams actually use.
Here's what I typically get hired to do:
→ Build greenfield data platforms on AWS or GCP from the ground up
→ Design and ship production ETL/ELT pipelines (Airflow, Dagster, dbt)
→ Set up scalable warehouses and governance (Snowflake, BigQuery, Redshift)
→ Implement real-time streaming pipelines (Kafka, Spark Streaming, CDC)
→ Build AI-powered data applications (RAG, LLMs, LangChain, vector DBs)
→ Fix broken or unreliable pipelines and make them production-grade
→ Architect cloud infrastructure on AWS, GCP, Azure (Terraform, Kubernetes)
Recent work includes:
- Led data platform engineering for a US e-commerce company processing
billions of events daily. I re-architected ingestion pipelines, built
Snowflake governance from scratch, introduced Prometheus monitoring and
CI/CD standards across the platform.
- Built a full data platform on GCP (BigQuery, Dataproc, Airflow) for a
music streaming company. Firebase, AppsFlyer, and app store data all
flowing into one warehouse within weeks.
- Designed an AWS data platform for a ride-hail company managing 500+
streaming and 700+ batch jobs — including a self-serve portal that
replaced multi-step CLI workflows for engineers.
- Built a legal AI search engine using LangChain, Pinecone, and RAG —
full pipeline from document ingestion to LLM-generated answers, deployed
on AWS with auto-scaling.
- Built an AI inventory insights agent for a US automotive company —
multi-source data pipelines, real-time APIs, conversational interface.
I work in English daily, communicate proactively, and deliver production-
ready code — not prototypes. I'm used to working directly with CTOs and
technical leads in US and European time zones.
Tools I work with regularly:
Python · SQL · Airflow · Dagster · dbt · Snowflake · BigQuery · Spark · Meltano ·
Kafka · AWS (S3, EMR, Glue, ECS, Lambda, EC2, EKS) · Databricks · GCP · Azure
Terraform · Docker · Kubernetes · LangChain · FastAPI · MLflow · Weaviate, Celery
If you're building a data platform, fixing one, or adding AI/ML
capabilities to your stack, let's talk.
Python
Google Cloud Platform
Microsoft Azure
Amazon Web Services
Data Engineering
Docker
DevOps
GitHub
BigQuery
Snowflake
Apache Airflow
Apache Spark
Terraform
ETL
Apache Kafka
Shahid B.
Islamabad, Pakistan
$15/hr
5.0
8 jobs
Messy data slowing your team down? I build scalable ETL/ELT pipelines and modern cloud architectures on Azure, Databricks, Fabric, and Snowflake that turn raw, chaotic data into clean, analytics-ready systems fast and reliably.
I bridge the gap between fragmented data sources and production-grade dashboards, seamlessly adapting to your existing infrastructure rather than forcing an expensive rebuild.
What I Can Help You With:
Data Warehouse & Lakehouse Architecture: Implementing Medallion design patterns (Bronze → Silver → Gold) using Delta Lake, Microsoft Fabric OneLake, and Snowflake.
Scalable ETL/ELT Ingestion: Building automated, metadata-driven pipelines via Azure Data Factory, Fabric Pipelines, Databricks (PySpark/SQL), and dbt.
Real-Time Data Streaming: Architecting low-latency workflows using Apache Kafka, Azure Event Hubs, and streaming engines.
Database Design & Optimization: Performance tuning, indexing, and data modeling for PostgreSQL, Azure SQL, and cloud warehouses.
Proven Project Highlights:
Microsoft Fabric Incremental Pipeline: Built a control-table pattern using Get Metadata, Lookup, and ForEach loops to orchestrate zero-duplicate, quarterly ingestion from SharePoint into OneLake via Dataflow Gen2.
Azure/Databricks Streaming: Developed a restaurant analytics platform processing 80,000+ events/day, cutting reporting lag from 6 hours to under 3 minutes.
Kafka/Snowflake Pipeline: Engineered a real-time stock market data pipeline tracking 120+ tickers with under 8 seconds end-to-end latency.
I write clean, documented code your team can maintain long-term and provide transparent daily updates.
Message me with your data challenge and I’ll walk you through exactly how to solve it.
Data Engineering
Data Modeling
Data Warehousing & ETL Software
Database Design
Microsoft Azure
Snowflake
Databricks Platform
Azure Service Fabric
Apache Kafka
PostgreSQL
SQL
Apache Spark
Python
Docker
Git
dbt
Ben R.
McKinney, Texas
$60/hr
5.0
1 jobs
I build the data plumbing that turns messy sources into decisions.
I'm a data engineer and analytics pro who owns the full modern data stack end to end — from ingestion to activation. I design and maintain ETL and reverse ETL pipelines that sync data reliably between warehouses, CRMs, and operational tools, so the people who need clean data actually get it where they work.
What I bring:
Pipelines & ingestion: ETL/ELT with Fivetran, Airbyte, and custom Python; orchestration via Airflow
Transformation & modeling: Advanced SQL and dbt — modular models, testing, documentation, and dimensional/star-schema design that scales
Reverse ETL: Operationalizing warehouse data back into Salesforce, HubSpot, and marketing tools with Hightouch/Census
Warehousing: BigQuery and GCP; performance tuning and cost control
Reliability: Data quality tests, CI/CD, and version-controlled workflows (git) so pipelines don't break silently
I get what it's like inside a small team: you don't need a data platform built for a Fortune 500, you need something that works, doesn't break silently, and doesn't cost a fortune to run. I've built a data function from the ground up and care as much about the business outcome as the code. Whether it's a one-off pipeline, a warehouse setup, or a dbt project that needs cleaning up, I'll deliver something clear, tested, and maintainable.
Let's talk about what you're trying to unlock in your data.
ETL
ETL Pipeline
Data Analysis
Task Automation
Business Analysis
Business Consulting
Information Analysis
Marcus G.
Johnson City, Tennessee
$121/hr
4.3
130 jobs
It's a pleasure to meet you!
I am Marcus Green, Ph.D. (っ◔◡◔)っ ♥ I put the soul in solution. ♥ Not ready to hire? Request a free fixed rate price quote from a professional on this job today!
𝐒𝐞𝐫𝐯𝐢𝐜𝐞. 𝐈𝐧𝐭𝐞𝐥𝐥𝐢𝐠𝐞𝐧𝐜𝐞. 𝐑𝐞𝐥𝐚𝐭𝐢𝐨𝐧𝐬𝐡𝐢𝐩𝐬. 𝐘𝐞𝐬, 𝐒.𝐈.𝐑.!
Service we are here to serve. Intelligence means we diagnose, build, test, and innovate. Relationships means we do not win one contract and disappear
I founded CTLR PLUS SOLUTIONS, a Tennessee-based marketing analytics, business intelligence, AI workflow, and MVP development lab.
We fix broken tracking, messy dashboards, disconnected tools, unclear revenue data, and custom technology problems.
Proof
I maintain a 100% Job Success Score (This is totally not easy!), Top Rated Plus status, and 120+ client engagements.
Review my Upwork feedback. plussolutions.ai. drgreencoconuttree on YouTube to see how I think, explain, and solve.
I am not an outsourcing agency hiding behind a generic logo. The work carries my name. 💠
What do we fix?
Most organizations are not short on tools. They are short on clarity.
You may already have GA4, GTM, Google Ads, Meta, Shopify, HubSpot, Salesforce, Looker Studio, Power BI, Airtable, Zapier, Make, BigQuery, Snowflake, or a CRM.
The problem is simple: the tools do not agree.
The website says one thing. GA4 says another. The CRM says something different. The ad platform takes credit for everything. The dashboard looks nice but does not answer the real business question. 🔹
We fix tracking, connect data, build dashboards in Looker Studio and Power BI, support BigQuery, Snowflake, and SQL workflows, and build custom AI/MVP tools.
The goal:
See what drives revenue, what wastes money, and what must happen next.
How your money works 🌐
Step 1: Audit
Most engagements start with a $99 Audit.
We review GA4, GTM, Google Ads, Meta, events, website tracking, attribution, reporting gaps, dashboards, and next steps.
The audit answers:
What is broken, missing, confusing, or underbuilt?
Step 2: Fix
We fix GA4 events, GTM setup, Google Ads conversions, Meta Pixel/CAPI, form tracking, click tracking, checkout tracking, CRM mapping, UTMs, attribution, and reporting mismatches.
If tracking is wrong, the dashboard is decoration.
We help you stop guessing. 🔷
Step 3: Build
Some clients need a $2,500 dashboard sprint.
Others need executive reporting, Power BI, Looker Studio, BigQuery/Snowflake architecture, AI workflows, MVPs, scrapers, Chrome extensions, internal tools, automations, or private business systems.
Larger AI and MVP builds can range from $15k to $30k+, depending on complexity.
The point is to build the right thing.
Step 4: Consult
Building the thing is not enough.
We help clients interpret results, explain performance, train teams, review funnels, analyze campaigns, understand attribution, improve dashboards, and make better decisions.
A dashboard without interpretation is just a pretty screen.
Step 5: Retain 🌀
Markets change. Platforms change. Tracking breaks. Campaigns launch. Teams ask new questions.
Many clients continue through advisory retainers, analytics support, development retainers, or ongoing business intelligence relationships.
Typical retainers range from $1,800 to $6,800+/month, depending on scope.
We are always in the solution business.
Are we alone?
No.
I lead strategy, diagnosis, business logic, client communication, and solution direction.
CTLR PLUS SOLUTIONS also uses trusted developers, data engineers, and automation specialists when needed.
You get senior strategy plus build support.
How do we interview?
Message me on Upwork and ask for a free thought-partnering session or $99 audit of your stack.
We clarify the problem, review the tools, identify the fastest path, and decide if we are the right fit.
Next steps 🔵♞💘 ˜”❀°•.˜”✿°• ɢʀᴇᴀᴛ! •°★”˜.•°✨”˜ 💙 🧭
Hire or Invite.
Fund the $99 Audit.
Send access to GA4, GTM, Google Ads, Meta, website, CRM, or dashboard tools.
𝙔𝙤𝙪𝙧 𝙛𝙧𝙞𝙚𝙣𝙙 𝙖𝙣𝙙 𝙖𝙣𝙖𝙡𝙮𝙨𝙩, -𝙈𝙖𝙧𝙘𝙪𝙨
If your business is tired of unclear tracking, disconnected tools, confusing dashboards, expensive guesswork, or half-built systems, you are in the right place. Tell your friends we do:
GA4 upgrade and event tracking completed successfully.
Power BI labor productivity and budget dashboard delivered.
Data Studio dashboards created with excellent communication.
Strategic consulting helped clients clarify testing, user experience, and performance.
MVP and Chrome extension projects moved clients in the right direction.
Marketing Analytics
Google Analytics 4
Looker Studio
Google Tag Manager
Google Ads
Google Search Console
HubSpot
SEMrush
Looker
LookML
Google Search
ETL
BigQuery
Google Analytics Report
Web Analytics
How it works
Post a job for freePost a job
Tell us what you need. Create your own job post or generate one with AI then filter talent matches.
Hire top talent fast
Consult, interview, and hire quickly, so you can meet the freelancers you're excited about.
Collaborate easily
Use Upwork to chat or video call, share files, and track project progress right from the app.
Payment simplified
Manage payments in one place with flexible billing options. Only pay for approved work, hourly or by milestone.
Don't just take our word for it
“Upwork provides an umbrella-level of security. I can see a talent’s work history and ratings. I can hold payments in escrow. I can communicate through Upwork Messages instead of working through my email address.”
KD
Kim Darling
Emerald Tiger
“Upwork is the best platform to hire skilled professionals when we're not looking for a full-time employee. All the companies in our portfolio use Upwork to find talent across a wide range of fields.”
DM
David Merry
Kinetic Investments
“Our very specific requirements can be a challenge—With Upwork, we’re able to access a bigger community to ensure the success of our projects.”
KK
Katja Krohn
Summa Linguae
How do I hire a Merrill DataSite Specialist on Upwork?
You can hire a Merrill DataSite Specialist on Upwork in four simple steps:
Create a job post tailored to your Merrill DataSite Specialist project scope. We’ll walk you through the process step by step.
Browse top Merrill DataSite Specialist talent on Upwork and invite them to your project.
Once the proposals start flowing in, create a shortlist of top Merrill DataSite Specialist profiles and interview.
Hire the right Merrill DataSite Specialist for your project from Upwork, the world’s largest work marketplace.
At Upwork, we believe talent staffing should be easy.
How much does it cost to hire a Merrill DataSite Specialist?
Rates charged by Merrill DataSite Specialists on Upwork can vary with a number of factors including experience, location, and market conditions. See hourly rates for in-demand skills on Upwork.
Why hire a Merrill DataSite Specialist on Upwork?
As the world’s work marketplace, we connect highly-skilled freelance Merrill DataSite Specialists and businesses and help them build trusted, long-term relationships so they can achieve more together. Let us help you build the dream Merrill DataSite Specialist team you need to succeed.
Can I hire a Merrill DataSite Specialist within 24 hours on Upwork?
Depending on availability and the quality of your job post, it’s entirely possible to sign up for Upwork and receive Merrill DataSite Specialist proposals within 24 hours of posting a job description.