Talent badge filter
Skills filter
Select talent location
Select talent time zones
$15/hr
100%
Job Success
$100+ earned
Start of list.
End of list.
Drowning in fragmented data and manual reporting? I build automated data pipelines, rigorous statistical models, and objective dashboards that replace the chaos. My goal is to deliver a single source of truth, giving your team reliable, real-time insights you can confidently act on.
━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━
𝗪𝗵𝗮𝘁 𝗜 𝗗𝗲𝗹𝗶𝘃𝗲𝗿
→ ⚙️ Data Engineering & Automation
⟶ Architecture & Pipelines (ETL/ELT): Designing end-to-end automated pipelines using Microsoft Fabric, Python, Apache Spark, and dbt to integrate messy data from ERPs, CRMs, APIs, and flat files.
⟶ Warehousing & Orchestration: Building scalable Lakehouses/Warehouses and automating workflows to eliminate manual data entry and reduce human error.
→ 📊 Business Intelligence & Dashboards
⟶ Enterprise Reporting: Building high-performance, user-focused dashboards in Power BI, Tableau, and Apache Superset that prioritize clean, objective data visualization.
⟶ Performance Optimization: Auditing and speeding up slow, bloated Power BI reports using DAX Studio, Tabular Editor, and Measure Killer.
→ 📈 Statistical Analysis & Research
⟶ Advanced Analytics: Moving beyond "what happened" to "why it happened" using R and Python for hypothesis testing, A/B analysis, and predictive modeling.
⟶ Exploratory Data Analysis (EDA): Deep-diving into complex datasets to uncover hidden patterns, delivering findings via reproducible reports (RMarkdown) or interactive web apps (R Shiny).
━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━
🔗 𝗧𝗼𝗼𝗹𝘀 & 𝗽𝗿𝗼𝗴𝗿𝗮𝗺𝗺𝗶𝗻𝗴 𝗹𝗮𝗻𝗴𝘂𝗮𝗴𝗲𝘀
Power BI (advanced DAX, Power Query, semantic modeling, DAX Studio, Tabular Editor, Bravo, Measure Killer) · Microsoft Fabric · Python · R · SQL · dbt · Soda Core · Elementary · Apache Airflow · Dagster · Airbyte · Prefect · Containers & Orchestrators (Docker, Git/GitHub, CI/CD) · Tableau · Azure · ClickHouse · PostgreSQL · MySQL · APIs · AI Agentic Development · Microsoft Excel / Google Sheets · PySpark, SparkSQL, SparkR
━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━
𝗣𝗿𝗼𝗳𝗲𝘀𝘀𝗶𝗼𝗻𝗮𝗹 𝗖𝗲𝗿𝘁𝗶𝗳𝗶𝗰𝗮𝘁𝗶𝗼𝗻𝘀
✅ Microsoft Fabric Data Engineer (DP-700)
✅ Microsoft Fabric Analytics Engineer (DP-600)
✅ Microsoft Power BI Data Analyst (PL-300)
━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━
I combine data modeling best practices with a commitment to objective, truth-revealing visualization. I deliver projects on time, meticulously documented, and ensure clear communication from day one.
I would love to hop on a quick call to discuss your project. The first consultation is on me!
$120/hr
$0 earned
Start of list.
End of list.
Are your data systems quietly killing your marketing? Disconnected CRM, ad platforms optimising on the wrong signals, sales teams not knowing who to chase? Most businesses don't have a marketing problem, they have a data problem underneath it. Book a free 30-minute data audit and I'll show you exactly where your setup is blocking growth and what to fix first.
I've seen problematic data infrastructure across ecommerce brands, growth-stage startups, and marketing agencies: broken tracking, pipelines patched together, and nobody trusts the numbers. So decisions get made on gut feel while the data sits there, unused.
I build the data layer that fixes that: clean, scalable, and designed to power both the humans and the AI systems that sit on top of it.
𝗪𝗵𝗮𝘁 𝘁𝗵𝗮𝘁 𝗹𝗼𝗼𝗸𝘀 𝗹𝗶𝗸𝗲 𝗶𝗻 𝗽𝗿𝗮𝗰𝘁𝗶𝗰𝗲:
Behavioural tracking that actually captures what users do on your websites and apps. Not just page views, but intent signals, session sequences, and identity stitching across anonymous and known users
Data pipelines built on tools you own: Snowplow, ClickHouse, Airflow, dbt. No vendor lock-in, no black box
Predictive models that push intelligence where it's used: propensity scores into your CRM, LTV forecasts into your ad platforms, lead signals to your sales team
A real time signals API so your marketing automation and AI agents can query behavioural context on demand
𝗥𝗲𝘀𝘂𝗹𝘁𝘀 𝗜'𝘃𝗲 𝗱𝗲𝗹𝗶𝘃𝗲𝗿𝗲𝗱:
For a high-ticket retail brand spending £250k/month on PPC, I built a behavioural scoring system that identified which site visitors were worth calling and which would convert without intervention. The result: 9-10% incremental conversion lift on outbound calls, and a clear model for where sales effort actually moves the needle.
Across other engagements I've helped brands secure grant funding through data-driven AI development, built the analytics foundation for an AI team from scratch, solved complex data challenges in online shopping aggregation, and integrated CRM and on-site analytics to optimise campaigns for high-value leads.
𝗪𝗵𝗮𝘁 𝗰𝗹𝗶𝗲𝗻𝘁𝘀 𝘀𝗮𝘆:
• "Mitch's expertise in the AI and data space was instrumental in helping us secure grant funding, which allowed us to expand our team." — Pangea
• "Always made an effort to understand our business and contributed heavily towards the creation of our AI team." — RedBrain
• "A driving force in solving the data challenges of online shopping aggregation, delivering remarkable results even on a shoestring budget." — Aisle3
• "Got our team going with several quick wins and set a strong platform for us to grow our in-house data analytics ops." — Hampr
• "Partnering closely with us to integrate our marketing activities with our on-site analytics and CRM, enabling us to optimise campaigns for high-value leads." — Marketcheck
𝗧𝗵𝗶𝘀 𝗶𝘀 𝘁𝗵𝗲 𝗿𝗶𝗴𝗵𝘁 𝗳𝗶𝘁 𝗶𝗳:
• Your ad platforms are missing conversion data they need to optimise
• You're sitting on behavioural data but can't act on it fast enough
• You want a modern data stack without hiring a full team
• You're building AI-powered workflows and need a reliable signal layer underneath
𝗡𝗼𝘁 𝘁𝗵𝗲 𝗿𝗶𝗴𝗵𝘁 𝗳𝗶𝘁 𝗶𝗳:
• You need a quick dashboard fix with no interest in the foundations
• You're looking for the cheapest option rather than the right one
𝗥𝗲𝗮𝗱𝘆 𝘁𝗼 𝗺𝗮𝗸𝗲 𝘆𝗼𝘂𝗿 𝗱𝗮𝘁𝗮 𝗮𝗰𝘁𝘂𝗮𝗹𝗹𝘆 𝘄𝗼𝗿𝗸?
Message me and we'll spend 20 minutes working out whether there's something here worth building together.
$25/hr
$0 earned
Start of list.
End of list.
Hi, I’m PandaJune, a Senior Java Backend & Data Processing Engineer with over 9 years of professional experience designing and building high-performance backend systems and large-scale data workflows.
I help clients turn raw, complex, or fragmented data into clean, reliable, and actionable insights through efficient backend architecture and automation.
💻 What I Do Best
Backend Development: Java (8–17), Spring Boot, RESTful API design, multithreading, and microservices
Data Processing / ETL: Build and optimize pipelines handling 100K–300K TPS; real-time ingestion and transformation
Database Design & Optimization: MySQL, PostgreSQL, ClickHouse, StarRocks, Redis, Kafka
System Integration: Custom TCP/UDP/WebSocket protocols, API integrations, data synchronization between systems
DevOps & Deployment: Linux scripting, Jenkins automation, component packaging and delivery
⚙️ Why Clients Work With Me
9 + years of hands-on experience across telecom, data analytics, and security industries
Proven ability to design end-to-end data systems — from architecture to deployment
Strong communication, clear documentation, and on-time delivery
I’m currently open to remote, long-term collaborations focused on backend engineering, ETL automation, or data-driven systems.
If you need a developer who can make your data flow smarter and faster — let’s connect!
$75/hr
$0 earned
Start of list.
End of list.
Self-managed data platforms on infrastructure you control. 25 years of production Linux - 13 of them running multi-terabyte ClickHouse, PostgreSQL and Apache NiFi on bare metal and private virtualisation for tier-1 telecom operators (Airtel, MTN, Orange, Cell C, Glo, Etisalat) - including the storage and hardware underneath.
Highlights:
• Cut multi-server rollouts from ~4 weeks to ~3 days with a custom self-installing Rocky Linux ISO.
• Standardised company-wide PostgreSQL backups on pgBackRest with tested point-in-time recovery - not just "backups".
• Managed multi-terabyte ClickHouse databases for telecom-grade analytics.
• Designed MongoDB databases with high availability and backups.
• Delivered CIS-compliant server hardening with automated security baselines.
Core Expertise
✅ Linux Systems Engineering – Red Hat, Rocky, Ubuntu, CentOS, SLES
✅ Automation & Scripting – Bash, Ansible, Python
✅ Databases & Data Pipelines – PostgreSQL, ClickHouse, MongoDB, Redis, Apache NiFi
✅ Infrastructure & Virtualization – VMware, Proxmox, Docker, HP ProLiant, Supermicro
✅ Monitoring & Security – Zabbix, CIS hardening, server performance tuning
✅ Web & Application Services – Nginx, reverse proxies, API integration
Proven Experience
Multiple infrastructure deployments across Africa, Asia, and the Middle East.
How I Work
I’m passionate about:
• Solving critical infrastructure issues under pressure
• Automating complex, repetitive processes
• Delivering secure, resilient systems that scale
• Implementing new ideas from scratch
Let's Work Together
If you need someone who can:
• Stabilise and secure your Linux environment
• Deliver automation that saves time and reduces errors
• Ensure stability across Linux, cloud, and hybrid environments
• Implement a new idea
👉 Then I’m here to help. Let’s make your infrastructure faster, safer, and more reliable.
Working remotely, UTC+8. Core hours 09:00–18:00; UK/EU afternoons and evenings by arrangement.
$20/hr
$0 earned
Start of list.
End of list.
Data Engineer specializing in SQL, ClickHouse, Python, and Excel. I help businesses improve data quality through data validation, profiling, and analysis.
$25/hr
$300+ earned
Available now
Start of list.
End of list.
📊 10+ years of experience. Clean data. Dashboards that get used. Pipelines that don't break.
I help startups and small businesses turn raw, messy data into clear insights and automated systems — so leadership can make faster, better decisions without drowning in spreadsheets.
I cover the full analytics stack: from defining the right metrics and structuring data, to building reliable ETL pipelines, interactive dashboards, and ML models that drive real business outcomes.
💬 "Highly professional, technically strong, and reliable... His expertise in Apache Superset, Airflow, and data visualization was evident from day one. He helped to structure workflows efficiently and built meaningful dashboards." — Upwork Client
🔧 WHAT I DO
📈 Dashboards & BI — Interactive dashboards in Superset, Tableau, Metabase and Looker Studio. Built 40+ dashboards for product, ops, finance and marketing teams at VK and beyond.
🔄 ETL & Data Pipelines — End-to-end pipelines with Apache Airflow and Python. Automated ingestion, transformation, alerting and data quality monitoring.
🧹 Data Cleaning & Transformation — SQL and Python (Pandas) to turn messy exports into analysis-ready datasets.
📊 ML & Predictive Modeling — Scoring models, clustering, regression and forecasting. Built and deployed models in production at Credit Bank of Moscow and VK.
🔍 Ad-hoc Analysis & Reporting — Deep-dive analysis, A/B testing, funnel analysis, cohort analysis and weekly performance reports.
🤖 AI-Assisted Development — I use Claude and GPT daily via VS Code to accelerate Python scripting, write and debug SQL, and speed up Airflow DAG development. AI is part of my workflow — not a service I sell separately, but a tool that makes my delivery faster and cleaner.
🚀 FEATURED PROJECT — Telegram News Aggregator (Production)
Built a fully automated news intelligence platform that collects, processes, and distributes content from 500+ Telegram channels across 4 languages.
Hourly ETL pipeline collecting ~1,000 posts/day with AI-generated summaries and automated Telegram distribution
Hybrid search (full-text + vector RAG) via FastAPI with multi-LLM response aggregation
Telegram Mini App (React + TypeScript) for end-user content discovery
BI analytics layer in Apache Superset over PostgreSQL
12 interconnected microservices orchestrated via Airflow (CeleryExecutor)
Multi-threaded LLM processing with API key rotation — 3-5x speed improvement
Diagnosed and fixed production PostgreSQL connection exhaustion incident
Stack: Python · Airflow · PostgreSQL · OpenAI GPT-4o-mini · LangChain · Weaviate · Elasticsearch · FastAPI · Docker · Redis · React/TypeScript · Superset
🛠 TOOLS & STACK
BI & Visualization: Apache Superset · Tableau · Metabase · Looker Studio · Dash
Data & SQL: ClickHouse · PostgreSQL · MS SQL
Python: Pandas · NumPy · Scikit-learn · CatBoost
Pipelines: Apache Airflow · Docker · Git · Redis
AI: Claude API · GPT-4o · LangChain · Weaviate · Elasticsearch
ML: Logistic Regression · CatBoost · Clustering · Forecasting
🎯 MY PROCESS
Understand your business goal → Define the right metrics → Clean and structure the data → Build reliable, scalable output → Hand over with documentation.
No overengineering. No scope creep. Clean work, clear communication, fast turnaround.
📩 Send me your project — I'll come back with a clear plan and realistic timeline.
$35/hr
$20 earned
Start of list.
End of list.
I’ve migrated a 1.2 billion-row data warehouse and cut its infrastructure costs by 70%. I build the systems that make data trustworthy — ETL pipelines, data-quality frameworks, and LLM evaluation harnesses that catch problems before they reach production.
Over 5 years I’ve worked across the full data lifecycle for companies in the US, Germany, and UAE:
DATA ENGINEERING — Designed ClickHouse enrichment pipelines processing 760M+ B2B rows. Migrated a 1.2B-row warehouse off ClickHouse Cloud to self-hosted infrastructure (~70% cost reduction). Built Airflow-orchestrated Spark/ETL jobs with idempotent transforms and retry logic across 10+ pipelines. Implemented sub-minute ingestion with ClickPipes.
AI / LLM ENGINEERING — Built an automated evaluation system for chatbot and voicebot agents (LangChain, LangGraph) measuring intent accuracy, retrieval quality, and hallucination rate on every release — lifting client success rate 70%. Integrated LLMs into data tooling, improving data quality 35%.
QUALITY ENGINEERING — 2.5 years as a Senior SQA Engineer supporting global e-commerce rollouts: functional, regression, and performance testing, Python/Selenium automation integrated into CI/CD, and root-cause analysis that drove a 23% traffic increase. Built fill-rate and match-rate regression checks on a 760M-row dataset, catching distribution drift before release.
Tools I work with daily: Python (FastAPI, Flask, Pytest, pandas), Advanced SQL, ClickHouse, PostgreSQL, Apache Spark, Airflow, Selenium, AWS (S3, EC2, Glue), Docker, Terraform, Grafana/Prometheus/Loki, LangChain, LangGraph.
How I work: I overlap with US and EU business hours, reply within a couple of hours, and send progress updates without being asked. I’d rather flag a risk early than surprise you late.
Message me with your data, AI, or quality problem — I’ll respond with an honest assessment of whether I’m the right fit and how I’d approach it.
$28/hr
$0 earned
Start of list.
End of list.
I help teams turn complex business requirements into scalable software, data, and automation solutions. My focus is on backend systems, data architecture, workflow automation, and performance optimization, especially where reliability and maintainability matter.
I bring experience in software architecture, PostgreSQL, ClickHouse, Redis, Temporal, Kafka, Django, Python, Go, and system integration. I have also led technical teams and helped modernize legacy systems, build data-driven products, and improve operational efficiency.
If your project needs someone who can go beyond implementation and contribute to architecture, planning, and technical decisions, I’m a strong fit.
$55/hr
100%
Job Success
$40K+ earned
Available now
Offers consultations
Start of list.
End of list.
Experienced Senior Data Engineer | AWS, Big Data, and Cloud Solutions Specialist
With over 8 years of experience as a Senior Data Engineer, I specialize in designing and implementing scalable data solutions across cloud platforms such as AWS and GCP. I have a proven track record in optimizing complex data pipelines, schema enforcement, and managing large-scale ETL processes. Skilled in Clickhouse, AWS Glue, Athena, Lambda, BigQuery, and Snowflake, I bring expertise in data architecture, cost reduction, and workflow automation. I'm passionate about delivering data-driven insights and leveraging cloud technologies to streamline processes for businesses.
Let's collaborate to build efficient, scalable, and cost-effective data solutions!
$30/hr
44%
Job Success
$600+ earned
Start of list.
End of list.
Data Analyst with 10+ years driving measurable growth across GameDev, EdTech, E-commerce, and FinTech.
Recently completed A/B Testing Bootcamp at Yandex School of Data Analysis (2025) and hold GenAI certifications from Google AI Studio & DeepLearning.AI (2024). Also co-founded data science community in Istanbul (200+ attendees, 100+ lessons).
What I deliver:
• Marketing Mix Models (Robyn) → Optimized budget allocation at Innova (2022-2024)
• Propensity/Predictive Models → 60% campaign efficiency boost at Yandex ($700K/month spend clients)
• A/B Testing & Experiments → 7.5% ARPU & 30% MAU growth at Uchi.ru
• AI-Enhanced Forecasting → Reduced revenue variance at Nexters Hero Wars (2025)
• End-to-End BI Dashboards → Tableau, Looker, Power BI, Python
Technical Stack: SQL (BigQuery/ClickHouse/PostgreSQL/MySQL), Python (pandas, NumPy, Plotly), GenAI tools, Airflow, AppsFlyer, SKAN, GTM.
Languages: English (advanced), Russian (native).
My analytical approach combining MMM, propensity models, and experiments has consistently driven double-digit growth. I think in commercial impact and prioritize issues that directly affect revenue.
Available for remote roles. Let's turn your data into growth-driving insights.