I'm a Python developer based in Bogotรก with three years of experience building automation and data pipelines for enterprise billing operations, including EDI processing systems currently running in production.
Most of my work falls into a few categories:
- Data extraction: pulling lead lists, pricing data, or directory listings from public sources, working within each site's access rules and rate limits.
- Spreadsheet automation: replacing manual, repetitive Excel work with Python scripts that update on their own.
- Dashboards and reporting: Power BI or custom Python dashboards built around the numbers that actually matter to your business.
- ETL pipelines: recurring data collection and transformation jobs that run reliably without ongoing supervision.
I write code meant to be maintained, not just to work once, I also document what I build so someone else can pick it up if needed. Happy to start with a small paid test task before committing to a bigger project.
Comfortable working across US time zones.
Data Analysis
Automation
Data Transformation
ETL
Data Extraction
Python
Business Intelligence
B2B Lead Generation
SQL
Web Scraping
Data Cleaning
API Integration
Shipping & Order Fulfillment Software
AI Bot
Aleksei A.
Novi Sad, Serbia
$25/hr
5.0
2 jobs
๐ 10+ years of experience. Clean data. Dashboards that get used. Pipelines that don't break.
I help startups and small businesses turn raw, messy data into clear insights and automated systems โ so leadership can make faster, better decisions without drowning in spreadsheets.
I cover the full analytics stack: from defining the right metrics and structuring data, to building reliable ETL pipelines, interactive dashboards, and ML models that drive real business outcomes.
๐ฌ "Highly professional, technically strong, and reliable... His expertise in Apache Superset, Airflow, and data visualization was evident from day one. He helped to structure workflows efficiently and built meaningful dashboards." โ Upwork Client
๐ง WHAT I DO
๐ Dashboards & BI โ Interactive dashboards in Superset, Tableau, Metabase and Looker Studio. Built 40+ dashboards for product, ops, finance and marketing teams at VK and beyond.
๐ ETL & Data Pipelines โ End-to-end pipelines with Apache Airflow and Python. Automated ingestion, transformation, alerting and data quality monitoring.
๐งน Data Cleaning & Transformation โ SQL and Python (Pandas) to turn messy exports into analysis-ready datasets.
๐ ML & Predictive Modeling โ Scoring models, clustering, regression and forecasting. Built and deployed models in production at Credit Bank of Moscow and VK.
๐ Ad-hoc Analysis & Reporting โ Deep-dive analysis, A/B testing, funnel analysis, cohort analysis and weekly performance reports.
๐ค AI-Assisted Development โ I use Claude and GPT daily via VS Code to accelerate Python scripting, write and debug SQL, and speed up Airflow DAG development. AI is part of my workflow โ not a service I sell separately, but a tool that makes my delivery faster and cleaner.
๐ FEATURED PROJECT โ Telegram News Aggregator (Production)
Built a fully automated news intelligence platform that collects, processes, and distributes content from 500+ Telegram channels across 4 languages.
Hourly ETL pipeline collecting ~1,000 posts/day with AI-generated summaries and automated Telegram distribution
Hybrid search (full-text + vector RAG) via FastAPI with multi-LLM response aggregation
Telegram Mini App (React + TypeScript) for end-user content discovery
BI analytics layer in Apache Superset over PostgreSQL
12 interconnected microservices orchestrated via Airflow (CeleryExecutor)
Multi-threaded LLM processing with API key rotation โ 3-5x speed improvement
Diagnosed and fixed production PostgreSQL connection exhaustion incident
Stack: Python ยท Airflow ยท PostgreSQL ยท OpenAI GPT-4o-mini ยท LangChain ยท Weaviate ยท Elasticsearch ยท FastAPI ยท Docker ยท Redis ยท React/TypeScript ยท Superset
๐ TOOLS & STACK
BI & Visualization: Apache Superset ยท Tableau ยท Metabase ยท Looker Studio ยท Dash
Data & SQL: ClickHouse ยท PostgreSQL ยท MS SQL
Python: Pandas ยท NumPy ยท Scikit-learn ยท CatBoost
Pipelines: Apache Airflow ยท Docker ยท Git ยท Redis
AI: Claude API ยท GPT-4o ยท LangChain ยท Weaviate ยท Elasticsearch
ML: Logistic Regression ยท CatBoost ยท Clustering ยท Forecasting
๐ฏ MY PROCESS
Understand your business goal โ Define the right metrics โ Clean and structure the data โ Build reliable, scalable output โ Hand over with documentation.
No overengineering. No scope creep. Clean work, clear communication, fast turnaround.
๐ฉ Send me your project โ I'll come back with a clear plan and realistic timeline.
Data Analysis
SQL
Python
Git
LLM Prompt
Big Data
Tableau
Apache Superset
Apache Airflow
ClickHouse
Machine Learning
Kevin D.
Toledo, Ohio
$75/hr
5.0
3 jobs
Hi, my name is Kevin,
I help B2B service companies, SaaS teams, agencies and operations-heavy businesses turn messy data, manual workflows and disconnected systems into cleaner, automated business operations.
My work sits at the intersection of data science, analytics engineering, intelligent automation and business systems. I use SQL, Python, Power BI, ETL/ELT workflows, APIs, CRM automation and workflow automation tools to help teams clean up their data, trust their reporting and reduce the manual work slowing them down.
Core areas I can help with:
โข SQL, advanced queries, joins, CTEs, window functions, data modeling and query optimization
โข Python, Pandas, API integrations, automation scripts, data cleaning, transformation and validation
โข ETL/ELT pipelines, data migration, data ingestion, workflow orchestration and data quality checks
โข Power BI dashboards, reporting layers, DAX logic, KPI tracking and business intelligence workflows
โข CRM automation, lead routing, pipeline cleanup, lifecycle stages, follow-up workflows and task automation
โข HubSpot, Salesforce, Pipedrive, Zapier, Make, n8n, ClickUp, Google Sheets, Airtable and API-based integrations
โข Data QA, reconciliation, row count checks, threshold checks, deduplication, reporting accuracy and process documentation
โข RevOps, GTM analytics, marketing analytics, lead management automation, client onboarding and recurring service workflows
The problems I usually solve look like this:
โข Leads come in from multiple channels, but routing, follow-up, scoring and nurturing are inconsistent.
โข The CRM has useful data, but the pipeline is messy, lifecycle stages are unclear, duplicates keep showing up and the team does not fully trust the system.
โข Reports exist, but the numbers do not match across Salesforce, HubSpot, Stripe, Google Sheets, Power BI, Looker, or the source database.
โข Dashboards look fine on the surface, but the SQL logic, data model, or ETL process underneath is weak.
โข Managers spend too much time chasing updates, checking task status, reminding people what comes next, or manually moving work between tools.
โข Client onboarding, project delivery, recurring service management, handoffs, reminders and renewal workflows depend on too many manual steps.
โข Marketing, sales and operations teams need better visibility into KPIs, conversion rates, revenue, pipeline, client delivery and team performance, but the data is scattered across too many systems.
I look at the full business workflow, not just one dashboard, script, or automation. I trace where the data comes from, how it moves, where it breaks and what the team actually needs to make faster decisions.
Data Analysis
Data Science
Python
SQL
Data Engineering
Data Processing
Data Visualization
HighLevel
Snowflake
Scripting
Business Process Automation
Business Analysis
CRM Automation
Marketing Automation
API Integration
Salesforce
HubSpot
Airtable
Zapier
Make.com
Vianca T.
Sagรฑay, Philippines
$10/hr
5.0
7 jobs
Are you looking for a dependable freelancer who can handle detailed data, research, e-commerce, and administrative tasks accurately and on time?
I am a Top Rated Upwork freelancer with a 100% Job Success Score and more than 1,000 completed hours. I have experience supporting e-commerce businesses, startups, and global organizations with data entry, product research, listing management, CRM updates, reporting, customer support, and everyday business operations.
I can assist with:
โข Data entry, cleanup, validation, and organization
โข Excel and Google Sheets tasks
โข Product research, listing updates, and image sourcing
โข E-commerce catalog and marketplace support
โข Web research and data collection
โข CRM and database maintenance
โข PDF and document editing or conversion
โข Administrative and virtual assistance
โข Customer support and order processing
โข Reports, dashboards, and data visualization
โข Basic graphic design and content creation
Tools I have worked with include: Microsoft Excel, Google Sheets, Power BI, Looker Studio, Salesforce, HubSpot, Shopify, Canva, Adobe Photoshop, Google Workspace, Microsoft 365, ClickUp, Jira, and AI tools such as ChatGPT, Claude, and CoPilot.
My clients value my attention to detail, reliability, responsiveness, and willingness to take ownership of the work. I am comfortable following established procedures, learning new systems, and handling repetitive or detail-heavy projects.
I am available for both short-term tasks and ongoing support. Send me the project details, expected output, and deadline, and I will let you know how I can help.
Data Entry
Data Cleaning
Data Management
Online Research
Product Listings
Microsoft Excel
Google Sheets
Virtual Assistance
Administrative Support
Executive Support
CRM Software
Customer Support
Data Analysis
Data Collection
Data Quality Assessment
Looker Studio
Graphic Design
Jared J.
Lehi, Utah
$90/hr
5.0
1 jobs
AI systems are easy to demo and hard to trust. Retrieval drifts, automation turns flaky, data quietly goes wrong, and it's usually production that finds out first. I build AI and automation that stays reliable after launch, backed by 20+ years of quality and data engineering.
๐ Senior AI & Automation Engineer with Two Decades of Experience:
Across enterprise, federal, and high-security systems, including 8+ years at the IRS. My strongest areas are AI development and automation, supported by deep experience in QA, data/ETL, and API engineering. I pair modern LLM work (RAG, AI agents, workflow pipelines) with the testing and data rigor that keeps real systems dependable.
I can take you from idea to production: architecture, build, AI integration, test automation, data validation, and the guardrails that keep it all running long after the demo.
Dive into my services:
โจ AI & Automation โจ RAG & LLM apps, AI agents, Voice AI agents, workflow automation (n8n, Make, Zapier), AI-assisted test generation, chatbots & document processing
โจ Test Automation โจ Playwright, Selenium, cross-browser sharded CI, flaky-test quarantine, visual regression, accessibility & security checks
โจ QA & Data/ETL โจ functional, regression, performance & load testing, ETL & data validation (Python + SQL), data-migration & pipeline integrity
โจ Python & API โจ REST/SOAP API automation, Postman, SoapUI, Python scripting, backend validation, CI/CD (GitHub Actions)
Recent Projects:
โถ RAG Knowledge Assistant
Built a retrieval-augmented assistant grounded in a client's own documents: ingestion, chunking, embeddings, vector search, and fallback handling that keep answers accurate and inside the source material.
โถ AI Agent & Workflow System
Developed AI agents that plan and execute multi-step tasks with tool-calling, memory, and orchestration, automating real business workflows instead of one-off chat responses.
โถ Voice AI Agent
Built a real-time voice assistant combining speech-to-text, LLM reasoning, and streaming responses for natural, hands-free conversation.
โถ AI Workflow Automation
Automated end-to-end business processes with n8n, Make, and Zapier: event-triggered pipelines that classify, route, and act on data across tools with no manual steps.
โถ AI Log & Report Intelligence
Used OpenAI APIs to summarize logs, generate test plans, and draft docs across microservice environments, with reusable prompt templates that cut manual write-up time.
โถ AI-Assisted Test Automation Framework
Used LLMs (OpenAI API, Copilot) to generate suites, scaffold flows, and expand edge cases on a Playwright framework, accelerating test development ~30% while keeping human review on every line.
โถ Enterprise Test Automation & API Validation
Built cross-browser Playwright suites with sharding and flaky-test quarantine that cut CI time 30-40%, and led REST/SOAP and JMeter performance testing across on-prem and cloud for a large federal system, cutting manual effort 30%.
โถ Enterprise Data & ETL Validation
Validated end-to-end data across a mainframe to Java to Oracle modernization and high-volume payment systems (ACH, Wire, SWIFT), verifying transformation accuracy, integrity, and lineage using Python and SQL.
โถ Power Apps & Salesforce Test Automation
Authored UI test flows and Power Fx assertion suites in Power Apps Test Studio, plus Salesforce automation, complemented by Copilot-generated test plans and automated regression integration.
My Skills:
- AI Engineering: LLM applications, RAG, AI agents, embeddings, vector search, prompt engineering, workflow automation (n8n, Make, Zapier), voice AI, OpenAI, Claude, Gemini, Pinecone/FAISS.
- Automation & QA: Playwright, Selenium, PyTest, functional/regression/performance/load testing, JMeter, visual regression, axe-core, OWASP ZAP.
- Data & ETL: Python, SQL, ETL and data validation, data integrity, SQL Server, MySQL, Oracle, data pipelines.
- Backend & API: REST/SOAP APIs, Postman, SoapUI, API automation, backend validation, integration testing.
- DevOps: GitHub Actions, CI/CD, Docker, deterministic pipelines, reporting (Allure/HTML).
If you need someone who can take an AI or automation project past the demo and make it hold up in production with real testing, clean data, and dependable pipelines behind it, send me a message. ๐ง
I'll help you plan the right approach and build it properly.
Artificial Intelligence
Large Language Model
Retrieval Augmented Generation
AI Agent Development
AI App Development
OpenAI API
Vector Database
Prompt Engineering
Automation
n8n
Zapier
Python
API Development
Test Automation
Selenium
Data Engineering
Data Analysis
ETL
SQL
CI/CD
Daniel Fabrico S.
Chaco Pora, Argentina
$30/hr
5.0
1 jobs
ยธยธโฌยทยฏยทโชยทยฏยทโซยธยธ ๐ช๐ฒ๐น๐ฐ๐ผ๐บ๐ฒ ๐๐ผ ๐บ๐ ๐ฝ๐ฟ๐ผ๐ณ๐ถ๐น๐ฒ! ยธยธโซยทยฏยทโชยธโฉยทยฏยทโฌยธยธ
I'm a Senior Data Engineer & Cloud Data Architect. I bridge the gap between fragmented raw data and high-performance, analytics-ready infrastructure. Whether you need to build a scalable data warehouse from scratch, transition from legacy ETL to modern dbt/Databricks stack, optimize costly cloud queries, or power real-time AI/ML applications, I specialize in architecting reliable, zero-downtime data pipelines across AWS, GCP, Azure, Snowflake, and BigQuery.
โก ๐๐จ๐ซ๐ ๐๐๐ซ๐ฏ๐ข๐๐๐ฌ
1. End-to-End Modern Data Stack (MDS) & ETL/ELT Pipelines
Designing and deploying automated, resilient pipelines that extract, clean, transform, and load petabyte-scale data into centralized analytics hubs.
โพ Batch & Stream Ingestion: Building automated ingestion jobs from SaaS applications, REST APIs, webhooks, and legacy DBs using Fivetran, Airbyte, Kafka, and Debezium (Change Data Capture - CDC).
โพ Analytics Engineering: Modular, version-controlled transformations with dbt (Data Build Tool), custom SQL, and PySpark-complete with automated documentation and lineage tracking.
โพ Workflow Orchestration: Designing DAGs, automated retries, and monitoring alerts using Apache Airflow, Prefect, Dagster, and AWS Step Functions.
2. Cloud Data Warehousing & Lakehouse Architecture (Snowflake, Databricks, BigQuery)
Structuring high-efficiency, cost-optimized databases designed for instant analytical querying and BI dashboard performance.
โพ Warehouse Optimization: Clustering keys, partitioning, materialization, micro-partitioning, and query tuning to cut monthly cloud compute/storage costs by 30%โ60%.
โพ Lakehouse & Open Table Formats: Architecting Delta Lake, Apache Iceberg, and Hudi layers on AWS S3/GCP Cloud Storage using Medallion Architecture (Bronze -> Silver -> Gold).
โพ Data Modeling: Dimensional modeling (Kimball methodology), Star/Snowflake Schemas, Data Vault 2.0, and Wide Flat Tables (OBT) optimized for Looker, Tableau, and PowerBI.
3. Real-Time Data Streaming & Event-Driven Systems
Enabling millisecond-latency processing for live dashboards, fraud detection, dynamic pricing, and real-time operational metrics.
โพ Event Streaming: Setting up Apache Kafka clusters, AWS Kinesis, GCP Pub/Sub, and RabbitMQ with event serialization (Avro, Protobuf).
โพ Real-Time Analytics: Developing continuous stream-processing engines using Apache Flink, Spark Streaming, and ClickHouse/RisingWave for immediate insight delivery.
4. Data Quality, Governance, MLOps & AI Infrastructure
Ensuring every byte of data entering your reporting systems is accurate, secure, compliant, and ready for advanced analytics or LLM applications.
โพ Data Quality & Observability: Automated schema validation, anomaly detection, and data testing using Great Expectations, Soda, and dbt test suites.
โพ AI/ML Infrastructure: Vector database setup (Pinecone, Weaviate, Qdrant, Milvus), RAG pipeline data ingestion, and feature store integration (Feast) for AI model training.
โพ Governance & Compliance: Role-Based Access Control (RBAC), Column/Row-level masking, PII obfuscation, and automated lineage mapping for GDPR/HIPAA compliance.
โก ๐๐๐๐ก๐ง๐จ๐ฅ๐จ๐ ๐ข๐๐ฌ & ๐ ๐ซ๐๐ฆ๐๐ฐ๐ผ๐ฟ๐ค๐ฌ ๐ ๐๐ฎ๐๐ ๐๐๐ฌ๐ญ๐๐ซ๐๐
- ๐ช๐ผ๐ฟ๐ธ๐ณ๐น๐ผ๐ ๐ข๐ฟ๐ฐ๐ต๐ฒ๐๐๐ฟ๐ฎ๐๐ถ๐ผ๐ป: Apache Airflow, Prefect, Dagster, Mage, AWS Step Functions, MWAA
- ๐๐ฎ๐๐ฎ ๐ช๐ฎ๐ฟ๐ฒ๐ต๐ผ๐๐๐ฒ๐ & ๐๐ป๐ด๐ถ๐ป๐ฒ๐: Snowflake, Google BigQuery, AWS Redshift, ClickHouse, Trino/Presto, DuckDB
- ๐๐ฎ๐๐ฎ ๐๐ฎ๐ธ๐ฒ / ๐๐ฎ๐ธ๐ฒ๐ต๐ผ๐๐๐ฒ: Databricks, Apache Iceberg, Delta Lake, Apache Hudi, AWS Glue, PySpark, Apache Spark
- ๐๐ง๐ / ๐๐๐ง & ๐ง๐ฟ๐ฎ๐ป๐๐ณ๐ผ๐ฟ๐บ๐ฎ๐๐ถ๐ผ๐ป: dbt (Core & Cloud), Airbyte, Fivetran, Kafka Connect, Debezium, Meltano
- ๐ฆ๐๐ฟ๐ฒ๐ฎ๐บ๐ถ๐ป๐ด & ๐ ๐ฒ๐๐๐ฎ๐ด๐ถ๐ป๐ด: Apache Kafka, AWS Kinesis, GCP Pub/Sub, Apache Flink, Spark Streaming, RabbitMQ
- ๐๐ฎ๐๐ฎ๐ฏ๐ฎ๐๐ฒ๐ (๐ก๐ผ๐ฆ๐ค๐ & ๐ฅ๐๐๐ ๐ฆ): PostgreSQL, MySQL, MongoDB, Redis, Cassandra, DynamoDB, Pinecone, Qdrant
- ๐๐ฎ๐ป๐ด๐๐ฎ๐ด๐ฒ๐ & ๐ฆ๐พ๐น: Python (Pandas, Polars, PySpark, SQLAchemy), SQL (Advanced Dialects), Scala, Bash, Go
- ๐๐ป๐ณ๐ฟ๐ฎ๐๐๐ฟ๐๐ฐ๐๐๐ฟ๐ฒ & ๐๐ฒ๐๐ข๐ฝ๐: Terraform, Docker, Kubernetes, AWS (S3, EC2, ECS, Lambda), GCP, Azure, GitHub Actions, CI/CD
- ๐ค๐๐ฎ๐น๐ถ๐๐ & ๐ข๐ฏ๐๐ฒ๐ฟ๐๐ฎ๐ฏ๐ถ๐น๐ถ๐๐: Great Expectations, Soda, Monte Carlo, OpenLineage, Datahub
โก ๐๐ผ๐ ๐ ๐ช๐ผ๐ฟ๐ธ:
I am hired to design production-grade data pipelines, modernize legacy data stacks, fix slow analytics queries, and bring software engineering best practices (Git, CI/CD, unit testing, modular code) into data infrastructure. I prioritize clean lineage, cost efficiency, ironclad data security, and zero-downtime migrations.
โก ๐ช๐ต๐ ๐๐น๐ถ๐ฒ๐ป๐๐ ๐๐ต๐ผ๐ผ๐๐ฒ ๐ ๐ฒ:
โ๏ธ Pipelines built to scale
โ๏ธ Massive cloud bill reduction
โ๏ธ Production-grade reliability
โ๏ธ Software engineering rigor
โ๏ธ Clear communication
๐ ๐๐น๐ถ๐ฐ๐ธ ๐ ๐ฒ๐๐๐ฎ๐ด๐ฒ - ๐น๐ฒ๐'๐ ๐๐ฎ๐น๐ธ.
SQL
Python
ETL Pipeline
Data Mining
Data Integration
Data Analysis
ETL
Big Data
Data Engineering
Data Warehousing & ETL Software
Database Architecture
Database Design
Machine Learning
BigQuery
Apache Spark
Data Warehousing
Amazon Web Services
Data Scraping
Data Migration
dbt
How it works
Post a job for freePost a job
Tell us what you need. Create your own job post or generate one with AI then filter talent matches.
Hire top talent fast
Consult, interview, and hire quickly, so you can meet the freelancers you're excited about.
Collaborate easily
Use Upwork to chat or video call, share files, and track project progress right from the app.
Payment simplified
Manage payments in one place with flexible billing options. Only pay for approved work, hourly or by milestone.
Don't just take our word for it
โUpwork provides an umbrella-level of security. I can see a talentโs work history and ratings. I can hold payments in escrow. I can communicate through Upwork Messages instead of working through my email address.โ
KD
Kim Darling
Emerald Tiger
โUpwork is the best platform to hire skilled professionals when we're not looking for a full-time employee. All the companies in our portfolio use Upwork to find talent across a wide range of fields.โ
DM
David Merry
Kinetic Investments
โOur very specific requirements can be a challengeโWith Upwork, weโre able to access a bigger community to ensure the success of our projects.โ
KK
Katja Krohn
Summa Linguae
How do I hire a Knorr Associates DataPipe Specialist on Upwork?
You can hire a Knorr Associates DataPipe Specialist on Upwork in four simple steps:
Create a job post tailored to your Knorr Associates DataPipe Specialist project scope. Weโll walk you through the process step by step.
Browse top Knorr Associates DataPipe Specialist talent on Upwork and invite them to your project.
Once the proposals start flowing in, create a shortlist of top Knorr Associates DataPipe Specialist profiles and interview.
Hire the right Knorr Associates DataPipe Specialist for your project from Upwork, the worldโs largest work marketplace.
At Upwork, we believe talent staffing should be easy.
How much does it cost to hire a Knorr Associates DataPipe Specialist?
Rates charged by Knorr Associates DataPipe Specialists on Upwork can vary with a number of factors including experience, location, and market conditions. See hourly rates for in-demand skills on Upwork.
Why hire a Knorr Associates DataPipe Specialist on Upwork?
As the worldโs work marketplace, we connect highly-skilled freelance Knorr Associates DataPipe Specialists and businesses and help them build trusted, long-term relationships so they can achieve more together. Let us help you build the dream Knorr Associates DataPipe Specialist team you need to succeed.
Can I hire a Knorr Associates DataPipe Specialist within 24 hours on Upwork?
Depending on availability and the quality of your job post, itโs entirely possible to sign up for Upwork and receive Knorr Associates DataPipe Specialist proposals within 24 hours of posting a job description.
Find more freelancers
Similar Knorr Associates DataPipe Specialist Skills