You will get a data pipeline / ETL system built and automated


Project details
You get a data pipeline that runs on its own - no more manual exports, no more 'someone forgot to update the spreadsheet.' I build the ingestion, cleaning, and loading logic, wire it into your database or warehouse, and automate it to run on schedule.
Why me: I built big-data ETL pipelines and customer-facing data features from scratch at Outbrain, handling large-scale data flows end-to-end - cleaning, validating, and moving data reliably at production traffic levels. Full-stack background (TypeScript, Node.js, PostgreSQL, AWS, Azure) means the pipeline fits directly into your existing stack, not a separate system you have to babysit.
Typical builds: API-to-database pipelines, multi-source data consolidation, scheduled batch jobs, data cleaning and validation layers, migration from spreadsheets/manual processes to a real pipeline.
Starter: one source, cleaned and loaded into your database. Standard: multiple sources, scheduled runs, error handling so failures don't fail silently. Advanced: full warehouse setup with orchestration and monitoring/alerting when something breaks.
You stop babysitting data manually; the pipeline does it for you, reliably.
Why me: I built big-data ETL pipelines and customer-facing data features from scratch at Outbrain, handling large-scale data flows end-to-end - cleaning, validating, and moving data reliably at production traffic levels. Full-stack background (TypeScript, Node.js, PostgreSQL, AWS, Azure) means the pipeline fits directly into your existing stack, not a separate system you have to babysit.
Typical builds: API-to-database pipelines, multi-source data consolidation, scheduled batch jobs, data cleaning and validation layers, migration from spreadsheets/manual processes to a real pipeline.
Starter: one source, cleaned and loaded into your database. Standard: multiple sources, scheduled runs, error handling so failures don't fail silently. Advanced: full warehouse setup with orchestration and monitoring/alerting when something breaks.
You stop babysitting data manually; the pipeline does it for you, reliably.
Data Tool
SQLWhat's included
| Service Tiers |
Starter
$500
|
Standard
$1,400
|
Advanced
$2,800
|
|---|---|---|---|
| Delivery Time | 4 days | 9 days | 16 days |
Number of Revisions | 1 | 2 | 3 |
About Amir
Data Scientist and AI Engineer
Ramat Gan, Israel - 9:18 pm local time
Former Chief Data and AI Scientist at Skyhawk Security, where I led the first-ever integration of Generative AI into cloud threat detection products. Published researcher in AI security with coverage in TechTarget and industry media.
Inventor on 12 patents in AI/ML, fraud detection, anomaly detection, and cloud security (NICE Ltd, Skyhawk Security, Outbrain).
What I deliver:
- LLM integration pipelines (Claude, GPT, Gemini) - prompt engineering, structured output, RAG, evaluation frameworks
- Production ML systems - classification, NLP, computer vision, recommendation engines
- Data engineering - ETL pipelines, data processing, analytics dashboards
- Cloud infrastructure - Azure, AWS, Docker, CI/CD
Education: MSc Machine Learning and Data Science (Tel Aviv University), dual BSc Computer Science + Mathematics (Technion).
Previously: Data Engineer at Outbrain, ML systems at IBM, Teaching Assistant at Technion (Calculus, Linear Algebra, Image Processing).
I specialize in taking AI from prototype to production - fixing output quality, optimizing inference pipelines, and building reliable systems that handle real-world data. Available for short-term sprints and ongoing engagements.
Steps for completing your project
After purchasing the project, send requirements so Amir can start the project.
Delivery time starts when Amir receives requirements from you.
Amir works on your project following the steps below.
Revisions may occur after the delivery date.
Source and destination mapping
Pipeline build and testing