Design and deploy AI pipelines that transform image, video, and audio data into reliable real-world applications, from computer vision analytics and 3D perception to audio (speech, music, sound) analysis and generative AI systems.
My work focuses on production AI systems, not research prototypes. I help companies move from early feasibility studies and PoC development to scalable deployments running in real environments.
Typical projects include:
• computer vision pipelines for detection, tracking, and segmentation
• real-time video analytics and edge AI deployments
• 3D vision and spatial perception systems
• speech processing and audio AI solutions
• generative AI pipelines for image, video, and audio
Computer Vision Systems
Design and development of advanced computer vision pipelines for image and video analysis.
Typical solutions include:
• object detection and multi-object tracking
• semantic and instance segmentation
• pose estimation and motion analysis
• OCR and document understanding
• visual search and recognition systems
These systems are used in industries such as manufacturing, sports analytics, retail, healthcare, and security.
3D Vision & Spatial AI
Development of AI systems that understand spatial structure and depth.
Experience includes:
• structure-from-motion (SfM)
• photogrammetry pipelines
• depth estimation models
• NeRF and neural rendering
• point cloud processing and 3D reconstruction
Applications include robotics, AR/VR, construction analytics, and digital twins.
Edge AI & On-Device ML
I specialize in deploying ML models on mobile and embedded devices where latency, memory, and power constraints are critical.
Typical optimization techniques include:
• model quantization and pruning
• architecture optimization
• real-time inference pipelines
• deployment on mobile and embedded hardware
Technologies include:
TensorRT, TensorFlow Lite, CoreML, ONNX Runtime.
Many deployed systems operate with 50–100 ms inference latency depending on hardware.
Generative AI for Vision & Video
Development of generative pipelines for media processing and synthetic data generation.
Typical solutions include:
• image and video generation pipelines
• diffusion-based editing and enhancement
• synthetic dataset generation for model training
These tools help accelerate AI training and improve model robustness.
Audio & Speech AI
Development of AI systems for speech processing, audio analysis, and voice technologies.
Examples include:
• phoneme segmentation and pronunciation analysis
• speech recognition pipelines
• voice feature extraction and audio analytics
• generative audio and music models
These systems are used in:
• language learning platforms
• speech therapy tools
• voice biometrics systems
• music AI applications
Technical Stack
Frameworks & Models
PyTorch, TensorFlow, OpenCV, Detectron2, MediaPipe, YOLO, DINO, SAM, CLIP
Deployment
TensorRT, TensorFlow Lite, CoreML, Docker, ONNX Runtime, FastAPI
Programming
Python, C, C++
3D Vision
NeRF, SLAM, Dust3r, point clouds
Leadership & R&D
I lead an R&D-focused AI team at It-Jim, an AI consulting company with 30+ engineers and 10+ PhDs specializing in:
• Computer Vision
• Generative AI
• Audio & Speech AI
• Edge AI systems
We help companies solve technically challenging AI problems and build reliable production systems.
If you are looking for experienced AI engineers to design, prototype, or deploy advanced machine learning solutions: feel free to reach out!
Deep Learning
Computer Vision
Machine Learning
Artificial Intelligence
Image Processing
Video Processing
OpenCV
PyTorch
TensorFlow
Edge AI
AI Model Development
AI App Development
Python
Solution Architecture
AI Audio Generation
Automatic Speech Recognition
Digital Signal Processing
Generative AI
Object Detection & Tracking
AI Video Generation
Andrii P.
Zaporizhzhia, Ukraine
$55/hr
5.0
46 jobs
Experienced in:
1. Computer vision
• C++, Python, OpenCV, CUDA, Git, Linux, Qt, Boost, OpenGl, PCL, Strong math background, Neural Networks
- road segmentation for unmanned vehicles (ENet, Caffe, OpenCV, C++, Linux)
- car tracking (Yolo v3, OpenCV, C++, Linux)
- wagon number identification (Yolo v4, Python)
- implementation of real time 360°/perspective camera transformation on Cuda (C++,
Cuda, OpenCV, Linux, Jetson Nano)
- distance calculation to point on 2D camera frame (C++, OpenCV, Linux)
- automate grading system for handwritten answer sheets (computer vision part, OpenCV,
Java, Android
Key stack: Linux, C++ (Qt), Python, Java, OpenCV, Yolo (darknet)
2. Machine Learning
Machine learning research projects in the following domains:
- person segmentation (ModNet, RVM, TDNet, UCTransNet, XMem etc.)
- image inpainting (Pen-Net, Deepfillv2, Shift-Net, ViNet etc.)
- image upscale (RDN, RRDN, Stable Diffusion, ISR etc.)
- image relighting (Total Relighting, DPR, RelightNet etc.)
- road segmentation for unmanned vehicles (ENet, Caffe, OpenCV, C++, Linux)
- car tracking (Yolo v3, OpenCV, C++, Linux)
- wagon number identification (Yolo v4, Python)
- implementation of real time 360°/perspective camera transformation on Cuda (C++,
Cuda, OpenCV, Linux, Jetson Nano)
- distance calculation to point on 2D camera frame (C++, OpenCV, Linux)
- automate grading system for handwritten answer sheets (computer vision part, OpenCV,
Java - Android, IOS - Swift)
Key stack: Linux, Python, Pytorch, Tensorflow, OpenCV, Pillow, Numpy, C++, CUDA, Darknet, SegNet
3. Robotics and Embedded development:
- OCPP Protocol, Linux, Modbus, Raspberry Pi, CAN;
-ROS Robot operating system;
- Skilled in SLAM, localization, mapping
- Experienced in path planning algorithms, obstacle avoidance, holonomic, and non-holonomic motion planning, trajectory planning for robotics arms;
- Used to work with Bayesian/Kalman filters, and sensor fusion (LiDAR, IMU, Visual, Odometry, Radar, GPS).
4. Android development (Kotlin, Java, Android Studio, Eclipse, Firebase).
Has expert colleagues in:
• .Net Framework (C#, VB.Net, ASP.Net, .Net Core, WPF, UWP,WCF, ADO.Net)
• Java (j2se, j2ee, servlets, java beans, Maven)
• JavaScript (Node.JS, Express.js, Vue.js, Element.js, Angular.js, D3.js)
• C++ (TCP/IP, HTTP, HTTPS, WebSocket, Modbus)
• Python (MAVLink, WebSocket)
• Step7 (S7 Communication, OCPP, Modbus, CANOpen, ProfiNet)
Database
• PostgreSQL
• MySQL
• MongoDB
• Microsoft SQL
• Oracle Database
• Neo4J
Software development for mobile platforms
• Crossplatform React Native, Flutter, Xamarin,
• Android (Kotlin, Java, Android Studio, Eclipse)
• iOS (Objective C, Swift)
Mobile apps development:
• Crossplatform: Futter, React Native, Xamarin.
• Android (Java, Kotlin)
• iOS (Objective C, Swift)
1. Native Development
- Kotlin/Java
- Swift / Objective-C
- iOS/macOS/tvOS/watchOS
- Firebase, CloudKit, Coredata
2. Cross-Platform and Hybrid App Development
- React Native/React
- Flutter / Dart
- Xamarin.iOS / Xamarin.Android / Xamarin.Forms
tech stack
● Android Studio, Gradle, Kotlin DSL, KSP
● Kotlin, Java programming languages
● AndroidX, Android Jetpack libraries, Android Architecture Components
● Jetpack Compose
● Material Design Components
● Clean Architecture, SOLID design principles
● MVVM, MVI, GoF design patterns
● Modularization (multi-module projects)
● Kotlin Coroutines + Flow, RxJava, RxBinding
● REST API / Networking - OkHttp, Retrofit 2, Socket IO
● Room Database, SQLite, Datastore
● Kotlinx Serialization, Protobuf, Moshi, Gson
● Dependency Injection (Hilt, Dagger 2, Koin)
● Git
● Firebase Products, Google Cloud APIs, HMS Services
● Admob, Google Play Billing Library (in-app purchases), Samsung/Huawei IAP
● Unit / Instrumented (UI) tests
● Agile Scrum development methodology
● CI/CD (GitHub Actions)
AR/VR:
Vuforia
ARkit/ARcore/AR Foundation
Wikitude
Oculus Integration
OpenXR
XR Interaction toolkit
VR Walkthrough
UltimateXR
VR Interaction Framework
Hardware expert:
• Nvidia Jetson Nano, TX2, Xavier;
• Raspberry Pi;
• Arduino, STM32;
• Depth cameras Intel Realsense d435i , Zed Sterelabs.
• Lidars, Radars.
Charging stations for the Electric Vehicles development software for the managing stations and networks (server and user applications):
• C#, SQL, PostgreSQL, .Net Core, REST Api, WebSockets
• OCPP Protocol, Linux, Modbus, Raspberry Pi
Software development
• .Net Framework (C#, VB.Net, ASP.Net, .Net Core, WPF, UWP,WCF, ADO.Net)
• Java (j2se, j2ee, servlets, java beans, Maven)
• JavaScript (Node.JS, Express.js, Vue.js, Element.js, Angular.js, D3.js)
• C++ (TCP/IP, HTTP, HTTPS, WebSocket, Modbus)
• Python (MAVLink, WebSocket)
• Step7 (S7 Communication, OCPP, Modbus, CANOpen, ProfiNet)
Database
• PostgreSQL
• MySQL
• MongoDB
• Microsoft SQL
• Oracle Database
• Neo4J
Deep Learning
Computer Vision
Machine Learning
Data Science
OpenCV
YOLO
Object Detection
Neural Network
Android App Development
Mobile App Development
English
CUDA
DNN
OCR Algorithm
Kotlin
Xamarin
Front-End Development
Desktop Application
.NET Framework
Oleg P.
Kyiv, Ukraine
$80/hr
5.0
23 jobs
AI Developer specializing in Generative AI, Computer Vision, and Audio AI systems.
I design and deploy production-grade AI solutions built from scratch for image, video, audio, and multimodal applications.
My work focuses on system-level development: from architecture design and model training to optimization and deployment, delivering scalable AI components ready for real-world environments.
Core Expertise
Generative AI Systems
▪️ Diffusion-based image and video generation (SDXL, FLUX, ControlNet)
▪️ Controlled generation, inpainting, style transfer
▪️ Text-to-video and 3D generation pipelines
▪️ Neural speech synthesis and AI audio generation
▪️ Synthetic data generation for model training
Computer Vision
▪️ Object detection, tracking, and segmentation
▪️ OCR, pose estimation, face recognition
▪️ 3D reconstruction and depth estimation
▪️ Real-time and edge-optimized vision systems
Audio AI & DSP
▪️ Audio segmentation and source separation
▪️ Spectrogram-based modeling and transformer audio embeddings
▪️ Speech recognition and synthesis systems
▪️ Feature-driven audio analysis pipelines
▪️ Cross-modal (audio–vision–text) modeling
Machine Learning & Architecture
▪️ Custom neural network architecture design (CNNs, U-Nets, Transformers)
▪️ End-to-end ML lifecycle: data strategy → training → validation → deployment
▪️ Performance optimization for accuracy, latency, and hardware constraints
Deployment & Edge Systems
▪️ Cloud, on-premise, and edge deployment
▪️ Dockerized inference services
▪️ ONNX / TensorRT optimization
▪️ Mobile and embedded AI systems
What I Deliver
▪️ Custom AI system development (from scratch)
▪️ Structured R&D and technical validation
▪️ Production-ready ML / CV / multimodal pipelines
▪️ Scalable architecture aligned with product and business goals
Deep Learning
Deep Neural Network
Computer Vision
Machine Learning
Image Processing
Artificial Intelligence
Generative AI
AI Audio Generation
AI Image Generation
Stable Diffusion
Python
Digital Signal Processing
Neural Network
AI Consulting
AI Model Development
3D Modeling
AI Development
Object Detection & Tracking
Machine Learning Algorithm
AI Mobile App Development
Sergii G.
Kyiv, Ukraine
$43/hr
5.0
7 jobs
I represent an award-winning company called LITSLINK. We are a product-oriented development team (which consists of 250+ highly qualified software engineers, UI/UX designers, business analysts, project architects, project managers, and QA engineers) experienced in almost all popular areas:
✅ AI: Machine Learning, Computer Vision, Natural Language Processing, Big Data Analytics
✅No-code, Low-code: Bubble, WebFlow
✅ Blockchain: Solidity, Rust, Solana/Near blockchains, smart contracts
✅ Cross-Platform Mobile development React Native, Flutter
✅ Web development: Front-end (React.JS, Vue.JS, Angular.JS); Back-end (Node.JS, Java, Python)
✅ AR/VR: Untiy3D/C#, Hololens, Hololens 2, GearVR, MS Mixed Reality, Oculus, HTC Vive, iOS/Android (AIRKit, ARCore)
We provide smart and fast solutions on time and on budget.
By hiring us, you are getting the expertise of our software architect board and the whole team's knowledge. Whether you need a single developer or a whole team, we are here to help.
Deep Learning
Artificial Intelligence
Machine Learning
Python
Natural Language Processing
Chatbot Development
JavaScript
API
AI Agent Development
TensorFlow
Data Science
AI App Development
Automation
Computer Vision
LangChain
Artificial Neural Network
OCR Software
OCR Algorithm
LLM Prompt Engineering
PyTorch
Volodymyr K.
Lviv, Ukraine
$50/hr
5.0
122 jobs
AI Developer Top-rated Plus (Upwork Top 1%) AI Agents, Generative AI & Full Stack Development
I help companies as AI Developer design, validate, and ship AI products fast, with a strong focus on AI Agents, Generative AI, and proAI Developer (Upwork Top 1%) AI Agents, Generative AI & Full Stack Development
8+ Years in AI & Machine Learning — production systems, not prototypes
50+ AI Products Delivered — from MVP to enterprise scale
Top 1% AI Developer on Upwork
AI MVPs shipped in 4–8 weeks — structured delivery, no overengineering
HIPAA / SOC2 / GDPR-ready AI systems
Full-cycle delivery — architecture, development, deployment, and optimization
Trusted by teams in Healthcare, FinTech, Logistics, Real Estate, and E-commerce
As a Senior AI Developer and Full Stack Software Engineer, I build production-ready Generative AI applications, custom AI Agents, and scalable machine learning systems.
With 8+ years in Machine Learning and software architecture, I work as an AI Developer, AI Engineer, and Full Stack Developer. Together with Indeema, a full-cycle AI and custom software company, I lead teams that turn AI strategy into deployed systems, from data pipelines and LLM orchestration to scalable backend architecture and enterprise integrations.
Proven track record with a 100% Job Success Score
Consistently earning 5★ client feedback
8+ years building production-ready AI systems
Clean, scalable, and enterprise-grade architecture
From idea to deployed MVP in 4–8 weeks
My structured delivery approach takes products from idea to working MVP in 4–8 weeks.
AI Agent Consulting & Delivery
I consult and implement agentic AI systems that observe, reason, and act across business workflows:
AI Agents & Autonomous Workflows
LLM-powered Products and Copilots
RAG Systems and Knowledge AI
AI Orchestration Architectures
AI Integration into existing platforms
AI MVP validation and architecture design
My approach combines strategy, architecture, and delivery helping companies avoid costly overengineering and ship production-ready AI faster.
What I Deliver
AI Agent for Document Management
AI Agent for Real Estate
AI Agent for Inventory Management
AI Agent for Fintech and Financial Management
AI Agent for Medical Clinics and Image Analysis
AI System for Advanced Analytics
Agentic AI System Design
Generative AI Development
Machine Learning and Predictive Analytics
AI App Development and Automation
NLP, Computer Vision, and Data Science
Enterprise AI Integration and Modernization
I work across Healthcare, FinTech, Logistics, Real Estate, and E-commerce, building systems that improve operations, reduce manual work, and unlock new revenue streams.
AI MVPs in 4–8 weeks using a structured delivery methodology
HIPAA / SOC2 / GDPR-ready solutions
AI integration with ERP, CRM, and legacy systems
Dedicated AI teams and fractional CTO leadership
⚙️ Technologies
Python, TensorFlow, PyTorch, LangChain, LLM APIs (OpenAI, Claude, Gemini), AWS SageMaker, Vertex AI, Docker, Kubernetes, MLOps pipelines, vector databases, real-time AI systems.
If you’re looking for a partner who can design the right AI solution and actually ship it, let’s talk.
Deep Learning
Machine Learning
Artificial Intelligence
Python
Amazon Web Services
React
Data Science
Node.js
Natural Language Processing
Artificial Neural Network
SQL
JavaScript
Git
CSS 3
PHP
Computer Vision
TensorFlow
Neural Network
Supervised Learning
Unsupervised Learning
Vadym S.
Kharkiv, Ukraine
$75/hr
5.0
28 jobs
🏅 Expert-Vetted Top 1% on Upwork
🏆 Top 10 Machine Learning Agency on Upwork
🎖 7+ Years Delivering Production AI Systems
Struggling to turn video or image data into reliable production AI? With 7+ years as a Computer Vision Engineer, I’ve delivered 20+ production systems and helped cut manual inspection time by 35% for sports, industrial, satellite, and healthcare teams. Let’s build detection, tracking, segmentation, or edge AI that works in the real world.
🎯 HOW I CAN HELP YOU
COMPUTER VISION SYSTEMS
Turn visual data into accurate, production-ready outputs that improve decisions, automate manual review, and scale across real-world conditions. I build systems for detection, tracking, segmentation, pose estimation, and video analysis that stay robust under occlusion, low light, and multi-camera setups.
SPORTS ANALYTICS AI
Convert broadcast or field video into player tracking, ball trajectories, and performance metrics that help teams measure more and react faster. My sports systems have covered football, tennis, pickleball, and 3D ball tracking with calibration and real-time inference optimized for usable match insights.
INDUSTRIAL INSPECTION AUTOMATION
Reduce defects, speed up quality control, and catch issues before they hit production. I have delivered defect detection and segmentation systems for manufacturing use cases, including fabric defects, roof inspection, and image stitching pipelines for precision analysis.
SATELLITE AND AERIAL IMAGE ANALYSIS
Extract actionable intelligence from drone and satellite imagery to support infrastructure monitoring, mapping, and large-area inspection. I use YOLO-based detection and segmentation pipelines that handle sparse labels, variable resolution, and geospatial complexity.
REAL-TIME EDGE AI DEPLOYMENT
Get low-latency AI running on NVIDIA Jetson, cloud APIs, or embedded devices without sacrificing accuracy. I optimize models with TensorRT, ONNX, Docker, and C++ when needed, helping clients meet strict speed and hardware constraints in production.
🤝 𝗚𝗘𝗧 𝗜𝗡 𝗧𝗢𝗨𝗖𝗛
Message me on Upwork if you need a Computer Vision Engineer or Machine Learning Engineer to turn your idea into a production-ready AI system. I can help with a free consultation, scope the best approach, and move fast on your next build.
I typically respond within 4 hours and can prioritize urgent projects.
𝗛𝗼𝘄 𝗜 𝗪𝗼𝗿𝗸
I build computer vision systems with OpenCV and PyTorch, productionize machine learning models in TensorFlow, and deploy Python pipelines for real-time video analysis and image segmentation.
Deep Learning
Deep Neural Network
Computer Vision
Machine Learning
Python
Artificial Intelligence
OpenCV
PyTorch
Data Science
Image Processing
TensorFlow
Automation
C++
Natural Language Processing
Keras
Data Entry
Neural Network
Image Recognition
3D Modeling
Photogrammetry
How it works
Post a job for freePost a job
Tell us what you need. Create your own job post or generate one with AI then filter talent matches.
Hire top talent fast
Consult, interview, and hire quickly, so you can meet the freelancers you're excited about.
Collaborate easily
Use Upwork to chat or video call, share files, and track project progress right from the app.
Payment simplified
Manage payments in one place with flexible billing options. Only pay for approved work, hourly or by milestone.
Don't just take our word for it
“Upwork provides an umbrella-level of security. I can see a talent’s work history and ratings. I can hold payments in escrow. I can communicate through Upwork Messages instead of working through my email address.”
KD
Kim Darling
Emerald Tiger
“Upwork is the best platform to hire skilled professionals when we're not looking for a full-time employee. All the companies in our portfolio use Upwork to find talent across a wide range of fields.”
DM
David Merry
Kinetic Investments
“Our very specific requirements can be a challenge—With Upwork, we’re able to access a bigger community to ensure the success of our projects.”
Top interview questions to help you hire the right Deep Learning Experts, faster.
How do I hire a Deep Learning Expert in Ukraine on Upwork?
You can hire a Deep Learning Expert in Ukraine on Upwork in four simple steps:
Create a job post tailored to your Deep Learning Expert project scope. We'll walk you through the process step by step.
Browse top Deep Learning Expert talent on Upwork and invite them to your project.
Once the proposals start flowing in, create a shortlist of top Deep Learning Expert profiles and interview.
Hire the right Deep Learning Expert for your project from Upwork, the world's largest work marketplace.
At Upwork, we believe talent staffing should be easy.
How much does it cost to hire a Deep Learning Expert?
Rates charged by Deep Learning Experts on Upwork can vary with a number of factors including experience, location, and market conditions. See hourly rates for in-demand skills on Upwork.
Why hire a Deep Learning Expert in Ukraine on Upwork?
As the world's work marketplace, we connect highly-skilled freelance Deep Learning Experts and businesses and help them build trusted, long-term relationships so they can achieve more together. Let us help you build the dream Deep Learning Expert team you need to succeed.
Can I hire a Deep Learning Expert in Ukraine within 24 hours on Upwork?
Depending on availability and the quality of your job post, it's entirely possible to sign up for Upwork and receive Deep Learning Expert proposals within 24 hours of posting a job description.