Hire the Best Cluster Computing Developers

Clients rate our Cluster Computing Developers
Rating is 4.9 out of 5.
4.9/5
Based on 289 client reviews
Mile P.

Skopje, North Macedonia

$10/hr
5.0
4 jobs

I'm a Computer Science graduate from FCSE Skopje with a 9.3/10 GPA and proven experience building production systems across the full stack. Backend Development: I work daily with Python and FastAPI for high-performance services, have extensive experience with Spring Boot and Spring Security for enterprise applications, and build RESTful APIs with proper authentication and data persistence layers using Spring Data JPA. Frontend & Full-Stack: I've developed user interfaces with Angular, ReactJS and NextJS and have hands-on experience creating complete full-stack applications from database to UI. Databases & Data: Proficient with PostgreSQL for relational data, Neo4j for graph databases, pgvector for vector similarity search, and Redis for caching in distributed systems. Cloud & Infrastructure: I architect and deploy microservices on Microsoft Azure using Terraform (Infrastructure as Code), containerize applications with Docker and Docker Compose, and have experience with AWS services including Glue for ETL pipelines. AI & Machine Learning: Currently building production RAG (Retrieval-Augmented Generation) systems, developing agentic architectures with LangGraph and AutoGen, and implementing AI workflows with Langfuse for monitoring and observability. DevOps & Best Practices: Established CI/CD pipelines using GitHub Actions with pre-commit hooks and linters, designed distributed systems with RabbitMQ messaging queues, and implement comprehensive monitoring with Spring Boot Actuator. I've proven I can own projects end-to-end, from architecture to deployment, delivering systems that handle thousands of concurrent users. I can help optimize your infrastructure, accelerate your development process, and build scalable solutions that work reliably in production.

  • Java
  • MySQL
  • Git
  • DevOps
  • Redis
  • Microsoft Azure
  • LangChain
  • Spring Boot
  • Terraform
  • PostgreSQL
  • Python
  • FastAPI
  • Docker
  • Retrieval Augmented Generation
  • MongoDB
  • React
  • Next.js
  • TypeScript
  • Node.js
  • RabbitMQ
Rodrigo E.

Vicente Lopez, Argentina

$29/hr
4.9
137 jobs

Senior Linux/DevOps Engineer & AI Automation Specialist Senior Linux/DevOps Engineer and AI Automation Specialist with 15+ years of experience designing, building, and scaling cloud-native infrastructure and intelligent workflow systems. My background combines deep cloud architecture expertise (AWS-focused) with hands-on experience building AI-powered automation using n8n, LLM APIs (OpenAI, OpenRouter), vector databases, and secure cloud environments. Most recently, I've led DevOps and cloud-security engineering for a U.S.-based healthcare SaaS platform (physician scheduling) operating under HIPAA and SOC 2. The work spans secure AWS operations, fleet patch management across mixed Windows/Linux estates, EKS lifecycle management, and building the security tooling and audit evidence that underpin a SOC 2 program — turning compliance requirements into reliable, automated, production-grade controls. Cloud & Infrastructure Expertise Over 12 years of AWS experience designing secure, scalable, production-grade environments across healthcare, banking, and enterprise sectors. AWS Services EC2, RDS, EKS, VPC, IAM, S3, CloudFront, Route53, Elastic Beanstalk, SQS, EFS, ElastiCache, Redshift Systems Manager (SSM), GuardDuty, Security Hub, AWS Config, Inspector, CloudTrail Infrastructure as Code Terraform & CloudFormation — multi-account secure environments ADR-driven design and Strangler Fig migration strategy Containers & Orchestration Kubernetes (EKS), Docker, ECS, Docker Swarm EKS cluster lifecycle management — Kubernetes version upgrades across dev/staging/prod CI/CD Jenkins, GitLab CI/CD, CircleCI, CodePipeline, ArgoCD GitHub Actions with ARC (Actions Runner Controller), Azure DevOps self-hosted agents GitOps & Kubernetes Configuration Management ArgoCD (declarative continuous delivery, app-of-apps pattern, multi-cluster sync), Kustomize (base/overlay structuring, environment promotion, patch management), Helm (chart authoring, versioning, values management across environments). GitOps workflows with full audit trail and rollback capabilities. Designed and operated production GitOps pipelines for EKS clusters in regulated environments. High Availability HAProxy, Nginx, Load Balancers, Heartbeat Led infrastructure teams in banking environments and designed highly available Kubernetes production systems for U.S.-based healthcare companies. Security, Compliance & Governance (SOC 2 / HIPAA) Hands-on implementation and continuous evidence for compliance programs in regulated healthcare environments. SOC 2 control implementation and evidence collection across multiple AWS accounts (logical access, audit logging, encryption, change management, patch management) HIPAA-aware architecture: least-privilege IAM, encryption at rest and in transit, end-to-end audit trails, and EDR coverage across the fleet AWS-native security stack: GuardDuty, Security Hub, AWS Config, Inspector, CloudTrail Prowler (soc2_aws framework) and Vanta (GRC) for continuous control mapping and posture monitoring AWS IAM Identity Center (SSO): MFA enforcement, permission sets, SAML federation, CC6.1 remediation EDR migration: CrowdStrike Falcon → Microsoft Defender, with continuity of endpoint protection enforced as a hard gate Immutable S3 evidence stores, fleet-wide backup-policy review, and EBS encryption rollout for compliance Multi-account delegated-administrator security posture (Security Hub / Config) Architecture Decision Records (ADRs) for traceable, auditable change governance Windows Server & Mixed-Fleet Operations Operated a ~70-instance EC2 fleet (Windows Server + Linux) plus ~35 EKS worker nodes AWS Systems Manager: Patch Manager (AWS-RunPatchBaseline), Session Manager, Fleet Manager, Run Command Established a formal monthly patch cadence; raised Windows patch compliance from a low baseline to ~84% Windows Server lifecycle / EOL planning: Server 2016 → 2022 upgrade and decommission roadmaps aligned to patch-management policy KB-level risk triage (e.g., hostname-length reboot risks on domain members) with documented remediation Active Directory domain controllers and SQL Server hosts (HA production pairs, RDS member servers, performance tuning) Safe-patch workflow with pre-patch EBS snapshots and post-reboot service validation Observability & Monitoring Datadog, New Relic, Zabbix — agent consolidation and cost optimization across EKS and production hosts Networking & Secure Access Tailscale (subnet routers, zero-trust access), VPC design, Route53, Load Balancers AI Automation & Workflow Engineering Beyond infrastructure, I specialize in AI-powered workflow automation: n8n workflow design (cloud & self-hosted) LLM integration (OpenAI, OpenRouter, local models) RAG architectures (Qdrant, vector search, embeddings) AI-driven CRM automation Intelligent document processing Automated client intake & communication systems Secure AI pipelines for regulated industries I design workflows that are secure, easy to maintain, scalable

  • Python
  • NGINX
  • MySQL
  • MariaDB
  • DevOps
  • Amazon EC2
  • Google Cloud Platform
  • Linux System Administration
  • Amazon S3
  • Network Administration
  • Bash Programming
  • n8n
  • Microsoft Azure
  • Claude
  • OpenAI API
  • PostgreSQL
  • Amazon RDS
  • Docker
  • Microsoft Windows PowerShell
  • Windows Server
Francisco S.

Valparaiso, Chile

$73/hr
5.0
4 jobs

Hi, I'm Fran 👋 I architect and ship production AI systems and cloud infrastructure that actually scale. → Architected & shipped a production AI copilot (agentic, RAG-grounded, human-in-the-loop) now serving customers → Sr DevOps running cloud infra for a NASDAQ-listed biotech, supporting Twist Bioscience (NASDAQ: TWST, $2B+) → Cut report generation time 50% at IBM ($150B+ market cap) with Python microservices → Cut cloud infrastructure costs 30% for a US biotech SaaS company using GCP rightsizing + autoscaling → 2× release velocity at a US biotech SaaS by streamlining CI/CD → Built recurring AI consulting from $0 to $1,500+/client serving LATAM tech professionals 8+ years building production systems that move real money 24/7. CKA + CKAD certified (Linux Foundation / Cloud Native Computing Foundation). 💼 What I do: → Cloud architecture (AWS, GCP, Azure) — Kubernetes, Terraform, ArgoCD → AI Agents, RAG & Automation — Claude / OpenAI, Python, production-grade → Infrastructure cost optimization — typical 25-40% savings → CI/CD pipeline acceleration — typical 3-5× speedup → Production systems on your existing stack (no rip-and-replace) 🎯 Best fit for: → B2B SaaS with infrastructure scaling challenges → Legal/professional firms needing AI document automation → Marketing agencies needing content automation systems → Teams needing senior engineering on fractional/project basis Stack: Kubernetes · Terraform · ArgoCD · AWS · GCP · Azure · Python · Go · Claude Code · GitHub Actions Let's chat 👇

  • Python
  • DevOps
  • Kubernetes
  • Docker
  • Terraform
  • AI Agent Development
  • CI/CD
  • Cloud Architecture
  • Google Cloud Platform
  • Infrastructure as Code
  • Amazon Web Services
  • Microsoft Azure
  • Prometheus
  • Grafana
  • Bash
  • HighLevel
  • n8n
  • Make.com
Ugochukwu O.

Abuja, Nigeria

$30/hr
4.9
23 jobs

Project Experience Blazestack - Senior AWS DevOps, MLOps and AI Evaluation Engineer. Dittofi UK - Senior AWS & Azure Cloud Solutions Architect. Kenyan Central Bank - Azure Cloud and Full-Stack Engineer - building and provisioning a Learning Management System on Azure. Rose Digital New York - AWS Cloud Engineer and Senior full-stack developer -optimising the New York Lottery services on AWS. Asksopi.ai - Senior Azure DevOps and LLM Expert. Tissimo, Musitech LLC - Lead Web Developer and GCP cloud engineer. StartupBuilder - Senior full-stack engineer.

  • Amazon Web Services
  • AWS Lambda
  • Golang
  • Node.js
  • Docker
  • DevOps Engineering
  • MLOps
  • Terraform
  • CI/CD
  • Serverless Stack
  • Cloud Architecture
  • Generative AI
  • Amazon API Gateway
  • Cybersecurity Management
  • Prompt Engineering
Richard M.

Bani, Dominican Republic

$100/hr
4.7
126 jobs

Leveraging over two decades of experience in Technology and High-Frequency Trading (HFT), I've worked in premier institutions including Aviva, FTSE, Merrill Lynch, Bank of New York, Kingdon Capital, Syneos Health Services, HSBC, Dresdner, Royal Bank Of Canada, Mitsubishi, Credit Suisse, and Insch Kintore. I bring to the table proficiency in: Trading Systems & Market Infrastructure ✅ HFT & Market Making: Ultra-low-latency systems, FPGA acceleration, co-location (NY4/TY3), FIX/ITCH/OUCH. ✅ Forex ECNs: Deep expertise in Cboe FX, EBS, and institutional FX market microstructure. ✅ Hedge Fund Infrastructure: OMS/PMS, real-time risk engines, portfolio analytics, multi-asset connectivity. ✅ Prediction Markets: Oracle integration, automated settlement, and market resolution protocols. ✅ Trading Expertise: Order book microstructure, AMM strategies, CEX dynamics, ICT/SMC (MT5). Exchanges, Blockchain & Tokenization ✅ Crypto Exchange: Matching engines, custody integration, KYC/AML, fiat on/off ramps. ✅ DeFi & DEX: AMMs, liquidity pools, cross-chain bridges, decentralized trading infrastructure. ✅ Blockchain & L1/L2: Custom chains, consensus mechanisms, rollups, validator infrastructure. ✅ Tokenization & RWA: Security tokens, real-world asset tokenization, regulatory frameworks. ✅ Crypto & Web3: Solana, Ethereum, QuickNode, Jupiter, Fireblocks. AI, Quant & Advanced Modelling ✅ AI/ML for Finance: Quant pipelines, alpha signal generation, feature engineering, and modeling across Forex and crypto markets. ✅ LLM Systems: LangChain, LangGraph, OpenAI MCP, agentic workflows, observability, and data reconciliation. ✅ Frameworks: PyTorch, TensorFlow, Transformers (low-rank, factorized), LLMs (GPT, BERT, PaLM). ✅ Advanced Methods: Deep Q-Learning, Reinforcement Kalman Filters, Mean Field Games. Engineering, Data & Low-Latency Stack ✅ Languages & Systems: C, C++, C#, .NET, Python, Rust, Go, TypeScript, Solidity, CUDA, Verilog/VHDL. ✅ Data & BI: SQL, MDX, XML, XBRL, SSAS, Power BI, star schema design. ✅ Real-Time Data: Kdb+/q, in-memory columnar streaming systems. ✅ Low-Latency Tech: FPGAs, SmartNICs, FIX, CME, NASDAQ, Trading Technologies, QuantConnect. Cloud, DevOps & Platforms ✅ Containerization: Docker, Kubernetes. ✅ Cloud: AWS (ECS, EMR, Glue, Firehose, SageMaker), Azure, Azure ML, DevOps. ✅ Data Platforms: Databricks. ⭐ I pride myself on my stellar reviews and feedback which are a testament to my dedication to excellence and client satisfaction. ⭐ My robust portfolio and strong work history demonstrate the breadth and depth of my expertise.

  • Cluster Computing
  • C#
  • C++
  • Python
  • Artificial Intelligence
  • Ethereum
  • Quantitative Finance
  • Deep Neural Network
  • Data Science
  • X86 Assembly Language
  • Mathematical Modeling
  • FPGA
  • CUDA
  • NumPy
  • Performance Optimization
Abad N.

Islamabad, Pakistan

$25/hr
5.0
3 jobs

𝗪𝗵𝘆 𝗸𝗲𝗲𝗽 𝗳𝗶𝗴𝗵𝘁𝗶𝗻𝗴 𝗳𝗮𝗶𝗹𝗲𝗱 𝗱𝗲𝗽𝗹𝗼𝘆𝗺𝗲𝗻𝘁𝘀, 𝗵𝗶𝗴𝗵 𝗰𝗹𝗼𝘂𝗱 𝗯𝗶𝗹𝗹𝘀, 𝗮𝗻𝗱 𝗽𝗿𝗼𝗱𝘂𝗰𝘁𝗶𝗼𝗻 𝗶𝘀𝘀𝘂𝗲𝘀? I help startups, SaaS companies, and engineering teams build, migrate, automate, and stabilize production platforms - reducing infrastructure costs, improving reliability, and making deployments faster and safer. How I Create Real Business Impact with Cloud & DevOps ➤ 𝗖𝗹𝗼𝘂𝗱 𝗜𝗻𝗳𝗿𝗮𝘀𝘁𝗿𝘂𝗰𝘁𝘂𝗿𝗲 & 𝗠𝗶𝗴𝗿𝗮𝘁𝗶𝗼𝗻 Design, migrate, and troubleshoot production infrastructure across AWS, GCP, VPS, and self-hosted environments. I handle networking, databases, storage, security, backups, DNS, and migration planning with minimal downtime. ➤ 𝗞𝘂𝗯𝗲𝗿𝗻𝗲𝘁𝗲𝘀, 𝗗𝗼𝗰𝗸𝗲𝗿 & 𝗚𝗶𝘁𝗢𝗽𝘀 Build reliable container platforms using Docker, Kubernetes, EKS, k3s, Helm, and ArgoCD. I have deployed GitOps environments supporting 10+ microservices with automated synchronization, self-healing, scaling, and monitoring. ➤ 𝗖𝗜/𝗖𝗗 & 𝗗𝗲𝗽𝗹𝗼𝘆𝗺𝗲𝗻𝘁 𝗔𝘂𝘁𝗼𝗺𝗮𝘁𝗶𝗼𝗻 Create and repair CI/CD pipelines using GitHub Actions, Bitbucket Pipelines, Jenkins, and self-hosted runners. I automate builds, testing, security checks, Docker releases, deployments, validations, and rollback workflows. ➤ 𝗢𝗯𝘀𝗲𝗿𝘃𝗮𝗯𝗶𝗹𝗶𝘁𝘆, 𝗦𝗥𝗘 & 𝗣𝗿𝗼𝗱𝘂𝗰𝘁𝗶𝗼𝗻 𝗦𝘂𝗽𝗽𝗼𝗿𝘁 Implement monitoring, logging, dashboards, and alerts using Prometheus, Grafana, Loki, CloudWatch, and GCP Cloud Monitoring. I investigate incidents, identify root causes, improve uptime, and create clear operational runbooks. ➤ 𝗖𝗹𝗼𝘂𝗱 𝗖𝗼𝘀𝘁 & 𝗣𝗲𝗿𝗳𝗼𝗿𝗺𝗮𝗻𝗰𝗲 𝗢𝗽𝘁𝗶𝗺𝗶𝘇𝗮𝘁𝗶𝗼𝗻 Analyze infrastructure usage, identify oversized or unused resources, optimize databases and compute workloads, and redesign inefficient architecture. I reduced AWS costs by approximately 30% for a production platform serving 1M+ users. ➤ 𝗗𝗮𝘁𝗮 𝗣𝗹𝗮𝘁𝗳𝗼𝗿𝗺 & 𝗙𝘂𝗹𝗹-𝗦𝘁𝗮𝗰𝗸 𝗗𝗲𝗽𝗹𝗼𝘆𝗺𝗲𝗻𝘁 Support BigQuery, ETL/ELT pipelines, scheduled jobs, data ingestion, and production data issues. I also deploy Java, Node.js, Python, React, and Next.js applications with PostgreSQL, MySQL, Redis, Nginx, systemd, and Docker. Why Clients Trust Me: * Production-Level Delivery: I build stable, secure, documented systems—not temporary fixes. * End-to-End Expertise: I work across applications, databases, cloud infrastructure, data pipelines, networking, and security. * Root-Cause Investigation: I identify the real source of failures instead of applying random workarounds. * Outcome-Focused Execution: Every improvement is connected to uptime, performance, cost, deployment speed, or operational visibility. Recent Platform Work * SaaS Infrastructure: Supported production platforms serving 1M+ users with 99.9% uptime * Cost Optimization: Reduced AWS infrastructure costs by approximately 30% * Cloud Migration: Migrated an IoT SaaS platform from AWS to self-hosted infrastructure with minimal downtime * Kubernetes: Built GitOps deployments for more than 10 microservices using Kubernetes and ArgoCD * Observability: Implemented Prometheus, Grafana, Loki, and CloudWatch monitoring environments * CI/CD: Built automated multi-service deployment pipelines using GitHub Actions * GCP & Data: Investigated BigQuery, ETL, ingestion, and production data pipeline issues * Application Recovery: Recovered and rebuilt production applications, databases, servers, and deployment workflows

  • Python
  • Git
  • Docker
  • Kubernetes
  • Azure DevOps
  • Linux
  • DevOps
  • CI/CD
  • Automation
  • Linux System Administration
  • AWS Server Migration
  • Server Virtualization
  • AI Chatbot
  • AI Agent Development
  • Platform Migration

How it works

Post a job for freePost a job

Tell us what you need. Create your own job post or generate one with AI then filter talent matches.

Hire top talent fast

Consult, interview, and hire quickly, so you can meet the freelancers you're excited about.

Collaborate easily

Use Upwork to chat or video call, share files, and track project progress right from the app.

Payment simplified

Manage payments in one place with flexible billing options. Only pay for approved work, hourly or by milestone.

Don't just take our word for it

How do I hire a Cluster Computing Developer on Upwork?

You can hire a Cluster Computing Developer on Upwork in four simple steps:

  • Create a job post tailored to your Cluster Computing Developer project scope. We’ll walk you through the process step by step.
  • Browse top Cluster Computing Developer talent on Upwork and invite them to your project.
  • Once the proposals start flowing in, create a shortlist of top Cluster Computing Developer profiles and interview.
  • Hire the right Cluster Computing Developer for your project from Upwork, the world’s largest work marketplace.

At Upwork, we believe talent staffing should be easy.

How much does it cost to hire a Cluster Computing Developer?

Rates charged by Cluster Computing Developers on Upwork can vary with a number of factors including experience, location, and market conditions. See hourly rates for in-demand skills on Upwork.

Why hire a Cluster Computing Developer on Upwork?

As the world’s work marketplace, we connect highly-skilled freelance Cluster Computing Developers and businesses and help them build trusted, long-term relationships so they can achieve more together. Let us help you build the dream Cluster Computing Developer team you need to succeed.

Can I hire a Cluster Computing Developer within 24 hours on Upwork?

Depending on availability and the quality of your job post, it’s entirely possible to sign up for Upwork and receive Cluster Computing Developer proposals within 24 hours of posting a job description.