Hire the Best New Relic Specialists

Clients rate our New Relic Specialists
Rating is 4.7 out of 5.
4.7/5
Based on 107 client reviews
Rodrigo E.

Vicente Lopez, Argentina

$29/hr
4.9
138 jobs

Senior Linux/DevOps Engineer & AI Automation Specialist Senior Linux/DevOps Engineer and AI Automation Specialist with 15+ years of experience designing, building, and scaling cloud-native infrastructure and intelligent workflow systems. My background combines deep cloud architecture expertise (AWS-focused) with hands-on experience building AI-powered automation using n8n, LLM APIs (OpenAI, OpenRouter), vector databases, and secure cloud environments. Most recently, I've led DevOps and cloud-security engineering for a U.S.-based healthcare SaaS platform (physician scheduling) operating under HIPAA and SOC 2. The work spans secure AWS operations, fleet patch management across mixed Windows/Linux estates, EKS lifecycle management, and building the security tooling and audit evidence that underpin a SOC 2 program โ€” turning compliance requirements into reliable, automated, production-grade controls. Cloud & Infrastructure Expertise Over 12 years of AWS experience designing secure, scalable, production-grade environments across healthcare, banking, and enterprise sectors. AWS Services EC2, RDS, EKS, VPC, IAM, S3, CloudFront, Route53, Elastic Beanstalk, SQS, EFS, ElastiCache, Redshift Systems Manager (SSM), GuardDuty, Security Hub, AWS Config, Inspector, CloudTrail Infrastructure as Code Terraform & CloudFormation โ€” multi-account secure environments ADR-driven design and Strangler Fig migration strategy Containers & Orchestration Kubernetes (EKS), Docker, ECS, Docker Swarm EKS cluster lifecycle management โ€” Kubernetes version upgrades across dev/staging/prod CI/CD Jenkins, GitLab CI/CD, CircleCI, CodePipeline, ArgoCD GitHub Actions with ARC (Actions Runner Controller), Azure DevOps self-hosted agents GitOps & Kubernetes Configuration Management ArgoCD (declarative continuous delivery, app-of-apps pattern, multi-cluster sync), Kustomize (base/overlay structuring, environment promotion, patch management), Helm (chart authoring, versioning, values management across environments). GitOps workflows with full audit trail and rollback capabilities. Designed and operated production GitOps pipelines for EKS clusters in regulated environments. High Availability HAProxy, Nginx, Load Balancers, Heartbeat Led infrastructure teams in banking environments and designed highly available Kubernetes production systems for U.S.-based healthcare companies. Security, Compliance & Governance (SOC 2 / HIPAA) Hands-on implementation and continuous evidence for compliance programs in regulated healthcare environments. SOC 2 control implementation and evidence collection across multiple AWS accounts (logical access, audit logging, encryption, change management, patch management) HIPAA-aware architecture: least-privilege IAM, encryption at rest and in transit, end-to-end audit trails, and EDR coverage across the fleet AWS-native security stack: GuardDuty, Security Hub, AWS Config, Inspector, CloudTrail Prowler (soc2_aws framework) and Vanta (GRC) for continuous control mapping and posture monitoring AWS IAM Identity Center (SSO): MFA enforcement, permission sets, SAML federation, CC6.1 remediation EDR migration: CrowdStrike Falcon โ†’ Microsoft Defender, with continuity of endpoint protection enforced as a hard gate Immutable S3 evidence stores, fleet-wide backup-policy review, and EBS encryption rollout for compliance Multi-account delegated-administrator security posture (Security Hub / Config) Architecture Decision Records (ADRs) for traceable, auditable change governance Windows Server & Mixed-Fleet Operations Operated a ~70-instance EC2 fleet (Windows Server + Linux) plus ~35 EKS worker nodes AWS Systems Manager: Patch Manager (AWS-RunPatchBaseline), Session Manager, Fleet Manager, Run Command Established a formal monthly patch cadence; raised Windows patch compliance from a low baseline to ~84% Windows Server lifecycle / EOL planning: Server 2016 โ†’ 2022 upgrade and decommission roadmaps aligned to patch-management policy KB-level risk triage (e.g., hostname-length reboot risks on domain members) with documented remediation Active Directory domain controllers and SQL Server hosts (HA production pairs, RDS member servers, performance tuning) Safe-patch workflow with pre-patch EBS snapshots and post-reboot service validation Observability & Monitoring Datadog, New Relic, Zabbix โ€” agent consolidation and cost optimization across EKS and production hosts Networking & Secure Access Tailscale (subnet routers, zero-trust access), VPC design, Route53, Load Balancers AI Automation & Workflow Engineering Beyond infrastructure, I specialize in AI-powered workflow automation: n8n workflow design (cloud & self-hosted) LLM integration (OpenAI, OpenRouter, local models) RAG architectures (Qdrant, vector search, embeddings) AI-driven CRM automation Intelligent document processing Automated client intake & communication systems Secure AI pipelines for regulated industries I design workflows that are secure, easy to maintain, scalable

  • Python
  • NGINX
  • MySQL
  • MariaDB
  • DevOps
  • Amazon EC2
  • Google Cloud Platform
  • Linux System Administration
  • Amazon S3
  • Network Administration
  • Bash Programming
  • n8n
  • Microsoft Azure
  • Claude
  • OpenAI API
  • PostgreSQL
  • Amazon RDS
  • Docker
  • Microsoft Windows PowerShell
  • Windows Server
Ehteshamuddin M.

Phoenix, Arizona

$45/hr
4.5
81 jobs

I build and run Production Infrastructure on AWS, GCP, Azure and on-prem. Most of my work the last few years has been Kubernetes (EKS, GKE, AKS, and bare metal RKE2 with Rancher) and getting that infrastructure through compliance audits: HIPAA, SOC 2, PCI DSS, NIST 800-53. 10+ years in DevOps. Top Rated Plus, 100% Job Success. ๐™Ž๐™ค๐™ข๐™š ๐™ง๐™š๐™˜๐™š๐™ฃ๐™ฉ ๐™ฅ๐™ง๐™ค๐™Ÿ๐™š๐™˜๐™ฉ๐™จ ๐™ฉ๐™ค ๐™œ๐™ž๐™ซ๐™š ๐™ฎ๐™ค๐™ช ๐™–๐™ฃ ๐™ž๐™™๐™š๐™– ๐™ค๐™› ๐™ฌ๐™๐™–๐™ฉ ๐™„ ๐™–๐™˜๐™ฉ๐™ช๐™–๐™ก๐™ก๐™ฎ ๐™™๐™ค: Built an on-prem RKE2 cluster from scratch for a company preparing for SOC 2 and HIPAA. Rancher for management, Cilium with WireGuard for encrypted pod traffic, GPU nodes for their ML workloads, Longhorn for storage, Velero for backups, ArgoCD for deployments. Also enforced SSO through Entra ID on every cluster and SSH access point because their auditors required it. Took a healthcare data platform through HIPAA hardening on AWS. ECS Fargate, Aurora PostgreSQL, private subnets, TLS everywhere, proper secrets handling. Phase one was making it work, phase two was making it pass an audit. Different skill sets, I do both. Modernized AWS infrastructure for a holding group running four brands. Moved everything to ECS Fargate and Aurora Multi-AZ, set up blue-green deployments, CloudFront in front. Midway through we found a security breach in their GitHub org and I handled the full remediation across every repo. Fixed a production Aurora MySQL performance problem that had their app timing out during peak hours. Root cause was tmp table spill to disk. Tuned it, added a reader, no downtime. I also do a lot of unglamorous work that keeps things running: CI/CD pipelines (GitHub Actions, GitLab CI, ArgoCD), Terraform for everything, monitoring with Prometheus and Grafana, cost cleanup, VPN and networking issues, database migrations. Industries I know well: Healthcare (HIPAA is second nature at this point), Fintech and Payments (PCI DSS), and B2B SaaS. ๐™’๐™๐™–๐™ฉ ๐™ฎ๐™ค๐™ช ๐™œ๐™š๐™ฉ ๐™ฌ๐™ค๐™ง๐™ ๐™ž๐™ฃ๐™œ ๐™ฌ๐™ž๐™ฉ๐™ ๐™ข๐™š: Infrastructure that is written down in code, documented, and reproducible. I answer messages fast and I will tell you upfront if something is a bad idea or if I am not the right person for it. Send me your project details and current setup. I will give you an honest read on scope and effort before you spend anything. Senior DevOps Engineer, DevOps Engineer, Site Reliability Engineer, Cloud Engineer, AWS, GCP, Azure, Kubernetes, Amazon EKS, Google GKE, Azure AKS, Docker, Terraform, Infrastructure as Code, CI/CD, GitHub Actions, GitLab CI, Jenkins, ArgoCD, GitOps, Helm, Rancher, RKE2, Linux, Cloud Infrastructure, Cloud Migration, AWS ECS, AWS Fargate, EC2, VPC, IAM, Networking, Load Balancing, Prometheus, Grafana, Monitoring, Observability, Automation, Production Infrastructure, High Availability, Disaster Recovery, Security Hardening, Cloud Security, DevSecOps, Kubernetes Security, HIPAA Compliance, SOC 2 Compliance, PCI DSS, NIST 800-53, Healthcare Infrastructure, FinTech, SaaS Infrastructure, Platform Engineering, Infrastructure Automation.

  • Kubernetes
  • Terraform
  • Docker
  • DevOps
  • Google Cloud Platform
  • Amazon Web Services
  • Cloud Computing
  • Cloud Architecture
  • CI/CD
  • Microsoft Azure
  • DigitalOcean
  • HIPAA
  • SOC 2
  • Ansible
  • Python
  • Data Warehousing & ETL Software
  • Django
  • Linux System Administration
  • AWS CloudFormation
  • Git
Alex L.

Budapest, Hungary

$35/hr
4.9
774 jobs

๐Ÿ“Œ About Us Led by a DevOps and SRE veteran with 20+ years of industry experience, we are a powerhouse team of 100+ top-tier DevOps, SecOps, Cloud, and Site Reliability Engineers. We specialize in designing, building, and managing infrastructures that empower innovation while guaranteeing stability, security, and cost-efficiency. With over 1,000 successful projects delivered for startups, enterprises, and global brands across fintech, e-commerce, health tech, media, and AI/ML, we take end-to-end ownershipโ€”from initial architecture design to 24x7x365 L1/L2/L3 support. Core Capabilities & Business Impact 1. Platform Engineering, DevOps & CloudOps We build multicloud architectures and scalable systems that enable daily releases through automated "golden paths" and predictable environments. Infrastructure as Code (IaC) & Automation: Terraform, Pulumi, CloudFormation, Ansible, Chef, Puppet. Containerization & Orchestration: Kubernetes, Docker, EKS, Rancher. CI/CD & GitOps: GitHub Actions, GitLab CI, Jenkins, CircleCI, Argo CD, Flux. Web & Database Tuning: LAMP and LEMP stacks setup, database tuning, and high-availability website speed optimization. 2. 24/7 Site Reliability Engineering (SRE) & IT Support We provide true 24x7x365 L1, L2, and L3 support for applications, servers, and end-users, ensuring SLA/SLO-driven operations. Reliability & Incident Management: On-call rotations, escalation trees, postmortems, root cause analysis, and error budgets. Observability: Transparent SLI/SLO dashboards using Prometheus, Grafana, OpenTelemetry, ELK, Datadog, and New Relic. IT Consultancy & Workstation Support: End-user maintenance (B2B/B2C), patching, updates, backup management, and policy enforcement via NinjaOne. Email Deliverability: DNS, DKIM, SPF, DMARC, and MX configuration. 3. FinOps & Cloud Cost Optimization We turn infrastructure into a strategic advantage by reducing cloud costs by 20โ€“60%. Cost Control: Budgeting, forecasting, rightsizing, cost governance, and waste reduction on AWS, Azure, and GCP. FinOps Tooling: CloudHealth, AWS Cost Explorer, GCP Billing. 4. Security, Networking & Compliance We ensure robust security postures and continuous compliance readiness. Security & Identity: Entra ID, HashiCorp Vault, SOPS, RBAC audits, WAF/Shield. Policy & Compliance: Policy-as-Code (OPA/Gatekeeper), CIS Benchmarks, SOC 2, and ISO 27001 readiness. Network Hardware & Integrations: CISCO remote networks, Mikrotik remote devices, FortiGate, and QNAP configuration. Server Hardening: Deep optimization and security hardening for Linux, UNIX, and Windows Server systems. 5. MLOps & AI Infrastructure We partner with ML engineers to build reliable, reproducible production workflows. ML Pipelines: MLflow, Kubeflow, SageMaker. Infrastructure: GPU-enabled infrastructure setups, data versioning, and model tracking. Comprehensive Technology Stack - Cloud Providers: AWS (Certified Partner), Microsoft Azure (Certified Partner), GCP, OCI. - Operating Systems: Linux (20+ years expertise), UNIX, Windows Server. - Containers & Orchestration: EKS, Kubernetes, Docker, Rancher. - IaC & Configuration: Terraform, Pulumi, CloudFormation, Ansible, Chef, Puppet. - CI/CD & GitOps: Argo CD, Flux, GitHub Actions, GitLab CI, Jenkins, CircleCI. - Observability & Monitoring: Prometheus, Grafana, OpenTelemetry, ELK Stack, Datadog, New Relic. - Security & Policy: OPA, Gatekeeper, Vault, SOPS, Entra ID, WAF/Shield. - FinOps & MLOps: CloudHealth, AWS Cost Explorer, GCP Billing, MLflow, Kubeflow, SageMaker. - Networking & IT Management: FortiGate, CISCO, Mikrotik, QNAP, NinjaOne. - Web & Email: LAMP, LEMP, DNS, DKIM, SPF, DMARC, MX. Who We Work With Startups: Building scalable, secure infrastructures from day one. Enterprises: Driving modernization, automation, and DevOps transformations. Finance & SaaS: Implementing rigorous security, compliance, and FinOps practices. ML/AI Teams: Deploying robust frameworks for model scaling. Ready to transform your infrastructure, optimize costs, and secure 24/7 reliability? Letโ€™s talk.

  • Apache Tomcat
  • Ansible
  • Docker
  • Amazon Web Services
  • Windows Administration
  • Service Cloud Administration
  • Linux System Administration
  • Technical Support
  • Apache Administration
  • DevOps
Nouman I.

Islamabad, Pakistan

$30/hr
4.5
101 jobs

Need AWS infrastructure that is reliable, automated, secure, and cost-efficient? I help SaaS, fintech, healthtech, and high-growth product teams design, migrate, automate, and operate production cloud platforms using AWS, Terraform, Kubernetes, Docker, and CI/CD. Top Rated Plus | $200K+ earned | 91 Upwork jobs | 6,900+ hours | 97% Job Success AWS Certified Solutions Architect | AWS Developer Associate | HashiCorp Terraform Associate WHAT I CAN HELP YOU WITH โ€ข AWS architecture & migration: VPC, EC2, ECS/EKS, RDS, S3, Route 53, CloudFront, IAM, WAF, Secrets Manager โ€ข Infrastructure as Code: Terraform modules, multi-account AWS Organizations, CloudFormation, reusable environments โ€ข Kubernetes & containers: EKS, ECS, Docker, Helm, autoscaling, production cluster operations โ€ข CI/CD & release automation: GitHub Actions, GitLab CI/CD, Jenkins, Azure DevOps, blue/green and zero-downtime deployments โ€ข Cloud cost optimization: rightsizing, Savings Plans, Reserved/Spot capacity, scheduling and storage lifecycle policies โ€ข SRE & observability: CloudWatch, Prometheus, Grafana, Datadog, New Relic, PagerDuty, incident response and RCA โ€ข Security & resilience: IAM, KMS, WAF, GuardDuty, Security Hub, backups, disaster recovery and multi-region failover SELECTED RESULTS โ€ข Migrated 200+ enterprise customer deployments from EC2 to EKS while maintaining 99.99% availability. โ€ข Reduced AWS spend by 40% for production SaaS/healthcare platforms through architecture changes, rightsizing and capacity planning. โ€ข Built a 5-account AWS Organizations landing zone with Terraform, VPC isolation/peering, SSO/IAM and MFA-secured private access. โ€ข Migrated a 200+ customer SaaS platform from EC2 to ECS with zero downtime and built active-passive multi-region disaster recovery. โ€ข Built production GKE/AKS/DOKS platforms for real-time and microservice workloads with Terraform, Helm, and automated CI/CD. I work hands-on from architecture and Terraform through deployment, monitoring, security hardening, and incident readiness. If you need an AWS DevOps engineer to stabilize, migrate, automate, scale, or reduce the cost of your infrastructure, send me the current architecture and the result you want to achieve.

  • Amazon Web Services
  • DevOps
  • Terraform
  • Kubernetes
  • Cloud Architecture
  • CI/CD
  • Docker
  • Amazon ECS
  • Amazon EC2
  • Amazon RDS
  • Amazon S3
  • GitHub
  • Jenkins
  • Ansible
  • AWS CloudFormation
  • Cloud Migration
  • Disaster Recovery
  • Microsoft Azure
  • Google Cloud Platform
  • Linux
Roshan N.

Pune, India

$40/hr
4.7
78 jobs

I am Roshan Nagekar, a DevOps and infrastructure engineer with 13+ years of experience across cloud platforms, technical support, team leadership, and customer-facing operations. I have worked with startups, enterprises, and global teams across every layer of technology โ€” from writing runbooks and managing cloud budgets to onboarding customers and teaching DevOps to university students. What makes me different from a typical technical hire is range. I am equally comfortable debugging a Kubernetes cluster at 2am, writing a PRD for a new feature, managing a global DevOps team, or qualifying a lead over web chat. I bring both depth and breadth, and I use AI tools daily to move faster across all of it. What I can help you with: DevOps and Cloud Engineering AWS, GCP, Azure, Kubernetes, Docker, Terraform, Ansible, CI/CD pipelines (Jenkins, GitHub Actions, GitLab CI), Prometheus, Grafana, Datadog, infrastructure as code, cost optimization, and 99.9% uptime at scale. System and Cloud Administration Linux and Windows server management, IAM and access control, DNS, networking, VPNs, backups, monitoring, on-call operations, and incident response. Technical Support and Customer Success Customer onboarding, troubleshooting complex technical issues, CRM management (HubSpot), SLA-driven support, escalation handling, and translating technical problems into plain language for non-technical users. QA and Documentation End-to-end QA testing, API validation (Postman, curl), staging environment management, runbook and SOP writing, and documentation that non-engineers can actually follow. Team Leadership and Management Director-level experience leading global DevOps and SRE teams, hiring, onboarding, performance management, KPI tracking, incident coordination, and stakeholder communication. Product and Operations Sprint grooming, requirements discovery, bridging technical feasibility with product value, AI-assisted workflow design, and building operational systems from scratch in fast-moving environments. AI-Native Operator I use Claude, ChatGPT, Gemini, and CrewAI as genuine daily work tools. I also build AI agents and MCP servers, so I understand this technology at a level that goes well beyond prompt-and-paste. Recent highlights: Reduced AWS costs from $45K to $18K per month (60% savings) while improving uptime from 97% to 99.9% Migrated 50+ microservices to Kubernetes, cutting deployment time from 6 hours to 15 minutes Led a global DevOps team as Director of Infrastructure at Aigent through ISO 27001 and GDPR compliance Visiting lecturer at two universities teaching DevOps Engineering and Cloud Platforms Organizer of DevOps Pune meetups with 3000+ engineers I take on clients where I can make a real difference. If you need someone who can wear multiple hats, communicate clearly, deliver without hand-holding, and bring genuine technical depth to whatever the role requires, let us talk.

  • Amazon Web Services
  • DevOps
  • Kubernetes
  • Docker
  • Terraform
  • CI/CD
  • Jenkins
  • Infrastructure as Code
  • Ansible
  • Cloud Architecture
  • Python
  • Microsoft Azure
  • Google Cloud Platform
  • AIOps
  • MLOps
  • LAMP Administration
  • Google App Engine
  • Technical Writing
  • Unix
  • Customer Service
  • MySQL
  • WordPress
  • Software Testing
Fendri F.

Sfax, Tunisia

$59/hr
5.0
65 jobs

๐Ÿ† ๐•‹๐• ๐•ก 1% - ๐•‹๐• ๐•ก โ„๐•’๐•ฅ๐•–๐•• โ„™๐•๐•ฆ๐•ค ๐Ÿง‘๐Ÿปโ€๐Ÿคโ€๐Ÿง‘๐Ÿป 60+ โ„๐•’๐•ก๐•ก๐•ช ๐•”๐•๐•š๐•–๐•Ÿ๐•ฅ๐•ค โฑ๏ธ 6000+ ๐•‹๐• ๐•ฅ๐•’๐• โ„๐• ๐•ฆ๐•ฃ๐•ค ๐ŸŽญ 100% ๐•๐•Š๐•Š) Hi ๐Ÿ‘‹, I specialize in designing and deploying ๐™ž๐™ฃ๐™ฃ๐™ค๐™ซ๐™–๐™ฉ๐™ž๐™ซ๐™š, ๐™จ๐™˜๐™–๐™ก๐™–๐™—๐™ก๐™š, ๐™–๐™ฃ๐™™ ๐™จ๐™š๐™˜๐™ช๐™ง๐™š cloud-native solutions๐Ÿ’ก. With more than ๐Ÿ– years of experience as ๐˜พ๐™ก๐™ค๐™ช๐™™ ๐˜ผ๐™ง๐™˜๐™๐™ž๐™ฉ๐™š๐™˜๐™ฉ, ๐™†๐™ช๐™—๐™š๐™ง๐™ฃ๐™š๐™ฉ๐™š๐™จ ๐˜ผ๐™™๐™ข๐™ž๐™ฃ๐™จ๐™ฉ๐™ง๐™–๐™ฉ๐™ค๐™ง ๐™–๐™ฃ๐™™ ๐˜ฟ๐™š๐™ซ๐™Ž๐™š๐™˜๐™Š๐™ฅ๐™จ ๐™€๐™ฃ๐™œ๐™ž๐™ฃ๐™š๐™š๐™ง. Tฬฒhฬฒiฬฒsฬฒ ฬฒiฬฒsฬฒ ฬฒmฬฒyฬฒ ฬฒPฬฒrฬฒoฬฒfฬฒeฬฒsฬฒsฬฒiฬฒoฬฒnฬฒaฬฒlฬฒ ฬฒCฬฒeฬฒrฬฒtฬฒiฬฒfฬฒiฬฒcฬฒaฬฒtฬฒiฬฒoฬฒnฬฒ ฬฒ:ฬฒ ๐Ÿฅ‡ ๐‘ฒ๐’–๐’ƒ๐’†๐’“๐’๐’†๐’•๐’†๐’” Administrator CKA ๐Ÿฅ‡ ๐‘ฒ๐’–๐’ƒ๐’†๐’“๐’๐’†๐’•๐’†๐’” Security specialist CKS ๐Ÿฅ‡ ๐‘ฒ๐’–๐’ƒ๐’†๐’“๐’๐’†๐’•๐’†๐’” Application Developer CKAD ๐Ÿฅ‡ ๐‘ฒ๐’–๐’ƒ๐’†๐’“๐’๐’†๐’•๐’†๐’” ๐’‚๐’๐’… ๐‘ช๐’๐’๐’–๐’… ๐‘ต๐’‚๐’•๐’Š๐’—๐’† Associate (KCNA) ๐Ÿฅ‡ ๐‘ฒ๐’–๐’ƒ๐’†๐’“๐’๐’†๐’•๐’†๐’” ๐’‚๐’๐’… ๐‘ช๐’๐’๐’–๐’… ๐‘ต๐’‚๐’•๐’Š๐’—๐’† ๐‘บ๐’†๐’„๐’–๐’“๐’Š๐’•๐’š Associate (KCSA) ๐Ÿฅ‡ ๐‘จ๐‘พ๐‘บ Solution Architect ๐Ÿฅ‡ ๐‘จ๐’›๐’–๐’“๐’† DevOps Expert๐Ÿฅ‡ ๐‘ป๐’†๐’“๐’“๐’‚๐’‡๐’๐’“๐’Ž Associate๐Ÿฅ‡ ๐‘จ๐’“๐’ˆ๐’ ๐‘ช๐‘ซ, GitOps At Scale ๐Ÿฅ‡ Certified ๐‘ฐ๐’”๐’•๐’Š๐’ and ๐‘ฌ๐’๐’—๐’๐’š Service Mesh With my expertise in cloud-native design, I can help you build scalable, high-performance applications using microservices, deployed seamlessly in containers on private, public, or hybrid cloud platforms. This will be realized using this ๐Ÿ†ƒ๐Ÿ…พ๐Ÿ…พ๐Ÿ…ป๐Ÿ†‚ ๐Ÿ› ๏ธ following the ๐™Ž๐™š๐™ซ๐™š๐™ฃ ๐™ˆ๐™ค๐™™๐™š๐™ก๐™จ ๐™ค๐™› ๐˜พ๐™ก๐™ค๐™ช๐™™ ๐™‰๐™–๐™ฉ๐™ž๐™ซ๐™š ๐˜ฟ๐™š๐™จ๐™ž๐™œ๐™ฃ ๐Ÿ”ฅ : ๐Ÿ‘จโ€๐Ÿ’ป ๐Ÿ. ๐Œ๐จ๐๐ž๐ซ๐ง ๐ƒ๐ž๐ฌ๐ข๐ ๐ง & ๐ƒ๐ž๐ฏ๐ž๐ฅ๐จ๐ฉ๐ฆ๐ž๐ง๐ญ ๐Œ๐จ๐๐ž - Programming Languages: ๐™‚๐™ค, ๐™‹๐™ฎ๐™ฉ๐™๐™ค๐™ฃ, ๐™‰๐™ค๐™™๐™š ๐™Ÿ๐™จ, ๐˜ฝ๐™–๐™จ๐™ - Containerization Tools: ๐˜ฟ๐™ค๐™˜๐™ ๐™š๐™ง, ๐˜ฟ๐™ค๐™˜๐™ ๐™š๐™ง-๐˜พ๐™ค๐™ข๐™ฅ๐™ค๐™จ๐™š, ๐™†๐™–๐™ฃ๐™ž๐™ ๐™ค, ๐™‹๐™ค๐™™๐™ข๐™–๐™ฃ - Container Image Registry: ๐™ƒ๐™–๐™ง๐™—๐™ค๐™ง, ๐˜ฟ๐™ค๐™˜๐™ ๐™š๐™ง๐™๐™ช๐™— ๐™‚๐™ž๐™ฉ๐™ก๐™–๐™— ๐™๐™š๐™œ๐™ž๐™จ๐™ฉ๐™ง๐™ฎ, ๐™‚๐™ž๐™ฉ๐™๐™ช๐™— ๐™๐™š๐™œ๐™ž๐™จ๐™ฉ๐™ง๐™ฎ - AฬฒPฬฒIฬฒ ฬฒDฬฒrฬฒiฬฒvฬฒeฬฒnฬฒ: ๐˜ผ๐™ฅ๐™ž๐™จ๐™ž๐™ญ, ๐™„๐™จ๐™ฉ๐™ž๐™ค ๐™‚๐™–๐™ฉ๐™š๐™ฌ๐™–๐™ฎ, ๐™†๐™ค๐™ฃ๐™œ, ๐™๐™ง๐™–๐™š๐™›๐™ž๐™ , ๐˜ผ๐™ข๐™—๐™–๐™จ๐™จ๐™–๐™™๐™ค๐™ง, ๐™๐™ฎ๐™  - Mฬฒoฬฒdฬฒeฬฒrฬฒnฬฒ ฬฒDฬฒaฬฒtฬฒaฬฒbฬฒaฬฒsฬฒeฬฒsฬฒ: ๐˜พ๐™–๐™จ๐™จ๐™–๐™ฃ๐™™๐™ง๐™–, ๐™€๐™ก๐™–๐™จ๐™ฉ๐™ž๐™˜๐™จ๐™š๐™–๐™ง๐™˜๐™, ๐˜พ๐™ค๐™ช๐™˜๐™๐™™๐™—, ๐™๐™š๐™™๐™ž๐™จ, ๐™‹๐™ค๐™จ๐™ฉ๐™œ๐™ง๐™š๐™Ž๐™Œ๐™‡, ๐™‘๐™ž๐™ฉ๐™š๐™จ๐™จ - Eฬฒvฬฒeฬฒnฬฒtฬฒ-ฬฒDฬฒrฬฒiฬฒvฬฒeฬฒnฬฒ ฬฒDฬฒeฬฒsฬฒiฬฒgฬฒnฬฒ:ฬฒ ๐™‰๐˜ผ๐™๐™Ž, ๐™๐™–๐™—๐™—๐™ž๐™ฉ๐™ˆ๐™Œ, ๐™†๐™–๐™›๐™ ๐™–, ๐˜ผ๐™˜๐™ฉ๐™ž๐™ซ๐™š๐™ˆ๐™Œ, ๐˜ฝ๐™š๐™–๐™ฃ๐™จ๐™ฉ๐™–๐™ก๐™ ๐™™ ๐Ÿ—๏ธ ๐Ÿ. ๐Œ๐จ๐๐ž๐ซ๐ง ๐ˆ๐ง๐Ÿ๐ซ๐š๐ฌ๐ญ๐ซ๐ฎ๐œ๐ญ๐ฎ๐ซ๐ž/๐ƒ๐ž๐ฏ๐Ž๐ฉ๐ฌ - ๐‚๐ˆ/๐‚๐ƒ - CI/CD Tools: ๐™‚๐™ž๐™ฉ๐™ก๐™–๐™— ๐˜พ๐™„, ๐™‚๐™ž๐™ฉ๐™๐™ช๐™— ๐˜ผ๐™˜๐™ฉ๐™ž๐™ค๐™ฃ๐™จ, ๐™…๐™š๐™ฃ๐™ ๐™ž๐™จ, ๐˜ฟ๐™–๐™œ๐™œ๐™š๐™ง, ๐™๐™ง๐™–๐™ซ๐™ž๐™จ ๐˜พ๐™„, ๐˜พ๐™ž๐™ง๐™˜๐™ก๐™š ๐˜พ๐™„, ๐™๐™š๐™ ๐™ฉ๐™ค๐™ฃ. - Security Scanning Tools: ๐™๐™ง๐™ž๐™ซ๐™ฎ, ๐˜ฟ๐™ค๐™˜๐™ ๐™š๐™ง ๐™Ž๐™š๐™˜๐™ช๐™ง๐™ž๐™ฉ๐™ฎ ๐™Ž๐™˜๐™–๐™ฃ๐™ฃ๐™š๐™ง, ๐™Ž๐™ฃ๐™ฎ๐™  - Secret Management: ๐™Ž๐™ค๐™ฅ๐™จ, ๐™‘๐™–๐™ช๐™ก๐™ฉ, ๐™†๐™š๐™ฎ๐™—๐™–๐™จ๐™š, ๐™‘๐™–๐™ช๐™ก๐™ฉ ๐™Ž๐™š๐™˜๐™ง๐™š๐™ฉ๐™จ ๐™Š๐™ฅ๐™š๐™ง๐™–๐™ฉ๐™ค๐™ง, ๐™„๐™ฃ๐™›๐™ž๐™จ๐™ž๐™˜๐™–๐™ก - Dฬฒeฬฒcฬฒlฬฒaฬฒrฬฒaฬฒtฬฒiฬฒvฬฒeฬฒ ฬฒAฬฒPฬฒIฬฒ ฬฒ(ฬฒIฬฒaฬฒaฬฒCฬฒ)ฬฒ: ๐™๐™š๐™ง๐™ง๐™–๐™›๐™ค๐™ง๐™ข, ๐™‹๐™ช๐™ก๐™ช๐™ข๐™ž, ๐˜ผ๐™ฃ๐™จ๐™ž๐™—๐™ก๐™š, ๐™‹๐™ช๐™ฅ๐™ฅ๐™š๐™ฉ, ๐˜พ๐™๐™š๐™›, ๐™†๐™ช๐™—๐™š๐™ซ๐™š๐™ก๐™– - Service Discovery & Service Mesh: ๐™„๐™จ๐™ฉ๐™ž๐™ค, ๐™€๐™ฃ๐™ซ๐™ค๐™ฎ, ๐˜พ๐™ค๐™ฃ๐™จ๐™ช๐™ก, ๐™‡๐™ž๐™ฃ๐™ ๐™š๐™ง๐™™, ๐™•๐™ค๐™ค๐™ ๐™š๐™š๐™ฅ๐™š๐™ง, ๐™ˆ๐™š๐™จ๐™๐™š๐™ง๐™ฎ - SSL: ๐˜พ๐™š๐™ง๐™ฉ ๐™ˆ๐™–๐™ฃ๐™–๐™œ๐™š๐™ง, ๐˜พ๐™š๐™ง๐™ฉ๐™—๐™ค๐™ฉ, ๐™‡๐™š๐™ฉโ€™๐™จ ๐™€๐™ฃ๐™˜๐™ง๐™ฎ๐™ฅ๐™ฉ - Applications Platforms: ๐™†๐™ช๐™—๐™š๐™ง๐™ฃ๐™š๐™ฉ๐™š๐™จ, ๐˜ฟ๐™ค๐™˜๐™ ๐™š๐™ง ๐™Ž๐™ฌ๐™–๐™ง๐™ข, ๐™Š๐™ฅ๐™š๐™ฃ๐™จ๐™๐™ž๐™›๐™ฉ, ๐™๐™–๐™ฃ๐™˜๐™๐™š๐™ง, ๐™Š๐™ฅ๐™š๐™ฃ๐™จ๐™ฉ๐™–๐™˜๐™  - Internal Developer Platforms: ๐˜ฝ๐™–๐™˜๐™ ๐™จ๐™ฉ๐™–๐™œ๐™š, ๐™†๐™ง๐™–๐™ฉ๐™ž๐™ญ, ๐™‹๐™ค๐™ง๐™ฉ - Kubernetes Package Managers: ๐™ƒ๐™š๐™ก๐™ข , ๐™ƒ๐™š๐™ก๐™ข๐™›๐™ž๐™ก๐™š, ๐™ƒ๐™š๐™ก๐™ข๐™จ๐™ข๐™–๐™ฃ โš™๏ธ ๐Ÿ‘. ๐๐ฎ๐ข๐ฅ๐ & ๐ƒ๐ž๐ฉ๐ฅ๐จ๐ฒ๐ฆ๐ž๐ง๐ญ ๐Œ๐จ๐๐ž๐ฅ - Worker Nodes Scaling: ๐™†๐™–๐™ง๐™ฅ๐™š๐™ฃ๐™ฉ๐™š๐™ง, ๐˜พ๐™ก๐™ช๐™จ๐™ฉ๐™š๐™ง ๐˜ผ๐™ช๐™ฉ๐™ค๐™จ๐™˜๐™–๐™ก๐™š๐™ง - Pods Replication Scaling: ๐™ ๐™ฃ๐™–๐™ฉ๐™ž๐™ซ๐™š, ๐™ƒ๐™ค๐™ง๐™ž๐™ฏ๐™ค๐™ฃ๐™ฉ๐™–๐™ก ๐™‹๐™ค๐™™ ๐˜ผ๐™ช๐™ฉ๐™ค๐™จ๐™˜๐™–๐™ก๐™š๐™ง (๐™ƒ๐™‹๐˜ผ), ๐™‘๐™š๐™ง๐™ฉ๐™ž๐™˜๐™–๐™ก ๐™‹๐™ค๐™™ ๐˜ผ๐™ช๐™ฉ๐™ค๐™จ๐™˜๐™–๐™ก๐™š๐™ง (๐™‘๐™‹๐˜ผ) - GฬฒiฬฒtฬฒOฬฒpฬฒsฬฒ: ๐™๐™ก๐™ช๐™ญ๐˜พ๐˜ฟ, ๐˜ผ๐™ง๐™œ๐™ค๐˜พ๐˜ฟ, ๐™๐™–๐™ฃ๐™˜๐™๐™š๐™ง ๐™๐™ก๐™š๐™š๐™ฉ, Argo Rollout , Canary, Blue/Green ๐Ÿ”ญ ๐Ÿ’. ๐‚๐ฅ๐จ๐ฎ๐ ๐Ž๐›๐ฌ๐ž๐ซ๐ฏ๐š๐›๐ข๐ฅ๐ข๐ญ๐ฒ - Monitoring: ๐™‹๐™ง๐™ค๐™ข๐™š๐™ฉ๐™๐™ช๐™š๐™จ, ๐™‚๐™ง๐™–๐™›๐™–๐™ฃ๐™– - Tracing: ๐™๐™š๐™ข๐™ฅ๐™ค, ๐™š๐˜ฝ๐™‹๐™, ๐™Š๐™ฅ๐™š๐™ฃ๐™๐™š๐™ก๐™š๐™ข๐™š๐™ฉ๐™ง๐™ฎ,๐™…๐™–๐™š๐™œ๐™š๐™ง, ๐™•๐™ž๐™ฅ๐™ ๐™ž๐™ฃ - APM: ๐™‰๐™š๐™ฌ ๐™๐™š๐™ก๐™ž๐™˜, ๐™Ž๐™ ๐™ฎ๐™ฌ๐™–๐™ก๐™ ๐™ž๐™ฃ๐™œ, ๐˜ฟ๐™–๐™ฉ๐™–๐™™๐™ค๐™œ - Continuous Profiling & Analysis: ๐™‹๐™–๐™ง๐™ฆ๐™–, ๐™๐™š๐™ฉ๐™ง๐™–๐™œ๐™ค๐™ฃ, ๐™‹๐™ฎ๐™ง๐™ค๐™จ๐™˜๐™ค๐™ฅ๐™š - Logging: ๐™‡๐™ค๐™ ๐™ž, ๐™€๐™‡๐™†, ๐™๐™ก๐™ช๐™š๐™ฃ๐™ฉ๐˜ฝ๐™ž๐™ฉ - Observability Platforms: ๐™‹๐™ž๐™ญ๐™ž๐™š, ๐™๐™ฅ๐™ฉ๐™ง๐™–๐™˜๐™š, ๐™‹๐™ž๐™ฃ๐™ฅ๐™ค๐™ž๐™ฃ๐™ฉ, ๐˜พ๐™ค๐™ง๐™ค๐™ค๐™ฉ - Security Policies: ๐™†๐™š๐™ฎ๐™˜๐™ก๐™ค๐™–๐™˜๐™ , ๐™๐™–๐™ก๐™˜๐™ค, ๐™†๐™ช๐™—๐™š๐˜ผ๐™ง๐™ข๐™ค๐™ง ๐Ÿ” ๐Ÿ“. ๐Ÿ’๐‚'๐ฌ ๐จ๐Ÿ ๐‚๐ฅ๐จ๐ฎ๐ ๐’๐ž๐œ๐ฎ๐ซ๐ข๐ญ๐ฒ - Container Security | Cluster Security | Cloud Security | Container Image Security โ˜๏ธ ๐Ÿ”. ๐‚๐ฅ๐จ๐ฎ๐ ๐๐ฅ๐š๐ญ๐Ÿ๐จ๐ซ๐ฆ๐ฌ - Private: ๐˜ผ๐™’๐™Ž ๐™‚๐™Š๐™‘ ๐˜พ๐™ก๐™ค๐™ช๐™™ , ๐˜ผ๐™ฏ๐™ช๐™ง๐™š ๐™‚๐™Š๐™‘ ๐˜พ๐™ก๐™ค๐™ช๐™™ - Public: ๐˜ผ๐™’๐™Ž , ๐˜ผ๐™ฏ๐™ช๐™ง๐™š, ๐™‚๐˜พ๐™‹, ๐˜ฟ๐™ž๐™œ๐™ž๐™ฉ๐™–๐™ก๐™Š๐™˜๐™š๐™–๐™ฃ, ๐™‘๐™ˆ๐™ฌ๐™–๐™ง๐™š, ๐™‡๐™ž๐™ฃ๐™ค๐™™๐™š, Scaleway, Hertzner, OVH Cloud ๐Ÿ” ๐Ÿ•. ๐€๐ฎ๐ญ๐จ๐ฆ๐š๐ญ๐ข๐จ๐ง - Chaos Engineering: ๐™‡๐™ž๐™ข๐™ž๐™ฉ๐™ช๐™จ, ๐˜พ๐™๐™–๐™ค๐™จ ๐™ˆ๐™š๐™จ๐™ , ๐˜พ๐™๐™–๐™ค๐™จ ๐™†๐™ช๐™—๐™š - MLOps Tools: ๐™ˆ๐™‡๐™๐™ก๐™ค๐™ฌ, ๐™‹๐™š๐™ง๐™›๐™š๐™˜๐™ฉ, ๐™ˆ๐™š๐™ฉ๐™–๐™›๐™ก๐™ค๐™ฌ, ๐™†๐™ช๐™—๐™š๐™๐™ก๐™ค๐™ฌ, ๐™๐™–๐™ฎ, ๐™ˆ๐™š๐™ฉ๐™–๐™๐™ก๐™ค๐™ฌ ๐Ÿ’ก What Clients Say: โญโญโญโญโญ"Firas has become a reliable resource for our company whenever we face any issues with our devops needs from cloud architecture, or cloud native app" โญโญโญโญโญ"Firas went above and beyond in delivering the work! His dedication and skillset are commendable. Highly recommended!!" โญโญโญโญโญ"Firas was like a GINUIS that solved a long persisting issues"

  • DevOps
  • DevOps Engineering
  • Kubernetes
  • Docker
  • Terraform
  • Amazon Web Services
  • Google Cloud Platform
  • Microsoft Azure
  • CI/CD
  • Infrastructure as Code
  • Cloud Engineering
  • Jenkins
  • Ansible
  • Linux
  • Prometheus
  • Grafana
  • Amazon ECS for Kubernetes
  • Python
  • System Administration
  • Network Administration

How it works

Post a job for freePost a job

Tell us what you need. Create your own job post or generate one with AI then filter talent matches.

Hire top talent fast

Consult, interview, and hire quickly, so you can meet the freelancers you're excited about.

Collaborate easily

Use Upwork to chat or video call, share files, and track project progress right from the app.

Payment simplified

Manage payments in one place with flexible billing options. Only pay for approved work, hourly or by milestone.

Don't just take our word for it

What does a New Relic specialist do?

A New Relic specialist configures and operates observability capabilities to monitor system health and support incident response. This role focuses on ingesting telemetry data from applications and infrastructure to create clear visibility into performance bottlenecks and errors. The specialist builds dashboards and alert policies that help engineering teams detect issues before they impact users. They also automate workflows to trigger specific actions when monitoring thresholds are breached.

  • Deploy and configure New Relic agents or integrations across hosts and containers to ingest telemetry data in context. This setup ensures that logs, metrics, and traces align with application performance monitoring signals for accurate troubleshooting. The specialist verifies that all required services send data correctly to the platform.
  • Create NRQL queries to explore data patterns and build custom dashboards for key operational indicators. These views allow stakeholders to visualize error rates, response times, and throughput without manual data aggregation. The specialist designs these interfaces to highlight critical service health metrics at a glance.
  • Define alert conditions and policies based on specific NRQL thresholds to notify teams of anomalies. Configure alert destinations to route notifications to communication channels or ticketing systems when breaches occur. This process reduces noise by ensuring only relevant incidents trigger immediate attention from on-call engineers.
  • Set up workflow automation to execute predefined actions when alert conditions are met. These automated steps might include creating incident tickets, sending detailed reports, or triggering remediation scripts. The specialist maps these workflows to ensure rapid response times during outages or performance degradation events.
  • Document monitoring configurations and provide guidance on how to investigate issues using New Relic data. This documentation helps internal teams understand the logic behind alert policies and dashboard metrics. It also serves as a reference for maintaining and updating instrumentation as the infrastructure evolves over time.

How to hire a New Relic specialist on Upwork

Step 1: Post a job

Describe your observability needs in a few sentences and let Job Post Generator powered by Umaโ„ข, Upwork's Mindful AI draft a complete job post for the role. You can write a new post, update a saved draft, or reuse an existing post to start hiring.

  • Specify which telemetry sources require instrumentation, such as application performance monitoring signals, infrastructure metrics, or log data from specific hosts and containers.
  • List the exact dashboards and alert policies you need, including the key performance indicators and error rates that must trigger notifications to your team.
  • Define the required proficiency with NRQL for building custom queries and the experience needed to configure workflow automation for incident response actions.

Step 2: Evaluate candidates

Look for portfolios that demonstrate configured New Relic environments and clear documentation of monitoring setups. Uma can run instant video interviews and build shortlists with side-by-side comparisons to help you assess technical fit quickly.

  • Review examples of NRQL-based dashboards that show how the candidate visualizes complex service dependencies and operational health metrics for stakeholders.
  • Check for evidence of alert policy configuration that maps specific conditions to notification destinations without creating excessive noise or false positives.
  • Verify experience integrating New Relic agents with external systems to ensure telemetry data flows correctly into the platform for accurate troubleshooting.

Step 3: Interview your top choices

Discuss specific scenarios where the candidate used New Relic data to triage performance issues or reliability problems. Interviews can be scheduled and conducted within Upwork Messages with an immediate transcript and summary after each one.

  • Ask how they instrument applications to ingest logs in context with APM data to speed up root cause analysis during live incidents.
  • Request examples of workflow automation they built to trigger specific actions when alert conditions are breached in production environments.
  • Evaluate their approach to defining alert thresholds using NRQL to balance sensitivity with operational stability for critical services.

Step 4: Agree on scope and begin work

Define clear deliverables such as configured integrations, custom dashboards, and documented alerting rules. Use Upwork Messages and the contract workroom for communication and project management, plus identity verification, payment protection, hourly tracking, and project funds for security.

  • Set milestones for deploying agents, building initial dashboards, and testing alert workflows to verify telemetry coverage across your infrastructure.
  • Require documentation that explains how to troubleshoot common issues using the New Relic data views and queries the specialist builds.
  • Establish a schedule for reviewing alert policies and refining NRQL queries based on real-world usage patterns and incident feedback.

Upwork is not affiliated with and does not sponsor or endorse any of the tools or services discussed in this article. These tools and services are provided only as potential options, and each reader and company should take the time needed to adequately analyze and determine the tools or services that would best fit their specific needs and situation.

The rates and information provided in this article are based on current data and industry sources available at the time of publication. Freelance rates can vary depending on factors such as experience, location, project scope, and market conditions. Readers are encouraged to conduct their own research to confirm current rates and trends, as this information may change over time.

How much does hiring a New Relic specialist cost?

$500-$1,500 per project is a typical range for focused New Relic specialist work. Final pricing depends on scope, technical complexity, required integrations, source-material quality, revision needs, and the freelancer's experience level.

Instrumentation setup

$500-$1,200/project

Entry-level to mid-level
  • Installed agents for application and infrastructure telemetry
  • Verified ingestion of logs and metrics in context
  • Documentation of installed integrations and coverage

Dashboard creation

$1,200-$2,500/project

Mid-level
  • Written queries to extract key performance signals
  • Built views for operational visibility and troubleshooting
  • Instructions for interpreting dashboard data

Alert policy configuration

$2,500-$4,500/project

Mid-level to senior-level
  • Defined thresholds for errors and performance breaches
  • Mapped alerts to specific communication destinations
  • Recorded logic for each alert condition and destination

Workflow automation

$4,500-$7,000/project

Senior-level
  • Configured actions triggered by alert breaches
  • Validated end-to-end incident response workflows
  • Steps for managing automated incident actions

Full observability implementation

$7,000-$12,000/project

Expert-level
  • Connected APM, infrastructure, and logs for full context
  • Built advanced NRQL views for complex system triage
  • Assessment of monitoring gaps and remediation plan

Frequently asked questions

Is hiring a New Relic specialist worth it?

For most businesses, yes: hiring a New Relic specialist is worthwhile. This expert configures agents and writes NRQL queries to turn raw telemetry into actionable dashboards and alert policies. They connect APM and infrastructure data so your team can triage incidents faster without manual log digging.

How do I evaluate New Relic specialist candidates?

Ask candidates to write an NRQL query that isolates high-latency transactions for a specific service and explain how they would set up an alert policy for it. Look for examples where they linked infrastructure metrics to application performance to reduce noise in alert notifications.

What deliverables should I expect from a New Relic specialist?

You should receive configured instrumentation for your applications and infrastructure along with custom NRQL dashboards. The specialist also sets up alert policies with defined notification destinations and documents the monitoring strategy for your team.

How does a New Relic specialist improve incident response?

They build workflow automations that trigger specific actions when alert conditions breach defined thresholds. This setup routes context-rich notifications to your communication tools so engineers can start troubleshooting immediately.