Caris MPI, Inc.
Senior DevOps Engineer (EKS/Kubernetes)
Rol remoto de DevOps Engineer con fit claro de ubicación del candidato.
Publicado25 jul 2026
Países elegibles1 país aceptado
Señal de senioritySenior
Modelo de trabajoRemoto
Ubicaciones aceptadas para candidatos
Estados Unidos
Resumen del rol
Senior DevOps Engineer (EKS/Kubernetes)
Requisitos y responsabilidades
Contenido del rol extraído en secciones para revisar más rápido.
Job Responsibilities
- Design, deploy, and maintain Linux infrastructure in on-premises and cloudenvironments.
- Automate infrastructure provisioning and configuration using tools such as Terraform, Ansible, or CloudFormation.
- Manage and optimize AWS environments with a focus on performance, scalability, security, and cost efficiency.
- Implement and maintain monitoring, logging, and alerting solutions (e.g., Datadog, Prometheus, Grafana, ELK, CloudWatch).
- Architect, deploy, and operate production Kubernetes/AWS EKS clusters, including node group strategy, cluster upgrades, multi-tenant workload isolation, and cross-region disaster recovery (DR) architecture and build outs.
- Define and lead cluster upgrade, security hardening, and disaster recovery strategies for production Kubernetes/AWS EKS environments at scale, while serving as a senior technical resource for complex production incidents.
- Manage Kubernetes networking, including VPC CNI configuration and ingress controllers (ALB/NGINX/Traefik).
- Implement IAM Roles for pod security standards, and network policies to secure EKS workloads.
- Configure and tune cluster autoscaling (Cluster Autoscaler or Karpenter) and workload autoscaling (HPA/VPA) to optimize performance and cost.
- Build and maintain Helm charts and GitOps-based deployment pipelines (e.g., ArgoCD, Flux) for Kubernetes workloads.
- Manage Docker container builds and registries in support of EKS-based application deployment.
- Deploy, scale, and maintain GitLab Runners (including Kubernetes executor runners on EKS) to support CI/CD pipeline throughput and reliability.
- Support and help operate database platforms on AWS RDS (MySQL, PostgreSQL), collaborating with data owners on performance and reliability.
- Ensure systems meet security and compliance requirements, including SOX and SOC 2initiatives.
- Execute and maintain Linux patching strategies, addressing security updates and CVEs in a timely manner.
- Participate in incident response, root cause analysis, and recovery efforts.
- Collaborate with development, QA, and cross-functional teams to improve reliability, release processes, and operational standards.
- Participate in on-call rotations and provide after-hours support as required.
Required Qualifications
- Bachelor’s degree in computer science, Information Technology or related field.
- 8+ years of experience in Linux Systems Administration, DevOps, or Site Reliability
- Engineering roles.
- 5+ years of experience with AWS services, including EC2, VPC, IAM, RDS, S3, and
- CloudWatch.
- 5+ years of hands-on experience designing and operating production workloads on Kubernetes/AWS EKS, including cluster upgrades, networking, and autoscaling.
- Proficiency in scripting and automation using Python and Bash.
- Strong hands-on experience with Infrastructure as Code using Terraform, and with CI/CD pipelines (GitLab CI/CD), including running CI/CD workloads on Kubernetes/EKS.
- Proficiency with Docker, Helm, and Kubernetes troubleshooting in a production environment.
- Solid understanding of networking fundamentals and cloud security best practices.
- CKA (Certified Kubernetes Administrator) certification expected or actively in progress;
Preferred Qualifications
- CKAD and AWS certifications (e.g., AWS Certified DevOps Engineer, Solutions Architect) a plus.
- Experience with Karpenter, Kyverno, OPA/Gatekeeper, Falco, and multi-cluster/multitenant EKS environments.
- Experience using AI and automation tools including Claude, Cursor, and OpenAI to streamline DevOps workflows through AI-assisted CI/CD, self-healing operations, and automated incident response.
- Experience with microservices, serverless architectures, and DevSecOps practices.
Physical Demands
- Ability to sit for extended periods while working on a computer.
Training
- All job specific, safety, and compliance training are assigned based on the job functions associated with this employee.
Other
- This position may require periodic travel and some evenings, weekends, and/or holidays.
- The job may require after-hours response to emergency issues and on-call availability as required.
- Willingness to pursue ongoing professional development and stay current with emerging technologies in the field.
- Job responsibilities may be modified or expanded at the discretion of management to meet changing business needs and organizational requirements.
Description of Benefits
- Highly competitive and inclusive medical, dental and vision coverage options
- Health Savings Account for medical expenses and dependent care expenses
- Flexible Spending Account to pay for certain out-of-pocket expenses
- Paid time off, including: vacation, sick time and holidays
- 401k match and Financial Planning tools
- LTD and STD insurance coverages, as well as voluntary benefit options
- Employee Assistance Program
- Pet Insurance
- Legal Assistance
- Tuition Assistance
Roles similares
Mantén una lista de respaldo.
CI/CD, Docker 1 país aceptado
Senior Software Engineer (Java)AppianVer rol AWS, Kubernetes 1 país aceptado
Lead Platform EngineerAppianVer rol Docker, PostgreSQL 5 países aceptados
Lead Full Stack EngineerKepler GroupVer rol AWS, CI/CD 8 países aceptados
Lead Full Stack EngineerFlint Group Packaging SolutionsVer rol Stack
Usa estas tags para comparar roles remotos similares.
Elegibilidad de ubicación
Candidatos deberían aplicar solo cuando el país del perfil aparece aquí.
Tu perfilPaís no definidoInicia sesión para comparar tu país con este rol.
Flujo de contratación
WithMira muestra el rol y luego envía candidatos a la aplicación de la empresa.
1Revisa fit del rol, stack y elegibilidad de ubicación en WithMira.
2Abre la página de aplicación de la empresa desde el link rastreado.
3Guarda el rol o suscríbete a oportunidades similares antes de salir.