Gorilla Logic
Senior Platform DevOps Engineer - AA, Remote: Colombia - Costa Rica, Fulltime
Remote
Get past the screening software and onto a recruiter's desk
hirly rewrites your resume for this job — matching the keywords and skills in the posting, moving your most relevant experience to the top, and writing a cover letter to fit. About 30 seconds.
- Keywords matched to this posting
- Fit score before you apply
- Cover letter included
Matched against 2.3M live jobs from 200,000+ employers in 200+ countries.
Tailor my resume for this job →Apply from your AI assistant
Connect hirly to Claude and ask it to apply to this job. hirly tailors your resume, fills the employer’s form and asks before sending. ChatGPT: manual setup today.
Some employer sites stop an application at a CAPTCHA or sign-in and hand it back with a link. Applying needs a paid plan. Works with any assistant that supports MCP.
hirly's read of this role
- Role family
- Engineering
- Seniority
- Senior
- Work mode
- Remote-friendly
- First seen by hirly
- 2 Oct 2026
Derived automatically from the posting. Upload your resume above to see how the role scores against it.
the posting
- This position is open to candidates located in Colombia or Costa Rica only -
Gorilla Logic is looking for a Senior Platform / DevOps Engineer with strong hands-on experience in Kubernetes, Terraform, and Python to join our team and support a production platform running complex, highly available workloads. In this role, you will take ownership of day-to-day platform operations, deployments, infrastructure automation, reliability, and production support.
You will work in an AWS-first Kubernetes environment and collaborate closely with engineering teams to ensure systems are scalable, reliable, secure, and maintainable. We are looking for someone who combines strong technical expertise with ownership, sound engineering judgment, and a quality-first mindset.
The ideal candidate is comfortable challenging decisions when necessary, protecting engineering standards, and making thoughtful trade-offs rather than sacrificing long-term quality for short-term speed.
What you'll do
Manage, operate, and troubleshoot production Kubernetes environments and workloads.
Build, maintain, and improve infrastructure using Terraform and Infrastructure as Code practices.
Manage application deployments, upgrades, configuration changes, and complex deployment lifecycles.
Work with GitOps-based deployment processes and tools such as Argo CD.
Develop and maintain Python scripts and tooling to support platform automation and operational workflows.
Troubleshoot and resolve production incidents across infrastructure, applications, and platform services.
Improve platform reliability, scalability, observability, and operational efficiency.
Support Kubernetes scaling and autoscaling strategies for production workloads.
Collaborate closely with development and platform teams to identify and resolve infrastructure and deployment challenges.
Participate in technical decisions and proactively identify risks, reliability concerns, and opportunities for improvement.
Maintain high engineering and quality standards, providing technical pushback when necessary to ensure reliable and maintainable solutions.
Take ownership of platform initiatives and drive issues through resolution with minimal supervision.
Required Qualifications
Strong hands-on experience managing Kubernetes in production environments.
Experience with Kubernetes deployments, scaling, troubleshooting, and operational management.
Hands-on experience with Terraform for provisioning and managing cloud infrastructure.
Ability to read and write Python for scripting, automation, troubleshooting, and platform tooling.
Experience with AWS cloud infrastructure, ideally including EKS or similar managed Kubernetes environments.
Experience with GitOps practices and deployment tools such as Argo CD.
Experience managing complex application deployment and upgrade lifecycles.
Proven experience troubleshooting, triaging, and supporting production incidents.
Strong understanding of infrastructure reliability, scalability, and operational best practices.
Strong problem-solving skills and the ability to independently investigate complex production issues.
Strong sense of ownership and accountability, with the ability to operate effectively with limited supervision.
Quality-first mindset with the judgment to balance delivery speed, reliability, and long-term maintainability.
Strong communication and collaboration skills.
Preferred Qualifications
Experience with Helm and Kubernetes package/deployment management.
Familiarity with PyTorch and Hugging Face Transformers.
Experience supporting GPU-based workloads or ML inference platforms.
Familiarity with NVIDIA Triton Inference Server.
Experience with Chainguard, distroless container images, Trivy, or container vulnerability reduction.
Experience implementing or improving Kubernetes autoscaling solutions.
Familiarity with streaming or messaging platforms such as Apache Kafka or similar technologies.
Experience with Elasticsearch or ArangoDB.
Experience troubleshooting complex service-to-service networking.
Exposure to OpenShift, IL5, FedRAMP, or similarly constrained environments.
Familiarity with AI/ML or agentic AI development environments.
Similar jobs
- Senior DevOps Engineer- Google CloudDevoteam · Madrid, MD, SpainFirst seen today
- Senior. Platform DevOps Engineer (Onsite)Globalhr · US-CO-AURORA-S75 ~ 16800 E Centretech Pkwy ~ BLDG S75First seen today
- Sr. DevOps Engineer- Terraform/PulumiJobgether · IndiaFirst seen todayremote
- Senior Full Stack DevOps EngineerJobgether · USFirst seen todayremote
- Senior DevOps EngineerBounteous · MexicoFirst seen todayremote
Browse similar roles
Want this one?
Upload your resume and hirly rewrites it for this job and writes the cover letter — in about thirty seconds, before you sign up.
Tailor my resume for this job