Senior DevOps & Infrastructure Engineer

Turing · Colombia, Huila, Colombia
full-time senior Posted 13 hours ago

Before you apply

Build my evidence-backed draft — free Apply on company site →

Paste your relevant resume section or 2–4 true bullets. See supported requirements and honest gaps. No account and no application sent.

Get weekly job alerts like this →

About this role

About Turing Turing’s mission is to accelerate superintelligence to drive real economic progress. Headquartered in San Francisco, Turing works with frontier AI labs to generate high-quality datasets, reinforcement learning environments, and frontier research benchmarks that improve model capabilities in software engineering, enterprise knowledge work, and advanced STEM reasoning. In software engineering, Turing is the largest and longest-running data provider in the category. Turing also works with Fortune 500 enterprises across financial services, life sciences, healthcare, retail, automotive, and CPG to build and deploy end-to-end agentic AI systems inside mission-critical workflows. By operating on both sides, Turing closes the loop between frontier research and enterprise deployment, turning real-world deployment signals into better data, evaluations, and more capable models. Learn more at www.turing.com .    Join Turing’s global Infrastructure Engineering team as a Senior DevOps & Infrastructure Engineer based in Brazil. You’ll play a critical role in building and evolving the cloud infrastructure and developer platform that powers our business, helping make it more reliable, secure, scalable, and efficient as we grow. This is a hands-on opportunity for an engineer who enjoys solving complex infrastructure challenges, turning manual processes into automation, and raising the bar for operational excellence. You’ll work across cloud infrastructure, developer tooling, reliability, and security while partnering closely with application and security teams to build systems that enable engineers to move faster with confidence. *]:pointer-events-auto R6Vx5W_threadScrollVars scroll-mb-[calc(var(--scroll-root-safe-area-inset-bottom,0px)+var(--thread-response-height))] scroll-mt-[calc(var(--header-height)+min(200px,max(70px,20svh)))]" data-turn-id="request-WEB:60f59bc7-12b5-4f21-a1f7-b4f8e3f49239-36" data-turn-id-container="request-WEB:60f59bc7-12b5-4f21-a1f7-b4f8e3f49239-36" data-testid="conversation-turn-58" data-turn="assistant"> The role’s regular working hours will generally overlap with U.S. Pacific Time , and the position will also participate in the shared global Infrastructure team’s on-call rotation.   What You’ll Work On Own Cloud Reliability at Scale: Operate and continuously improve production infrastructure across GCP, AWS, and Azure . Participate in a global on-call rotation , lead incident response, troubleshoot complex issues, drive root-cause remediation, and improve monitoring, logging, alerting, performance, capacity, and cloud cost efficiency. Use AI-assisted engineering tools to support log analysis, debugging, troubleshooting, and optimization. Build the Developer Platform: Build and support engineering platforms, release pipelines, and deployment automation using GitHub Actions, Google Cloud Build, Jenkins , and related tooling. Design and maintain secure, scalable Terraform-based IaC , develop reusable modules, automate workflows with Python, Go, Shell , or similar technologies, and operate production Kubernetes/GKE environments across deployment, scaling, networking, security, observability, and troubleshooting. Secure and Shape the Future of Infrastructure: Design and troubleshoot cloud networking, including VPCs, firewalls, load balancers, routing, DNS, and WAFs . Partner with security and engineering teams to implement secure-by-default infrastructure and meet requirements across identity and access management, network security, secrets management, and infrastructure security . Collaborate with engineering teams on scalable architectures, deployments, production issues, and operational best practices; mentor engineers and lead initiatives that improve reliability, security, and developer productivity. What We’re Looking For We expect this level of expertise to be consistent with 5+ years of experience in cloud, infrastructure, DevOps, SRE, or platform engineering, including hands-on work in GCP production environments . Independent experience with Terraform/IaC, CI/CD , using tools such as GitHub Actions, Jenkins, Cloud Deploy, or equivalent, and Kubernetes in production . Strong understanding of cloud architecture, distributed systems, and networking , including VPCs, firewalls, load balancers, and WAFs, along with experience securing cloud environments, preferably in GCP. Strong scripting or programming skills in Python, Go, Shell , or equivalent. Experience responding to production incidents, troubleshooting complex issues, and conducting root-cause analyses. Strong communication, ownership, and problem-solving skills. Willingness to participate in an on-call rotation alongside a globally distributed infrastructure team.   Bonus Points Ability to independently operate AWS or Azure environments. Ability to independently use Ansible, Chef, or Puppet . Ability to independently secure containers across Docker

Similar Jobs

Related searches:

On-site Jobs Senior Jobs On-site Senior Jobs Senior Healthcare AISenior Machine LearningSenior AI InfrastructureSenior Robotics & AutonomySenior AI Agents & RAGSenior Backend & Systems distributed-systemsagentshealthcarereinforcement-learningcloudinfrastructuredevops

Get jobs like this delivered weekly

Free AI jobs newsletter. No spam.