Senior Cloud Networking SRE
Hace 4 días
Chile
SoftServe
Jornada completa
Gratis con email o Google
Guarda esta oferta y sigue tu búsqueda
Crea una cuenta gratis para guardar empleos, crear alertas y volver a esta oferta desde tu panel.
Gratis con email o Google
Al continuar, aceptas nuestros Términos & Política de Privacidad.
About The Role
In this role, you will own the reliability, evolution, and automation of cloud networking foundations supporting internal services and engineering productivity for a fast-growing US-based cloud data platform company. You will provide senior technical ownership for cloud networking and site reliability, making connectivity predictable, observable, recoverable, secure, and easier to operate. Responsibilities Define cloud networking standards across cloud providers, covering routing, segmentation, firewalls, NAT, DNS, ingress, egress, and private connectivity Own the technical direction of VPN and private-access connectivity, improving the reliability and operability of critical network services Collaborate with Cloud Platform, DevPlat, Systems Engineering, IT, and InfoSec teams on connectivity architecture and technical decisions Establish SLOs for critical network services and implement actionable monitoring and alerting using Grafana, Mimir, Loki, Pingdom, and PagerDuty Lead diagnosis and mitigation of significant connectivity incidents, driving permanent remediation and maintaining runbooks, failure-mode documentation, and escalation paths Develop infrastructure automation using Terraform or OpenTofu, Terragrunt, Python, and Go, with safe validation, deployment, and rollback practices Guide VPN, bastion, DNS, device-trust, IPAM, and cloud-to-data-center connectivity patterns, including technologies such as Tailscale, Headscale, and NetBox Mentor engineers and distribute operational knowledge through standards, architecture decisions, and operational practices that strengthen coverage across time zones Requirements Significant experience with designing and operating production cloud networks across both GCP and AWS Deep practical knowledge of routing, segmentation, firewalls, DNS, VPNs, bastion access, private connectivity, and cloud-network troubleshooting Strong SRE experience with SLOs, observability, incident response, post-incident remediation, capacity planning, and disaster recovery Hands-on experience with Infrastructure as Code using Terraform or OpenTofu and Terragrunt, plus infrastructure automation using Python, Go, or a comparable language Practical experience with Kubernetes networking or adjacent platform infrastructure Hands-on experience with FreeIPA and Keycloak, including practical understanding of cloud identity and access dependencies; familiarity with OIDC, SAML, SCIM, or JIT access is beneficial Solid understanding of security best practices for cloud infrastructure, including safe infrastructure changes and operational guardrails Hands-on experience with AI-driven infrastructure and workflows; experience with Tailscale, Headscale, WireGuard, NetBox, or mixed cloud and data-center environments is a plus Proven ability to lead cross-team technical decisions, document architecture and operational standards, and collaborate without relying on formal authority Demonstrated B2+ English proficiency, both written and spoken, for daily communication with the team and a US-based client, including technical ownership during US business hours and critical-incident escalation SoftServe is an equal opportunity employer. Qualified applicants will receive consideration regardless of race, color, ancestry, ethnicity, national origin, religion, sex, sexual orientation, gender identity or expression, age, citizenship, disability, health condition, marital or family status, veteran status, or any other characteristic protected by applicable law.