Site Reliability Engineer (Kubernetes)
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
TL;DR
Site Reliability Engineer (Kubernetes): Ensuring the security, stability, and scalability of cloud services with an accent on infrastructure automation, microservices management, and high-availability operations. Focus on designing robust tooling, optimizing cloud performance, and driving reliability improvements across complex distributed systems.
Location: Must be a USA Citizen or in process to become one by January 2027; Remote (US)
Salary: $130,000–$160,000 USD per year
Company
is a global leader in data integration, governance, and analytics, leveraging advanced AI/ML capabilities to transform complex data landscapes into actionable insights.
What you will do
- Ensure high reliability and availability of cloud platforms through proactive monitoring and optimization.
- Collaborate with architecture and domain teams to design and develop new infrastructure features.
- Define and evangelize cloud-related optimizations and best practices for scalability and performance.
- Analyze and resolve complex issues across infrastructure, systems, network, and application layers.
- Participate in 24/7 on-call rotations to maintain service performance and availability.
- Share expertise in cloud technologies, automation, and security with the broader engineering organization.
Requirements
- Must be a USA Citizen or in process to become one by January 2027.
- 3+ years of professional experience in cloud engineering (AWS/Azure).
- 3+ years of experience working with production systems running on Kubernetes.
- 3+ years of experience with Infrastructure as Code (IaC) tools like Terraform or Ansible.
- 3+ years of experience operating or building microservices.
- Proficiency in scripting and automation (Bash, Python, or Go).
- Experience with observability tooling (Prometheus, Open Telemetry, Splunk).
Nice to have
- Experience supporting FedRAMP or DoD IL4 certification initiatives.
- Experience with self-hosted Temporal workflow infrastructure.
- Relevant certifications such as CKS, CKD, or AWS Advanced Networking/Security.
Culture & Benefits
- Comprehensive benefits including medical, dental, vision, and 401(k) with company match.
- Flexible, diverse, and international work environment.
- Genuine career progression pathways and mentoring programs.
- Dedicated time for personal development and corporate responsibility programs.
- Culture focused on innovation, technology, and collaboration.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →