обновлено 5 часов назад
Senior DevOps Engineer (AI)
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Senior DevOps Engineer (AI): Designing and operating highly available production infrastructure handling millions of requests per second, with an accent on AWS multi-account environments, large-scale server and container operations, and AI-driven automation. Focus on building LLM and agentic workflow infrastructure, improving observability, and solving complex reliability and performance challenges.
Location: Tel Aviv, Israel
Company
provides digital intelligence data and solutions that help more than 5,000 customers make data-driven business decisions.
What you will do
- Design, build, and maintain highly available, scalable, and resilient production infrastructure handling millions of requests per second.
- Manage and optimize tens of AWS accounts in a multi-account cloud environment.
- Develop AI-driven automation frameworks and tools used by more than 400 R&D engineers.
- Design, implement, and support scalable LLM and agentic workflow infrastructure.
- Operate thousands of servers and containers while improving logging, monitoring, alerting, and observability.
- Troubleshoot production incidents, participate in on-call rotations, and collaborate with engineering teams to improve performance and deployment processes.
Requirements
- 5+ years of experience building and maintaining production infrastructure at scale.
- 5+ years of experience developing automation tools and server-side applications with Python, Go, Ruby, Java, or Node.js.
- Strong experience with AWS, Google Cloud, or similar cloud platforms.
- Experience deploying, managing, and leveraging large language models and agentic infrastructure.
- Deep knowledge of Linux systems, including troubleshooting, architecture, and system internals.
- Understanding of web servers, load balancers, caching systems, relational databases, and networking.
Culture & Benefits
- Work with a publicly traded company serving customers worldwide.
- Own critical infrastructure with a direct impact on platform reliability and scalability.
- Work on advanced infrastructure, AI, and production reliability challenges.
- Collaborate in an innovation-driven engineering environment.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →