E

Senior DevOps / Site Reliability Engineer - Healthcare / Voice / Observability

🔥 Senior DevOps / Site Reliability Engineer - Healthcare / Voice / Observability

elinext На сервисе с: 05.10.26 13:38

Зарплата не указанаПольшаКазахстанКипрУзбекистанГерманияНидерландыИспанияПортугалияГрузияЧехияЛюксембургЛатвияЛитваЭстонияВенгрияРумынияФинляндияФранцияАвстрияГрецияДанияИталияСловакия

🔥 Senior DevOps / Site Reliability Engineer - Healthcare / Voice / Observability

We are currently looking for a Middle+ / Senior DevOps / Site Reliability Engineer to join an international healthcare project focused on workflow and voice services for US health-system use.
The main goal of the role is to make production services observable, reliable, recoverable, and safe to operate.

📍 Locations: Poland, EU countries, Georgia, Uzbekistan, Kazakhstan
💼 Seniority: Middle+ / Senior
🇬🇧 Language: English

About the Project
The platform supports medical voice services and workflow automation for healthcare systems in the US.
The role focuses on production reliability: monitoring, incident response, SLOs, deployment safety, recovery, and uptime evidence.
❗️ MUST HAVE
— Strong production experience in SRE / DevOps / Backend Operations
— Hands-on experience with monitoring, logging, alerting, and tracing
— Real experience with Incident Response / Incident Management
— Clear responsibility for uptime, SLA / SLO / service reliability
— Experience defining or working with SLOs and production dashboards
— Strong cloud infrastructure background
— Experience with containers and deployment automation
— Experience with Infrastructure as Code (IaC)
— Strong CI/CD knowledge
— Python or comparable scripting skills for operational troubleshooting and automation
— Experience diagnosing and resolving production failures
— Understanding of backup, recovery, and capacity testing
— Ability to communicate clearly during incidents and provide concise written updates

Important: the CV should clearly show that you were not only configuring infrastructure, but were directly involved in incident response and accountable for uptime / SLA / SLO.

Responsibilities
— Define service-level objectives for production services
— Build and maintain useful reliability dashboards
— Instrument metrics, logs, and traces
— Monitor workflow, voice, and third-party service dependencies
— Lead incident response during agreed coverage hours
— Write clear post-incident reviews
— Improve deployment safety with rollback and staged releases
— Maintain change records and release processes
— Test capacity, backups, and disaster recovery
— Coordinate handover and escalation with engineering teams
— Provide accurate uptime and reliability evidence
— Support continuous improvement of platform resilience
Strong Advantage
— Healthcare or another regulated production environment
— WebRTC / real-time media / voice infrastructure
— LLM-backed services
— Experience handling third-party provider outages
— Grafana / Prometheus / Datadog or similar observability tools
— Experience with customer-facing incident communication

⏰ Working Hours - Important
The team works with US-based stakeholders, and calls are expected during Arizona time (MST, UTC-7).

Candidates should be comfortable with the possibility of regular communication and incident-related coverage during US / Arizona business hours.

💚 What We Offer
— Small-company feel within a fast-growing international environment
— Friendly, collaborative, and mission-driven team
— 25 calendar days of vacation + 5 additional paid sick days
— Medical insurance
— Corporate English courses
— Corporate events and team-building activities
— Support with professional certifications
— Reimbursement for professional courses and training
— Long-term international projects
— Opportunities for professional and technical growth

Похожие вакансии DevOps

hh.ru
eyes of wonder software llc

Senior DevOps Engineer

eyes of wonder software llcНа сервисе с: 20.08.26 10:51↑ Вакансия с автоподнятием
Зарплата не указанаРоссияМоскваОфис

Хэллоу!

Мы ищем Senior DevOps Engineer для Multilogin. Наша DevOps-практика сейчас в стадии трансформации: мы модернизируем инструменты, процессы и подходы к работе по всей команде, и у тебя будет ключевая роль в том, как это будет развиваться.

Multilogin создает цифровые аватары — браузерные и мобильные профили с уникальными цифровыми отпечатками, локацией и параметрами, которые выглядят как отдельные устройства. Они позволяют управлять десятками аккаунтов в соцсетях, e-commerce и других онлайн-платформах, создавая отдельные цифровые личности для людей, компаний и AI-агентов. Multilogin создает аватары для всех. И для всего.

Чем предстоит заниматься:

График работы: офис, 5/2, с 9-00 до 18-00 (по МСК)

  • Разрабатывать автоматизированные пайплайны для сборки, тестирования и деплоя приложений; создавать, поддерживать и документировать CI/CD-окружения
  • Обеспечивать надёжность и доступность платформы, отказоустойчивость и быстрое восстановление после сбоев
  • Устанавливать стандарты и гайдлайны по конфигурации инфраструктуры и ПО для гладких CI/CD-процессов
  • Предлагать и внедрять улучшения системы: установка, апгрейды/патчинг, разрешение проблем, конфигурейшн-менеджмент и безопасность
  • Участвовать в развитии и масштабировании инфраструктуры нашей delivery-платформы
  • Обеспечивать прозрачность и оптимизацию облачных затрат, встраивать cost-awareness в ежедневную работу вместе с инженерами и финансами
  • Проектировать и эксплуатировать serverless-архитектуры (AWS Lambda, Step Functions, API Gateway или аналоги)
  • Участвовать в трансформации DevOps-практики и обсуждениях трейд-оффов между производительностью, масштабируемостью, стоимостью и time-to-market

Мы предлагаем:

  • Трансформация зрелого продукта с сильной репутацией под реалии мира AI

  • Сильная продуктовая и инженерная экспертиза: над продуктом работает команда с большим опытом автоматизации браузеров и мобильных профилей

  • Доступ к лучшим AI инструментам для повышения эффективности работы

  • Компенсация расходов на профессиональное развитие (профильные курсы, обучение, книги, конференции) в рамках политики обучения компании

  • 28 рабочих дней оплачиваемого отпуска + оплачиваемые больничные

Что мы ожидаем:

  • Продакшн-опыт на одном из крупных облаков (предпочтительно AWS) с Kubernetes, Helm и Docker
  • Уверенное владение Terraform: инфраструктура как версионируемый код
  • Построенные и поддерживаемые CI/CD-пайплайны (GitHub Actions или аналог) и понимание GitOps-доставки
  • Опыт эксплуатации стеков мониторинга и алертинга (Prometheus/Grafana или аналог) и понимание, как выглядит хорошая надёжность
  • Практический опыт с distributed tracing (OpenTelemetry, Jaeger или аналог) и умение внедрять их для более глубокой наблюдаемости
  • Системное мышление: понимание основ сетей, распределённых вычислений и асинхронного обмена сообщениями; автоматизация на Bash и Python вместо повторения рутины
  • Ownership: ты берёшь ответственность за пределами своей задачи и докапываешься до первопричины проблемы в стеке
  • AI-forward подход: ты используешь AI-инструменты руками и имеешь позицию, в чём они сильны, а где пока пасуют — воспринимаешь их как усилитель для senior-инженера, а не как игрушку
  • Cвободный английский от B2

Multilogin - продукт компании Eyes of Wonder, стартап-студии, более 10 лет создающей успешные IT продукты в сфере браузерных и мобильных технологий, веб-автоматизации и AI. Сейчас у нас 3 активных инструмента, которыми пользуются более 20 000 клиентов по всему миру

Другие площадки
wheely

Site Reliability Engineer

wheelyНа сервисе с: 16.06.26 04:06↑ Вакансия с автоподнятием
от 5 000 € (≈ от 426k ₽)КипрNicosiaОфис

Wheely is a high-end ride-hailing service redefining premium transportation across major cities in Europe and the Middle East. We combine technology with the art of five-star chauffeuring to deliver a consistently exceptional experience.

As a profitable, fast-growing scale-up with $43M raised, we're expanding rapidly across EMEA and the US. We're adding exceptional talent to drive our next phase of growth.

Over the past few years, we have rebuilt our infrastructure almost from scratch. And now we are expanding the infrastructure team.

As a part of the team, you will work to minimize the number of incidents (availability, performance, security), speed up time-to-delivery of new features to customers, building flexible, high availability, and protected infrastructure to ensure every customer’s journey or request runs smoothly.

Responsibilities

  • Work with alerts and incidents that arise in the system.
  • Assist development teams with emerging issues.
  • Automate work and write documentation.
  • Participate in the creation of application pipelines.
  • Participate in the migration of infrastructure and application services.
  • Develop, maintain and ensure system resilience.
  • Develop dynamic environments.

Requirements

  • Proven experience with cloud-based solutions (AWS/GCP).
  • Knowledge of Linux troubleshooting, including networking and file systems.
  • Experience in infrastructure automation using Ansible, Terraform, Bash, Python, etc.
  • Experience with containers and platforms such as Docker, Kubernetes, etc.
  • Experience with administration RabbitMQ/Kafka.
  • Experience with administration of SQL/NoSQL databases.
  • Good writing and verbal communication skills to ensure efficient communication within and outside the team and to explain your decisions.
  • Ability to focus on what matters most, manage your time, and get things done.
  • To be a team player, ready to help engineers investigate issues or teach them new things.
  • Love to automate manual work and try new modern technology/approaches.

What we Offer

  • Office-based role in Nicosia, four days a week with flexible start and finish times, plus one remote day of your choice.
  • Competitive salary.
  • Employee stock options plan.
  • Private medical and dental insurance.
  • Daily Lunch allowance.
  • Latest-generation MacBook Pro and 4k display.
  • Professional development stipend.
  • Relocation support, including visa sponsorship and allowance.
Соц.сети
0

Senior DataOps Engineer

01_techНа сервисе с: 05.10.26 14:05
Зарплата не указанаНе указана странаЛокация не указанаУдалёнка

🔼 This stack needs a Senior

В •••••••• открыта позиция Senior DataOps Engineer

Distributed systems, MPP, Python / Go, AWS, Karpenter, Istio, database infrastructure и задачи, в которых можно влиять на то, как все работает вместе

Remote или офис с релокацией за счет компании, поддержка профессионального развития и много свободы в технических решениях

➡️ ••••••••

HireSeeker собирает вакансии со всех площадок и присылает только релевантные. Бесплатно.