ironquill.tech/board

$ cat jobs/senior-sre-cloud-engineer-miratech-a56387b9710e.json

Senior SRE / Cloud Engineer

miratech·Worldwide·Ukraine·senior
awsgcpazuresre
Apply on himalayas → Get AI match score →
We are looking for an SRE / Cloud Engineer with strong hands-on experience supporting and improving production cloud infrastructure. This role is ideal for someone who understands reliability engineering, cloud operations, incident response, infrastructure automation, and observability in real production environments. The SRE / Cloud Engineer will help operate, monitor, and improve cloud platforms and services with a focus on uptime, performance, scalability, alerting quality, and operational excellence. This person should be comfortable troubleshooting live systems, improving monitoring coverage, building automation, and partnering with engineering teams to make production systems more reliable. Key Responsibilities Operate, maintain, and improve production cloud infrastructure across AWS, Azure, GCP, Windows, or hybrid environments. Build and maintain monitoring, logging, metrics, tracing, dashboards, and alerting for production services. Improve observability coverage across infrastructure, applications, databases, queues, and network dependencies. Tune alerts to reduce noise, improve signal quality, and ensure actionable incident response. [ Participate in incident response, production troubleshooting, root cause analysis, and post-incident remediation. Build automation for infrastructure operations, deployments, health checks, runbooks, and recovery workflows. Partner with engineering teams to define SLOs, SLIs, error budgets, and operational readiness standards. Support cloud networking, DNS, TLS, load balancing, IAM, storage, compute, and managed service operations. Improve reliability, availability, performance, and scalability of cloud-hosted systems Maintain Infrastructure as Code and configuration management practices for repeatable environments. Create and maintain runbooks, operational documentation, and escalation procedures. Identify production risks and drive remediation through automation, architecture improvements, and platform standards. Hands-on

Similar remote roles

Site Reliability Engineer (Guardicore AI Platform) - Remote
akamai technologies · EU · mid
Principal Machine Learning Engineer
armis · Canada · senior
Cybersecurity Architect
Smithsgroup2 · Worldwide · senior
Solutions Architect - Digital Native Business, Named Accounts
databricks · US · senior
Consulting - Consultant Confirmé / Senior – Stratégie Data & IA Générative - Paris (H/F)
Talan · Worldwide · senior
ENGINEERING MANAGER
black financial consult · US · mid
QA INFRASTRUCTURE ENGINEER
black financial consult · US · mid
Senior DevOps Engineer
Dealpath · US · senior