ironquill.tech/board

$ cat jobs/senior-sre-production-reliability-engineer-symphony-solutions-c496b1d81728.json

Senior SRE / Production Reliability Engineer

symphony solutions·Worldwide·Ukraine·senior
typescriptreactscalakafkapostgreselasticsearchdockerkuberneteshelmgcplinuxdevopssreci/cdgit
Apply on himalayas → Get AI match score →
We're building a multi-brand iGaming platform that helps operators run their business in regulated markets worldwide. It brings together three main products — Player Account Management (PAM), Sportsbook, and Casino — all built on a modern microservices architecture using Scala, with a React/TypeScript back office. We're looking for a Senior SRE to join our team and help us build and maintain reliable, observable, and scalable infrastructure. You'll work closely with developers, own reliability practices, and contribute to the team's DevOps culture. Requirements Must Have: Technical Skills Linux — confident troubleshooting in terminal, logs, processes, networking basics, resource usage. Kubernetes / GKE — strong hands-on understanding of workloads, pods, services, ingress/gateway, probes, RBAC, resources, autoscaling, and troubleshooting. GCP — practical experience with cloud infrastructure, especially around GKE, IAM, networking, load balancing, artifact registry, and production diagnostics. Docker / Containers — images, registries, runtime debugging, container lifecycle. Helm — understands Helm releases, values, deployment state, and rollback. GitOps / FluxCD — able to understand and troubleshoot GitOps deployment flow, drift, image automation, and Git-based rollback. CI/CD understanding — can investigate failed pipelines and deployment issues; does not need to be the main pipeline builder if DevOps owns that. Observability — Prometheus, Alertmanager, Grafana; understands metrics, logs, traces, dashboards, alerting, and production monitoring. Networking fundamentals — DNS, load balancers, ingress, gateways, TLS, routing, firewall/security rules. SLI / SLO / SLA — can define and apply reliability targets, not just explain the terms. Incident response — experience with production incidents, rollback, service recovery, RCA/postmortem. Databases / messaging operational basics — PostgreSQL, Couchbase, Kafka, Elasticsearch/ELK or similar; enough to monitor health, migrat

Similar remote roles

DevOps & SRE Engineer
bright vision technologies · US · mid
Cloud DevOps Engineer (AWS, Terraform, K8s, Automation)
Renesaselectronics · Worldwide · mid
Cloud DevOps Engineer (AWS, Terraform, K8s, Automation)
Renesaselectronics · Worldwide · mid
Cloud DevOps Engineer (AWS, Terraform, K8s, Automation)
Renesaselectronics · Worldwide · mid
Cloud DevOps Engineer (AWS, Terraform, K8s, Automation)
Renesaselectronics · Worldwide · mid
Node.js / React Developer
Noir · Worldwide · mid
DevOps Engineer (Google Cloud)
Ncsaustralia · Worldwide · mid
Backend Developer – Go
bright vision technologies · US · mid