$ cat jobs/senior-site-reliability-engineer-hardrockdigital-9aa372243e4a.json
Senior Site Reliability Engineer
What are we building? Hard Rock Digital is a team focused on becoming the best online sportsbook, casino, and social gaming company in the world. We’re building a team that resonates passion for learning, operating, and building new products and technologies for millions of consumers. We care about each customer interaction, experience, behavior, and insight and strive to ensure we’re always acting authentically. Rooted in the kindred spirits of Hard Rock and the Seminole Tribe of Florida, Hard Rock Digital taps a brand known the world over as the leader in gaming, entertainment, and hospitality. We’re taking that foundation of success and bringing it to the digital space - ready to join us? What’s the position? We are looking for a Senior Site Reliability Engineer who combines deep infrastructure expertise with a forward-thinking approach to AI-driven operations. In this role you will maintain and improve the reliability, scalability, and performance of our Java-based applications while pioneering the use of large language models (LLMs), agentic workflows, and intelligent automation to transform how we monitor, respond to, and prevent incidents. You will design and build autonomous and semi-autonomous AI agents that consume observability data, triage alerts, generate runbooks, automate incident response steps, and surface actionable insights—reducing toil and accelerating mean time to resolution. This is a hands-on engineering role for someone who is equally comfortable tuning a JVM, writing PromQL, and prototyping an agentic pipeline with tool-calling LLMs. Key Responsibilities Application Reliability & Performance Ensure the availability, reliability, and performance of high-traffic Java-based applications in a distributed environment. Troubleshoot and resolve complex issues across production and non-production environments. Participate in pre- and post-deployment performance testing and monitoring to continuously improve application performance. Optimize Java ap
Similar remote roles
Senior Site Reliability Engineer
Hardrockdigital · Worldwide · senior
Senior SRE/DevOps Engineer (Remote, Colombia/ Brazil) - Min.
exceptionly · US · senior
DevOps Engineer
Jobsforhumanity · Worldwide · mid
Associate Distinguished Engineer (Agentic AI Architect)
Nagarro1 · APAC · senior
Desarrollador IA / Agentes LLM
Grupo TECDATA Engineering · Worldwide · mid
Software Engineer, Frontend (React, NextJS) - All Genders
Leap Ahead · Worldwide · mid
AI developer (Remote 100%)
Grupo TECDATA Engineering · Worldwide · mid
Site Reliability Engineer, Enterprise Technology Services
Apple · Worldwide · mid