Site Reliability Engineer
Posted 1ds ago
Employment Information
Report this job
Job expired or something wrong with this job?
Job Description
Site Reliability Engineer securing and scaling InRule’s cloud decision-intelligence SaaS platform. Automating Azure infrastructure, observability, reliability, and incident response.
Responsibilities:
- Design, implement, and manage tools used to develop, deploy, monitor, and support cloud infrastructure and services
- Define and track SLIs and SLOs, using error budgets to balance reliability work against feature velocity
- Monitor and troubleshoot cloud resources and applications using the observability stack
- Participate in an on-call rotation and resolve incidents with development and IT teams
- Automate and optimize processes to improve cloud operations efficiency and reliability
- Monitor and analyze cloud resource usage to identify cost-saving opportunities
- Build and maintain infrastructure-as-code for the Azure environment using Bicep/OpenTofu
- Support the shift from manual portal-driven changes to GitOps-based workflows
- Support deployment and configuration of InRule products and solutions
- Recommend application architecture or design changes to improve security, performance, and efficiency
- Implement Azure security controls, including access management, secure baselines, and monitoring
- Support least-privilege access, just-in-time elevation with Azure PIM, and infrastructure drift detection initiatives
- Assist with onboarding or tuning SAST, DAST, and SCA security scanning in CI/CD pipelines
- Design and deploy agentic workflows for release change tracking, vulnerability triage, and incident response
Requirements:
- 5+ years of experience in cloud operations, site reliability engineering, or DevOps
- Strong knowledge of core cloud services, including compute, serverless, managed databases, storage, and secrets management, on AWS, Azure, or GCP
- Experience with PowerShell and other scripting or development languages such as SQL, Python, C#, or JavaScript
- Experience with infrastructure-as-code tools such as OpenTofu, Bicep/ARM, or Pulumi
- Direct experience with at least one major observability platform: Datadog, New Relic, ELK, Prometheus/Loki/Grafana
- Experience with OpenTelemetry-based instrumentation
- Strong DevOps fundamentals and CI/CD experience with GitHub Actions, Azure DevOps, or equivalent
- Experience with Linux operating systems and tooling such as Docker and Kubernetes
- Willingness to take ownership of security-adjacent tasks
- Experience with Azure Cloud, including Functions, Container Solutions, SQL Database, Storage, and Key Vault
- Hands-on experience with Azure security tooling, including Entra ID, Defender for Cloud, Sentinel, and Azure Policy
- Background operating in a regulated environment such as SOC 2, ISO 27001, or HIPAA
- Experience with networking, DNS, VPC, and database operations and concepts
- Exposure to Windows Server and IIS configuration and maintenance
- Experience with configuration management tools such as Ansible, Puppet, or Chef
- Must be authorized to work in the United States without current or future sponsorship, or be based in Sweden for the Sweden position
Benefits:
- Competitive compensation and benefits
- Flexible work environment
- Professional growth within a scaling SaaS organization
- Collaborative culture with close partnership across Support, Engineering, Product, and Customer Success
- Opportunity to build and shape a premium support function with measurable customer impact

















