Site Reliability Engineer

Posted 1ds ago

Employment Information

Education
Salary
Experience
Job Type

Report this job

Job expired or something wrong with this job?

Job Description

Site Reliability Engineer securing and scaling InRule’s cloud decision-intelligence SaaS platform. Automating Azure infrastructure, observability, reliability, and incident response.

Responsibilities:

  • Design, implement, and manage tools used to develop, deploy, monitor, and support cloud infrastructure and services
  • Define and track SLIs and SLOs, using error budgets to balance reliability work against feature velocity
  • Monitor and troubleshoot cloud resources and applications using the observability stack
  • Participate in an on-call rotation and resolve incidents with development and IT teams
  • Automate and optimize processes to improve cloud operations efficiency and reliability
  • Monitor and analyze cloud resource usage to identify cost-saving opportunities
  • Build and maintain infrastructure-as-code for the Azure environment using Bicep/OpenTofu
  • Support the shift from manual portal-driven changes to GitOps-based workflows
  • Support deployment and configuration of InRule products and solutions
  • Recommend application architecture or design changes to improve security, performance, and efficiency
  • Implement Azure security controls, including access management, secure baselines, and monitoring
  • Support least-privilege access, just-in-time elevation with Azure PIM, and infrastructure drift detection initiatives
  • Assist with onboarding or tuning SAST, DAST, and SCA security scanning in CI/CD pipelines
  • Design and deploy agentic workflows for release change tracking, vulnerability triage, and incident response

Requirements:

  • 5+ years of experience in cloud operations, site reliability engineering, or DevOps
  • Strong knowledge of core cloud services, including compute, serverless, managed databases, storage, and secrets management, on AWS, Azure, or GCP
  • Experience with PowerShell and other scripting or development languages such as SQL, Python, C#, or JavaScript
  • Experience with infrastructure-as-code tools such as OpenTofu, Bicep/ARM, or Pulumi
  • Direct experience with at least one major observability platform: Datadog, New Relic, ELK, Prometheus/Loki/Grafana
  • Experience with OpenTelemetry-based instrumentation
  • Strong DevOps fundamentals and CI/CD experience with GitHub Actions, Azure DevOps, or equivalent
  • Experience with Linux operating systems and tooling such as Docker and Kubernetes
  • Willingness to take ownership of security-adjacent tasks
  • Experience with Azure Cloud, including Functions, Container Solutions, SQL Database, Storage, and Key Vault
  • Hands-on experience with Azure security tooling, including Entra ID, Defender for Cloud, Sentinel, and Azure Policy
  • Background operating in a regulated environment such as SOC 2, ISO 27001, or HIPAA
  • Experience with networking, DNS, VPC, and database operations and concepts
  • Exposure to Windows Server and IIS configuration and maintenance
  • Experience with configuration management tools such as Ansible, Puppet, or Chef
  • Must be authorized to work in the United States without current or future sponsorship, or be based in Sweden for the Sweden position

Benefits:

  • Competitive compensation and benefits
  • Flexible work environment
  • Professional growth within a scaling SaaS organization
  • Collaborative culture with close partnership across Support, Engineering, Product, and Customer Success
  • Opportunity to build and shape a premium support function with measurable customer impact