About the Role
Client is hiring a Site Reliability Engineer (Cloud) to keep customer environments running on their ARC and Fusion platforms reliable, secure, and performant across Microsoft Azure infrastructure and Cloudflare edge services. You'll own production incident response, drive root cause analysis, and feed improvements back into the platform — with real visibility into uptime, performance, and customer SLA outcomes.
What You'll Do
- Provide technical support for customer environments deployed on Dataweavers ARC and Fusion within Microsoft Azure
- Own and resolve support requests through Zendesk — thorough investigation, clear communication, efficient resolution
- Respond rapidly to production incidents affecting customer websites and infrastructure
- Troubleshoot across Azure infrastructure, networking, deployments, and application environments
- Perform root cause analysis and contribute preventative improvements
- Collaborate with customer development teams on application-level issues
- Maintain operational documentation, troubleshooting guides, and knowledge base articles
- Identify and drive automation opportunities to reduce operational overhead
- Partner with Product and Development teams on platform feedback and continuous improvement
- Support proof-of-concept work and hotfixes where operational expertise is needed
What You'll Bring
Required experience:
- 3+ years in DevOps, cloud operations, infrastructure support, or SRE roles
- Hands-on production experience with Microsoft Azure
- Experience supporting enterprise web platforms or digital experience platforms (e.g., Sitecore or similar .NET-based solutions)
Technical skills:
- Strong troubleshooting across Azure infrastructure, networking, and application platforms
- PowerShell scripting, particularly Az PowerShell modules
- Azure resource management: RBAC, App Services, Function Apps, Azure SQL PaaS, Azure Front Door, Virtual Networks
- Cloudflare DNS, security, and caching configuration
- Git and enterprise branching strategies
- Azure Key Vault and EntraID for securing applications and pipelines
- CI/CD pipeline management with Azure DevOps
- Infrastructure as Code: ARM templates, YAML, Bicep, or Terraform
- Monitoring platforms: Azure Monitor, Application Insights, or Grafana
- Operational tooling: Zendesk, Confluence, Lucid
Nice to have:
- Microsoft AZ-400 DevOps Engineer certification
Who Thrives Here
- Strong, proactive communicator who can multi-task and prioritize independently
- Takes ownership and provides consistent proactive updates to team and customers
- Genuinely passionate about cloud infrastructure and platform reliability
- Treats uptime and performance as mission-critical, not a checkbox
- Enjoys complex technical problem-solving
- Automation-minded — looks to eliminate repeat work, not just resolve tickets