Search jobs now Find the right job type for you Create a job alert Explore how we help job seekers Contract talent Permanent talent Learn how we work with you Executive search Finance and Accounting Technology Marketing and Creative Legal Administrative and Customer Support Technology Risk, Audit and Compliance Finance and Accounting Digital, Marketing and Customer Experience Legal Operations Human Resources 2026 Salary Guide Demand for Skilled Talent Report Job Market Outlook Press Room Tech insights Labor market overview AI in recruiting Navigating the AI era Staffing for small businesses Cost of a bad hire Browse jobs Find your next hire Our locations

Add your latest resume to match with open positions.

1 result for Systems Engineer in Toledo, OH

Site Reliability Engineer (SRE)
  • Maumee, OH
  • remote
  • Temporary / Contract
  • 0 - 0 USD / Yearly
  • <p>We are looking for an experienced Site Reliability Engineer (SRE) to strengthen observability and operational resilience across a Microsoft Azure environment. This long-term Contract role will work closely with DevOps and engineering teams to establish monitoring standards, expand telemetry coverage, and improve service reliability across cloud-based platforms. The ideal candidate brings deep expertise in Azure operations, modern observability tooling, and production support, with the ability to turn data into actionable insight for faster troubleshooting and stronger system performance.</p><p><br></p><p>Responsibilities:</p><p>• Create and advance an observability framework for Azure-hosted systems and integrated third-party platforms, ensuring scalable monitoring coverage.</p><p>• Develop meaningful dashboards, alerting rules, log analysis views, and distributed tracing to provide actionable insight into application and infrastructure behavior.</p><p>• Utilize Azure services such as Azure Monitor, Log Analytics, Application Insights, Managed Prometheus, and Azure Managed Grafana to expand end-to-end visibility.</p><p>• Work alongside DevOps and software engineering teams to strengthen platform stability, incident response readiness, and service performance.</p><p>• Assess existing monitoring practices to uncover blind spots, reduce unnecessary alert volume, and support quicker root-cause identification.</p><p>• Improve insight into the health of applications, infrastructure components, and dependent services across production environments.</p><p>• Support reliability-focused engineering efforts by applying SRE principles such as service measurement, alert strategy refinement, and operational readiness improvements.</p>
  • 2026-08-24T00:00:00Z