×
Register Here to Apply for Jobs or Post Jobs. X

Site Reliability Engineer (Level 2

Job in Farmington Hills, Oakland County, Michigan, USA
Listing for: Peterson Technology Partners
Full Time position
Listed on 2026-08-15
Job specializations:
  • IT/Tech
    Cloud Computing: Infrastructure & Operations, IT Support, Systems Administrator
Salary/Wage Range or Industry Benchmark: 50 - 55 USD Hourly USD 50.00 55.00 HOUR
Job Description & How to Apply Below
Position: Site Reliability Engineer (Level 2)

The Site Reliability Engineer (Level
2) is responsible for operating and enhancing the performance, availability, and reliability of cloud and on-premises infrastructure. This individual consults on more complex observability scenarios, streamlining server and batch operations, and contributing to the efficient management of data center resources. Consults with cross-functional teams to identify opportunities for process improvements, implement best practices, and support critical business operations.

Essential Tasks/Major Duties:

  • Develop, implement, and maintain observability tools to monitor cloud and on-premises systems.
  • Create dashboards, alerts, and reports to track system health, performance, and availability.
  • Proactively leverage observability tools and identify opportunities.
  • Analyze metrics and logs to identify trends, prevent potential issues, and optimize system performance.
  • Collaborate with Fin Ops teams to monitor resource utilization and ensure cost-effective operations across cloud environments.
  • Support the lifecycle of cloud and on-premises servers, including provisioning, patching, configuration, and decommissioning.
  • Troubleshoot and resolve server-related issues, ensuring minimal downtime. Implement and enforce server security policies and compliance requirements.
  • Schedule, monitor, and manage batch processes to ensure timely execution of critical tasks.
  • Identify and resolve batch failures or delays, coordinating with relevant teams to ensure smooth operations.
  • Optimize batch jobs for improved performance and resource utilization.
  • Manage on-site and remote data center operations, ensuring proper functioning of hardware, power, cooling, and network infrastructure.
  • Coordinate with vendors and service providers for hardware maintenance, replacements, and upgrades.
  • Maintain accurate inventory of data center assets and ensure compliance with organizational standards.
  • Participate in on-call rotations to address system incidents and outages promptly.
  • Conduct root cause analysis and implement solutions to prevent recurrence of issues. Document and communicate incident resolution processes to relevant stakeholders.
  • Work closely with cross-functional teams, including Dev Ops, Networking, and Application Development, to implement and maintain system integrations.
  • Maintain and create comprehensive documentation for configurations, processes, and incident resolutions.
  • Provide training and support to team members and other departments.

Knowledge, Skills & Abilities:

  • Bachelor’s degree in computer science, Information Technology, or a related field, or equivalent experience.
  • 3 years of experience working with monitoring and observability tools (e.g., Datadog, Pager Duty).
  • Certified Datadog Fundamentals or equivalent experience required.
  • Certified Pager Duty Administrator or equivalent experience required.
  • 3 years of experience in cloud operations or server management roles.
  • Certified AWS Sys Ops Administrator or equivalent experience required.
  • 3 years of progressive server administration experience (Windows, Linux).
  • 3 years of experience in designing, implementing, and managing IT workload automation solutions to optimize scheduling, orchestration, and execution of enterprise workflows across on-prem and cloud environments.
  • Experience leveraging artificial intelligence to drive innovation and solve complex problems. Demonstrated ability to utilize AI-driven solutions that optimize processes, enhance decision-making, or create transformative business outcomes.
  • 3 years working with cloud platforms (AWS, Azure, OCI).
  • Certified AWS Cloud Practitioner or equivalent experience required.
  • Strong experience with data center infrastructure and knowledge of best practices.
  • Proficiency in scripting and automation tools (Python, Bash, Power Shell).
  • Strong understanding of networking and identity management in cloud environments.
  • Working knowledge of security best practices and compliance standards.
  • Working knowledge of agile methodologies.
  • Excellent troubleshooting, problem-solving, and communication skills.

Salary/Rate: $50-$55/HR (depends on experience level). This is a contract position with candidates expected to work 40 hours/ week.

To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary