Site Reliability Engineer
Job in
Frisco, Collin County, Texas, 75034, USA
Listed on 2026-07-27
Listing for:
Infinite Computer Solutions
Full Time
position Listed on 2026-07-27
Job specializations:
-
IT/Tech
Systems Engineer, Cloud Computing: Infrastructure & Operations, SRE/Site Reliability
Job Description & How to Apply Below
Job Title:
Senior Cloud Infrastructure / SRE Engineer
Location:
Frisco, TX
Type:
Full Time/W2 with Infinite Computer Solutions
Required Skills & Experience
- Strong experience provisioning and managing Azure infrastructure using Terraform.
- Hands-on experience with Ansible for configuration management and automation.
- Strong Linux administration and operating system troubleshooting skills.
- Experience supporting both on-premises and Azure cloud-based enterprise applications and infrastructure.
- Hands-on experience with Azure services including VMSS, App Gateway, Load Balancers, Key Vault, Storage, Networking, and Monitoring.
- Proven experience in cloud transformation and migration initiatives from on-premises environments to Azure.
- Strong Site Reliability Engineering (SRE) experience supporting business-critical production systems.
- Experience managing production incidents, problem management, root cause analysis (RCA), and service restoration.
- Hands-on experience supporting Java/J2EE applications and middleware platforms in production environments.
- Experience with Dev Ops practices and CI/CD pipelines using Azure Dev Ops, Git Hub Actions, Jenkins, or similar tools.
- Strong scripting and automation skills using Bash, Python, or Shell scripting.
- Solid understanding of Layer 4 and Layer 7 networking concepts, load balancing, DNS, SSL/TLS, and firewall technologies.
- Hands-on experience with containerization technologies including Docker, Kubernetes, and Azure AKS.
- Experience implementing observability, monitoring, alerting, and logging using tools such as Dynatrace, Splunk, or Azure Monitor.
- Experience with Azure security services, RBAC, Managed Identities, Service Principals, and Key Vault integration.
- Strong understanding of high availability, disaster recovery, backup, and resiliency architectures.
- Knowledge of release management, change management, and ITIL service management processes.
- Strong troubleshooting, analytical, and problem-solving skills across infrastructure, application, and network layers.
- Excellent verbal and written communication skills with the ability to drive technical discussions and operational excellence.
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
Search for further Jobs Here:
×