×
Register Here to Apply for Jobs or Post Jobs. X

Infrastructure Reliability Operations Administrator

Job in Boston, Suffolk County, Massachusetts, 02210, USA
Listing for: PeopleSERVE, Inc.
Full Time position
Listed on 2026-07-22
Job specializations:
  • IT/Tech
    Unix/Linux, Systems Administrator, Systems Engineer, IT Infrastructure
Job Description & How to Apply Below
Position: Infrastructure Reliability Operations Administrator- 181480
We are seeking a talented Infrastructure Reliability Operations engineer to join our Compute Operations team. In this role, you will collaborate with senior technology professionals, peers across operations teams, and represent Compute Operations during incident calls. You will also work closely with teams based in our India offices on a daily basis. As a key member of the operations team, you will continuously evaluate opportunities for automation, promote Dev Ops and Eng Ops practices, and ensure the stability and security of our environment.

The Role
  • Manage and coordinate Linux, Unix, and Windows operating systems
  • Oversee hypervisors and hardware infrastructure
  • Ensure security compliance and execute patching activities
  • Maintain environment stability and support critical server operations
  • Participate in incident management and change execution
  • Contribute to operational KPIs , metrics & observability
  • Drive automation initiatives and promote Dev Ops/Eng Ops work
  • Collaborate with global teams and vendors to resolve issues and implement solutions
  • Identify process improvements to improve operational stability
  • Communicate effectively with engineering, operations leaders, and partners
This is a shift-based position:
  • Schedule:

    Wed Sat or Sun Wed (10 hours per day)
  • Flexibility:
    Occasional coverage on other days may be required
Required Expertise & Skills
  • 5+ years of IT experience across a broad range of technologies, with a focus on server and storage infrastructure
  • Strong knowledge of Linux (Redhat Linux 7, 8, and
    9)
  • Disk storage management expertise
  • Experience with virtualization technologies (preferably OLVM)
  • On-call coverage and incident management experience
  • Troubleshooting skills for OS, hardware, and storage issues
  • Shell scripting and Python proficiency; coding experience using AI is a plus
  • Experience working with enterprise-level customers and application teams
  • Windows, VMware, and OLVM (Oracle Linux Virtualization Manager) experience is highly desirable
  • Knowledge of storage subsystems and SAN/NAS/CAS infrastructure is a significant plus
  • Experience crafting and maintaining logging, monitoring, and alerting capabilities using Observability tools
Additional Skills
  • Proven experience in server operations and infrastructure teams
  • Expertise in managing high-severity incident calls
  • Strong understanding of Agile methodology and IT Service Management
  • Comprehensive knowledge of infrastructure tech stack:
    Linux, AIX, Windows, VMware, Open Stack, Client, Dell, IBM AIX server hardware
  • Ability to identify gaps and drive process improvements
  • Collaboration with vendors (Redhat, Client, Dell, IBM AIX) for root cause analysis and solution implementation
  • Passion for automation, self-service, and self-healing infrastructure
  • Excellent communication and relationship-building skills
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary