Core Infrastructure Engineer
Listed on 2026-08-26
-
IT/Tech
Systems Engineer
Job Description
Implements and optimizes components within existing distributed systems under guidance. Applies basic scalability requirements, conducts performance/load testing, and configures resiliency features (retries, circuit breakers, timeouts) to handle network variability. Builds telemetry, alerts, and runbook-driven procedures; delivers scoped features and fault-injection tests; and helps implement basic data replication and synchronization. Assists with on-call rotations, uses automation/IaC scripts to troubleshoot, and follows change, security, and compliance procedures (encryption, access controls, remediation plans) while escalating complex issues to senior engineers.
Implements and optimizes components within existing distributed systems under guidance. Applies basic scalability requirements, conducts performance/load testing, and configures resiliency features (retries, circuit breakers, timeouts) to handle network variability. Builds telemetry, alerts, and runbook-driven procedures; delivers scoped features and fault-injection tests; and helps implement basic data replication and synchronization. Assists with on-call rotations, uses automation/IaC scripts to troubleshoot, and follows change, security, and compliance procedures (encryption, access controls, remediation plans) while escalating complex issues to senior engineers.
ResponsibilitiesKey Responsibilities System Design & Architecture
- System Scalability:
- –Assist in the implementation of components of distributed systems that support horizontal and vertical scaling under the guidance of senior engineers.
- –Optimize code segments and/or systems for large-scale data processing with oversight from senior engineers.
- –Implement scalability requirements for assigned components.
- –Learn about the use of data plane platforms for large-scale data retrieval, storage, and processing.
- –Execute performance and load testing, with guidance.
- System Reliability Design:
- –Collaborate with the team to build fault-tolerant components capable of withstanding in-service updates by learning about redundancy, replication, and automatic failover mechanisms.
- –Learn about recovery oriented computing principles and assist in applying them to component designs.
- –Configure and test retry mechanisms, circuit breakers, and timeouts to help handle network unreliability, with guidance
- System Reliability Performance:
- –Implement testing and alarming configurations to detect issues/failures.
- –Support efforts to recover from failures by drafting and executing runbooks and operational procedures, under guidance.
- –Help build dashboards, telemetry systems, and alerting mechanisms to monitor component health.
- Correctness / Availability:
- –Implement functional requirements and testing for assigned features within an existing system.
- –Implement test scenarios (e.g., fault-injection, brown-out) to evaluate system correctness, under guidance.
- –Help implement basic data replication and synchronization techniques to maintain data integrity and availability.
- –Assist in diagnosing and debugging issues in system components to support ongoing operation, under supervision.
- –Follow protocols to prevent interruptions, ensuring no maintenance windows are required for customers and users when resolving issues.
- –Run basic automation scripts and tooling to troubleshoot
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).