Staff DevOps Engineer
Job in
Bloomington, Monroe County, Indiana, 47401, USA
Listed on 2026-08-19
Listing for:
Tazapay
Full Time
position Listed on 2026-08-19
Job specializations:
-
IT/Tech
Systems Engineer, Cloud Computing: Infrastructure & Operations, SRE/Site Reliability
Job Description & How to Apply Below
- Make architectural tradeoffs, design and deliver highly scalable, reliable, secure and fault tolerant cloud infrastructure.
- Be a role model for Dev Ops and platform engineers, mentor engineers.
- Demonstrate technical leadership and create impact across teams.
- Lead AI strategy and adoption for productivity, infrastructure automation and operational efficiency.
- Drive infrastructure and deployment practices while being secure and compliant (PCI-DSS, SOC2 and others).
- Participate in infrastructure, architecture and security design reviews to maintain our high engineering standards.
- Partner with engineering and product management teams to define and execute the platform roadmap.
- Translate business and engineering requirements into scalable and extensible infrastructure design.
- Proactively manage stakeholder communication related to deliverables, risks, changes and dependencies.
- Coordinate with cross functional teams (Backend, Frontend, Data, Security, QA etc.) on planning and execution.
- Continuously improve platform reliability, developer experience, deployment velocity and operational excellence.
- Engage in capacity and demand planning, system performance analysis, tuning and cost optimization.
- Lead production outages, incident response, post-mortems and drive SRE practices across the engineering organisation.
- Make architectural tradeoffs, design and deliver highly scalable, reliable, secure and fault tolerant cloud infrastructure.
- Be a role model for Dev Ops and platform engineers, mentor engineers.
- Demonstrate technical leadership and create impact across teams.
- Lead AI strategy and adoption for productivity, infrastructure automation and operational efficiency.
- Drive infrastructure and deployment practices while being secure and compliant (PCI-DSS, SOC2 and others).
- Participate in infrastructure, architecture and security design reviews to maintain our high engineering standards.
- Partner with engineering and product management teams to define and execute the platform roadmap.
- Translate business and engineering requirements into scalable and extensible infrastructure design.
- Proactively manage stakeholder communication related to deliverables, risks, changes and dependencies.
- Coordinate with cross functional teams (Backend, Frontend, Data, Security, QA etc.) on planning and execution.
- Continuously improve platform reliability, developer experience, deployment velocity and operational excellence.
- Engage in capacity and demand planning, system performance analysis, tuning and cost optimization.
- Lead production outages, incident response, post-mortems and drive SRE practices across the engineering organisation.
- Degree in Computer Science or equivalent (B.E/B.Tech or higher) with 10+ years of experience in Dev Ops, SRE, Platform or Cloud Engineering roles for large distributed systems in reputed organizations.
- Hands-on experience in designing, building and operating cloud infrastructure for large scale production systems on AWS.
- Deep knowledge of Linux as a production environment.
- Strong knowledge of distributed systems, networking, systems internals and asynchronous architectures.
- Expert in at least 1 of the following languages for automation and tooling : go, python, shell scripting.
- Extensive experience with AWS services (ECS, EC2, VPC, IAM, RDS, Kinesis, Secrets Manager, SSM, WAF and more).
- Infrastructure as Code expertise with Terraform, CDK or Cloud Formation.
- Strong experience with CI/CD tooling such as Git Hub Actions, Git Lab CI, Jenkins, and Git Ops tools.
- Hands-on experience with observability stacks
- Prometheus, Grafana, Open Telemetry, distributed tracing and log aggregation. - Experience with container infrastructure
- Docker, container runtimes and image security. - Experience administering and operating RDBMS/No
SQL systems at scale, such as Postgres, MongoDB and Redis. - Ability to design and operate low latency services behind load balancers and API gateways.
- Strong understanding of system performance, scaling and reliability engineering.
- Possess excellent communication, sharp analytical abilities with proven design skills, able to think critically of the current platform in terms of growth and stability.
- Experience with microservice architecture and service-to-service communication patterns.
- Continuously refactor infrastructure and tooling to ensure high-quality design.
- Ability to plan, prioritize, estimate and execute platform releases with good degree of predictability.
- Ability to scope, review and refine user stories for technical completeness and to alleviate dependency risks.
- Passion for learning new things, solving challenging problems.
- Prior experience with fintech and payments.
- Expert level proficiency with Kubernetes in production.
- Prior experience operating infrastructure for stable coins, blockchain or cryptocurrency platforms.
- AWS Cloud Certifications.
- Familiarity with security standards and compliance - PCI-DSS, SOC2, OWASP, static code…
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
Search for further Jobs Here:
×