More jobs:
Cloud Platform Developer
Job in
Denver, Denver County, Colorado, 80285, USA
Listed on 2026-07-05
Listing for:
Veriipro
Full Time
position Listed on 2026-07-05
Job specializations:
-
IT/Tech
SRE/Site Reliability, Cloud Computing: Infrastructure & Operations, IT Infrastructure, Data Engineering
Job Description & How to Apply Below
ONLY USC/GC – W2 ROLE Must Have Skills
- Must have deep, hands‑on experience running Kafka in large‑scale production environments, including cluster operations, upgrades, patches, and migrations.
- Should understand Kafka internals such as partitions, replication, retention/compaction, and rebalance strategies.
- Kafka Administration
- Platform / SRE / Dev Ops Experience
- Kafka Ecosystem Tools
- Linux + Networking
- Automation / Scripting
- Monitoring / Observability
- Disaster Recovery
- AWS MSK / Apache Kafka Cloud:
Experience with MSK operations and cloud‑aligned Kafka environments. - Helpful for cross‑environment consistency between on‑prem and cloud.
- Hardware Refresh
Experience:
Prior work leading Kafka hardware refreshes or cluster rebuilds.
Job Description
We’re seeking a senior contract Kafka/Confluent administrator to own and evolve our on‑prem event streaming platform, with a primary focus on Confluent Platform. You will lead planning and execution of a hardware refresh for our on‑prem clusters, drive reliability and performance, and embed Dev Ops/automation across provisioning, deployment, observability, and incident response. Experience with Apache Kafka and AWS MSK is desired for secondary support and cross‑environment alignment.
Comprehensive documentation and runbooks are required deliverables.
Key Responsibilities
- Design, deploy, and operate highly available Kafka clusters (on‑prem, cloud, and/or managed services such as Confluent Cloud or AWS MSK).
- Manage topics, partitions, quotas, retention policies, and consumer group strategies for performance and cost.
- Own upgrades, patches, and migrations.
- Implement and manage Kafka components:
Kafka Connect, Schema Registry, Mirror Maker/Confluent Replicator, REST Proxy; familiarity with Kafka Streams and ksqlDB is a plus. - Performance tuning (producers/consumers, batching, compression, acks, ISR, controller health), throughput testing, and benchmarking.
- Capacity planning, partitioning strategy, and cluster right‑sizing.
- Hardware refresh plan: capacity model, sizing, architecture diagrams, migration/cutover strategy, risk register
- Implement and validate on‑prem clusters on refreshed hardware with performance benchmarks
- Operational documentation: standards, runbooks, monitoring/alerts configuration, backup/restore and DR playbooks.
- Knowledge transfer sessions and documentation handoff at milestones and project close.
- 5+ years in systems/platform engineering, SRE, or Dev Ops; 4+ years operating Kafka in production at scale.
- Deep knowledge of Kafka internals: partitions, replication, retention/compaction, rebalance strategies.
- Hands‑on with Kafka Connect, Schema Registry, Mirror Maker/Confluent Replicator.
- Strong Linux fundamentals; networking (TCP, DNS, load balancing), and performance analysis.
- Proficiency in automation/scripting.
- Monitoring/observability:
Data Dog, Grafana, JMX exporters, and log aggregation. - Experience with DR, multi‑region design, and incident management.
- Proven ability to produce clear, comprehensive documentation
- Experience with Apache Kafka and AWS MSK operations and integration.
- Experience executing hardware refreshes and major cluster rebuilds/migrations with minimal downtime.
5+ years
Certifications NeededNone
#J-18808-LjbffrTo View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
Search for further Jobs Here:
×