← Back to jobs
Chicago, IL, USA
No related jobs found
>> Strong expertise in Microservices architecture with practical experience designing, deploying, and supporting distributed systems in production environments. >> Deep hands-on knowledge of Kubernetes (deployment management, scaling, upgrades, troubleshooting, cluster operations) with a focus on reliability, resilience, and performance. >> Working proficiency with API Gateway platforms such as Azure API Management (APIM), Kong, and IBM API Connect (APIC) for traffic management, rate limiting, routing, and API observability. >> Solid experience with observability tooling, including Splunk, AppDynamics, Instana, or similar solutions covering log analytics, metrics, traces, dashboards, alerting, and SLO-based monitoring. >> Ability to diagnose and resolve complex production issues, perform root cause analysis (RCA), and implement preventative measures. >> Familiarity with Site Reliability Engineering (SRE) best practices, including error budgets, SLIs/SLOs, incident response, postmortems, automation, and continuous improvement. >> Experience with performance tuning, capacity planning, and improving system reliability through scalable architectures and elimination of toil. >> As this is production support, the 2 resources should be able to support off-hour incidents, releases, and maintenance. Requirements: >> 4-6 years of experience >> Experience in Microservices architecture >> Practical experience in designing and deploying distributed systems in production environments >> Hands-on knowledge of Kubernetes >> Proficiency with API Gateway platforms such as Azure API Management (APIM), Kong, and IBM API Connect (APIC) >> Experience with observability tooling including Splunk
Bachelor's degree
No related jobs found
← Back to jobs