What are the top 3 skills required for this role? 1. Kubernetes 2. SRE 3. AWS
Job Description:
Partner with business stakeholders across Prosperity to understand workflows and pain points, and proactively identify areas where agentic AI can drive meaningful improvement.
Monitor and maintain production environments running on AWS EKS and Linux/Unix servers.
Ensure platform availability, reliability, scalability, and operational stability.
Troubleshoot application, infrastructure, networking, and deployment-related issues.
Investigate alerts, incidents, and performance bottlenecks; drive resolution within SLA timelines.
Perform Root Cause Analysis (RCA) and implement preventive actions.
Support Kubernetes cluster operations, upgrades, patching, and capacity management.
Manage and troubleshoot Kubernetes resources including Pods, Deployments, Services, Ingress, ConfigMaps, and Secrets.
Configure and maintain monitoring, alerting, logging, and observability solutions.
Develop automation scripts and operational runbooks to reduce manual effort and improve system reliability.
Support CI/CD deployments, release validations, rollbacks, and production change activities.
Collaborate with development, infrastructure, cloud, and support teams during incidents and deployments.
Participate in on-call and rotational support activities.
Continuously identify opportunities for reliability improvements, automation, and operational excellence.
Applicant Notices & Disclaimers
For information on benefits, equal opportunity employment, and location-specific applicant notices, click here
At SPECTRAFORCE, we are committed to maintaining a workplace that ensures fair compensation and wage transparency in adherence with all applicable state and local laws.This position's pay range is $40.00/hr - $45.00/hr.