Site Reliability Engineer
Indotronix International Corporation
8 days ago
Contract
On-site
Greenwood Village, Colorado, United States
Site Reliability Engineer – Greenwood Village, CO (Hybrid)
About the Role:
Join a forward-thinking engineering team as a Site Reliability Engineer, specializing in enterprise-scale experimentation and configuration management platforms. In this operations-focused role, you’ll drive platform reliability, optimize cloud architecture, and take ownership of mission-critical systems within a dynamic, hybrid work environment based in Greenwood Village, Colorado.
Responsibilities:
- Maintain and enhance Terraform modules to define and audit AWS infrastructure, ensuring state consistency and resolving configuration drift.
- Operate and optimize AWS services and resources such as EKS, Helm, Istio, Aurora, DocumentDB, Redis, Amazon MQ, Route53, WAFv2, CloudFront, and S3.
- Own and enforce deployment standards using GitLab CI/CD pipelines, including progressive promotion and pipeline-only deployment.
- Build, deploy, and validate software releases across multiple environments; document and manage detailed release notes.
- Right-size and scale resources to meet stringent SLAs while optimizing for cost efficiency.
- Collaborate closely with developers and test engineers to elevate application performance.
- Lead end-to-end monitoring and alerting using Datadog, Splunk, and related observability tools.
- Serve as the first responder for incidents, handling mitigation, recovery, and root-cause analysis under SLA obligations.
- Act as the subject-matter expert for infrastructure and pipeline issues, answering team questions and escalating architectural decisions.
- Work cross-functionally with onshore and offshore teams to ensure platform stability and continuous improvement.
Required Skills and Experience:
- 6+ years of DevOps experience in large-scale, complex environments.
- Proficiency with AWS cloud infrastructure and Terraform.
- Strong Kubernetes expertise, including hands-on experience with containerized microservice applications.
- Proven experience deploying with GitLab CI/CD or similar tools.
- Advanced skills in observability and monitoring (e.g., Datadog, Splunk), including dashboard creation and alert tuning.
- Demonstrated success in production incident triage, mitigation, and root-cause analysis under SLA constraints.
- Solid knowledge of Git-based source control workflows.
- Bachelor’s degree in Computer Science, Engineering, or related field, or equivalent professional experience.
Preferred Skills:
- Familiarity with Python, Node.js, React, TypeScript, GraphQL application stacks.
- Experience with both SQL and NoSQL/document data stores.
- Working knowledge of Kubernetes internals, Helm, and Istio service mesh.
- Experience with blue/green or canary/progressive deployment strategies.
- Hands-on infrastructure cost optimization and AWS multi-account architecture exposure.
- Master’s degree or higher in a related field.
Benefits:
- Hybrid work environment offering flexibility and work-life balance.
- Opportunity to work on innovative, enterprise-scale platforms with cutting-edge cloud and DevOps technologies.
- Collaborate with highly skilled engineering professionals in a supportive, growth-oriented culture.
- Expand your expertise in AWS, Kubernetes, CI/CD, observability, and infrastructure automation.
- Contract position with competitive compensation.
How to Apply:
If you are passionate about platform reliability and optimization, and thrive in high-impact technical roles, submit your application today to join our Greenwood Village team. Face-to-face interviews required; candidates must be authorized to work in the U.S. without sponsorship.
(JSON format):
About the Role:
Join a forward-thinking engineering team as a Site Reliability Engineer, specializing in enterprise-scale experimentation and configuration management platforms. In this operations-focused role, you’ll drive platform reliability, optimize cloud architecture, and take ownership of mission-critical systems within a dynamic, hybrid work environment based in Greenwood Village, Colorado.
Responsibilities:
- Maintain and enhance Terraform modules to define and audit AWS infrastructure, ensuring state consistency and resolving configuration drift.
- Operate and optimize AWS services and resources such as EKS, Helm, Istio, Aurora, DocumentDB, Redis, Amazon MQ, Route53, WAFv2, CloudFront, and S3.
- Own and enforce deployment standards using GitLab CI/CD pipelines, including progressive promotion and pipeline-only deployment.
- Build, deploy, and validate software releases across multiple environments; document and manage detailed release notes.
- Right-size and scale resources to meet stringent SLAs while optimizing for cost efficiency.
- Collaborate closely with developers and test engineers to elevate application performance.
- Lead end-to-end monitoring and alerting using Datadog, Splunk, and related observability tools.
- Serve as the first responder for incidents, handling mitigation, recovery, and root-cause analysis under SLA obligations.
- Act as the subject-matter expert for infrastructure and pipeline issues, answering team questions and escalating architectural decisions.
- Work cross-functionally with onshore and offshore teams to ensure platform stability and continuous improvement.
Required Skills and Experience:
- 6+ years of DevOps experience in large-scale, complex environments.
- Proficiency with AWS cloud infrastructure and Terraform.
- Strong Kubernetes expertise, including hands-on experience with containerized microservice applications.
- Proven experience deploying with GitLab CI/CD or similar tools.
- Advanced skills in observability and monitoring (e.g., Datadog, Splunk), including dashboard creation and alert tuning.
- Demonstrated success in production incident triage, mitigation, and root-cause analysis under SLA constraints.
- Solid knowledge of Git-based source control workflows.
- Bachelor’s degree in Computer Science, Engineering, or related field, or equivalent professional experience.
Preferred Skills:
- Familiarity with Python, Node.js, React, TypeScript, GraphQL application stacks.
- Experience with both SQL and NoSQL/document data stores.
- Working knowledge of Kubernetes internals, Helm, and Istio service mesh.
- Experience with blue/green or canary/progressive deployment strategies.
- Hands-on infrastructure cost optimization and AWS multi-account architecture exposure.
- Master’s degree or higher in a related field.
Benefits:
- Hybrid work environment offering flexibility and work-life balance.
- Opportunity to work on innovative, enterprise-scale platforms with cutting-edge cloud and DevOps technologies.
- Collaborate with highly skilled engineering professionals in a supportive, growth-oriented culture.
- Expand your expertise in AWS, Kubernetes, CI/CD, observability, and infrastructure automation.
- Contract position with competitive compensation.
How to Apply:
If you are passionate about platform reliability and optimization, and thrive in high-impact technical roles, submit your application today to join our Greenwood Village team. Face-to-face interviews required; candidates must be authorized to work in the U.S. without sponsorship.
(JSON format):