T

DevOps & Site Reliability Engineer โ€“ Digital-(java spring boot,POS, payment systems)

Tcs Usa
19 hours ago
Full-time
On-site
Deerfield, Illinois, United States
$140,000 - $150,000 USD yearly

DevOps & Site Reliability Engineer โ€“ Digital-(java spring boot,POS, payment systems)

Must Have Technical/Functional Skills

Technology and Programming (Expert Level)

  • Strong proficiency in Java full stack developerย 

  • Object-Oriented programming principles and concepts

  • Hands-on experience on Observability platform Dynatrace

  • Hands-on experience with Spring Framework (Spring Boot, Spring MVC, Spring Security)

  • Knowledge if RESTful API development

  • Experience with database like Oracle, DB2, MySQLย 

  • Proficiency in Payment Switch BASE24 EPS, C++, AS400 and Python is also added advantageย 

  • Domain, Cloud & Platform Engineeringย 

  • Must have domain experience on Retail Point of Sale/Payment Systems/Merchandising/Inventory/Logistics area

  • Expertise in Microsoft Azure, including:ย 

  • Compute (VMs, App Services, Azure Container Apps)

  • Containers & Orchestration (AKS, Docker)

  • Storage, Azure Key Vault, Azure Monitor, Log Analytics

  • Proven experience designing enterprise grade, highly available cloud platforms

DevOps & Engineering Excellence

  • Advanced experience with Azure DevOps and CI/CD pipeline architecture

  • Strong scripting skills (PowerShell, Bash)

  • GitOps concepts, branching strategies, release orchestration

Site Reliability Engineering:

  • Ownership of platform reliability, resiliency, and performance

  • Definition and governance of:ย 

  • SLIs, SLOs, SLAs

  • Error budgets and reliability metrics

Advanced observability strategy, designing and implementation:ย 

  • Metrics, logs, traces, alerts, dashboards using Dynatrace

  • Incident response leadership, RCA facilitation, and long term remediation planning

  • Experience operating 99.9%โ€“99.99% availability systems

Security, Compliance & Cost

  • Secure cloud design using Key Vault, managed identities, RBAC

  • Cost optimization (FinOps mindset) across cloud infrastructure

Roles & Responsibilities

  • ย Act as SRE Technical Architect (Should be interested to work on Implementations) for clients Retail platforms, owning reliability and stability outcomes

  • Define and enforce SRE standards, best practices, and operating models

  • Architect and govern highly available, scalable cloud platforms

  • Lead the design and implementation of CI/CD and IaC strategies

  • Establish proactive monitoring, alerting, and incident prevention mechanisms

  • Own major incident leadership, RCA execution, and corrective action tracking

  • Partner with application, security, and architecture teams to build reliability by design

  • Drive automation to reduce toil and improve operational efficiency

  • Mentor and coach SRE and DevOps engineers across teams

  • Influence roadmap decisions with a reliability, scalability, and cost lens