Description
GlobalLogic estimates the starting base salary for the Senior Site Reliability Engineer position in Minneapolis, MN to be between $120000 to $130000 and reflects base salary only and does not include additional performance-linked variable compensation, benefits etc that may be applicable for the role. This pay range is provided as a good faith estimate and the amount offered may be higher or lower. GlobalLogic takes many factors into consideration in making an offer, including candidate qualifications, work experience, operational needs, travel and onsite requirements, internal peer equity, prevailing wage, responsibilities, and other market and business considerations.
Requirements
Experience in SRE, DevOps, or related application delivery & infrastructure roles
• AWS Platform: Hands on experience with AWS services – in particular serverless architectures (CloudWatch, S3, RDS, Lamba, IAM, API Gateway, etc…) supporting API development in a microservices architecture
• Infrastructure as Code: Experience with Terraform and CloudFormation – Proven ability to write and manage Infrastructure as Code (IaC)
• Programming Languages: Strong programming skills in Python
• Data Formats: Experience with JSON, XML and other relevant data formats
• CI/CD Tools: experience setting up and managing CI/CD pipelines using GitLab CI, Jenkins, or similar tools
• Scripting and automation: experience in scripting language such as Python, TypeScript, etc…
• Monitoring and Logging: Familiarity with monitoring & logging tools like Splunk, Dynatrace, Elastic, Prometheus, Grafana, Datadog, New Relic
• Source Code Management: Expertise with git commands and associated VCS (Gitlab, Github, Gitea or similar)
• Documentation: Experience with markup syntax and Docs as Code for creating technical documentation
Job responsibilities
System Reliability & Monitoring:
Contribute to the design, build, and maintenance of highly available and scalable infrastructure in AWS based cloud ecosystem
Implement and manage observability tools (e.g., Splunk, Dynatrace, Elastic, Prometheus, Grafana, Datadog, New Relic) for real-time monitoring, logging, and alerting.
Perform proactive performance tuning and capacity planning across various system components, including web applications, APIs and data pipelines
Implement end-to-end observability for AI/ML model pipelines, ensuring continuous monitoring of data quality, training outputs, model inference accuracy, and operational metrics—critical for traceability under FDA SaMD and EU MDR.
Work with product teams to identify product performance indicators (KPIs) and develop dashboards and alerts
Incident Management & Root Cause Analysis
Participate in incident response efforts to ensure SLA adherence and continuous improvement of Digital Health platforms
Development and automation of solutions to improve system recovery time, recovery objectives, and resiliency
Manage cyber security monitoring capabilities and resolution of vulnerabilities
Security & Compliance
Work with compliance and cybersecurity teams to ensure systems meet regulatory requirements (HIPAA, GDPR, ISO standards)
Implement logging, access controls, and audit trails to support secure operations
Documentation
Develop and maintain infrastructure and operational documentation using “documentation as code”
Ensure all runbooks, incident response procedures, and troubleshooting guides are up-to-date, version-controlled, and easily accessible
Collaborate with cross-functional teams to document SLAs, SLOs, and system dependencies clearly and consistently
Integrate documentation with CI/CD pipelines to enable automated validation and change tracking of system configurations
Create and maintain automated onboarding and knowledge base resources to support team training and continuity
What we offer
Exciting Projects:Come take your place at the forefront of digital transformation! With clients across all industries and sectors, we offer an opportunity to work on market-defining products using the latest technologies.
Collaborative Environment: You can expand your skills by collaborating with a diverse team of highly talented people in an open, laidback environment — or even abroad in one of our global centers or client facilities!
Work-Life Balance:GlobalLogic prioritizes work-life balance, which is why we offer flexible work schedules and opportunities to work from home.
Professional Development:We provide continuing education classes, professional certification and training (technical, soft skills, language, and communication skills) to help you realize your professional goals. Being part of a global organization, there are additional learning opportunities through international knowledge exchanges.
Excellent Benefits:We provide our employees with competitive salaries, health and life insurance, short-term and long-term disability insurance, a matched contribution 401K plan, flexible spending accounts, and PTO and holidays
About GlobalLogic
GlobalLogic, a Hitachi Group Company, is a trusted digital engineering partner to the world’s largest and most forward-thinking companies. Since 2000, we’ve been at the forefront of the digital revolution – helping create some of the most innovative and widely used digital products and experiences. Today we continue to collaborate with clients in transforming businesses and redefining industries through intelligent products, platforms, and services.



