Job code
IRC302637
Published on 18 August 2026

DevOps Lead IRC302637

Function

Software Product Engineering

Experience

10-15 years

Location

India - Hyderabad

Skills

CI/CD Tools, Cloud Infrastructure, Cloud Platforms: Google Cloud Platform (Primary), Databases, Monitoring, Networking, Scripting, Security, Version Control

Work Model

Hybrid

Apply

Description

Lead the end-to-end DevOps and infrastructure function for an advanced analytics platform deployed on GCP. The role owns infrastructure-as-code, CI/CD pipeline design and governance, multi-environment promotion strategy, MLOps infrastructure, multi-tenant provisioning, and production reliability – ensuring the platform scales securely across clients and geographies.

Requirements

Domain
Requirement
Cloud Platform
Deep hands-on expertise with GCP – Compute, Networking, IAM, Cloud Run, Cloud SQL, BigQuery, Cloud Storage, Artifact Registry, Vertex AI, Cloud Composer
Infrastructure as Code
Production-grade Terraform – modules, workspaces, remote state (GCS), state locking, Terragrunt wrappers; experience managing 4+ environments from a single codebase
CI/CD
GitLab CI/CD – pipeline design, YAML templating, multi-stage deployment, manual gates, Kubernetes executor runners, pipeline optimization
Containers
Docker containerization, image lifecycle management, Artifact Registry, Cloud Run deployment patterns
Networking
VPC design, private connectivity, HTTPS Load Balancers, SSL/TLS, DNS management, CDN/WAF integration (Akamai or Cloudflare)
Databases
Operational management of Cloud SQL (PostgreSQL) and BigQuery – provisioning, backups, replication, access control
Security
GCP IAM, service accounts, workload identity, secret management, least-privilege enforcement, Cloud Armor
Monitoring
Setting up observability stacks — GCP Cloud Monitoring/Logging, alerting policies, SLO tracking
Scripting
Proficiency in Python and Bash for automation, glue scripts, and pipeline tooling
Version Control
Advanced Git workflows, branch protection, MR-based collaboration, GitOps principles

Analytics & Data Platform Domain Skills
Multi-Tenant Infrastructure: Experience provisioning isolated per-client resources (datasets, databases, storage buckets, service accounts) from parameterized IaC templates
Data Pipeline Infrastructure: Managing Cloud Composer/Airflow environments, DAG deployment, and orchestration of BigQuery ↔ PostgreSQL sync pipelines
MLOps Infrastructure: Provisioning Vertex AI Workbenches, model pipeline compute (including GPU), and model artifact storage with governed access
Environment Parity: Maintaining functional parity across dev/qa/uat/prd with differences limited to scaling – ensuring consistent data architecture, networking, and IAM across tiers
Client Onboarding Automation: Automating tenant provisioning via Terraform – dual-database strategy (workspace DB + published DB), BigQuery datasets, GCS buckets, and IAM bindings per client

Good-to-Have
Experience with Helm charts for Kubernetes-based auxiliary services
Familiarity with Akamai CDN/WAF configuration and integration with GCP Load Balancers
Exposure to WPP Open OS or similar micro-frontend hosting platforms (DevHub publishing, CSP configuration)
Experience with Cloud Functions for lightweight serverless glue logic
Knowledge of Redis cache provisioning and management
Exposure to Pub/Sub for event-driven pipeline triggers
Experience managing GitLab seat licensing and repository access governance at an organizational level
Familiarity with data governance, retention policies, and GDPR-aligned infrastructure design
Certification: GCP Professional Cloud DevOps Engineer or equivalent

Soft Skills
Strong ownership mindset for platform reliability and production uptime
Ability to create and maintain clear documentation – runbooks, ADRs, environment specs, and deployment guides
Effective cross-functional collaborator – comfortable working with backend engineers, data scientists, QA, and product stakeholders
Proactive problem solver who anticipates infrastructure bottlenecks and builds for scale
Calm under pressure – able to lead incident response and post-incident reviews

 

Job responsibilities

Key Responsibilities
Infrastructure & IaC
Own and evolve the Terraform codebase for provisioning all GCP resources – VPC, Cloud SQL (PostgreSQL), Cloud Run, Load Balancers, Artifact Registry, BigQuery, Cloud Composer (Airflow), Cloud Storage, and Vertex AI Workbench
Implement modular, parameterized Terraform modules using Terragrunt wrappers aligned with the Measure Project Factory templates
Manage Terraform state (GCS backend with state locking) and ensure consistent, drift-free environments across dev → qa → uat → prd
Provision and manage per-client infrastructure (datasets, databases, buckets, service accounts) as part of the multi-tenant onboarding process
CI/CD Pipeline Design & Governance
Design, maintain, and optimize GitLab CI/CD pipelines for frontend, backend (FastAPI/Python), infrastructure, and data engineering deployments
Implement environment promotion workflows with appropriate manual gates, MR approval policies, and branch protection rules
Enforce quality gates – build-fail enforcement on unit tests, SonarQube, code coverage thresholds, and configuration-policy validation
Manage GitLab Kubernetes executor runners and pipeline caching strategies
Prevent environment misconfigurations through automated validation of environment-specific variables and credentials
Cloud Platform & Networking
Architect and manage GCP networking — dedicated VPCs, subnets, firewall rules, private IP connectivity, and VPC connectors for secure service-to-service communication
Configure and maintain HTTPS Load Balancers with SSL/TLS termination, Cloud Armor security policies, and Akamai CDN/WAF integration
Manage DNS, certificate provisioning, and edge routing for production traffic
MLOps & Data Infrastructure
Provision and manage Vertex AI Workbench instances for data scientists and analysts with appropriate IAM bindings
Support MLOps pipeline infrastructure – GPU-enabled Vertex AI compute, model artifact storage (GCS), pipeline orchestration, and Artifact Registry for container images
Manage Cloud Composer (Airflow) environments for data pipeline orchestration – DAG deployment, monitoring, and health validation
Support BigQuery-to-PostgreSQL replication pipelines and data sync automation
Security & Access Management
Implement least-privilege IAM policies with per-client service accounts and role-based access control
Manage workload identity federation, secret management, and service account lifecycle
Ensure OIDC/OAuth2 integration for application authentication and platform-level authorization
Maintain compliance with data governance requirements including data retention policies and GDPR alignment
Observability & Reliability
Establish monitoring, alerting, and observability for infrastructure and application services — including pipeline failures, data sync gaps, and service health
Define and track SLOs for environment availability, deployment success rates, and recovery objectives (RPO/RTO)
Implement automated health checks, self-healing configurations, and incident response procedures
Production Readiness & Operations
Lead production deployments with zero-downtime strategies and validated rollback procedures
Conduct infrastructure readiness reviews before UAT handover and production go-live
Document deployment runbooks, disaster recovery procedures, and operational playbooks
Team Leadership
Mentor and guide junior DevOps/cloud engineers, conduct code reviews for IaC changes, and establish engineering standards

What we offer

Exciting Projects: We focus on industries like High-Tech, communication, media, healthcare, retail and telecom. Our customer list is full of fantastic global brands and leaders who love what we build for them.

Collaborative Environment: You Can expand your skills by collaborating with a diverse team of highly talented people in an open, laidback environment — or even abroad in one of our global centers or client facilities!

Work-Life Balance: GlobalLogic prioritizes work-life balance, which is why we offer flexible work schedules, opportunities to work from home, and paid time off and holidays.

Professional Development: Our dedicated Learning & Development team regularly organizes Communication skills training(GL Vantage, Toast Master),Stress Management program, professional certifications, and technical and soft skill trainings.

Excellent Benefits: We provide our employees with competitive salaries, family medical insurance, Group Term Life Insurance, Group Personal Accident Insurance , NPS(National Pension Scheme ), Periodic health awareness program, extended maternity leave, annual performance bonuses, and referral bonuses.

Fun Perks: We want you to love where you work, which is why we host sports events, cultural activities, offer food on subsidies rates, Corporate parties. Our vibrant offices also include dedicated GL Zones, rooftop decks and GL Club where you can drink coffee or tea with your colleagues over a game of table and offer discounts for popular stores and restaurants!

About GlobalLogic

GlobalLogic, a Hitachi Group Company, is a trusted digital engineering partner to the world’s largest and most forward-thinking companies. Since 2000, we’ve been at the forefront of the digital revolution – helping create some of the most innovative and widely used digital products and experiences. Today we continue to collaborate with clients in transforming businesses and redefining industries through intelligent products, platforms, and services.

Apply Now

The gender information on this form helps us understand the makeup of our applicant pool in this key area, and to continuously improve our efforts to make our workforce more inclusive.

Drag and drop your file here or click here to upload

Only .docx, .rtf, .pdf formats allowed to a max size of 5 MB.

Alternately you can include your Linkedin profile