DevOps Lead
Apply for this position
All fields marked * are required
What You'll Do:
Design, implement, and manage highly available, scalable, and secure systems on Google Cloud Platform
Automate infrastructure provisioning using Infrastructure as Code tools such as Terraform and Deployment Manager
Build and maintain CI/CD pipelines using Jenkins, GitHub Actions, GitLab CI, or Cloud Build
Implement monitoring, logging, and alerting solutions using Google Cloud Monitoring / Logging and tools like Prometheus, Grafana, ELK (any one)
Manage Kubernetes clusters using Google Kubernetes Engine
Ensure system reliability through capacity planning, performance tuning, disaster recovery planning, and chaos engineering practices
Participate in on-call rotations and incident response processes
Drive root cause analysis (RCA) and conduct post-incident reviews
Collaborate with development teams to improve deployment processes and application reliability
Implement cloud security best practices across environments
What You Know:
Must have 9+ years of experience as an SRE, DevOps Engineer, or Cloud Engineer
Strong hands-on experience with GCP services, including:
Google Compute Engine
Google Kubernetes Engine
Google Cloud Functions / Google Cloud Run
Google Cloud Storage
Google BigQuery
Google Cloud Pub/Sub
Google Cloud IAM & Networking
Experience with containerization and orchestration (Docker, Kubernetes)
Proficiency in Infrastructure as Code tools such as Terraform, Helm, Ansible, or similar
Strong scripting skills in at least one language (Python, Go, Bash, or Java)
Experience building and managing CI/CD pipelines and automation frameworks
Solid understanding of networking concepts, Linux systems, and security best practices
Experience with monitoring and observability tools
Strong troubleshooting and problem-solving skills
Preferred Qualifications
Experience working in multi-cloud environments (GCP)
Knowledge of service mesh technologies such as Istio or Anthos
Experience with GCP databases such as Cloud SQL, Spanner, or Firestore
Familiarity with SLO / SLI / SLA frameworks
Experience supporting high-traffic production environments
Mandatory Skills
Experience managing infrastructure on Google Cloud Platform (GCP)
Strong hands-on expertise in Compute Engine, GKE, Cloud Functions / Cloud Run
Experience implementing CI/CD pipelines and Infrastructure as Code (Terraform)
Kubernetes cluster management experience
Monitoring, logging, and alerting implementation
Incident management, RCA, and reliability engineering practices
Strong understanding of networking, Linux, and cloud security
Education:
Bachelor’s degree in computer science, Information Systems, Engineering, Computer Applications, or related field.
Required
Preferred