New · Job Board MCP
Connect Claude, Codex, Copilot, Cursor, or Windsurf to InterviewStack. It finds roles that genuinely fit your resume and saves the best ones to your tracker, running free on a local open model, with your resume never leaving your machine.
Every posting is de-duped, the real hiring company identified, and salary, location, work mode and seniority normalized, then mapped to 80+ defined roles. Your AI filters clean, classified data, not keyword noise.
Run the matching on a local open model (Ollama + Qwen, Llama): no API keys, no per-token bills, and your resume stays on your machine. Or use the assistant you already pay for.
Set it once: every morning your assistant finds fresh matches and saves the best to your applications, each with a concrete note on why it fits. You review and apply.
PHTN.ai
Senior GCP Cloud Engineer
Role Overview
We are seeking a highly skilled Senior GCP Cloud Engineer to design, implement, govern, and support scalable, secure, highly available, and automated cloud platforms on Google Cloud Platform (GCP). The role requires deep technical expertise in GCP services, Infrastructure as Code (IaC), cloud reliability engineering, networking, containerization, and production operations.
The ideal candidate will act as a senior technical leader responsible for defining reference architectures, driving platform reliability, governing cloud standards across multiple projects, and mentoring engineering teams while collaborating closely with architects, development teams, and stakeholders.
Key Responsibilities
Cloud Infrastructure & Platform Engineering
Design, implement, and govern scalable, secure, and highly available cloud infrastructure across GCP environments (Development, Testing, UAT, Pre-Production, Production).
Define and maintain reusable reference architectures and platform standards for GCP services and cloud-native deployments.
Build enterprise-grade architectures using GCP services such as:
GKE, Cloud Run, Cloud Functions
BigQuery, Bigtable, Cloud SQL, Spanner
Cloud Storage, Pub/Sub, Memorystore, Artifact Registry
Configure and manage networking components including:
VPCs, Shared VPC, Private Service Connect, VPC Peering
Load Balancers, Cloud DNS, Cloud Armor, Firewall Rules
VPN and Interconnect connectivity solutions
Drive standardization and consistency of cloud platform implementations across multiple projects and teams.
Reliability Engineering & Production Operations
Lead reliability engineering initiatives to improve platform stability, resiliency, scalability, and fault tolerance.
Establish and enforce operational best practices for monitoring, alerting, incident management, and disaster recovery.
Conduct production readiness reviews (PRRs) for new applications, services, and platform changes prior to production deployments.
Perform capacity planning and infrastructure sizing to ensure platform scalability and performance under varying workloads.
Define and track platform SLIs, SLOs, and operational health metrics.
Lead root cause analysis (RCA) and resolution of critical production incidents and complex infrastructure issues.
Drive continuous improvement initiatives focused on reliability, availability, performance, and operational excellence.
Infrastructure as Code (IaC)
Develop, enhance, and maintain Terraform modules and reusable infrastructure templates.
Ensure consistency, reliability, and compliance across environments using Infrastructure as Code best practices.
Drive automation for infrastructure provisioning, configuration management, and deployment processes.
Collaborate with platform and Terraform engineering teams to improve reusable IaC standards and governance.
Containerization & Microservices
Build and manage containerized applications using Docker.
Deploy and manage workloads on Kubernetes (GKE) and Cloud Run platforms.
Support cloud-native and microservices-based application architectures.
Collaborate with application teams to optimize container orchestration, deployment reliability, and scalability.
Monitoring, Troubleshooting & Incident Management
Monitor platform health using Cloud Monitoring, Cloud Logging, Dynatrace, and related observability tools.
Troubleshoot and resolve complex infrastructure, networking, deployment, and performance issues.
Act as the senior technical escalation point for critical GCP platform related challenges.
Support application teams in diagnosing platform-related production issues and performance bottlenecks.
Security, Governance & Compliance
Implement and enforce IAM policies, security controls, and governancestandards across GCP environments.
Ensure secure networking, data protection, and compliance best practices are consistently followed.
Govern cloud resource usage, security posture, and platform standards across multiple projects and environments.
Collaborate with security and compliance teams to support audits, risk remediation, and governance initiatives.
Technical Leadership
Provide technical leadership and mentorship to junior cloud and DevOps engineers.
Guide engineering teams on GCP best practices, cloud-native architectures, automation strategies, and operational excellence.
Participate actively in technical planning discussions.
Maintain and improve technical documentation, architectural standards, and operational runbooks.
Collaboration & Agile Delivery
Participate actively in Agile/Scrum ceremonies (stand-ups, sprint planning, retrospectives).
Collaborate with:
Platform leads and architects
Development and QA teams
Security and networking teams
Onsite and offshore stakeholders
Maintain and update technical documentation in Confluence and other tools.
Continuous Improvement & Innovation
Stay up to date with the latest GCP innovations, cloud-native technologies, DevOps practices, and industry trends.
Identify opportunities for automation, optimization, performance tuning, and cost efficiency.
Drive innovation in cloud platform engineering, observability, reliability engineering, and deployment automation.
Minimum Qualifications
Bachelor’s or Master’s degree in Computer Science, Information Technology, or a related field.
6+ years of overall IT experience with minimum 4+ years of hands-on GCP experience.
Strong hands-on experience with:
Google Cloud Platform (GCP)
Terraform (Infrastructure as Code)
Docker & containerization
Kubernetes (GKE)
CI/CD tools such as Jenkins and SonarQube
Scripting (Bash/Shell, Groovy)
Strong understanding of:
GCP architecture and cloud-native services
Reliability engineering and production operations
Networking (VPC, VPN, Interconnect, PSC)
Security, IAM, and governance
Monitoring, logging, and observability
Familiarity with Atlassian tools (JIRA, Confluence, Bitbucket).
Strong analytical, troubleshooting, communication, and leadership skills.
Preferred / Nice-to-Have Skills
GCP certifications:
Professional Cloud Architect
Professional Cloud DevOps Engineer
Experience with:
Apigee API management
Microservices architecture
Hybrid cloud environments (On-Premises + GCP)
Reliability engineering / SRE practices
Production readiness review processes
Capacity planning and performance engineering
Exposure to:
Akamai integrations
Dynatrace and advanced observability platforms
GitOps and deployment automation tools
Knowledge of build tools such as NPM and Gradle.
Key Competencies
Strong hands-on technical expertise
Cloud architecture and platform engineering leadership
Reliability engineering and operational excellence mindset
Advanced troubleshooting and incident management skills
Automation and DevOps orientation
Technical mentorship and cross-team collaboration
Strategic thinking and problem-solving ability
Clear communication and documentation skills
Adaptability in fast-paced enterprise environments
This job is found at InterviewStack.io