|
Ampcus Inc. is a certified global provider of a broad range of Technology and Business consulting services. We are in search of a highly motivated candidate to join our talented Team.
CTH/FTE
Austin/Southlake, TX
Detailed Job Description
We are seeking a highly skilled and visionary Lead GCP SRE / DevOps Engineer to spearhead the reliability, scalability, and automation of our cloud infrastructure. In this role, you will bridge the gap between development and operations, leading a team to design and maintain robust CI/CD pipelines, optimize Google Cloud Platform (GCP) infrastructure, and enforce high availability across all production environments.
Key Responsibilities
Infrastructure & Architecture
- Design and maintain scalable, secure, and fault-tolerant infrastructure entirely within Google Cloud Platform (GCP).
- Architect Infrastructure as Code (IaC) templates using Terraform to manage multi-environment setups.
- Optimize cloud spend and maximize performance across GCP services (GKE, Compute Engine, BigQuery, Cloud SQL).
CI/CD & Automation
- Own the deployment lifecycle by building and optimizing automated CI/CD pipelines (using tools like GitHub Actions, GitLab CI, or Jenkins).
- Drive containerization strategies using Docker and orchestration via Google Kubernetes Engine (GKE).
- Automate repetitive operational tasks ("toil") using scripting languages like Python, Go, or Bash.
Observability & SRE Practices
- Define, measure, and report critical reliability metrics, including SLIs, SLOs, and Error Budgets.
- Implement comprehensive monitoring, logging, and alerting systems using GCP Cloud Monitoring/Logging, Prometheus, Grafana, or Datadog.
- Lead incident response rotations and facilitate constructive, blameless post-mortems to prevent recurrence.
Required Skills & Qualifications
Technical Essentials
- Experience: 8 years of experience in DevOps, Systems Engineering, or SRE roles, in GCP Platform.
- Cloud Platform: Deep production-level expertise with Google Cloud Platform (GCP).
- Orchestration: Advanced hands-on experience managing production workloads on Kubernetes (specifically GKE).
- Infrastructure as Code: Proficient with Terraform for state management and modular design.
- Programming: Strong coding skills in Python or Go, alongside excellent shell scripting capabilities.
- CI/CD: Proven track record building enterprise-grade pipelines.
- Experience migrating legacy workloads from on-premise or AWS/Azure into GCP.
Soft Skills & Leadership
- Strong communication skills with the ability to explain complex architectural concepts to non-technical stakeholders.
- Proven ability to manage high-pressure live production incidents calmly and methodically.
Preferred Qualifications (Nice to Have)
- Certifications: GCP Certified Professional Cloud Architect or GCP Certified Professional Cloud DevOps Engineer.
- Familiarity with service mesh technologies (e.g., Istio, Anthos Service Mesh).
Ampcus is an Equal Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, age, protected veterans or individuals with disabilities.
|