TD Cloud Infrastructure Engineer
Summary
Design, build, and maintain secure, scalable Google Cloud Platform infrastructure for what appears to be a government-sector engagement, using Terraform, Docker/Kubernetes (GKE), CI/CD pipelines, and Python/Go automation. Requires an active Top Secret/SCI clearance with polygraph and includes on-call rotations.
- Design, deploy, and manage GCP infrastructure using Infrastructure as Code (IaC) tools like Terraform.
- Develop and implement automation for provisioning, configuration, and management of cloud resources.
- Ensure the security, compliance, and reliability of cloud environments, adhering to Google's internal standards and best practices.
- Collaborate with software engineering teams to define infrastructure requirements and integrate cloud services into application development pipelines.
- Implement and manage networking, compute, storage, and database services within GCP.
- Monitor, troubleshoot, and resolve issues related to cloud infrastructure and services.
- Contribute to and maintain documentation, including architecture diagrams, runbooks, and playbooks.
- Participate in on-call rotations as needed to support critical infrastructure.
- Bachelor's degree in Computer Science, a related technical field, or equivalent practical experience.
- 2+ years of experience in cloud infrastructure engineering, with a focus on Google Cloud Platform (GCP).
- Proficiency with Infrastructure as Code (IaC) tools, especially Terraform.
- Experience with containerization technologies (e.g., Docker, Kubernetes) and orchestration (e.g., Google Kubernetes Engine - GKE).
- Strong understanding of networking concepts, including VPC, subnets, firewalls, and load balancing.
- Familiarity with scripting languages (e.g., Python, Go) for automation.
- Experience with CI/CD pipelines and DevOps practices.
- Master's degree in a technical field.
- Google Cloud certifications (e.g., Professional Cloud Architect, Professional Cloud DevOps Engineer).
- Experience with large-scale distributed systems and microservices architectures.
- Knowledge of Site Reliability Engineering (SRE) principles and practices.
- Familiarity with monitoring and logging tools (e.g., Cloud Monitoring, Cloud Logging).
- Ability to work effectively in a collaborative, fast-paced environment.