Data Center Cloud & Automation SME Infrstructure
Culver, CA - USA
Job Summary
Data Center / Cloud & Automation SME - AWS / Linux / Terraform / Ansible
Experience: 10 Years
Location: Culver City CA
Work Mode: 100% Onsite from Day 1
Role Level: Systems Engineer - L3 Infrastructure Support
We are seeking an experienced Data Center / Cloud & Automation SME to provide L3 infrastructure support across Linux systems AWS cloud infrastructure and automation platforms.
The ideal candidate will have strong hands-on expertise in Linux Administration AWS Terraform Ansible and infrastructure automation with additional experience in Windows Citrix VMware and Nutanix environments.
- Linux Administration
- AWS Cloud
- Terraform
- Ansible
- Infrastructure Automation
- Bash/Shell Scripting
- Git/GitHub
- Windows Administration
- Citrix Administration
- VMware
- Nutanix
- Provide L3 infrastructure support for Linux servers and enterprise infrastructure.
- Install configure patch monitor and troubleshoot RHEL Amazon Linux and Ubuntu servers.
- Manage Linux users permissions services processes storage and system performance.
- Provision configure and maintain AWS infrastructure.
- Manage AWS services including EC2 VPC Subnets IAM S3 Certificate Manager Load Balancers Auto Scaling and CloudWatch.
- Implement Infrastructure as Code (IaC) using Terraform.
- Create and maintain reusable Terraform modules variables configurations and state files.
- Automate OS configuration application deployment patching and operational tasks using Ansible.
- Develop and maintain Ansible playbooks and roles.
- Manage infrastructure code using Git/GitHub.
- Monitor servers and cloud infrastructure and respond to incidents.
- Perform troubleshooting root cause analysis (RCA) and preventive actions.
- Implement security access-control encryption patching and compliance best practices.
- Collaborate with application network security and other infrastructure teams.
- Support application deployments infrastructure changes and production releases.
- Maintain SOPs runbooks architecture documentation and operational procedures.
- Identify opportunities to improve automation reliability scalability and operational efficiency.
- Strong hands-on experience with RHEL Amazon Linux and/or Ubuntu.
- Server installation configuration patching and upgrades.
- User and permission management.
- System troubleshooting and performance monitoring.
- Bash/shell scripting.
Hands-on experience with:
- EC2
- VPC
- Subnets
- IAM
- S3
- Certificate Manager
- Load Balancers
- Auto Scaling
- CloudWatch
Understanding of AWS security availability scalability and cost optimization.
- Hands-on experience developing and maintaining Terraform configurations.
- Experience creating Terraform modules and variables.
- Understanding of Terraform state management.
- Experience provisioning and maintaining AWS infrastructure through IaC.
- Ability to develop reusable and standardized infrastructure components.
- Strong hands-on experience with Ansible.
- Ability to develop and maintain playbooks and roles.
- Experience automating OS configuration patching deployment and operational tasks.
- Ansible Tower/AWX experience is a plus.
- Working knowledge of Git/GitHub.
- Experience managing infrastructure code and collaborating through source control.
- Strong understanding of networking fundamentals.
- Understanding of VPCs subnets routing security groups load balancing DNS and connectivity.
- Experience with infrastructure and cloud monitoring.
- Incident management and production troubleshooting.
- Root cause analysis and preventive action implementation.
- Understanding of:
- Access control
- IAM
- Encryption
- Patch management
- Secure configuration
- Vulnerability management
- Compliance requirements
- Windows Server Administration
- Citrix:
- StoreFront
- Delivery Controllers
- NetScaler
- PAM
- XenApp
- VMware
- Nutanix
- Jenkins / GitHub Actions / GitLab CI
- Docker / Kubernetes
- AWS RDS Lambda DynamoDB Route 53 EKS
- Prometheus / Grafana / ELK / CloudTrail
- Ansible Tower / AWX
- Advanced Terraform
- Multi-account AWS environments
- Terraform workspaces and complex module design
- Python scripting
- ITIL / ITSM
- ServiceNow
- Technical documentation and knowledge sharing
- 10 years of overall infrastructure / systems engineering experience.
- Strong experience in L3 infrastructure support.
- Strong hands-on experience with Linux Administration.
- Strong AWS infrastructure experience.
- Hands-on Terraform and Ansible automation experience.
- Experience supporting enterprise production environments.
- Strong troubleshooting and problem-solving capabilities.
- Ability to work independently and collaborate with cross-functional teams.