Senior Cloud DevOps Engineer
Department:
Job Summary
About Kinaxis
Are you looking to join an innovative market-leading company where you can truly elevate your career At Kinaxis we are serious about culture we are serious about technology we are serious about customers and we are serious about not taking ourselves too seriously. If you are looking to be part of an incredible growth story then we might just be the place for you!
In 1984 we started out as a team of three engineers. Today we have grown to become a global organization with over 2000 employees around the world 6 global office and a best-in-class HQ in Ottawa Canada. As winners of several Top Employer awards globally we are proud to work with our customers and employees towards solving some of the biggest challenges facing supply chains today.
Kinaxis is a global leader in modern supply chain orchestration powering complex global supply chains and supporting the people who manage them. Our powerful AI infused platform provides full transparency and visibility across end-to-end supply chains enabling our customers to make faster better decisions. We are trusted by renowned global brands to provide the agility and predictability needed to navigate todays volatility and disruption. With more than 40000 users in over 100 countries we are expanding our team as we continue to innovate and revolutionize how we support our customers.
Location
Chennai India
About the team
The Senior Cloud Engineer as a fully proficient seasoned professional contributes to the software development lifecycle process for the cloud platform. This includes design implementation testing and troubleshooting features of advanced complexity that span multiple teams. The incumbent will support solutions delivered in production environments as required.
The successful candidate actively mentors junior engineers serves as a technical advisor to managers and team leads and acts as a Subject Matter Expert (SME) in one or more domains. The role requires close collaboration with internal engineering and product teams to develop maintain and enhance cloud platform solutions.
What you will do
- Lead collaborate and coordinate across multiple teams as a technical primary on assigned features to clarify problems and proactively develop long-term solutions. Ensure requirements are met and implementation execution and testing are delivered with excellence.
- Design build and operate secure scalable cloud infrastructure using Infrastructure as Code and automation frameworks.
- Contribute to and support the production SaaS ecosystem ensuring solutions are designed and implemented with operability security performance scalability and availability in mind.
- Follow the software development lifecycle to translate business and technical requirements into infrastructure as code.
- Build and enhance CI/CD platforms deployment pipelines and self-service capabilities to improve developer productivity platform reliability and operational efficiency.
- Implement code of advanced complexity that spans multiple areas of functionality with performance and testability in mind. Apply software engineering principles to operational challenges with a focus on automation security self-healing and monitoring solutions.
- Troubleshoot issues within the cloud platform identify defects provide explanations for platform behavior across multiple components and recommend solutions and workarounds for engineering teams. Resolve and test moderate to advanced defects.
- Contribute to the development and maintenance of dashboards reports and metrics to track platform health cloud usage and costs. Drive process improvement initiatives focused on reliability and cost optimization.
- Operate in a security and regulated environment by applying best practices for IAM role-based access controls network security vulnerability management environment isolation and compliance requirements.
Technologies we use
- GCP AzureAWSTerraformAnsibleGitHub
- Actions DockerPackerGKEAKSKubernetesHelmHashiCorp
- Vault PythonGoShell/BashJFrog ArtifactoryDatadogPrometheusELKWindows
- Server Argo CDCapsuleWorkload Identity FederationAuth0Entra
- ID OktaPatchmanRundeckAptlyJIRAServiceNow
What we are looking for
- Bachelors degree/diploma in Computer Science Engineering or an equivalent related discipline.
- 7 years of hands-on experience in cloud engineering platform engineering infrastructure automation or a related field.
- Strong experience designing building and operating cloud infrastructure using Infrastructure as Code.
- Experience working with public cloud platforms such as GCP AWS or Azure with GCP preferred.
- Proficiency with Terraform Kubernetes Docker GitHub Actions and automation tooling.
- Experience developing automation solutions and tooling using Python Bash or similar scripting languages.
- Solid understanding of identity and access management networking security and cloud governance best practices.
- Experience troubleshooting complex platform and infrastructure issues in production environments.
- Excellent communication and collaboration skills with the ability to work effectively across engineering product and security teams.
- Experience mentoring engineers and contributing to technical knowledge sharing within a team environment.
- Familiarity with Agile methodologies and Scrum practices.
- Working knowledge of Windows Server administration automation and lifecycle management is considered an asset.
- Kinaxis product knowledge is considered an asset.
Nice to Have
- GCP Landing Zone or Azure enterprise-scale architecture experience
- NetApp CVS / cloud storage volume configuration
- Capsule multi-tenancy for Kubernetes namespace management
- Image security scanning Trivy SBOM cosign signing
- RapidResponse or supply chain platform exposure
- Windows Server enterprise environment experience
- Aptly Debian package repository management
- Shared services administration WSUS EFT Syncovery Postfix
Required Skills
Cloud Infrastructure:
- 6 years hands-on experience with public cloud platforms GCP Azure or AWS console and API (GCP preferred)
- Strong Terraform modules layered architecture state management remote backends
- Ansible roles playbooks Windows (WinRM) and Linux targets
- GKE / AKS cluster operations node pools Workload Identity Federation namespace management
- Experience automating cloud solutions applications infrastructure storage platforms and data
Networking & Security:
- GCP VPC networking hub-and-spoke shared VPC NCC NAT DNS VPN
- Azure networking VNet peering private endpoints NSGs
- IAM design service accounts role bindings least-privilege patterns
- Identity providers Auth0 Entra ID Okta
- HashiCorp Vault secrets management dynamic credentials
- Strong knowledge of system design to manage operational and reliability trade-offs
CI/CD & Image Automation:
- GitHub Actions reusable workflows self-hosted ARC runners on GKE release automation
- Docker Kubernetes Helm container orchestration and application delivery
- Packer image builds provisioners multi-region publishing
- Argo CD / GitOps application delivery and environment onboarding
- Artifact management JFrog Artifactory or equivalent registries
- Windows Server automation prerequisites patching hardening via Ansible
Identity & Platform Tooling:
- Workload Identity Federation service account provisioning GCP and Azure IAM bindings
- GKE namespace management Capsule multi-tenancy tenant isolation
- YAML-based environment configuration deployment manifests globals environment-specific overrides
- Development in Go Python and Shell/Bash CLI tooling and platform automation
- Practical understanding of Agile methodologies and Scrum
Leadership & Collaboration:
- Experience leading or mentoring a team of platform or infrastructure engineers
- Ability to define technical direction break down epics into stories and drive delivery across a squad
- Comfortable working across teams security product customer success to align on platform requirements
- Strong written and verbal communication ADRs runbooks and technical documentation for multiple audiences
- Advanced analytical problem-solving and critical thinking skills
- Agile and adaptable excels at managing and prioritising workloads in an environment of ongoing urgency and change
Operations:
- Linux patching automation Patchman Rundeck Aptly package repository management
- Monitoring and alerting Datadog Prometheus ELK Log Analytics GCP Cloud Monitoring
- Incident response runbooks escalation post-mortems
#Intermediate #LI-RJ1 #Fulltime
Work With Impact: Our platform directly helps companies power the worlds supply chains. We see the results of what we do out in the world every day when we see store shelves stocked when medications are available for our loved ones and so much more.
Work with Fortune 500 Brands: Companies across industries trust us to help them take control of their integrated business planning and digital supply chain. Some of our customers include Lockheed Martin Unilever P&G ExxonMobil Cisco and more.
Social Responsibility at Kinaxis: Our Diversity Equity and Inclusion Committee weighs in on hiring practices talent assessment training materials and mandatory training on unconscious bias and inclusion fundamentals. Sustainability is key to what we do and were committed to a long-term net-zero operations strategy. We are involved in our communities and support causes where we can make the most impact.
People matter at Kinaxis and here are some of the perks and benefits we offer which may vary by location and employee:
- Flexible vacation and Kinaxis Days (company-wide days off)
- Flexible work options
- Physical and mental well-being programs
- Regularly scheduled virtual fitness classes
- Mentorship programs training and career development
- Recognition programs and referral rewards
- Hackathons
For more information visit the Kinaxis website at or the companys blog at .
Kinaxis welcomes candidates to apply to our inclusive community. We provide accommodations upon request to ensure fairness and accessibility throughout our recruitment process for all candidates including those with specific needs or disabilities. If you require an accommodation please reach out to us at This contact information is for accessibility requests only and cannot be used to inquire about the status of applications.
Kinaxis is committed to ensuring a fair and transparent recruitment process. We use artificial intelligence (AI) tools in the initial step of the recruitment process to compare submitted resumes against the job description to identify candidates whose education experience and skills most closely match the requirements of the role. After the initial screening all subsequent decisions regarding your application including final selection are made by our human recruitment team. AI does not make any final hiring decisions.
Required Experience:
Senior IC
About Company
Revolutionize supply chain management with Kinaxis. Get end-to-end transparency to make fast, collaborative decisions with the power of concurrency.