Infrastructure & Observability Engineer
Lugano - Switzerland
Job Summary
FORFIRM is providing solutions to real business challenges for our clients through innovation and deep industry understanding. We pride ourselves on being a knowledge-based company with no barriers or pre-built solutions we listen to our clients and solve their unique problems.
At FORFIRM we are creating a culture where each person can define their own role parameters and speak their mind without any hesitation. We are a true meritocracy where individual results define each persons career path.
We are looking for an experienced and motivated Infrastructure & Observability Engineer to join our lively international team and work on projects for Europes leading brands.
- Manage maintain and evolve enterprise Linux infrastructures based on Red Hat Enterprise Linux and Debian platforms.
- Support the administration and optimization of containerized environments based on Docker and Kubernetes.
- Ensure system reliability availability and performance through proactive infrastructure management and continuous improvement initiatives.
- Contribute to patch management system hardening and infrastructure lifecycle activities.
- Design implement and manage enterprise monitoring and observability solutions.
- Develop dashboards alerts and reports to provide actionable insights into system performance availability and operational health.
- Manage and optimize monitoring platforms including Splunk Icinga Grafana Prometheus Dynatrace and related technologies.
- Implement automated monitoring solutions for infrastructure applications and business-critical services.
- Support the evolution of observability frameworks through the adoption of industry best practices.
- Employ DevOps and Infrastructure as Code best practices to automate deployment configuration monitoring and operational processes.
- Design and maintain automation workflows using Ansible Terraform and Python.
- Integrate monitoring and observability solutions within CI/CD pipelines and operational processes.
- Promote operational excellence through automation and continuous improvement initiatives.
- Develop and maintain scripts to automate operational tasks and system integrations.
- Integrate monitoring CMDB and ITSM platforms through APIs and automated workflows.
- Support data consistency and process standardization across multiple infrastructure platforms.
- Support the administration and maintenance of configuration management databases (CMDB) preferably Device42.
- Ensure accurate asset inventory dependency mapping and infrastructure documentation.
- Contribute to configuration governance and operational best practices.
- Support the implementation and management of solutions for the centralized administration of secrets credentials keys and digital certificates.
- Collaborate with security teams to ensure infrastructure monitoring complies with internal standards and regulatory requirements.
- Support auditing reporting and compliance-related initiatives.
- Act as a subject matter expert for infrastructure monitoring and observability technologies.
- Produce and maintain technical documentation procedures and operational guidelines.
- Collaborate with infrastructure development security and operations teams to ensure service excellence and operational resilience.
- Proven experience in infrastructure administration monitoring and observability in enterprise environments.
- Solid background in DevOps methodologies automation and operational excellence practices.
- Hands-on experience with infrastructure monitoring logging and alerting platforms.
- Experience working in complex and highly available environments.
- Strong knowledge of Linux system administration particularly Red Hat Enterprise Linux and Debian.
- Experience managing Docker and Kubernetes platforms.
- Strong expertise with Splunk for log management monitoring and dashboard development.
- Experience with Icinga Zabbix Dynatrace Prometheus Grafana and similar observability tools.
- Familiarity with monitoring concepts including metrics logs traces alerting and performance analysis.
- Experience with Ansible and Terraform for infrastructure automation and patch management.
- Knowledge of CMDB platforms preferably Device42.
- Experience with centralized solutions for managing credentials secrets and certificates.
- Proficiency in Python scripting and automation.
- Good knowledge of Git and version control practices.
- Experience designing and implementing automated web monitoring and application health checks.
- Strong problem-solving skills with a proactive and results-oriented mindset.
- Excellent communication collaboration and documentation abilities.
- Ability to work effectively in a dynamic and fast-paced environment.
- Strong analytical thinking and attention to detail.
- Continuous improvement mindset and passion for automation.
- Experience in financial services banking or other highly regulated industries.
- Previous experience in Site Reliability Engineering (SRE) Platform Engineering or Infrastructure Engineering roles.
- Experience defining enterprise monitoring and observability frameworks.
- Knowledge of ITIL processes and operational governance frameworks.
- Relevant certifications in Linux Kubernetes DevOps Monitoring or Observability technologies.
- Opportunity to work on international multidisciplinary projects for leading brands from day one.
- Highly meritocratic environment with clear growth opportunities.
- Continuous learning through internal and external training programs.
- Supportive dynamic and collaborative team environment.
- Exposure to modern infrastructure cloud-native and observability technologies.
FORFIRM is an equal opportunities employer that values diversity within the company. Qualified applicants will receive consideration for employment without discrimination based on race religion colour national origin gender sexual orientation age marital status veteran status or disability status.