Site Reliability Engineer
Job Summary
Were on the search for a Site Reliability Engineer (SRE) ready to join our SRE team in monitoring automating and alerting our systems. Our Development team is based in Malaga but if you code your best from the comfort of your own home or even from a different citythe location doesnt matter you do
What is Plytix
This is where youd typically see a long paragraph of boring text that no one reads peppered with corporate buzzwords that dont mean anything. But Plytix is no typical company so well spare you from that. Instead watch this video and if you like what you see keep scrolling.
We are looking for a skilled and experienced Site Reliability Engineer to join our team and to help us scale and keep our applications reliable under certain parameters of quality (low latencies low error rate and so on).
Our current technical stack is simple so the ideal candidate should have a strong background in:
Grafana/Prometheus: Monitoring.
Elastic: Our logging backend.
AWS: Where our infrastructure runs.
K8s: We run our applications in Kubernetes.
A good foundation in the following technologies will also be valued:
Kong: The gateway.
Python: All our backend is developed in Python.
MongoDB / Postgres: We save our data in these databases.
Redis: Cache and temporary result backend.
RabbitMQ: Our message broker.
What will you be doing
Maintain reliable and scalable infrastructure on AWS using Kubernetes and other tools.
Monitor and troubleshoot production systems and respond to incidents in a timely manner.
Design and implement automated deployment and testing pipelines to ensure quality and reliability.
Perform capacity planning and scaling of infrastructure to support growing demand.
Collaborate with development teams to ensure that applications are designed for reliability and scalability.
Develop and maintain monitoring and alerting systems to detect issues before they become critical.
Continuously improve the reliability and performance of our infrastructure and applications.
Be part of releases. Youll need to collaborate to determine the suitability of the deployments.
After 1 month:
Youll have a good grasp of our favorite tools and processes and your code will start flowing. Youll start learning our current architecture and services performing small tasks so that you can start to show off your engineering skills. Youll take part in any engineering discussions and production procedures.
After 3 months:
You should know almost everything about our infrastructure and have an idea on what things should be improved. At this point youll be ready to propose your ideas to improve our system:
- Observability: Improve our logging system and make it easier for other teams to see the status of the platform.
- Automation: Propose some automations for some processes.
- High-availability: Prepare and test our system to make it more efficient and reliable.
- Alerting: Improve our alerting system.
- Disaster recovery: Have procedures for disaster recovery.
6 months in:
Youre fully integrated into the team at this point. You know all our strengths and weaknesses and should be able to prepare a roadmap to improve all our systems. Youre also collaborating with other departments to find the best possible solutions for our customers to have the best quality system. Youre delivering high-quality fast-running solutions thatll be tested to iron out any potential errors. By now youre confident in your role performing tasks in development as well as production.
Who will you be working with
Youll be working with the Infrastructure team making sure that everything is running smoothly but they arent the only shining face youll be working with on a daily basis. Youll also be working with the QA Development and DevOps teams to keep customers smiling about how amazing their Plytix PIM is.
If youre curious about what its like working at our office take a peek into what Plytix has to offer (you know you want to ).
We expect you to:
8 years of experience in IT and 3 years of experience in a Site Reliability Engineering or DevOps role.
Strong experience with AWS Kubernetes and logging/monitoring systems.
Experience with automation and configuration management tools (e.g. Ansible Terraform).
Strong understanding of networking and security principles.
Excellent troubleshooting and problem-solving skills.
Ability to work independently and as part of a team in a fast-paced environment.
Strong communication and collaboration skills.
Nice to have:
Experience with other cloud providers (e.g. GCP Azure).
Experience with other container orchestration systems (e.g. Docker Swarm Mesos).
Experience with other messaging systems (e.g. Kafka).
Experience with other programming languages (e.g. Go Java).
Why work at Plytix
- Voted one of the top best workplaces by Great Place to Work
- Be a part of a welcoming culture with bright friendly people from around the world
- Have a purpose and plenty of opportunities to grow from day one
- Work from home or from our offices in the center of Malaga
- Dont work on your birthdayenjoy an extra day of paid time off
- Holidays: 23 work days 24th & 31st December off
- Enjoy free catered lunches and unlimited ice cream when working from the office in Malaga
- Competitive full-time salary
- Private health insurance covered by Plytix
- Gym discount
- Factorial Flexible Retribution (Restaurant Transport & Nursery Vouchers)
- And more!
About our culture:
If you love reading empowering phrases like vibrant collective and innovative individuals then be prepared to be disappointed
Thats because to get a typical result you have to do only typical things. But thats not really our style.
Were a community of chill and fun-loving humans who believe that work doesnt have to be boring. We make each day exciting by collaborating together and enjoying the occasional bout of silliness: like a pool competition potluck party or game of padel.
Were kind of the cool kids in the PIM playground (if we do say so ourselves) and we love to challenge the status quo to help anyone break free from f*#king boring spreadsheets.
PLYTIX is a company committed to diversity and equality. Therefore all our selection processes guarantee equal opportunities for all candidates regardless of race color religion sex origin nationality age disability sexual orientation and gender identity. All employment decisions are based on business needs job requirements merit and individual qualifications.
We are committed to managing your personal information with care and in accordance with our privacy policy. To learn more about how we handle your data please click HERE