We are looking for a Site Reliability Engineer (Monetization) who is pragmatic product-minded and equally comfortable writing code and running production systems to join our Infrastructure departments SRE team. The best candidate will be someone who thrives in a fast-paced highly collaborative and exceptionally dynamic setting and is excited to own the application-level infrastructure and reliability of a high-traffic commerce domain end to end - from deploy pipelines and Kubernetes manifests to SLOs capacity planning and production readiness.
Strong Kubernetes observability and software engineering skills are essential along with experience in operating production services in a cloud environment (GCP/GKE or comparable) and partnering closely with product development teams. The ability to hold a dual perspective - understanding both how developers ship features and what infrastructure needs to stay reliable - and to bring the reliability lens into design decisions early will be key to your success in this role.
This is a hybrid embedded role: you remain part of the SRE organization (practices standards duty rotation) while being functionally embedded into the Monetization product domain. Youll build long-term working relationships with the domains engineering teams own a meaningful share of their application infrastructure execution and co-author the reliability practices used company-wide.
If youre passionate about making complex distributed systems boringly reliable and love building the commerce and monetization backbone that lets game developers around the world get paid we would love to hear from you!
ABOUT US
Xsolla is a global commerce company with robust tools and services to help developers solve the inherent challenges of the video game industry. From indie to AAA companies partner with Xsolla to help them fund distribute market and monetize their games. Grounded in the belief in the future of video games Xsolla is resolute in the mission to bring opportunities together and continually make new resources available to creators. Headquartered and incorporated in Los Angeles California Xsolla operates as the merchant of record and has helped over 1500 game developers to reach more players and grow their businesses around the world. With more paths to profits and ways to win developers have all the things needed to enjoy the game.
Own the application-level infrastructure of the Monetization domain: Helm charts Terraform configurations Kubernetes deployments runtime configuration and service-level networking and integrations
Own the domains observability: design and implement SLOs/SLIs monitors alerts and dashboards for critical services on Datadog and OpenTelemetry-based tooling
Help to set up and evolve CI/CD pipelines for domain services (GitLab CI GitHub Actions) including deploy and rollback automation
Perform capacity planning and performance tuning ahead of expected load - product launches sales events and regional rollouts - including load testing and performance regression investigation
Run Production Readiness Reviews for new services and major changes; define and enforce what production-ready means for the domain
Support domain incident response: assist with deep investigation of complex incidents contribute to post-mortems drive follow-up reliability improvements and maintain runbooks
Maintain and drive a forward-looking reliability roadmap for the domain together with product engineering leads
Participate in product team planning refinements and architecture reviews bringing the reliability perspective before design decisions become expensive to change
Co-author company-wide SLO/SLI capacity and operational standards together with the broader SRE team; contribute improvements directly to shared SRE-operated subsystems
Participate in the SRE duty rotation supporting developers across the company
Qualifications & Skills
3 years of proven SRE DevOps or platform engineering experience: on-call or incident response duty SLO/monitoring ownership deploy pipeline and infrastructure work for production services
Software development background: you have built and shipped backend services not only operated them - comfortable reading application code during an investigation and writing production-quality automation in at least one language (e.g. Go PHP)
Hands-on Kubernetes experience:Helm manifests deploy strategies debugging application-level performance and networking issues (GKE or another managed Kubernetes)
Solid observability practice: building monitors dashboards and SLOs/SLIs on a modern platform (Datadog preferred; Prometheus/Grafana experience also relevant) familiarity with OpenTelemetry
Infrastructure as Code exposure (Terraform/Terragrunt) for collaboration with platform teams
GCP experience (IAM networking managed services)
Experience building and maintaining CI/CD pipelines (GitLab CI and/or GitHub Actions)
Programming/scripting proficiency sufficient to build automation and tooling (e.g. Python Go or Bash)
Practical experience with incident response post-mortems and driving reliability improvements from incidents
Strong collaboration and communication skills this role works embedded with product development teams daily
Experience in payments fintech e-commerce or gaming high-traffic transactional systems
Nice to Have:
Kubernetes certifications
Google Cloud Platform certifications
HashiCorp certifications
$120000 - $160000 a year
Salary varies depending on experience level and location.
Benefits
We are passionate about fostering a supportive environment for our team so we prioritize the physical mental and emotional well-being of our employees and their families through a comprehensive Benefits Program. This includes medical dental and vision PTO and a personalized career roadmap for each employee. By investing in professional development through training and educational opportunities we ensure that our team thrives both personally and professionally. Together were not just building a business; were cultivating a community that values creativity collaboration and the transformative power of play.
Equal Employment Opportunity Statement
Xsolla is an equal opportunity employer. We celebrate diversity and are committed to creating an inclusive environment for all employees. We do not discriminate based on race color religion sex national origin age disability sexual orientation gender identity or any other characteristic protected by law. We consider qualified applicants with criminal histories in accordance with the Fair Chance Act.
Criminal History Consideration
For the Site Reliability Engineer (Monetization) position we will conduct a background check that may include the following:
Criminal history check
Employment verification
Education verification
Relevance to Job Responsibilities
The background check is relevant to this position because of the following role responsibilities:
Accessing confidential company data
Handling infrastructure that processes sensitive financial transactions
Ensuring compliance with regulatory requirements
Rights Under the Fair Chance Act
Applicants are encouraged to inquire about their rights under the Fair Chance Act. If you have questions regarding our hiring practices please contact emailprotected.
We may use artificial intelligence (AI) tools to support parts of the hiring process such as reviewing applications analyzing resumes or assessing responses and identifying potential inconsistencies or verification signals in application materials based on available information. These tools assist our recruitment team but do not replace human judgment. Final hiring decisions are ultimately made by humans. If you would like more information about how your data is processed please contact us.
Required Experience:
IC
ABOUT YOUWe are looking for a Site Reliability Engineer (Monetization) who is pragmatic product-minded and equally comfortable writing code and running production systems to join our Infrastructure departments SRE team. The best candidate will be someone who thrives in a fast-paced highly collaborat...
ABOUT YOU
We are looking for a Site Reliability Engineer (Monetization) who is pragmatic product-minded and equally comfortable writing code and running production systems to join our Infrastructure departments SRE team. The best candidate will be someone who thrives in a fast-paced highly collaborative and exceptionally dynamic setting and is excited to own the application-level infrastructure and reliability of a high-traffic commerce domain end to end - from deploy pipelines and Kubernetes manifests to SLOs capacity planning and production readiness.
Strong Kubernetes observability and software engineering skills are essential along with experience in operating production services in a cloud environment (GCP/GKE or comparable) and partnering closely with product development teams. The ability to hold a dual perspective - understanding both how developers ship features and what infrastructure needs to stay reliable - and to bring the reliability lens into design decisions early will be key to your success in this role.
This is a hybrid embedded role: you remain part of the SRE organization (practices standards duty rotation) while being functionally embedded into the Monetization product domain. Youll build long-term working relationships with the domains engineering teams own a meaningful share of their application infrastructure execution and co-author the reliability practices used company-wide.
If youre passionate about making complex distributed systems boringly reliable and love building the commerce and monetization backbone that lets game developers around the world get paid we would love to hear from you!
ABOUT US
Xsolla is a global commerce company with robust tools and services to help developers solve the inherent challenges of the video game industry. From indie to AAA companies partner with Xsolla to help them fund distribute market and monetize their games. Grounded in the belief in the future of video games Xsolla is resolute in the mission to bring opportunities together and continually make new resources available to creators. Headquartered and incorporated in Los Angeles California Xsolla operates as the merchant of record and has helped over 1500 game developers to reach more players and grow their businesses around the world. With more paths to profits and ways to win developers have all the things needed to enjoy the game.
Own the application-level infrastructure of the Monetization domain: Helm charts Terraform configurations Kubernetes deployments runtime configuration and service-level networking and integrations
Own the domains observability: design and implement SLOs/SLIs monitors alerts and dashboards for critical services on Datadog and OpenTelemetry-based tooling
Help to set up and evolve CI/CD pipelines for domain services (GitLab CI GitHub Actions) including deploy and rollback automation
Perform capacity planning and performance tuning ahead of expected load - product launches sales events and regional rollouts - including load testing and performance regression investigation
Run Production Readiness Reviews for new services and major changes; define and enforce what production-ready means for the domain
Support domain incident response: assist with deep investigation of complex incidents contribute to post-mortems drive follow-up reliability improvements and maintain runbooks
Maintain and drive a forward-looking reliability roadmap for the domain together with product engineering leads
Participate in product team planning refinements and architecture reviews bringing the reliability perspective before design decisions become expensive to change
Co-author company-wide SLO/SLI capacity and operational standards together with the broader SRE team; contribute improvements directly to shared SRE-operated subsystems
Participate in the SRE duty rotation supporting developers across the company
Qualifications & Skills
3 years of proven SRE DevOps or platform engineering experience: on-call or incident response duty SLO/monitoring ownership deploy pipeline and infrastructure work for production services
Software development background: you have built and shipped backend services not only operated them - comfortable reading application code during an investigation and writing production-quality automation in at least one language (e.g. Go PHP)
Hands-on Kubernetes experience:Helm manifests deploy strategies debugging application-level performance and networking issues (GKE or another managed Kubernetes)
Solid observability practice: building monitors dashboards and SLOs/SLIs on a modern platform (Datadog preferred; Prometheus/Grafana experience also relevant) familiarity with OpenTelemetry
Infrastructure as Code exposure (Terraform/Terragrunt) for collaboration with platform teams
GCP experience (IAM networking managed services)
Experience building and maintaining CI/CD pipelines (GitLab CI and/or GitHub Actions)
Programming/scripting proficiency sufficient to build automation and tooling (e.g. Python Go or Bash)
Practical experience with incident response post-mortems and driving reliability improvements from incidents
Strong collaboration and communication skills this role works embedded with product development teams daily
Experience in payments fintech e-commerce or gaming high-traffic transactional systems
Nice to Have:
Kubernetes certifications
Google Cloud Platform certifications
HashiCorp certifications
$120000 - $160000 a year
Salary varies depending on experience level and location.
Benefits
We are passionate about fostering a supportive environment for our team so we prioritize the physical mental and emotional well-being of our employees and their families through a comprehensive Benefits Program. This includes medical dental and vision PTO and a personalized career roadmap for each employee. By investing in professional development through training and educational opportunities we ensure that our team thrives both personally and professionally. Together were not just building a business; were cultivating a community that values creativity collaboration and the transformative power of play.
Equal Employment Opportunity Statement
Xsolla is an equal opportunity employer. We celebrate diversity and are committed to creating an inclusive environment for all employees. We do not discriminate based on race color religion sex national origin age disability sexual orientation gender identity or any other characteristic protected by law. We consider qualified applicants with criminal histories in accordance with the Fair Chance Act.
Criminal History Consideration
For the Site Reliability Engineer (Monetization) position we will conduct a background check that may include the following:
Criminal history check
Employment verification
Education verification
Relevance to Job Responsibilities
The background check is relevant to this position because of the following role responsibilities:
Accessing confidential company data
Handling infrastructure that processes sensitive financial transactions
Ensuring compliance with regulatory requirements
Rights Under the Fair Chance Act
Applicants are encouraged to inquire about their rights under the Fair Chance Act. If you have questions regarding our hiring practices please contact emailprotected.
We may use artificial intelligence (AI) tools to support parts of the hiring process such as reviewing applications analyzing resumes or assessing responses and identifying potential inconsistencies or verification signals in application materials based on available information. These tools assist our recruitment team but do not replace human judgment. Final hiring decisions are ultimately made by humans. If you would like more information about how your data is processed please contact us.
Find out how you can launch, monetize and scale your video games worldwide, with no upfront costs, using Xsolla's comprehensive suite of tools and services.