[8BE] Data Scientist (AI + ML)
Buenos Aires - Argentina
Job Summary
We are seeking a Data Scientist with deep expertise in probabilistic AI and statistical machine learning to support a clients e-commerce platform built on a distributed microservices architecture. The platform roadmap includes a set of intelligence capabilities that require rigorous statistical modeling rather than standard supervised ML. This role is responsible for designing validating and productionizing probabilistic models and for working closely with backend engineering to translate those models into service-oriented production architecture within the platforms existing microservices ecosystem.
What you will do
- Bayesian modeling and inference: Design and implement Bayesian statistical models priors likelihoods and posterior inference to support decisioning under uncertainty across pricing segmentation and demand-related use cases.
- Markov chains and Hidden Markov Models: Build Markov chain and Hidden Markov Model formulations for sequential and behavioral patterns (e.g. customer lifecycle stages state transitions) producing outputs that downstream services can consume.
- MCMC and Metropolis-Hastings sampling: Apply Markov Chain Monte Carlo methods including Metropolis-Hastings sampling to estimate posterior distributions for models without closed-form solutions and validate convergence and sampling quality.
- Mixture modeling: Develop mixture models Gaussian Mixture Models in particular to support segmentation use cases identifying latent customer or product groupings from transactional and behavioral data.
- Expectation-Maximization: Implement Expectation-Maximization for latent-variable estimation underlying mixture models and related unsupervised learning tasks.
- Production translation: Work with backend engineering to translate statistical models into production service architecture defining APIs data contracts and integration points within the platforms existing microservices and event-driven pipelines.
- Model lifecycle management: Define the approach for model training validation versioning monitoring/drift detection and retraining cadence once models are in production.
- Roadmap collaboration: Partner with delivery and engineering leads to size sequence and estimate probabilistic/statistical modeling initiatives on the product roadmap.
- Documentation and handoff: Document modeling assumptions methodology and validation results and provide clear hand-off guidance so models remain maintainable by the engineering team after the engagement.
Qualifications :
- 90% English written and oral (at least B2 level) with excellent communication skills
- Strong demonstrable background in Bayesian statistics/Bayesian inference Markov chains Hidden Markov Models MCMC methods (including Metropolis-Hastings sampling) mixture models (ideally Gaussian Mixture Models) and Expectation-Maximization.
- Proven experience building and deploying statistical/ML models into production systems not just research notebooks or offline analysis.
- Proficiency in Python (or R) with standard probabilistic/statistical libraries (e.g. PyMC Stan scikit-learn NumPy/SciPy) for model development and validation.
- Ability to translate statistical/mathematical models into service-oriented production architecture defining APIs and data contracts and working directly with backend engineers to integrate them.
- Solid understanding of version control testing practices and CI/CD sufficient to collaborate effectively with an engineering team on production delivery.
- Strong written and verbal communication skills with the ability to explain model behavior assumptions and uncertainty to non-technical stakeholders.
Additional Information :
Preferred Qualifications
- Experience in e-commerce or retail domains particularly pricing optimization customer segmentation or demand forecasting.
- Experience integrating ML models with microservices architectures (REST/GraphQL) and event-driven systems (e.g. message queues/pub-sub) and deploying to cloud infrastructure.
- Familiarity with common backend service ecosystems (e.g. .NET Java or ) even if modeling itself is done in Python for a smoother handoff to the production engineering team.
- Experience with MLOps tooling such as model registries monitoring and feature stores.
- Background in pricing science recommendation systems or marketing analytics.
Our Benefits
- Educational resources
- Flexible schedule and Work From Anywhere
- Referral Program
- Supportive and chill atmosphere
- Trajectory recognition plan
We are accepting applications from LATAM countries
Remote Work :
Yes
Employment Type :
Full-time
About Company
Software Mind develops solutions that make an impact for companies around the globe. Tech giants & unicorns, transformative projects, emerging technologies and limitless opportunities these are a few words that describe an average day for us. Building cross-functional engineering te ... View more