As Site Reliability Engineer, you will be responsible for the reliability, scalability, observability, and operational excellence of Vinted’s ML Platform services including SLOs/alerting, incident response, capacity planning, automation, and on-call for critical services such as the Vinted Feature Store. You’ll join the ML Platform team, which owns the tooling for ML/LLM development, deployment and other platform capabilities that increase ML/AI delivery speed at Vinted. You will be working closely with ML platform users, Data Infrastructure and Production Engineering teams.
Here are some of the technologies we use: Kubernetes, Terraform, Chef, Google Cloud Platform, Go, Kafka, Vespa, Redis, Vitess.
Nice to have
Nuoroda į skelbimą bus pridėta automatiškai žinutės pabaigoje.