Die Stelle
03Your tasks:
- Build and operate the ML infrastructure and platforms powering A1-s AI products
- Design systems for model training, evaluation, deployment, inference, and experimentation
- Build and optimise model serving and inference infrastructure for high-throughput and low-latency workloads
- Improve reliability, scalability, latency, and cost efficiency of AI systems
- Develop reliable pipelines for data preparation, training, evaluation, model release, and continuous improvement
- Build platforms and tooling that enable AI engineers and researchers to experiment, evaluate, and ship models faster
- Develop evaluation and benchmarking infrastructure to measure model quality, performance, and regressions
- Build production observability, monitoring, tracing, and alerting for AI/ML workloads
- Improve AI systems across reliability, scalability, latency, throughput, and cost
- Identify bottlenecks across the ML stack and continuously improve system performance
- Work closely with AI engineers, researchers, and product teams to turn evolving model requirements into production-ready infrastructure
01Build and operate the ML infrastructure and platforms powering A1-s AI products
02Design systems for model training, evaluation, deployment, inference, and experimentation
03Build and optimise model serving and inference infrastructure for high-throughput and low-latency workloads
04Improve reliability, scalability, latency, and cost efficiency of AI systems
05Develop reliable pipelines for data preparation, training, evaluation, model release, and continuous improvement
06Build platforms and tooling that enable AI engineers and researchers to experiment, evaluate, and ship models faster
07Develop evaluation and benchmarking infrastructure to measure model quality, performance, and regressions
08Build production observability, monitoring, tracing, and alerting for AI/ML workloads
09Improve AI systems across reliability, scalability, latency, throughput, and cost
10Identify bottlenecks across the ML stack and continuously improve system performance
11Work closely with AI engineers, researchers, and product teams to turn evolving model requirements into production-ready infrastructure