Machine Learning Platform Engineer

Brak informacji o wynagrodzeniu
SeniorFull-time
#411414·Dodano 9 dni temu·0
Źródło: Bjak
Aplikuj teraz

Tech Stack / Keywords

Machine LearningAIPythonPyTorchLLMCloudDatabases

Firma i stanowisko

There are over 5 billion users using basic applications today such email, notes, tasks that are not AI-native. Our mission is to build a proactive smart assistant for everyday users to bring intelligence to conversations, errands, organising and workflows, with minimal prompting.

Our product focuses on achieving high reliability for long-running workflows, persistent context, and real-world task completion. The system must handle multi-step reasoning, interact with external tools, and remain reliable despite non-deterministic model behavior. Our objective is to help users complete tasks daily enjoyable with over ~90%* reduced time.

Wymagania

  • Strong software engineering fundamentals and experience building production systems
  • Experience building ML infrastructure, platforms, or production machine learning systems
  • Experience with model deployment, inference, evaluation, or data pipelines
  • Strong understanding of distributed systems and system reliability
  • Ability to write clean, maintainable, production-quality code
  • Comfortable working in ambiguous, fast-moving environments
  • Bias toward ownership, experimentation, and continuous improvement

Obowiązki

  • Build and operate the ML infrastructure and platforms powering A1’s AI products
  • Design systems for model training, evaluation, deployment, inference, and experimentation
  • Build and optimise model serving and inference infrastructure for high-throughput and low-latency workloads
  • Improve reliability, scalability, latency, and cost efficiency of AI systems
  • Develop reliable pipelines for data preparation, training, evaluation, model release, and continuous improvement
  • Build platforms and tooling that enable AI engineers and researchers to experiment, evaluate, and ship models faster
  • Develop evaluation and benchmarking infrastructure to measure model quality, performance, and regressions
  • Build production observability, monitoring, tracing, and alerting for AI/ML workloads
  • Improve AI systems across reliability, scalability, latency, throughput, and cost
  • Identify bottlenecks across the ML stack and continuously improve system performance
  • Work closely with AI engineers, researchers, and product teams to turn evolving model requirements into production-ready infrastructure
Bjak

Bjak

38 aktywnych ofert

Zobacz wszystkie oferty
Aplikuj teraz