A
Appier
MLOps

Staff/Senior Software Engineer, Machine Learning Platform (Ad Cloud) - Tokyo

On-siteStaffMLOpsposted 4mo ago
Role summaryAI-generated

This role focuses on designing and maintaining a high-performance, scalable machine learning platform for Appier’s agentic AI systems, ensuring seamless integration between model training, deployment, and real-time inference. You’ll architect solutions to handle ad cloud workloads while optimizing for latency, cost-efficiency, and reliability in a globally distributed environment.

Skills required

About this role

About Appier

Appier is an AI-native Agentic AI as a Service (AaaS) company that uses artificial intelligence (AI) to power business decision-making. Founded in 2012 with a vision of democratizing AI, Appier’s mission is turning AI into ROI by making software intelligent. Appier now has 17 offices across APAC, Europe and U.S., and is listed on the Tokyo Stock Exchange (Ticker number: 4180). Visit www.appier.com for more information.

The Impact You’ll Make at Appier

We’re looking for a Staff/Senior Machine Learning Platform Engineer to join our Machine Learning Platform Team, which powers end-to-end infrastructure for model training, evaluation, deployment, and monitoring at scale. Our platform supports daily execution of hundreds of ML models and processes billions of data records across batch and streaming pipelines.
In this role, you’ll shape the architecture and core components of our ML platform—covering batch (Spark), streaming (Flink), job orchestration (Argo on Kubernetes), and infrastructure tools—while ensuring the platform remains robust, scalable, and developer-friendly. You’ll also champion best practices and modern development tools including LLM-based programming assistants.

What You’ll Work On

  • Architect, implement, and scale batch (Spark) and streaming (Flink) pipelines that process billions of records daily for ML training and evaluation.
  • Design and operate robust ML job execution frameworks for training, inference, and post-processing.
  • Build and maintain internal API servers and developer tools to orchestrate ML jobs on Kubernetes (via Argo Workflows, Helm, Terraform).
  • Design and monitor data infrastructure using ClickHouse and PostgreSQL.
  • Ensure high availability and observability through monitoring tools like Prometheus and Grafana.
  • Collaborate with data scientists, product managers, and engineers to deliver reliable and efficient ML platform capabilities.
  • Actively adopt and promote the use of LLM-based tools (e.g., GitHub Copilot, ChatGPT) to accelerate development, documentation, and debugging.
  • Mentor junior engineers and help evolve team engineering culture and standards.

What We’re Looking For

  • Bachelor’s degree in Computer Science, Engineering, or a related field; Master’s preferred.
  • 4+ years of hands-on experience in data systems, machine learning infrastructure, or platform engineering.
  • Strong coding proficiency in Python and/or Java, with experience building large-scale production systems.
  • Practical experience with Spark, Flink, Kubernetes (GKE), and infrastructure-as-code tools such as Terraform and Helm.
  • Experience managing high-throughput data infrastructure using ClickHouse, PostgreSQL, or similar systems.
  • Deep understanding of ML pipelines and distributed job execution in production environments.
  • Proven ability to apply LLM-based tools (claude code, codex) to boost engineering productivity.
  • Strong ownership, architectural thinking, and ability to lead cross-functional platform projects.

#LI-AK1

Apply on AppierOpens in new tab

Similar open roles

M
NEW

Senior Engineer, ML Data Services

Motional·Boston, Massachusetts, United States; Las Vegas, Nevada, United States; Pittsburgh, Pennsylvania, United States; Remote U.S.
HybridSeniorMLOps
$149k – $199k USD
2d ago
Z
NEW

ML Infrastructure Engineer

Zipline·South San Francisco, California, USA
On-siteMidMLOps
$160k – $250k USD
2d ago
C
NEW

Software Engineer, Inference AI/ML

CoreWeave· Sunnyvale, CA / Bellevue, WA
HybridMidMLOps
$92k – $135k USD
2d ago
SK
Sponsored

Land 5x More Interviews - Resume & Strategy

Shaqeeq Khan·Built this board, coached engineers from Netflix, Google, IBM, Amazon. 100+ grads placed.
CV ReviewLive One on One CallsStrategy
Book A Call