Staff AI Engineer (AI Agents Development), CDAO Office, Tokyo

  • Tokyo
  • Partial Remote
  • Full-time
  • September 25, 2026
Conditions
yen-icon
¥11M ~ ¥20M /yr
location-icon
Apply from Anywhere 👍
visa-icon
Relocation to Japan 👍
(Overseas visa sponsorship supported)
Requirements
language-icon
Language Requirements
Japanese: Not Required 👍
English: Business Level
career-icon
Minimum Experience
Senior or above

Overview

Guided by Money Forward AI Vision 2026, Money Forward is driving company-wide AX (AI Transformation) to deliver "digital workers" — AI agents that carry out business operations autonomously. The CDAO Office leads the AI and data strategy that makes this possible across the entire group.

Join the Mepar (Money Forward Engineering Productivity AI Research) team as a Staff AI Engineer and lead the design of production-grade AI agents that power customer-facing products. In this role you'll own the hardest technical problems at the intersection of agent engineering and backend/infrastructure at scale — turning fast-moving prototypes into reliable, secure, cost-efficient systems that ship to real users.

This is a hands-on, high-leverage individual-contributor role. You'll set technical direction for how we build, evaluate, deploy, and operate agentic systems — including the durable, long-running execution runtime that lets agents complete real back-office work end to end — and raise the bar for engineering across the team while writing code yourself. You'll partner closely with product, business, backend, frontend, infrastructure/SRE, and QA teams to deliver impactful, responsible AI end to end.

 

Responsibilities and Duties

  • Agentic System Design: Architect, build, and scale multi-agent systems and LLM-powered services for customer-facing products — from prototype to production, built to sustain heavy real-world load.
  • Agent Engineering: Design reliable tool-use, function-calling, memory, multi-turn, protocols and modern orchestration frameworks; establish guardrails, evaluation, and safe fallback behavior.
  • Backend & Infrastructure at Scale: Own backend services, APIs, and the infrastructure that agents run on — high availability, low latency, secure secrets/credential handling, and horizontal scalability under production traffic.
  • Evaluation & Benchmarking (core focus): Own how we measure agent quality. Build the evaluation harness end to end — representative task sets, fixed inputs and reference outcomes, repeatable runners, and scoring across task completion, correctness, latency, cost, and human-intervention rate. Establish LLM-as-a-judge and offline/online eval, wire evaluation into CI as a regression gate, and grow from a pilot task set to a durable benchmark suite that product and QA can trust as an acceptance gate.
  • Durable & Long-Running Agent Execution: Design the runtime for agentic tasks that run for ten minutes or longer — durable task lifecycle (submit, status, timeout, cancel, complete, fail), isolated sandboxes for code and file execution, artifact generation and retrieval, and a harness for retries, back-off, checkpointing, resume, and self-recovery.
  • Reliability & Cost: Identify bottlenecks; optimize latency, throughput, token/compute cost, and reliability; instrument observability across the agent lifecycle.
  • Safety & Governance: Apply security and governance best practices (input validation, content filtering, PII handling, HITL escalation, risk-based sampling) appropriate for customer-facing systems.
  • Cross-Functional Technical Leadership & Mentorship: Partner day-to-day with product, business, backend, frontend, infrastructure/SRE, and QA to translate business goals into agent architecture; align standards across BE/AI/Infra, drive system-level decisions that span team boundaries, and mentor engineers on agent development, LLM integration, and evaluation methodology.

 

Required Skills and Experience

  • 7+ years of professional software engineering experience, with strong recent hands-on delivery (not purely managerial).
  • Deep backend engineering expertise — designing, building, and operating large-throughput , low-latency production systems and secure APIs.
  • Strong infrastructure skills: cloud ( AWS and/or Azure ), containers ( Docker ), orchestration ( Kubernetes ), Infrastructure as Code ( Terraform ), and CI/CD — including operating and troubleshooting production services.
  • Proven experience building AI agents / agentic orchestration for real products, with hands-on use of modern agent frameworks — especially the Claude Agent SDK (agentic loop, tool use, sessions, sandboxed execution, Skills). Comparable depth in LangGraph or similar orchestration frameworks is relevant, as is judgment on when a single agentic loop beats a multi-stage pipeline.
  • Strong command of Python or TypeScript for building production services (e.g., FastAPI / FastMCP , async request handling, dependency and lifecycle management, ASGI/Node runtimes)
  • Solid understanding of MCP, REST API, GraphQL protocols for tool integration and agent-to-agent communication.
  • Demonstrated experience building agent evaluation from scratch — not just consuming dashboards. You have designed task sets and scoring rubrics, run controlled baseline-versus-variant experiments, and used the results to drive architecture decisions.
  • Experience with LLM tracing and observability (OpenTelemetry / OpenLLMetry, Langfuse, or equivalent) and with performance and load testing of production services.
  • Experience building or operating durable / long-running execution infrastructure — job and task lifecycle management, isolated sandboxes (containers, microVMs, or managed sandbox services), artifact storage, and failure recovery.
  • Strong grasp of core CS fundamentals: data structures, algorithms, software design, and engineering best practices.

 

Preferred Skills and Experience

  • Experience shipping AI features in customer-facing / B2B SaaS products under real security and compliance constraints.
  • Experience with guardrails, safety, and governance frameworks for LLM/agent systems (e.g., OWASP-style threat modeling for agents).
  • Experience with cost optimization for LLM usage (context compression, model routing, non-frontier models, gateways).
  • Fluent use of AI-assisted development tools (Claude Code, Cursor, GitHub Copilot, Codex) with sound judgment on when to delegate to AI and when to verify.

 

Language Requirements

  • Japanese: Not required but nice to have
  • English: Business level (equivalent to TOEIC 700 or above)
    • If you do not have a qualification equivalent to TOEIC 700 or above, you may be required to take a company-designated test during the selection process.

 

Who We’re Looking For

  • Strong interpersonal and communication skills with a track record of leading across cross-functional teams.
  • Enthusiastic about mentoring and growing other engineers.
  • High level of ownership and accountability, comfortable driving ambiguous, high-impact initiatives.

 

Work Environment

At Money Forward, we provide an environment where we can create world-class services together, and we are looking forward to welcoming you.

  • Provided PC Specs: We provide PCs equipped with the latest CPUs (MacOS or Windows). Custom-made PCs tailored to business requirements and replacements with the latest OS are also possible.
  • Systems to Enhance the Development Environment: Peripheral devices necessary for work (such as displays, mice, keyboards) can be purchased as office supplies. Generally, you can choose from standard products (catalog), and if conditions are met, you can apply for non-standard products as well.
  • Money Forward Library: We have a library system where you can freely borrow books, ranging from technical books to management books. Desired books can be purchased at the company's expense.
  • Referral Driven: We cover the cost of recruitment meals. There is a referral reward system.
  • Conference Participation Support: The company partially covers participation in domestic and international conferences, such as RubyKaigi and Google I/O.

Money Forward, founded in 2012, strives to deliver exceptional value to users in various business domains. As a leading FinTech company, we offer over 40 services, ranging from personal finance management to B2B SaaS products.

We have been growing rapidly, and we are expanding our global hiring to help further expand the company. That means that we are open to hiring those with limited or no Japanese language proficiency.

Money Forward is one of Japan's hottest FinTech companies and it is now a great opportunity to be a part of one of our continued growths!

View Money Forward's company page

↑ Back to top ↑

Staff AI Engineer (AI Agents Development), CDAO Of... at Money Forward
APPLY NOW  ➜