Guided by Money Forward AI Vision 2026, Money Forward is driving company-wide AX (AI Transformation) to deliver “digital workers” — AI agents that carry out business operations autonomously. The CDAO Office leads the AI and data strategy that makes this possible across the entire group.
This is a dedicated SRE/infrastructure position within our MLOps Platform , focused specifically on the domain serving our Digital Bank and group-company fintech products — the shared credit evaluation model initiative known internally as “Financier”.
Our MLOps Platform is more than a machine learning platform. It also encompasses the serving interfaces, authentication mechanisms, and security layers used by connected products, making it a group-wide product platform. The team currently operates as a scrum unit of a Product Manager, Engineering Manager, Data Scientists, and ML Engineers, running prototype validation. With the Digital Bank launch scheduled for FY2027, standing up infrastructure that can withstand high traffic and high availability requirements — along with production operations and monitoring — has become our most urgent priority.
A separate dedicated SRE owns the standard infrastructure for the MLOps Platform as a whole. In this role, you will work alongside that shared foundation while focusing your efforts on infrastructure built specifically for Digital Bank and fintech products. The scope may look broad at first glance, but it is deliberately bounded: by making full use of AI development tools such as Claude Code and by collaborating closely with our application and ML engineers, this is designed as a role that one dedicated engineer can own end to end.
Success in this role is clearly defined: ensuring the FY2027 Digital Bank launch succeeds from an infrastructure and reliability standpoint. Everything — platform build-out, security architecture, and observability — works backward from that milestone.
Responsibilities
- Build infrastructure and codify it with IaC (Terraform)
- Design, build, and continuously improve the AWS infrastructure behind our serving interfaces, data pipelines, and ML development platform, managed as code with Terraform.
- Design and monitor financial-grade network security
- Architect the infrastructure for mTLS (mutual TLS authentication), build and continuously improve private network connectivity for Digital Bank integration (AWS PrivateLink, VPC Peering, and similar), and expand security monitoring coverage.
- Improve observability and stand up operations and on-call
- Design and introduce monitoring for logs, metrics, and traces using CloudWatch and related tooling. Define SLOs and SLAs, establish an on-call rotation, and standardize incident response processes.
- Design and support the data platform / MLOps interface
- Design and build the interface between the Databricks platform operated by our DRE organization and the ML Ops Platform (data processing and pipeline integration), and provide infrastructure support on the ML Ops side.
- Drive infrastructure collaboration across the group
- Partner with the SRE who owns platform-wide ML Ops standards, as well as data platform and product platform teams in other organizations and group companies, to expand Digital Bank and fintech infrastructure while keeping it consistent with our cross-group foundation.
Requirements
- Approximately 5+ years of professional experience as an SRE or infrastructure engineer
- Hands-on experience designing, building, and operating infrastructure on AWS using services such as ECS, API Gateway, VPC, and IAM
- Proven track record operating infrastructure using Terraform for infrastructure as code, module design, and CI/CD automation
- Solid grounding in advanced networking and security with experience designing architectures using VPC Peering, AWS PrivateLink, and TLS or mTLS authentication
- Experience building and operating observability systems including monitoring, alert design, and performance log analysis in cloud environments
- Working knowledge of CI/CD and container technologies with experience optimizing deployment pipelines using GitHub Actions and Docker
- Business level Japanese (equivalent to JLPT N2 or above)
- Please note that the interviews in the selection process will be conducted in Japanese.
- Basic business level English (equivalent to TOEIC 700 or above)
Nice to haves
While not specifically required, tell us if you have any of the following.
- Infrastructure support experience in MLOps or data platform domains using tools like Amazon SageMaker, Databricks, or Airflow
- Track record of standing up operations, defining SLOs and SLAs, and standardizing incident response using tools like Datadog and PagerDuty
- Security and audit experience in financial services or payments aligned with standards such as the FISC Security Guidelines
- Experience designing, building, and operating shared platforms in collaboration with data and product teams across multiple organizations
Compensation
¥7,008,000 ~ ¥11,004,000 annually.