Backend & Cloud Engineer
About this role
※日本語の募集要項は下部にございます。
About Citadel AI
Citadel AI builds software products to evaluate, monitor, and govern AI systems. Our mission is to build trust in AI.
We’re located in Tokyo, but we have an international team and global customers. Our goal is to build a world-class startup from Japan, and our team members have previously worked at leading companies such as Google, Paidy, Apple, Stripe, and Meta.
This is a unique opportunity to join a team in Tokyo building new products for a global market. As an early employee at Citadel AI, you’ll have significant impact and ownership over the growth of our startup. We’re supported by Japan’s best investors and a top-tier enterprise customer base, and are currently scaling our business in the 1 → 10 stage.
About the role
We’re hiring a backend engineer to work with our small, full-stack team to build cloud infrastructure and products to evaluate, monitor, and govern AI systems.
You’ll have a wide range of responsibilities across our cloud infrastructure and product backend. On the cloud infrastructure side, you’ll work on latency and cost optimization, scalability, monitoring, and deployments. On the product side, you’ll work with web servers, SQL databases, LLM/AI libraries, GPU inference, API design, integrating with external services (e.g. SSO), and general product functionality to govern AI models and agents.
We primarily run on Google Cloud for our hosted products (mostly Compute Engine, Cloud Run, Cloud Build, Artifact Registry, Vertex AI), but also support self-hosted deployments on Azure and AWS. Our backend is primarily in Python. Some technologies we use are: Docker, Helm, Tilt, OpenTofu, PostgreSQL, Redis, FastAPI, Flask.
Job requirements
Our company, technology, and market are new and changing rapidly, so a large part of the job is to adapt and learn new things. We value learning fast over pre-existing knowledge (“slope is more important than y-intercept”).
You should have at least 3 years of experience in cloud infrastructure and/or backend engineering. We don’t require experience in GCP or Python specifically, but we expect you to be an expert in the cloud platform or technical stack of your choice.
Business-level English and Japanese are required. We prioritize candidates in Tokyo, but may consider remote for exceptional candidates.
Nice-to-have: you've worked at a startup before, or you have AI/ML/LLM experience.
----------
〜AIシステムを支える、堅牢でスケーラブルなプロダクト基盤を構築する〜
役割・ミッション
Citadel AIは、AIへの信頼を構築することをミッションに、AIモデルやAIエージェントの品質、安全性、信頼性を継続的に評価・管理するためのソフトウェアを開発しています。
本ポジションでは、少数精鋭のフルスタックチームの一員として、クラウドインフラからプロダクトのバックエンド機能まで、幅広い領域を担当していただきます。インフラ領域では、レイテンシーやコストの最適化、スケーラビリティ、モニタリング、デプロイメントなどに取り組みます。プロダクト領域では、Webサーバーやデータベース、API、LLM・AIライブラリ、GPU、外部サービスとの連携などを通じて、AI システムやエージェントを適切に管理するための機能を開発します。
技術スタック
クラウド基盤には主に Google Cloud を利用しており、Compute Engine、Cloud Run、Cloud Build、Artifact Registry、Vertex AI などを活用しています。また、Azure および AWS 上でのセルフホスト型デプロイにも対応しています。
バックエンドは主にPythonで開発しています。現在使用している主な技術は、Docker、Helm、Tilt、OpenTofu、PostgreSQL、Redis、FastAPI、Flaskなどです。
なお、記載されているすべての技術に関する経験は必須ではありません。特定のクラウドプラットフォームや技術スタックにおける専門性と、新しい技術を素早く習得し、実務に適用できる能力を重視しています。
具体的な業務内容
クラウドインフラの設計・構築・運用および継続的な改善
システムのレイテンシー、コスト、可用性、スケーラビリティの最適化
モニタリング、障害対応、デプロイメント基盤の設計・改善
Webサーバー、API、SQLデータベースを利用したバックエンド機能の開発
LLM・AI関連ライブラリやGPU推論基盤を活用した機能開発
Azure および AWS 環境へのセルフホスト型デプロイメントの設計・サポート
必須スキル・経験(Must)
クラウドインフラ及びバックエンドシステムに関する3年以上の実務経験
ビジネスレベルの日本語および英語でのコミュニケーション能力
変化の速い技術や市場環境に適応し、新しい技術を自ら学びながら業務を進める能力
歓迎スキル・経験(Want)
AIプロジェクトの推進経験
主要クラウドサービスを利用したインフラの設計・構築・運用経験
英語での基礎的なコミュニケーション能力
この仕事で得られるもの
スタートアップでの勤務経験
AI、機械学習、LLMを利用したプロダクトの開発経験
顧客への導入支援の経験
