English summary for screening — check the original posting before applying.
The AI Engineer will be responsible for implementing AI models developed through research into production, building a stable and scalable operational foundation. This role involves MLOps/LLMOps for integrating speech recognition, speaker recognition, and LLMs into products, as well as developing private AI and solutions for clients.
Must-haves
- 2+ years of experience implementing machine learning models in production using Python
- Experience designing and developing server-side applications (e.g., Web APIs)
- Experience with containerization using Docker
- Experience with team development using Git
- Experience using cloud services (AWS/Azure/GCP)
- Business-level Japanese proficiency (understanding business documents)
Nice-to-haves
- Experience implementing speech processing and natural language processing models
- Experience developing applications using LLMs
- Experience optimizing model inference (quantization, TensorRT, ONNX, etc.)
- Experience with Kubernetes/ECS for container orchestration
- Experience building CI/CD pipelines
- Experience operating experiment management tools (MLflow, Weights & Biases)
- Experience operating LLM monitoring tools (Langfuse)
Tech stack
PythonFastAPIFlutterReactNext.jsTypeScriptGitDockerAWSAzureGCPECSSageMakerLambdaAzure OpenAI ServiceBigQueryLangfuseKubernetes
Work style
Hybrid (2 days remote per week: Tuesdays and Thursdays), Location: Tokyo, Japan
Other notes
Annual salary: ¥7,000,000 - ¥10,000,000. Includes 45 hours of fixed overtime per month.