English summary for screening — check the original posting before applying.
Join a startup specializing in AI x Movement Coaching, aiming to be the world's No. 1 healthcare company. This role focuses on developing real-time voice interaction experiences using speech recognition, speech synthesis, and generative AI to enhance motion analysis feedback.
Must-haves
- Development experience in machine learning, generative AI, voice processing, or backend development.
- Research and development experience with machine learning models or AI systems.
- Experience designing evaluation metrics and improving systems based on data.
- System development experience considering real-time performance, latency, and reliability.
- Experience independently driving technical investigations from research to operation.
- Ability to break down ambiguous problems and collaborate with stakeholders.
Nice-to-haves
- Development experience in speech recognition (ASR/STT), speech synthesis (TTS), or Conversational AI.
- Experience with WebRTC, streaming audio, and real-time media processing.
- Knowledge of quality evaluation and analysis for voice conversational systems.
- Development experience in NLP, LLM, RAG, or AI agents.
- Experience improving inference speed, cost, and stability of models or systems.
- Inference optimization experience using ONNX, TensorRT.
- Experience implementing inference on edge devices.
- Interest in technology development in sports, healthcare, or motion analysis domains.
Tech stack
AISpeech RecognitionSpeech SynthesisGenerative AILLMs3D Motion AnalysisONNXTensorRTWebRTC
Other notes
Salary: Senior: 8,000,000 - 12,000,000 JPY; Specialist: 12,000,000 - 20,000,000 JPY. Full-time employee. Probationary period: 3 months.