TechApr 28, 20266 minSarashina2.2-TTS Is a Japanese-First Zero-Shot Voice Synthesis ModelSB Intuitions released sarashina2.2-tts, an LLM-based TTS model focused on Japanese. It clones speaker voice and style from short reference audio without fine-tuning, and handles Japanese-English code-switching.AITTSVoice SynthesisLLMVoice Cloning