GeneralFoundational model capabilities for multi-industry, multi-scenario use
SpeechU2-ASR
Transcribe audio to text quickly and accurately for multiple scenarios.
SpeechU2-TTS
Text-to-speech with natural, expressive speech in multiple voices.
SpeechU2-TTS-Clone
All-in-one TTS and voice cloning: clone in seconds with high fidelity.
SpeechU2-TTS-Design
Freely tune voice style and emotion to produce custom voices quickly and meet diverse speech synthesis needs.
MedicalSpecialized model capabilities for healthcare scenarios
SpeechU2-ASR-MedComing Soon
Medical speech recognition model
SpeechU2-TTS-MedComing Soon
Medical speech synthesis model
VisionU1-OCR-Med
Medical OCR combining document classification, layout parsing, and professional information extraction.
VisionU2-RadiMedComing Soon
Medical imaging model
U2 SkillHub
AI Agent skill store with one-click install, a wide selection of curated skills, and custom Skill uploads.
U2 Agent
Native Agent LLM intelligent assistant with chat and expert Agent dual modes, suited for office, finance, and other scenarios.

U2Claw
Desktop AI Agent lobster tool that aggregates multi-domain AI experts.
Vision›OCR Extract
API Key not detected. Click "Configure API Key" in the top-right corner before using the playground.