A comprehensive suite of 260+ AI APIs covering optical character recognition, speech recognition and synthesis, natural language processing, and computer vision. RESTful APIs with SDKs for every major language — integrate AI in minutes.
260+ production-ready AI APIs across six domains
Multi-scenario, multi-language, high-accuracy text recognition for document digitization and content moderation — available as offline SDK and cloud API
On-device face SDK for offline identity verification; cloud face API in planning
Offline TTS SDK, on-device SDKs, and LLM-powered speech APIs — with proven Europe and North America node deployments
LLM-powered and classic translation APIs for cross-border e-commerce, product globalization, and smart hardware — 200+ languages with custom terminology
AI input method for smart terminals — 104 supported languages, delivered as APK, NRE + per-license pricing
| OCR SDK | OVERSEAS READY — offline, Windows / Android / iOS devices |
| OCR API | PLANNED — PaddleOCR recommended, 111 languages |
| Face Offline SDK | OVERSEAS READY — offline liveness & recognition |
| Face API | PLANNED — face search & liveness detection |
| Offline TTS SDK | OVERSEAS READY — on-device synthesis, no network |
| Speech SDK / API | OVERSEAS READY — EU & NA nodes proven, per-project evaluation |
| Translation APIs | OVERSEAS READY — 200+ languages, LLM & classic, custom terms |
| Input Method (Offline) | OVERSEAS READY — 104 languages, watch / car / TV / large screen |
| Total APIs | 260+ across 6 categories |
| Speech Languages | 20+ including EN, ZH, JA, KO, FR, DE, ES |
| Translation Pairs | 100+ language combinations |
| SDKs | Python, Java, Go, Node.js, PHP, C#, C++ |
| API Protocol | REST (JSON) + gRPC + WebSocket (streaming) |
| Rate Limits | Up to 5,000 QPM (enterprise tier) |
| Latency (p99) | Under 200ms for most vision APIs |
| SLA | 99.95% uptime guarantee |