Parlex is engineered for privacy-first, zero-latency cross-lingual intelligence. It runs 100% on-device neural models without network requests, subscriptions, or corporate data collection.
Balanced between microsecond UI responsiveness and deep neural linguistic precision.
Google ML Kit distilled neural translation engine for everyday interactions, instant clipboard monitoring, and real-time screen overlay translation.
Tencent's translation-specialized models running via llama.cpp C++17 JNI. Trained specifically for nuance, cross-lingual context, and idiom preservation.
Transparent memory consumption, quantization profiles, and license requirements.
| Model Name | Architecture | Parameters | Quant | Download Size | Target RAM | Runtime Engine |
|---|---|---|---|---|---|---|
| Hy-MT 1.5 | Encoder-Decoder LLM | 1.8B | Q4_K_M | 1.13 GB | ~1.8 GB | llama.cpp JNI (C++17) |
| Hy-MT2 Mobile | Specialized Translation | 1.8B | Q4_K_M | 1.13 GB | ~1.8 GB | llama.cpp / Adreno OpenCL |
| Google ML Kit NMT | Distilled Compact NMT | Edge NMT | INT8 | 35 MB | 0 MB | Google ML Kit Native |
| TranslateGemma 4B | Multimodal Edge | 4.0B | E2B / E4B | 2.40 GB | ~2.8 GB | Google LiteRT-LM GPU |
| PP-OCRv6 Tiny | Text Detection & OCR | Mobile Vision | FP16 | 18 MB | ~80 MB | Alibaba MNN Engine |
Built for real-world travel, speech dialogue, and script transliteration without network access.
Natural two-way voice conversations with real-time speech turn detection, acoustic echo cancellation (AEC), and Whisper Tiny neural acoustic recognition.
Comprehensive offline phonetic romanization engine. Instantly converts non-Latin scripts into readable phonetic transcriptions.
Zero internet required. Full bidirectional coverage across 1056 language pairs.
Automated manifest endpoints for package validation, download verification, and checksums.
Parlex client applications query the lightweight catalog manifest to verify model weights, SHA-256 hashes, and package updates: