arrow_back Back to freelances
F

RKNN Multilingual Voice Recognition

Freelancer

Share:
placeMY home_workRemote assignmentContract publicAggregated job · MY

eventPublished on Sep 17, 2026 · verifiedWe confirmed this when the job was aggregated

Is this your business?

US$ 250 – US$ 750 per project

About the job

I need a complete voice-to-text pipeline that not only recognises speech but also detects the spoken language on the fly, all running natively on the Rockchip NPU through RKNN. The target device is Android only, so everything—from the model optimisation to the demo app—must be tuned for that environment. Languages to be auto-identified and transcribed: English, Mandarin, Thai, Cantonese, Japanese, Korean, Hokkien and Malay. I am aiming for high identification accuracy; false detections must be the rare exception, not the rule. You are free to start from TensorFlow, PyTorch or ONNX models as long as you convert and fine-tune them with the RKNN Toolkit so that real-time performance is achieved on typical Rockchip boards (e.g., RK356x or similar). Quantisation awareness, mixed-precision tricks and any NPU-specific optimisations are all welcome, but latency must remain low enough for conversational use. Deliverables • An optimised RKNN model capable of streaming inference • An Android demo project (Java/Kotlin or C++ with the NDK) that records audio, detects the language, and outputs the transcription in UTF-8 • Clear build & integration notes so my team can reproduce results on fresh hardware • A short benchmark report showing word-error rate and language-ID accuracy on our provided test set The project is complete once the demo app hits the required accuracy thresholds and runs in real time on the reference Rockchip board.

Keep reading for free

Create a free account to see the full job and apply.

  • badgePortfolio visible to companies
  • notificationsNew job alert by email
  • favoriteAlways free, no catch