VoiVision AI

Speech Engine (Software Only)

Conformer E2E · real-time on CPU · fully offline

Built on NEU NLP Lab's in-house engine with Conformer + CTC/Attention E2E architecture. INT8/INT4 quantization, pruning, knowledge distillation, and operator fusion enable real-time ASR on CPU alone. Supports 30 languages + 22 dialects, LLM translation/summary across 100 languages, machine translation at ≥800 chars/s, and 10:1 file transcription. Native support for NVIDIA (T4/L4/A10/A100), Ascend (310P/910B), Cambricon (MLU370/590), Hygon DCU; Docker/K8s one-click deploy.

Highlights

  • Conformer + CTC/Attention E2E architecture
  • Real-time on CPU · no GPU needed
  • INT8/INT4 quantization + pruning + distillation
  • LM compute −70% · accuracy +10–20%
  • 30 languages + 22 dialects · LLM translation/summary for 100 languages
  • MT ≥ 800 chars/s · 10:1 file transcription
  • ASR×OCR multimodal fusion · term accuracy 94%+
  • NVIDIA / Ascend / Cambricon / Hygon · Docker/K8s · RESTful/WebSocket

Software Modules

Machine Translation

Text translation across 100 languages

  • Translation speed ≥ 800 chars/s
  • Multilingual translation incl. Chinese/English

Speech Recognition

Streaming, one-shot, and file transcription modes

  • Live streaming recognition with real-time text
  • One-shot utterance recognition
  • Audio file recognition with file transcription

LLM Minutes

LLM-powered structured meeting minutes

  • Chinese minutes: text and graphic formats
  • Automatic to-do extraction
  • Mind-map visualization

User-side API Integration

Bidirectional data flow with WeCom / OA

  • API for private minutes platform with WeCom and third-party OA systems
  • Display, query and manage full minutes inside WeCom / OA terminals

Book a Personalized Demo

Tell us your meeting scenario and compliance needs — get a tailored plan.

Book now
Live Chat