enginex-mr_series-sherpa-onnx

EngineX-Iluvatar/enginex-mr_series-sherpa-onnx

Archived

Author	SHA1	Message	Date
Fangjun Kuang	c2dcdabab1	Fix sherpa-onnx-node-version in node examples (#879 )	2024-05-15 14:32:30 +08:00
Fangjun Kuang	03c956a317	Add keyword spotting API for node-addon-api (#877 )	2024-05-14 20:26:48 +08:00
Fangjun Kuang	75630b986b	Support adding puncutations to text for node-addon-api (#876 )	2024-05-14 19:28:56 +08:00
Fangjun Kuang	d19f50b799	Add audio tagging APIs for node-addon-api (#875 )	2024-05-14 17:32:30 +08:00
Fangjun Kuang	388e6a98fc	Add speaker identification APIs for node-addon-api (#874 )	2024-05-14 13:28:50 +08:00
Fangjun Kuang	939fdd942c	Add spoken language identification for node-addon-api (#872 )	2024-05-13 20:26:11 +08:00
Fangjun Kuang	031134b4d4	Add TTS for node-addon-api (#871 )	2024-05-13 19:24:09 +08:00
Fangjun Kuang	697b960768	Add non-streaming ASR APIs for node-addon-api (#868 )	2024-05-13 16:03:34 +08:00
Fangjun Kuang	384f96c40f	Add streaming CTC ASR APIs for node-addon-api (#867 )	2024-05-13 11:58:25 +08:00
Fangjun Kuang	677bc1da3e	Add Speaker ID demo for C# (#862 )	2024-05-11 13:27:33 +08:00
Fangjun Kuang	46e4e5b7ac	Add C++ support for streaming NeMo CTC models. (#857 )	2024-05-10 16:26:43 +08:00
Fangjun Kuang	17cd3a5f01	Add C++ runtime for non-streaming faster conformer transducer from NeMo. (#854 )	2024-05-10 12:15:39 +08:00
Fangjun Kuang	5d8c35e44e	Add C++ support for non-streaming NeMo fast conformer hybrid transducer ctc (the ctc branch) (#848 )	2024-05-09 15:32:22 +08:00
Fangjun Kuang	dbaa26ff4b	Publish node-addon-api npm package for linux arm64 (#841 )	2024-05-07 23:05:40 +08:00
Fangjun Kuang	37a4135dd7	Publish npm package with node-addon-api for Windows (#838 )	2024-05-06 16:21:29 +08:00
Fangjun Kuang	4f758e6cd3	Publish node-addon-api wrapper for sherpa-onnx as npm packages (#829 )	2024-05-04 13:27:39 +08:00
Fangjun Kuang	612002da57	Fix C# to support Chinese tts models using jieba (#815 )	2024-04-26 11:50:07 +08:00
Fangjun Kuang	13730ecbd8	Add C API for punctuation (#768 )	2024-04-14 19:02:34 +08:00
Fangjun Kuang	68b8b88b5a	Add Python API for punctuation models. (#762 )	2024-04-13 13:28:17 +08:00
Fangjun Kuang	329fe1aa8b	Support adding punctuations to the speech recogntion result (#761 )	2024-04-13 12:15:57 +08:00
Fangjun Kuang	f204e62b44	Add C API for audio tagging (#754 )	2024-04-11 14:18:43 +08:00
Fangjun Kuang	34d70a259f	Add Python API and Python examples for audio tagging (#753 )	2024-04-11 11:12:48 +08:00
Fangjun Kuang	f20291cadc	Support audio tagging using zipformer (#747 )	2024-04-10 14:47:06 +08:00
Fangjun Kuang	6fb8ceda57	Add VAD examples using ALSA for recording (#739 )	2024-04-08 16:41:01 +08:00
Fangjun Kuang	a5f8fbc83f	Support heteronyms in Chinese TTS (#738 )	2024-04-08 11:01:30 +08:00
Fangjun Kuang	dbff2eaadb	Add C API for streaming HLG decoding (#734 )	2024-04-05 10:31:20 +08:00
Fangjun Kuang	db67e00c77	Add HLG decoding for streaming CTC models (#731 )	2024-04-03 21:31:42 +08:00
Fangjun Kuang	2e0bccad36	Add C API for speaker embedding extractor. (#711 )	2024-03-28 18:05:40 +08:00
Fangjun Kuang	305c373107	Add C# API for spoken language identification (#697 )	2024-03-25 18:45:09 +08:00
Fangjun Kuang	83a10a55a5	Add Swift API for spoken language identification. (#696 )	2024-03-25 16:22:25 +08:00
Fangjun Kuang	ab7cff2513	Add C API for spoken language identification. (#695 )	2024-03-25 15:16:47 +08:00
Fangjun Kuang	0d258dd150	Support spoken language identification with whisper (#694 )	2024-03-24 22:57:00 +08:00
Fangjun Kuang	24f437a6f1	Refactor github actions tests (#688 )	2024-03-22 21:22:42 +08:00
Wei Kang	734bbd91dc	Add Python API for keyword spotting (#576 ) * Add alsa & microphone support for keyword spotting * Add python wrapper	2024-03-01 09:31:11 +08:00
Wei Kang	2ff1049079	change modelscope link to github for build-kws-apki (#540 )	2024-01-24 16:40:14 +08:00
Wei Kang	626775e5e2	Change model url from modelscope to github (#538 )	2024-01-23 10:15:58 +08:00
Wei Kang	b6c020901a	decoder for open vocabulary keyword spotting (#505 ) * various fixes to ContextGraph to support open vocabulary keywords decoder * Add keyword spotter runtime * Add binary * First version works * Minor fixes * update text2token * default values * Add jni for kws * add kws android project * Minor fixes * Remove unused interface * Minor fixes * Add workflow * handle extra info in texts * Minor fixes * Add more comments * Fix ci * fix cpp style * Add input box in android demo so that users can specify their keywords * Fix cpp style * Fix comments * Minor fixes * Minor fixes * minor fixes * Minor fixes * Minor fixes * Add CI * Fix code style * cpplint * Fix comments * Fix error	2024-01-20 22:52:41 +08:00
Fangjun Kuang	2024e96639	Add C++ runtime for speaker verification models from NeMo (#527 )	2024-01-13 21:42:09 +08:00
Fangjun Kuang	68a525a024	Export speaker verification models from NeMo to ONNX (#526 )	2024-01-13 19:49:45 +08:00
Fangjun Kuang	afc81ec122	Add C++ runtime for models from 3d-speaker (#523 )	2024-01-11 19:10:30 +08:00
Fangjun Kuang	e475e750ac	Support streaming zipformer CTC (#496 ) * Support streaming zipformer CTC * test online zipformer2 CTC * Update doc of sherpa-onnx.cc * Add Python APIs for streaming zipformer2 ctc * Add Python API examples for streaming zipformer2 ctc * Swift API for streaming zipformer2 CTC * NodeJS API for streaming zipformer2 CTC * Kotlin API for streaming zipformer2 CTC * Golang API for streaming zipformer2 CTC * C# API for streaming zipformer2 CTC * Release v1.9.6	2023-12-22 13:46:33 +08:00
Fangjun Kuang	868c339e5e	Support distil-small.en whisper (#472 )	2023-12-08 11:59:20 +08:00
Fangjun Kuang	3ae984f148	Remove the 30-second constraint from whisper. (#471 )	2023-12-07 17:47:08 +08:00
Fangjun Kuang	62dc3c3e46	Use piper-phonemize to convert text to token IDs (#453 )	2023-11-30 23:57:43 +08:00
Fangjun Kuang	db41778e99	Support piper-phonemize (#452 )	2023-11-28 19:12:58 +08:00
Fangjun Kuang	8dc08a9b97	Fix nodejs on Windows (#450 )	2023-11-25 21:23:15 +08:00
Fangjun Kuang	2f22e6ed63	Add Swift API for TTS (#439 )	2023-11-22 16:04:26 +08:00
Fangjun Kuang	fe977b8e8e	support nodejs (#438 )	2023-11-21 23:20:08 +08:00
Fangjun Kuang	049fb9f451	Add Python APIs for WeNet CTC models (#428 )	2023-11-16 14:20:41 +08:00
Fangjun Kuang	fac4f6bc7c	Support streaming conformer CTC models from wenet (#427 )	2023-11-16 10:35:23 +08:00

1 2

78 Commits