enginex-mr_series-sherpa-onnx

EngineX-Iluvatar/enginex-mr_series-sherpa-onnx

Archived

Author	SHA1	Message	Date
Fangjun Kuang	d148860d2c	Add Kotlin and Java API for FireRedAsr AED model (#1870 )	2025-02-17 10:50:25 +08:00
Fangjun Kuang	69f489f0cd	Support scaling the duration of a pause in TTS. (#1820 )	2025-02-08 12:47:26 +08:00
Fangjun Kuang	4372a7a7b0	Add Java and Koltin API for Kokoro TTS 1.0 (#1798 )	2025-02-07 09:59:27 +08:00
Fangjun Kuang	c84a833863	Add C++ and Python API for Kokoro 1.0 multilingual TTS model (#1795 )	2025-02-06 22:57:13 +08:00
Fangjun Kuang	8b989a851c	Fix keyword spotting. (#1689 ) Reset the stream right after detecting a keyword	2025-01-20 16:41:10 +08:00
Fangjun Kuang	99cef4198b	Add Koltin and Java API for Kokoro TTS models (#1728 )	2025-01-17 17:36:13 +08:00
Fangjun Kuang	3422b9388d	Add Kotlin API for Matcha-TTS models. (#1668 )	2024-12-31 19:20:52 +08:00
Fangjun Kuang	e639c70d78	Support linking onnxruntime statically for Android (#1619 )	2024-12-14 09:53:44 +08:00
Fangjun Kuang	bd4b223920	Add Kotlin and Java API for Moonshine models (#1474 )	2024-10-26 22:30:29 +08:00
Fangjun Kuang	94b26ff07c	Android JNI support for speaker diarization (#1421 )	2024-10-12 13:03:48 +08:00
Fangjun Kuang	1ed803adc1	Dart API for speaker diarization (#1418 )	2024-10-11 21:17:41 +08:00
Fangjun Kuang	2d412b1190	Kotlin API for speaker diarization (#1415 )	2024-10-11 14:41:53 +08:00
Fangjun Kuang	e7ffcbd677	Add APIs about max speech duration in VAD for various programming languages (#1349 )	2024-09-14 12:30:13 +08:00
RGdevz	1f29e4a1a9	throw error instead exit (#1323 )	2024-09-06 09:59:21 +08:00
Fangjun Kuang	ca729faebf	Support reading multi-channel wave files with 8/16/32-bit encoded samples (#1258 )	2024-08-15 14:54:43 +08:00
Robin Zhong	62c4d4ab62	Add emotion, event of SenseVoice. (#1257 ) * Add emotion, event of SenseVoice. * Fix tokens size check and update java api. https://github.com/k2-fsa/sherpa-onnx/pull/1257	2024-08-14 15:50:13 +08:00
ivan provalov	9f06b059d7	Update offline-recognizer.cc (#1253 ) Adding setConfig method to JNI to support setting a config on the previously initialized offline-recognizer.	2024-08-13 23:04:51 +08:00
Fangjun Kuang	94e256244d	Add blank penalty for various language bindings. (#1234 )	2024-08-08 10:43:31 +08:00
Fangjun Kuang	dd300b1de5	Add Java and Kotlin API for sense voice (#1164 )	2024-07-22 14:08:40 +08:00
Fangjun Kuang	c2cc9dec58	Add Flush to VAD so that the last segment can be detected. (#1099 )	2024-07-09 16:15:56 +08:00
Fangjun Kuang	a25075101c	Build sherpa-onnx as a single shared library (#1078 ) When `-D BUILD_SHARED_LIBS=ON` is passed to `cmake`, it builds a single shared library. Specifically, - For C APIs, it builds `libsherpa-onnx-c-api.so` - For Python APIs, it builds `_sherpa_onnx.cpython-xx-xx.so` - For Kotlin and Java APIs, it builds `libsherpa-onnx-jni.so` There is no `libsherpa-onnx-core.so` any longer. Note it affects only shared libraries.	2024-07-06 16:41:54 +08:00
Manix	55decb7bee	Add config for TensorRT and CUDA execution provider (#992 ) Signed-off-by: manickavela1998@gmail.com <manickavela1998@gmail.com> Signed-off-by: manickavela1998@gmail.com <manickavela.arumugam@uniphore.com>	2024-07-05 15:18:37 +08:00
Fangjun Kuang	2f8c489698	Publish pre-built jni libs for windows and osx (#1056 )	2024-06-25 11:59:04 +08:00
Fangjun Kuang	9dd0e03568	Enable to stop TTS generation (#1041 )	2024-06-22 18:18:36 +08:00
Fangjun Kuang	6789c909d2	Inverse text normalization API of streaming ASR for various programming languages (#1022 )	2024-06-18 13:42:17 +08:00
Fangjun Kuang	6e09933d99	Inverse text normalization API for other programming languages (#1019 )	2024-06-17 17:02:39 +08:00
Fangjun Kuang	fd5a0d1e00	Add C++ runtime for Tele-AI/TeleSpeech-ASR (#970 )	2024-06-05 00:26:40 +08:00
Fangjun Kuang	f1cff83ef9	Add address sanitizer and undefined behavior sanitizer (#951 )	2024-05-31 13:17:01 +08:00
Fangjun Kuang	bcaa6df389	Add VAD demo for Java API (#928 )	2024-05-28 14:59:47 +08:00
Wei Kang	b012b78ceb	Encode hotwords in C++ side (#828 ) * Encode hotwords in C++ side	2024-05-20 19:41:36 +08:00
Fangjun Kuang	65635b09d8	Fix a typo in jni (#885 )	2024-05-16 14:31:45 +08:00
linziguan	d2745698c5	Support building JNI on Windows (#881 )	2024-05-16 06:25:53 +08:00
Fangjun Kuang	db85b2c1d8	Add Android APKs for NeMo CTC models. (#866 )	2024-05-12 14:58:36 +08:00
Fangjun Kuang	fcd6024200	Fix typos in JNI TTS (#824 )	2024-05-01 14:14:24 +08:00
Fangjun Kuang	5407f880c0	Add Java and Kotlin API for punctuation models (#818 )	2024-04-26 22:06:48 +08:00
Fangjun Kuang	f7b3735621	Add CTC HLG decoding for JNI (#810 )	2024-04-25 17:20:02 +08:00
Fangjun Kuang	c3a2e8a67c	Refactor Java API (#806 )	2024-04-24 18:41:48 +08:00
Fangjun Kuang	9b67a476e6	Refactor the JNI interface to make it more modular and maintainable (#802 )	2024-04-24 09:48:42 +08:00
Fangjun Kuang	7f3b9ffe5d	Refactor TTS Android code to support jieba for Chinese TTS models (#800 )	2024-04-22 17:21:05 +08:00
Fangjun Kuang	c1608b3524	Support CED models (#792 )	2024-04-19 15:20:37 +08:00
Fangjun Kuang	d97a283dbb	Add Android demo for spoken language identification using Whisper multilingual models (#783 )	2024-04-18 14:33:59 +08:00
Fangjun Kuang	3a43049ba1	Add JNI support for spoken language identification (#782 )	2024-04-17 19:27:15 +08:00
Fangjun Kuang	bcd9e48150	Add Android demo for audio tagging (#776 ) See https://k2-fsa.github.io/sherpa/onnx/audio-tagging/apk.html	2024-04-16 20:47:16 +08:00
Fangjun Kuang	5981adf454	Add Kotlin API for audio tagging (#770 )	2024-04-15 13:49:35 +08:00
Fangjun Kuang	a5f8fbc83f	Support heteronyms in Chinese TTS (#738 )	2024-04-08 11:01:30 +08:00
Fangjun Kuang	2e0bccad36	Add C API for speaker embedding extractor. (#711 )	2024-03-28 18:05:40 +08:00
Leo Huang	638f48f47a	Added progress for callback of tts generator (#712 ) Co-authored-by: leohwang <leohwang@360converter.com>	2024-03-28 17:12:20 +08:00
longshiming	de655e838e	delete incorrect logs (#714 ) Co-authored-by: longshiming <longshiming@greesoft.com>	2024-03-28 10:49:45 +08:00
Fangjun Kuang	4e040c596e	Support including TTS conditionally. (#699 )	2024-03-26 17:21:35 +08:00
GaryLaurenceauAva	ac43c2d7b6	Expose 'language' 'task' 'tailPaddings' in OfflineWhisperModelConfig (#643 ) Co-authored-by: Gary <gary.laurenceau@gmail.com>	2024-03-08 19:52:30 +08:00

1 2

77 Commits