Commit Graph

4680 Commits

Author SHA1 Message Date
游雁
03911c82ec python runtime 2024-07-22 17:16:54 +08:00
游雁
c2575f022d docs 2024-07-22 17:11:44 +08:00
游雁
37fc6ad946 v1.1.3 2024-07-22 17:04:29 +08:00
维石
2ae59b6ce0 ONNX and torchscript export for sensevoice 2024-07-22 16:58:27 +08:00
gaochangfeng
340c55838b
EMO_UNK禁用和Merge VAD修复 (#1940)
* 添加富文本解码约束

* special token

* bug fix

* fix

* 增加unk score的参数

* emobaned

* kwargs2cfg

* merge_vad bug fix

---------

Co-authored-by: 常材 <gaochangfeng.gcf@alibaba-inc.com>
2024-07-22 15:28:27 +08:00
游雁
f9c13d7f4b bugfix 2024-07-22 15:07:59 +08:00
游雁
e42f539eb8 bugfix 2024-07-22 15:03:35 +08:00
Shi Xian
e09d87193b
Merge pull request #1928 from liugz18/main
Rename 'res' in line 514 to avoid with naming conflict with line 365
2024-07-22 11:32:39 +08:00
kaixindelele
89b0f92fde
fix model name bug (#1934) 2024-07-22 09:44:03 +08:00
凪咲
85c1675e7c
fix: fix input download logic (#1929) 2024-07-19 10:26:58 +08:00
liugz18
d80ac2fd2d
Rename 'res' in line 514 to avoid with naming conflict with line 365 2024-07-18 21:34:55 +08:00
北念
bd352983c6 add default emo and event target for sensevoice 2024-07-18 15:05:28 +08:00
北念
a98550fdf5 fix sense_voice_datasets 2024-07-17 16:05:58 +08:00
游雁
beef97a2fc update 2024-07-17 10:38:08 +08:00
游雁
a836eca98e update 2024-07-17 10:16:19 +08:00
游雁
374998bd36 sensevoice 2024-07-16 14:30:16 +08:00
游雁
774caaf752 sensevoice 2024-07-16 14:27:13 +08:00
游雁
0248d50d2c Merge branch 'main' of github.com:alibaba-damo-academy/FunASR
merge
2024-07-16 14:23:01 +08:00
游雁
ee8b6e2d99 v1.1.2 2024-07-16 14:22:30 +08:00
游雁
f0eec4c4da sensevoice 2024-07-16 14:22:07 +08:00
游雁
8b74979bd3 sensevoice 2024-07-16 13:57:51 +08:00
Johntheprime
685550515e
fix possible memory leak of funasr-wss-client (#1913) 2024-07-16 11:11:22 +08:00
游雁
f097706c40 v1.1.1 2024-07-16 10:56:14 +08:00
Yuekai Zhang
584cfbdc43
Add triton server for SenseVoice (#1901)
* add triton server for SenseVoice

* fix formatting
2024-07-15 18:43:19 +08:00
彭震东
f2ed4b3856
fix progress bar for batch_size (#1917) 2024-07-15 17:54:27 +08:00
北念
0fe232fd7b add postprocess for sensevoice 2024-07-10 11:32:49 +08:00
北念
5448e926a2 add postprocess for sensevoice 2024-07-10 11:27:35 +08:00
游雁
c1df24ea95 wechat 2024-07-08 11:38:03 +08:00
游雁
d1c612189e v1.1.0 2024-07-05 20:36:50 +08:00
北念
ba64c946c0 update docs 2024-07-05 10:20:20 +08:00
wuhongsheng
3a4281f495
优化speakid和语句匹配逻辑,部分解决speakid不从0递增问题 (#1870) 2024-07-05 00:55:32 +08:00
游雁
0170f534b0 sensevoice 2024-07-05 00:17:06 +08:00
Dogvane Huang
dfcc5d4758
fix c# demo project to new onnx model files (#1689) 2024-07-02 12:24:13 +08:00
wuhongsheng
ba0325c004
修复增加标点后断句起点时间戳bug (#1865)
* 修复断句之间时间戳bug

* 修复增加标点后断句起点时间戳
2024-07-02 12:23:48 +08:00
雾聪
40427797c8 update funasr-runtime-sdk-gpu-0.1.1 2024-07-01 20:49:53 +08:00
雾聪
31bf3a88a0 update funasr-runtime-sdk-gpu-0.1.1 2024-07-01 20:43:06 +08:00
wuhongsheng
d8c1b46daf
修复断句之间时间戳bug (#1863) 2024-07-01 13:41:06 +08:00
游雁
92b14aaa2a update 2024-07-01 11:25:23 +08:00
游雁
05f8022500 update 2024-07-01 11:16:23 +08:00
游雁
a456ab57a8 update 2024-07-01 11:15:08 +08:00
游雁
e8f68b44dd v1.0.29 2024-07-01 11:09:01 +08:00
游雁
0650696dd0 update 2024-07-01 11:08:32 +08:00
wuhongsheng
810046e3df
优化merge segments 参数,解决新闻联播男女主持人“晚上好”合并一个speakid问题 (#1861) 2024-07-01 10:42:58 +08:00
zhifu gao
8c87a9d8a7
Dev gzf deepspeed (#1858)
* total_time/accum_grad

* fp16

* update with main (#1817)

* add cmakelist

* add paraformer-torch

* add debug for funasr-onnx-offline

* fix redefinition of jieba StdExtension.hpp

* add loading torch models

* update funasr-onnx-offline

* add SwitchArg for wss-server

* add SwitchArg for funasr-onnx-offline

* update cmakelist

* update funasr-onnx-offline-rtf

* add define condition

* add gpu define for offlne-stream

* update com define

* update offline-stream

* update cmakelist

* update func CompileHotwordEmbedding

* add timestamp for paraformer-torch

* add C10_USE_GLOG for paraformer-torch

* update paraformer-torch

* fix func FunASRWfstDecoderInit

* update model.h

* fix func FunASRWfstDecoderInit

* fix tpass_stream

* update paraformer-torch

* add bladedisc for funasr-onnx-offline

* update comdefine

* update funasr-wss-server

* add log for torch

* fix GetValue BLADEDISC

* fix log

* update cmakelist

* update warmup to 10

* update funasrruntime

* add batch_size for wss-server

* add batch for bins

* add batch for offline-stream

* add batch for paraformer

* add batch for offline-stream

* fix func SetBatchSize

* add SetBatchSize for model

* add SetBatchSize for model

* fix func Forward

* fix padding

* update funasrruntime

* add dec reset for batch

* set batch default value

* add argv for CutSplit

* sort frame_queue

* sorted msgs

* fix FunOfflineInfer

* add dynamic batch for fetch

* fix FetchDynamic

* update run_server.sh

* update run_server.sh

* cpp http post server support (#1739)

* add cpp http server

* add some comment

* remove some comments

* del debug infos

* restore run_server.sh

* adapt to new model struct

* 修复了onnxruntime在macos下编译失败的错误 (#1748)

* Add files via upload

增加macos的编译支持

* Add files via upload

增加macos支持

* Add files via upload

target_link_directories(funasr PUBLIC ${ONNXRUNTIME_DIR}/lib)
target_link_directories(funasr PUBLIC ${FFMPEG_DIR}/lib)
添加 if(APPLE) 限制

---------

Co-authored-by: Yabin Li <wucong.lyb@alibaba-inc.com>

* Delete docs/images/wechat.png

* Add files via upload

* fixed the issues about seaco-onnx timestamp

* fix bug (#1764)

当语音识别结果包含 `http` 时,标点符号预测会把它会被当成 url

* fix empty asr result (#1765)

解码结果为空的语音片段,text 用空字符串

* update export

* update export

* docs

* docs

* update export name

* docs

* update

* docs

* docs

* keep empty speech result (#1772)

* docs

* docs

* update wechat QRcode

* Add python funasr api support for websocket srv (#1777)

* add python funasr_api supoort

* change little to README.md

* add core tools stream

* modified a little

* fix bug for timeout

* support for buffer decode

* add ffmpeg decode for buffer

* libtorch demo

* update libtorch infer

* update utils

* update demo

* update demo

* update libtorch inference

* update model class

* update seaco paraformer

* bug fix

* bug fix

* auto frontend

* auto frontend

* auto frontend

* auto frontend

* auto frontend

* auto frontend

* auto frontend

* auto frontend

* Dev gzf exp (#1785)

* resume from step

* batch

* batch

* batch

* batch

* batch

* batch

* batch

* batch

* batch

* batch

* batch

* batch

* batch

* batch

* batch

* train_loss_avg train_acc_avg

* train_loss_avg train_acc_avg

* train_loss_avg train_acc_avg

* log step

* wav is not exist

* wav is not exist

* decoding

* decoding

* decoding

* wechat

* decoding key

* decoding key

* decoding key

* decoding key

* decoding key

* decoding key

* dynamic batch

* start_data_split_i=0

* total_time/accum_grad

* total_time/accum_grad

* total_time/accum_grad

* update avg slice

* update avg slice

* sensevoice sanm

* sensevoice sanm

* sensevoice sanm

---------

Co-authored-by: 北念 <lzr265946@alibaba-inc.com>

* auto frontend

* update paraformer timestamp

* [Optimization] support bladedisc fp16 optimization (#1790)

* add cif_v1 and cif_export

* Update SDK_advanced_guide_offline_zh.md

* add cif_wo_hidden_v1

* [fix] fix empty asr result (#1794)

* english timestamp for valilla paraformer

* wechat

* [fix] better solution for handling empty result (#1796)

* update scripts

* modify the qformer adaptor (#1804)

Co-authored-by: nichongjia-2007 <nichongjia@gmail.com>

* add ctc inference code (#1806)

Co-authored-by: haoneng.lhn <haoneng.lhn@alibaba-inc.com>

* Update auto_model.py

修复空字串进入speaker model时报raw_text变量不存在的bug

* Update auto_model.py

修复识别出空串后spk_model内变量未定义问题

* update model name

* fix paramter 'quantize' unused issue (#1813)

Co-authored-by: ZihanLiao <liaozihan1@xdf.cn>

* wechat

* Update cif_predictor.py (#1811)

* Update cif_predictor.py

* modify cif_v1_export

under extreme cases, max_label_len calculated by batch_len misaligns with token_num

* Update cif_predictor.py

torch.cumsum precision degradation, using float64 instead

* update code

---------

Co-authored-by: 雾聪 <wucong.lyb@alibaba-inc.com>
Co-authored-by: zhaomingwork <61895407+zhaomingwork@users.noreply.github.com>
Co-authored-by: szsteven008 <97944818+szsteven008@users.noreply.github.com>
Co-authored-by: Ephemeroptera <605686962@qq.com>
Co-authored-by: 彭震东 <zhendong.peng@qq.com>
Co-authored-by: Shi Xian <40013335+R1ckShi@users.noreply.github.com>
Co-authored-by: 维石 <shixian.shi@alibaba-inc.com>
Co-authored-by: 北念 <lzr265946@alibaba-inc.com>
Co-authored-by: xiaowan0322 <wanchen.swc@alibaba-inc.com>
Co-authored-by: zhuangzhong <zhuangzhong@corp.netease.com>
Co-authored-by: Xingchen Song(宋星辰) <xingchensong1996@163.com>
Co-authored-by: nichongjia-2007 <nichongjia@gmail.com>
Co-authored-by: haoneng.lhn <haoneng.lhn@alibaba-inc.com>
Co-authored-by: liugz18 <57401541+liugz18@users.noreply.github.com>
Co-authored-by: Marlowe <54339989+ZihanLiao@users.noreply.github.com>
Co-authored-by: ZihanLiao <liaozihan1@xdf.cn>
Co-authored-by: zhong zhuang <zhuangz@lamda.nju.edu.cn>

* sensevoice

* sensevoice

* sensevoice

* sensevoice

* sensevoice

* sensevoice

* sensevoice

* sensevoice

* sensevoice

* sensevoice

* sensevoice

* sensevoice

* sensevoice

* v1.0.28 (#1836)

* sensevoice

* sensevoice

* sensevoice

* sensevoice

* sensevoice

* update (#1841)

* v1.0.28

* version checker

* version checker

* rollback cif_v1 for training bug

* fixbug

* fixbug for cif

* fixbug

---------

Co-authored-by: 维石 <shixian.shi@alibaba-inc.com>

* update (#1842)

* v1.0.28

* version checker

* version checker

* rollback cif_v1 for training bug

* fixbug

* fixbug for cif

* fixbug

---------

Co-authored-by: 维石 <shixian.shi@alibaba-inc.com>

* inference

* inference

* inference

* requests

* finetune

* finetune

* finetune

* finetune

* finetune

* add inference prepare func (#1848)

* docs

* docs

* docs

* docs

* docs

---------

Co-authored-by: 雾聪 <wucong.lyb@alibaba-inc.com>
Co-authored-by: zhaomingwork <61895407+zhaomingwork@users.noreply.github.com>
Co-authored-by: szsteven008 <97944818+szsteven008@users.noreply.github.com>
Co-authored-by: Ephemeroptera <605686962@qq.com>
Co-authored-by: 彭震东 <zhendong.peng@qq.com>
Co-authored-by: Shi Xian <40013335+R1ckShi@users.noreply.github.com>
Co-authored-by: 维石 <shixian.shi@alibaba-inc.com>
Co-authored-by: 北念 <lzr265946@alibaba-inc.com>
Co-authored-by: xiaowan0322 <wanchen.swc@alibaba-inc.com>
Co-authored-by: zhuangzhong <zhuangzhong@corp.netease.com>
Co-authored-by: Xingchen Song(宋星辰) <xingchensong1996@163.com>
Co-authored-by: nichongjia-2007 <nichongjia@gmail.com>
Co-authored-by: haoneng.lhn <haoneng.lhn@alibaba-inc.com>
Co-authored-by: liugz18 <57401541+liugz18@users.noreply.github.com>
Co-authored-by: Marlowe <54339989+ZihanLiao@users.noreply.github.com>
Co-authored-by: ZihanLiao <liaozihan1@xdf.cn>
Co-authored-by: zhong zhuang <zhuangz@lamda.nju.edu.cn>
Co-authored-by: PerfeZ <90945395+PerfeZ@users.noreply.github.com>
2024-06-28 17:28:09 +08:00
雾聪
e78d649ddb update readme 2024-06-28 14:28:43 +08:00
雾聪
f170a8e07f update readme 2024-06-28 11:12:07 +08:00
维石
c3ec6b9b9e update cif export 2024-06-28 10:24:54 +08:00
lingji-yidong
c880db5364
Fix: Return tuple ('', []) when char_list is empty to prevent ValueError (#1857)
This commit fixes an issue where an empty char_list causes a ValueError due to insufficient values to unpack. The function now returns a tuple ('', []) when char_list is empty.
2024-06-28 01:28:24 +08:00
雾聪
3c50a034b2 update sdk_roadmap.jpg 2024-06-27 17:49:33 +08:00
Yabin Li
d9529818f5
Add files via upload 2024-06-27 17:44:55 +08:00