中国科学院机构知识库网格
Chinese Academy of Sciences Institutional Repositories Grid
首页
机构
成果
学者
登录
注册
登陆
×
验证码:
换一张
忘记密码?
记住我
×
校外用户登录
CAS IR Grid
机构
自动化研究所 [136]
心理研究所 [39]
计算技术研究所 [19]
半导体研究所 [11]
软件研究所 [9]
新疆理化技术研究所 [7]
更多
采集方式
OAI收割 [236]
iSwitch采集 [7]
内容类型
期刊论文 [131]
学位论文 [75]
会议论文 [35]
EI期刊论文 [1]
专著章节, 文集论文 [1]
发表日期
2024 [9]
2023 [15]
2022 [14]
2021 [17]
2020 [7]
2019 [15]
更多
学科主题
人工智能 [6]
心理语言学 [4]
Cognitive ... [2]
认知神经科学 [2]
Cognitive ... [1]
Psycholing... [1]
更多
筛选
浏览/检索结果:
共243条,第1-10条
帮助
条数/页:
5
10
15
20
25
30
35
40
45
50
55
60
65
70
75
80
85
90
95
100
排序方式:
请选择
发表日期升序
发表日期降序
题名升序
题名降序
提交时间升序
提交时间降序
作者升序
作者降序
DEFT: Data-Efficient Fine-Tuning Through Multi-Dimensional Data Selection
期刊论文
OAI收割
IEEE TRANSACTIONS ON AUDIO SPEECH AND LANGUAGE PROCESSING, 2026, 卷号: 34, 页码: 352-363
作者:
Dai, Shaojie
;
Liu, Xin
;
Yu, Yue
  |  
收藏
  |  
浏览/下载:0/0
  |  
提交时间:2026/05/25
Data models
Complexity theory
Training data
Training
Speech processing
Semantics
Adaptation models
Tuning
Measurement
Large language models
Data-efficient fine-tuning
large language models (LLMs)
instruction fine-tuning
data selection
Dubbing Movies via Hierarchical Phoneme Modeling and Acoustic Diffusion Denoising
期刊论文
OAI收割
IEEE TRANSACTIONS ON PATTERN ANALYSIS AND MACHINE INTELLIGENCE, 2025, 卷号: 47, 期号: 11, 页码: 10361-10377
作者:
  |  
收藏
  |  
浏览/下载:0/0
  |  
提交时间:2025/12/03
Videos
Lips
Visualization
Acoustics
Cloning
Noise reduction
Motion pictures
Head
Adaptation models
Text to speech
Visual voice cloning
speech synthesis
hierarchical phoneme modeling
contrastive learning
acoustic diffusion denoising
Agent-SiMT: Agent-Assisted Simultaneous Translation With Large Language Models
期刊论文
OAI收割
IEEE TRANSACTIONS ON AUDIO SPEECH AND LANGUAGE PROCESSING, 2025, 卷号: 33, 页码: 2074-2083
作者:
  |  
收藏
  |  
浏览/下载:0/0
  |  
提交时间:2025/12/03
Translation
Hidden Markov models
Transformers
Training
Large language models
Collaboration
Vocabulary
Speech processing
Memory modules
Machine translation
Agent collaboration
large language models
simultaneous translation
Temporal neural dynamics of understanding communicative intentions from speech prosody
期刊论文
OAI收割
NEUROIMAGE, 2024, 卷号: 299, 页码: 14
作者:
Gao, Panke
;
Jiang, Zhufang
;
Yang, Yufang
;
Zheng, Yuanyi
;
Feng, Gangyi
  |  
收藏
  |  
浏览/下载:55/0
  |  
提交时间:2024/10/14
Communicative intention
Speech prosody
EEG
RSA
Social interactions
SceneFake: An initial dataset and benchmarks for scene fake audio detection
期刊论文
OAI收割
PATTERN RECOGNITION, 2024, 卷号: 152, 页码: 12
作者:
Yi, Jiangyan
;
Wang, Chenglong
  |  
收藏
  |  
浏览/下载:70/0
  |  
提交时间:2024/07/04
Scene manipulation
Fake audio detection
Speech enhancement
SceneFake dateset
Spatial reconstructed local attention Res2Net with F0 subband for fake speech detection
期刊论文
OAI收割
NEURAL NETWORKS, 2024, 卷号: 175, 页码: 11
作者:
Fan, Cunhang
;
Xue, Jun
;
Tao, Jianhua
;
Yi, Jiangyan
;
Wang, Chenglong
  |  
收藏
  |  
浏览/下载:85/0
  |  
提交时间:2024/07/04
ASVspoof
Fake speech detection
Fundamental frequency
Res2Net
SMART: Syntax-Calibrated Multi-Aspect Relation Transformer for Change Captioning
期刊论文
OAI收割
IEEE TRANSACTIONS ON PATTERN ANALYSIS AND MACHINE INTELLIGENCE, 2024, 卷号: 46, 期号: 7, 页码: 4926-4943
作者:
Tu, Yunbin
;
Li, Liang
;
Su, Li
;
Zha, Zheng-Jun
;
Huang, Qingming
  |  
收藏
  |  
浏览/下载:37/0
  |  
提交时间:2024/12/06
Semantics
Visualization
Transformers
Decoding
Switches
Syntactics
Image representation
Change captioning
multi-aspect relation learning
part-of-speech
visual switch
transformer
Compensation or Preservation? Different Roles of Functional Lateralization in Speech Perception of Older Non-musicians and Musicians
期刊论文
OAI收割
NEUROSCIENCE BULLETIN, 2024, 页码: 15
作者:
Jin, Xinhu
  |  
收藏
  |  
浏览/下载:46/0
  |  
提交时间:2024/07/09
Functional lateralization
Speech perception
Aging
Musical training experience
Emotion selectable end-to-end text-based speech editing
期刊论文
OAI收割
ARTIFICIAL INTELLIGENCE, 2024, 卷号: 329, 页码: 16
作者:
Wang, Tao
;
Yi, Jiangyan
;
Fu, Ruibo
;
Tao, Jianhua
;
Wen, Zhengqi
  |  
收藏
  |  
浏览/下载:69/0
  |  
提交时间:2024/07/03
Emotion selectable
Text-based speech editing
Emotion decoupling
Mask prediction
Few-shot learning
Text-to-speech
Overview of the Tenth Dialog System Technology Challenge: DSTC10
期刊论文
OAI收割
IEEE-ACM TRANSACTIONS ON AUDIO SPEECH AND LANGUAGE PROCESSING, 2024, 卷号: 32, 页码: 765-778
作者:
Yoshino, Koichiro
;
Chen, Yun-Nung
;
Crook, Paul
;
Kottur, Satwik
;
Li, Jinchao
  |  
收藏
  |  
浏览/下载:63/0
  |  
提交时间:2024/05/20
Task analysis
Internet
History
Oral communication
Measurement
Context modeling
Visualization
Dialog systems
natural language processing
speech processing
multimodal sensors