Tom Ko

3.4k citations
44 papers · 2.0k · 3 hit papers · h-index 13

Impact in

    • Speech and Audio Processing
    • Music and Audio Processing
    • Speech Recognition and Synthesis
    • Natural Language Processing Techniques
    • Topic Modeling
    • Speech and dialogue systems

Papers in

    • Speech Recognition and Synthesis 35
    • Natural Language Processing Techniques 19
    • Topic Modeling 13
    • Speech and dialogue systems 7
    • Domain Adaptation and Few-Shot Learning 4
    • Speech and Audio Processing 19
    • Music and Audio Processing 13

Tom Ko

44 papers receiving 1.8k citations

Tom Ko's Hit Papers

WavCaps: A ChatGPT-Assisted Weakly-Labelled Audio Captioning Dataset for Audio-Language Multimodal Research 2024 · 76 citations
760+3+7Years since publication250500750

Peers

Tom Ko
Comparison fields: 5 of 83
  • Signal Processing 1.4k
  • Artificial Intelligence 1.7k
  • Experimental and Cognitive Psychology 117
  • Computer Vision and Pattern Recognition 194
  • Developmental Biology 9
Replace Yoshihiko Nankaku with:
Yoshihiko Nankaku Japan
Vijayaditya Peddinti United States
Oliver Watts United Kingdom
Gregory Sell United States
Pavel Matějka Czechia
Chao Weng China
Rahim Saeidi Finland
Xuankai Chang United States
Shinji Takaki Japan
Mitchell McLaren United States
Tom Ko relative to Yoshihiko Nankaku Japan Yoshihiko Nankaku's profile →
Citations per field
00.5×8.7×
Yoshihiko Nankaku · 1×
Citations per year

Countries citing papers authored by Tom Ko

Since Specialization
Citations

This map shows the geographic impact of Tom Ko's research. It shows the number of citations coming from papers published by authors working in each country. You can also color the map by specialization and compare the number of citations received by Tom Ko with the expected number of citations based on a country's size and research output (numbers larger than one mean the country cites Tom Ko more than expected).

Fields of papers citing papers by Tom Ko

Since Specialization
Physical SciencesHealth SciencesLife SciencesSocial Sciences

This network shows the impact of papers produced by Tom Ko. Nodes represent research fields, and links connect fields that are likely to share authors. Colored nodes show fields that tend to cite the papers produced by Tom Ko. The network helps show where Tom Ko may publish in the future.

Co-authors

The 25 scholars most cited alongside Tom Ko, linked wherever they have co-authored with each other. Click a name or a connecting line to browse the papers they share.

Border = papers with Tom Ko Line = papers co-authored together Tom Ko links everyone, so they are left out of the graph.

All Works

20 of 20 papers shown

Showing the 20 most-cited of 44 papers — load more, or switch the sort, to bring in the rest.

#Work
1
Audio augmentation for speech recognition
Hit paper breakdown →
2015760
2
A study on data augmentation of reverberant speech for robust speech recognition
Hit paper breakdown →
2017548
3 2018160
4 202278
5
WavCaps: A ChatGPT-Assisted Weakly-Labelled Audio Captioning Dataset for Audio-Language Multimodal Research
Hit paper breakdown →
202476
6 201566
7 201656
8 202229
9 202225
10 201917
11 202316
12 202013
13 202412
14 202312
15 202212
16 202312
17 202312
18 202111
19 20229
20 20238

About Tom Ko

Tom Ko is a scholar working on Artificial Intelligence, Signal Processing, Computer Vision and Pattern Recognition, Information Systems and Computational Mechanics, having authored 44 papers that have together received 2.0k indexed citations. Recurring topics across this work include Speech Recognition and Synthesis (35 papers), Natural Language Processing Techniques (19 papers), Speech and Audio Processing (19 papers), Music and Audio Processing (13 papers), Topic Modeling (13 papers), Speech and dialogue systems (7 papers), Domain Adaptation and Few-Shot Learning (4 papers) and Multimodal Machine Learning Applications (2 papers). The work is most often cited by research in Signal Processing (1.4k citations), Artificial Intelligence (1.7k citations), Experimental and Cognitive Psychology (117 citations), Computer Vision and Pattern Recognition (194 citations) and Developmental Biology (9 citations). Tom Ko has collaborated with scholars based in China, Hong Kong and United States. Frequent co-authors include Daniel Povey, Vijayaditya Peddinti, Sanjeev Khudanpur, Michael L. Seltzer, Brian Mak, David Snyder, Qing Li, Long Zhou, Rui Wang and Wenwu Wang. Their work appears in journals such as Speech Communication, IEEE/ACM Transactions on Audio Speech and Language Processing, IEEE Transactions on Audio Speech and Language Processing, ICASSP 2022 - 2022 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) and Rare & Special e-Zone (The Hong Kong University of Science and Technology).

Rankless uses publication and citation data sourced from OpenAlex, an open and comprehensive bibliographic database. While OpenAlex provides broad and valuable coverage of the global research landscape, it—like all bibliographic datasets—has inherent limitations. These include incomplete records, variations in author disambiguation, differences in journal indexing, and delays in data updates. As a result, some metrics and network relationships displayed in Rankless may not fully capture the entirety of a scholar's output or impact.

Explore authors with similar magnitude of impact