Awesome Speech Audio
๐
Papers
๐งญ
Topics
๐ฅ
Trending
๐บ๏ธ
Map
๐
Leaderboards
๐
Learn
๐ค
Ask AI
โฏ
More
๐ฅ
Authors
๐
Reading Packs
๐
Datasets
๐ ๏ธ
Tools
๐ฐ
News
๐
Blogs
โ๏ธ
Newsletter
๐ฏ
Research Radar
๐
Saved
+ Add Paper
โพ
โ
โ authors
ยท
overview
Loading authorโฆ
๐ค
Ask AI
Hirokazu Kameoka โ most-cited papers & profile ยท Speech Audio
โ authors
ยท
overview
Hirokazu Kameoka
33
papers ยท
911
citations ยท
40
h-index
NTT (Japan) ยท NTT Medical Center ยท NTT Basic Research Laboratories
Google Scholar โ
Semantic Scholar โ
OpenAlex โ
Most-cited papers
CycleGAN-VC2: Improved CycleGAN-based Non-parallel Voice Conversion
2019 ยท 275 citations
Parallel-Data-Free Voice Conversion Using Cycle-Consistent Adversarial Networks
2017 ยท 179 citations
MaskCycleGAN-VC: Learning Non-parallel Voice Conversion with Filling in Frames
2021 ยท 72 citations
iSTFTNet: Fast and Lightweight Mel-Spectrogram Vocoder Incorporating Inverse Short-Time Fourier Transform
2022 ยท 63 citations
StarGAN-VC: Non-parallel many-to-many voice conversion with star generative adversarial networks
2018 ยท 48 citations
ACVAE-VC: Non-parallel many-to-many voice conversion with auxiliary classifier variational autoencoder
2018 ยท 48 citations
Voice Transformer Network: Sequence-to-Sequence Voice Conversion Using Transformer with Text-to-Speech Pretraining
2019 ยท 38 citations
Nonparallel Voice Conversion with Augmented Classifier Star Generative Adversarial Networks
2020 ยท 32 citations
ConvS2S-VC: Fully convolutional sequence-to-sequence voice conversion
2018 ยท 20 citations
WaveCycleGAN2: Time-domain Neural Post-filter for Speech Waveform Generation
2019 ยท 18 citations
StarGAN-VC2: Rethinking Conditional Methods for StarGAN-Based Voice Conversion
2019 ยท 17 citations
Many-to-Many Voice Transformer Network
2020 ยท 9 citations
VoiceGrad: Non-Parallel Any-to-Many Voice Conversion with Annealed Langevin Dynamics
2020 ยท 9 citations
FastS2S-VC: Streaming Non-Autoregressive Sequence-to-Sequence Voice Conversion
2021 ยท 8 citations
Wave-U-Net Discriminator: Fast and Lightweight Discriminator for Generative Adversarial Network-Based Speech Synthesis
2023 ยท 8 citations
Top co-authors
Kou Tanaka
ยท 21
Takuhiro Kaneko
ยท 20
Nobukatsu Hojo
ยท 13
Shogo Seki
ยท 4
Shoki Sakamoto
ยท 3
Tadahiro Taniguchi
ยท 3
Wen-Chin Huang
ยท 3
and Takuhiro Kaneko
ยท 2
Junichi Yamagishi
ยท 2
Tomoki Hayashi
ยท 2
Tomoki Toda
ยท 2
Yi-Chiao Wu
ยท 2
Topics
Audio Generation
Voice Cloning
Speech Enhancement
Speech Recognition
Text-to-Speech
Speaker Analysis
Speech Translation
Music Generation
Multimodal Audio