vix.ing · top · new · best · stats · spec

Sakti, Sakriani

  1. Listening while Speaking: Speech Chain by Deep Learning
    2017/07/16 by Andros Tjandra, Tjandra, Andros, Sakriani Sakti +3 · 3 citations
    Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #Topic Modeling
  2. The Zero Resource Speech Challenge 2019: TTS without T
    2019/04/25 by Dunbar, Ewan, Algayres, Robin, Karadayi, Julien +10 · 2 citations
    #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
  3. NusaCrowd: Open Source Initiative for Indonesian NLP Resources
    2022/12/19 by Samuel Cahyawijaya, Cahyawijaya, Samuel, Holy Lovenia +91 · 4 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Natural Language Processing Techniques #Speech Recognition and Synthesis #Speech and dialogue systems
  4. Make Skeleton-based Action Recognition Model Smaller, Faster and Better
    2019/07/23 by Yang, Fan, Sakti, Sakriani, Wu, Yang +1 · 2 citations
    #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences
  5. Simultaneous Speech-to-Speech Translation System with Neural Incremental ASR, MT, and TTS
    2020/11/10 by Katsuhito Sudoh, Takatomo Kano, Sudoh, Katsuhito +9 · 2 citations
    Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Natural Language Processing Techniques #Speech Recognition and Synthesis #Topic Modeling
  6. ReMOTS: Self-Supervised Refining Multi-Object Tracking and Segmentation
    2020/07/07 by Fan Yang, Xin Chang, Yang, Fan +11 · 2 citations
    Computer Science · #Advanced Neural Network Applications #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Image Enhancement Techniques #Video Surveillance and Tracking Methods
  7. Tensor Decomposition for Compressing Recurrent Neural Network
    2018/02/28 by Andros Tjandra, Sakriani Sakti, Tjandra, Andros +3 · 1 citation
    Computer Science · Mathematics · #FOS: Computer and information sciences #Machine Learning (cs.LG) #Parallel Computing and Optimization Techniques #Tensor decomposition and applications
  8. VQVAE Unsupervised Unit Discovery and Multi-scale Code2Spec Inverter for Zerospeech Challenge 2019
    2019/05/27 by Tjandra, Andros, Sisman, Berrak, Zhang, Mingyang +3 · 1 citation
    #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Sound (cs.SD) #electronic engineering #information engineering
  9. Transformer VQ-VAE for Unsupervised Unit Discovery and Speech Synthesis: ZeroSpeech 2020 Challenge
    2020/05/24 by Andros Tjandra, Tjandra, Andros, Sakriani Sakti +3 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  10. Instance-level Heterogeneous Domain Adaptation for Limited-labeled Sketch-to-Photo Retrieval
    2022/11/26 by Yang, Fan, Wu, Yang, Wang, Zheng +3 · 1 citation
    #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Machine Learning (cs.LG)
  11. Machine Speech Chain with One-shot Speaker Adaptation
    2018/03/28 by Andros Tjandra, Tjandra, Andros, Sakriani Sakti +3 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Topic Modeling #electronic engineering #information engineering
  12. End-to-End Feedback Loss in Speech Chain Framework via Straight-Through Estimator
    2018/10/31 by Tjandra, Andros, Sakti, Sakriani, Nakamura, Satoshi · 1 citation
    #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Sound (cs.SD) #electronic engineering #information engineering