Portfolio

Voice Assessment
This project is customized to use the latest ASR model Whisper-large, with additional implementation of a React UI and several ...

The 9th International Workshop on Vietnamese Language and Speech Processing (VLSP Hanoi 2022)
FastSpeechStyle : Vietnamese Emotional Speech Synthesis for VLSP 2022 Shared Task. VLSP is the most prestigious and quality con...

Synthetic Speech Attribution - 2022 IEEE Signal Processing Cup
Fake synthetic speech audio tracks can be generated through a wide variety of available methods. Given an audio recording repre...

Music Voice Conversion - 1st The Sound of AI Hackathon 2022
Sing any song without speaking the language

Voice-Preserving Speech Machine Translation - STL Hackathon (Quatar 2022)
Speech Translation for Low Resource Language with Voice I/O and Preserve the Characteristics of the Voice Input

Sleep Stage Classification with Multi-Scale Multi-Period CNN
Sleep stage classification refers to the process of categorizing different stages of sleep based on the patterns and characteri...

Viphoneme
Pypi Package Viphoneme: Phonetization, Convert Vietnamese Grapheme to IPA. I did this project in my 3rd year of college, which ...

Vinorm
Python - NSW package for Vietnamese: Normalization system to convert numbers, abbreviations, and words that cannot be pronounce...

Reading Comprehension Backend - AI Core
This is the AI Service Core system design for virtual assistants I made during my internship in my 3rd year of university.

Core AI Research and Statistic Report for RASA
This is the test statistics of the Rasa chatbot system I made during my 3rd year internship

Statistics and Frontend Improvements for Voice of Southern Speech Synthesis (AILAB)
This report aim to analyze the performance of Voice of Southern TTS system and take an overview about Vietnamese TTS. Some stat...

Lyric based approach for music emotion recognition using hierarchical attention networks (NLP)
We utilize the natural structure of a song which is words combine to lines, lines combine to segments, and segments combine to ...

Ethics in the voice recreating technology or Voice Cloning (AI Ethic)
Technology is growing rapidly, especially the explosion of artificial intelligence in recent years has raised many concerns abo...

CUDA programing language
CUDA is architecture and programming model developed by NVIDIA to run parallel computing on graphics processing units (GPUs) CU...