I am a Research Scientist at DubGuild specializing in speech LLMs. I worked on the training and evaluation of conversational speech models for synthetic dialogue generation, through full-parameter supervised fine-tuning and reinforcement-learning post-training (YouTube demo). I also helped scale an English-Japanese speech LLM (check our blogs).

I received my Ph.D. in Informatics from Nagoya University at Toda Laboratory under the supervision of Professor Tomoki Toda. My research on speech synthesis, voice conversion, and speech recognition has been published at ICASSP, Interspeech, EUSIPCO, ASRU, and in IEEE journals. I was the head organizer of the Singing Voice Conversion Challenge 2025 and a member of the organizing committee for the 2023 challenge, and I serve on the peer-review committees of conferences such as ASRU, SLT, ICASSP, Interspeech, and IJCNN, and journals like IEEE JSTSP.

I have a deep international background now based in Japan, having done my B.S. in the Philippines and a research exchange in France. Outside of research, I like bouldering (check out my instagram page) and learning Japanese.

News

Experience

Jul. 2025 to Present

Research Scientist — DubGuild

Focused on the training and evaluation of conversational speech LLMs for synthetic dialogue generation, including full-parameter SFT and RL post-training.

Nov. 2024 to Jul. 2025

Research Engineer — CoeFont

Developed real-time voice conversion models for the CoeFont Voice Changer and trained large-scale emotional TTS models.

Feb. 2024 to Nov. 2024

ML Engineer — Voice-Swap.AI

Developed singing voice conversion models for the Voice-Swap singing studio, used by music-industry clients.

Oct. 2023 to Mar. 2024

Research Assistant — Sony CSL Tokyo

Manager: Dr. Taketo Akama

Researched highly controllable, low-resource singing voice synthesis.

Mar. 2022

Research Intern — NTT Media Intelligence Laboratories

Manager: Dr. Atsushi Ando

Developed and analyzed speaker diarization systems using various encoders.

Jan. 2022 to Feb. 2022

Research Intern — Hitachi Ltd.

Manager: Dr. Takashi Sumiyoshi

Developed speech recognition systems for low-resource datasets.

Technical Focus

ML Frameworks

PyTorch, Megatron-Bridge, NeMo AutoModel, verl, TRL, vLLM

Research Areas

SFT, RL post-training (GRPO, DPO, KTO), reward design

Infrastructure

Multi-node GPU training (FSDP, expert parallelism, context parallelism)

Languages

English (native), Tagalog (native), Japanese (upper intermediate)

Publications

An Extensive Analysis of the Singing Voice Conversion Challenge 2025 Evaluation Results

In Review 2026

An Extensive Analysis of the Singing Voice Conversion Challenge 2025 Evaluation Results

Lester Phillip Violeta, Xueyao Zhang, Jiatong Shi, Yusuke Yasuda, Wen-Chin Huang, Zhizheng Wu, Tomoki Toda

Scaling Japanese Speech Foundation Models and Examining TTS Performance (in Japanese)

Technical Report 2026

Scaling Japanese Speech Foundation Models and Examining TTS Performance (in Japanese)

長谷川 直哉, 相田 優希, 廣岡 聖司, 林 春太朗, Lester Phillip Violeta, 大嶽 匡俊

The Singing Voice Conversion Challenge 2025: From Singer Identity Conversion To Singing Style Conversion

ICASSP 2026

The Singing Voice Conversion Challenge 2025: From Singer Identity Conversion To Singing Style Conversion

Lester Phillip Violeta, Xueyao Zhang, Jiatong Shi, Yusuke Yasuda, Wen-Chin Huang, Zhizheng Wu, Tomoki Toda

Serenade: A Singing Style Conversion Framework Based on Audio Infilling

EUSIPCO 2025

Serenade: A Singing Style Conversion Framework Based on Audio Infilling

Lester Phillip Violeta, Wen-Chin Huang, Tomoki Toda

Resolving Domain Mismatches in Electrolaryngeal Speech Enhancement With Linguistic Intermediates

IEEE JSTSP 2025

Resolving Domain Mismatches in Electrolaryngeal Speech Enhancement With Linguistic Intermediates

Lester Phillip Violeta, Wen-Chin Huang, Ding Ma, Ryuichi Yamamoto, Kazuhiro Kobayashi, Tomoki Toda

Electrolaryngeal Speech Intelligibility Enhancement through Robust Linguistic Encoders

ICASSP 2024

Electrolaryngeal Speech Intelligibility Enhancement through Robust Linguistic Encoders

Lester Phillip Violeta, Wen-Chin Huang, Ding Ma, Ryuichi Yamamoto, Kazuhiro Kobayashi, Tomoki Toda

Pretraining and Adaptation Techniques for Electrolaryngeal Speech Recognition

IEEE/ACM TASLP 2024

Pretraining and Adaptation Techniques for Electrolaryngeal Speech Recognition

Lester Phillip Violeta, Ding Ma, Wen-Chin Huang, Tomoki Toda

A Preliminary Investigation on Flexible Singing Voice Synthesis Through Decomposed Framework with Inferrable Features

Technical Report 2024

A Preliminary Investigation on Flexible Singing Voice Synthesis Through Decomposed Framework with Inferrable Features

Lester Phillip Violeta, Taketo Akama

The Singing Voice Conversion Challenge 2023

ASRU 2023

The Singing Voice Conversion Challenge 2023

Wen-Chin Huang, Lester Phillip Violeta, Songxiang Liu, Jiatong Shi, Tomoki Toda

An Analysis of Personalized Speech Recognition System Development for the Deaf and Hard-of-hearing

APSIPA 2023

An Analysis of Personalized Speech Recognition System Development for the Deaf and Hard-of-hearing

Lester Phillip Violeta, Tomoki Toda

Intermediate Fine-tuning Using Imperfect Synthetic Speech for Improving Electrolaryngeal Speech Recognition

ICASSP 2023

Intermediate Fine-tuning Using Imperfect Synthetic Speech for Improving Electrolaryngeal Speech Recognition

Lester Phillip Violeta, Ding Ma, Wen-Chin Huang, Tomoki Toda

Investigating Self-Supervised Pretraining Frameworks for Pathological Speech Recognition

Interspeech 2022

Investigating Self-Supervised Pretraining Frameworks for Pathological Speech Recognition

Lester Phillip Violeta, Wen-Chin Huang, Tomoki Toda

Education

Apr. 2023—Mar. 2026

Nagoya University, Japan

Ph.D. in Informatics

Advisor: Prof. Tomoki Toda

Thesis: Domain Adaptation Techniques for Electrolaryngeal Speech Recognition and Enhancement

Apr. 2021—Mar. 2023

Nagoya University, Japan

M.S. in Informatics

Advisor: Prof. Tomoki Toda

Thesis: Pretraining and Adaptation Techniques for Pathological Speech Recognition

2015—2020

Ateneo de Manila University, Philippines

B.S. Electronics Engineering

Aug. 2019—Feb. 2020

Institut catholique d'arts et métiers — Site de Paris-Sénart, France

Research Exchange Semester

Academic Service & Awards

Dec. 2024—Jan. 2026

Head Organizer

The Singing Voice Conversion Challenge 2025

ICASSP 2026

Dec. 2022—Dec. 2023

Organizing Committee

The Singing Voice Conversion Challenge 2023

ASRU 2023 Special Session

Aug. 2024—Present

Peer Review Committee

IEEE SLT, IEEE ICASSP, ISCA Interspeech, IEEE IJCNN, IEEE ASRU, IEEE JSTSP

Scholarship

Monbukagakusho Japanese Government Scholarship (Ph.D.)

Scholarship

Monbukagakusho Japanese Government Scholarship (Master’s)

Travel Grant

Interspeech 2022 Travel Grant

Fellowship

Nagoya University Interdisciplinary Frontier Fellowship