{"id":3675,"date":"2026-09-08T15:27:43","date_gmt":"2026-09-08T08:27:43","guid":{"rendered":"https:\/\/trivita.ai\/?p=3675"},"modified":"2026-09-30T16:00:50","modified_gmt":"2026-09-30T09:00:50","slug":"tuyen-dung-intern-junior-ai-engineer-speech-ai","status":"publish","type":"post","link":"https:\/\/trivita.ai\/en\/tuyen-dung-intern-junior-ai-engineer-speech-ai\/","title":{"rendered":"Hiring Intern\/Junior AI Engineer (Speech AI)"},"content":{"rendered":"<ul class=\"wp-block-list\">\n\n\n\n\n\n\n\n\n\n\n\n\n\n<\/ul>\n\n\n\n\n\n<ul class=\"wp-block-list\">\n\n\n\n\n\n\n\n<\/ul>\n\n\n\n\n\n<ul class=\"wp-block-list\">\n\n\n\n\n\n\n\n\n\n<\/ul>\n\n\n\n\n\n<ul class=\"wp-block-list\">\n\n\n\n\n\n\n\n<\/ul>\n\n\n\n\n\n<ul class=\"wp-block-list\">\n\n\n\n\n\n\n\n\n\n<\/ul>\n\n\n\n\n\n<ul class=\"wp-block-list\">\n\n\n\n<\/ul>\n\n\n\n<p>Trivita AI is hiring an Intern \/ Junior AI Engineer (Speech AI) to research and develop ASR, TTS, speech enhancement, and voice technologies.<\/p>\n\n\n\n<h2 class=\"wp-block-heading\"><strong>Job description<\/strong><\/h2>\n\n\n\n<ul class=\"wp-block-list\">\n<li>Research and develop Speech AI solutions related to:\n<ul class=\"wp-block-list\">\n<li>Speech-to-Text (ASR)<\/li>\n\n\n\n<li>Text-to-Speech (TTS)<\/li>\n\n\n\n<li>Speech Enhancement \/ Denoising<\/li>\n\n\n\n<li>Voice Activity Detection (VAD)<\/li>\n\n\n\n<li>Speaker Diarization \/ Speaker Recognition<\/li>\n<\/ul>\n<\/li>\n\n\n\n<li>Build pipelines for audio data processing and evaluation.<\/li>\n\n\n\n<li>Preprocess, label, validate, and analyze speech data.<\/li>\n\n\n\n<li>Experiment with open-source or commercial models such as Whisper, wav2vec 2.0, MMS, NVIDIA NeMo, CosyVoice, and similar models.<\/li>\n\n\n\n<li>Fine-tune and optimize models for Vietnamese and real-world use cases.<\/li>\n\n\n\n<li>Evaluate model quality using metrics such as WER, CER, MOS, latency, and throughput.<\/li>\n\n\n\n<li>Collaborate with Backend and MLOps teams to deploy models as APIs or production services.<\/li>\n\n\n\n<li>Read technical documentation and research papers, and stay up to date with emerging Speech AI technologies.<\/li>\n<\/ul>\n\n\n\n<h2 class=\"wp-block-heading\"><strong>Intern requirements<\/strong><\/h2>\n\n\n\n<ul class=\"wp-block-list\">\n<li>Third-year or final-year students, or recent graduates in Computer Science, Artificial Intelligence, Data Science, Electrical\/Computer Engineering, or related fields.<\/li>\n\n\n\n<li>Basic knowledge of Python, Machine Learning, and Deep Learning.<\/li>\n\n\n\n<li>Knowledge of signal processing or audio processing is a plus.<\/li>\n\n\n\n<li>Experience using frameworks such as PyTorch or TensorFlow.<\/li>\n\n\n\n<li>Strong self-learning mindset with the initiative to experiment and read technical documentation.<\/li>\n<\/ul>\n\n\n\n<h2 class=\"wp-block-heading\"><strong>Junior requirements<\/strong><\/h2>\n\n\n\n<ul class=\"wp-block-list\">\n<li><strong>6 months to 2 years of experience<\/strong> in AI\/ML or Speech AI.<\/li>\n\n\n\n<li>Proficiency in Python and at least one Deep Learning framework, preferably PyTorch.<\/li>\n\n\n\n<li>Understanding of one or more areas such as ASR\/TTS, Audio Classification, Speech Enhancement, Voice Conversion, or Speaker Diarization.<\/li>\n\n\n\n<li>Experience working with audio datasets, feature extraction, and model evaluation.<\/li>\n\n\n\n<li>Familiarity with Git, Docker, and Linux.<\/li>\n\n\n\n<li>Experience deploying models as APIs is a plus.<\/li>\n<\/ul>\n\n\n\n<h2 class=\"wp-block-heading\"><strong>Preferred qualifications<\/strong><\/h2>\n\n\n\n<ul class=\"wp-block-list\">\n<li>Personal projects or research experience related to Speech AI.<\/li>\n\n\n\n<li>Experience with Whisper, wav2vec 2.0, NVIDIA NeMo, Hugging Face, or ONNX.<\/li>\n\n\n\n<li>Understanding of Mel-spectrograms, MFCC, FFT, sampling rates, and noise reduction.<\/li>\n\n\n\n<li>Experience with inference optimization techniques such as quantization, batching, streaming, or GPU serving.<\/li>\n\n\n\n<li>Published articles, research papers, GitHub projects, or contributions to open-source projects.<\/li>\n<\/ul>\n\n\n\n<h2 class=\"wp-block-heading\"><strong>Benefits<\/strong><\/h2>\n\n\n\n<ul class=\"wp-block-list\">\n<li>Work on real-world Speech AI products serving actual users.<\/li>\n\n\n\n<li>Receive guidance from engineers experienced in AI and production systems.<\/li>\n\n\n\n<li>Work with Vietnamese speech data and large-scale Speech AI problems.<\/li>\n\n\n\n<li>Gain hands-on experience in model training, evaluation, deployment, and performance optimization.<\/li>\n\n\n\n<li>Work in an environment that encourages research, experimentation, and new ideas.<\/li>\n\n\n\n<li>Opportunity to become a full-time employee after the internship.<\/li>\n<\/ul>\n\n\n\n<h2 class=\"wp-block-heading\">Contact information<\/h2>\n\n\n\n<ul class=\"wp-block-list\">\n<li>Address: No. 01, Street 104, Quarter 3, Binh Trung Ward, Ho Chi Minh City<\/li>\n\n\n\n<li>Phone: 0909797699<\/li>\n\n\n\n<li>Email: hr@trivita.ai<\/li>\n<\/ul>","protected":false},"excerpt":{"rendered":"<p>Trivita AI is hiring an Intern \/ Junior AI Engineer (Speech AI) to research and develop ASR, TTS, speech enhancement, and voice technologies. Job description Intern requirements Junior requirements Preferred qualifications Benefits Contact information<\/p>","protected":false},"author":1,"featured_media":3811,"comment_status":"closed","ping_status":"closed","sticky":false,"template":"","format":"standard","meta":{"footnotes":""},"categories":[4],"tags":[],"class_list":["post-3675","post","type-post","status-publish","format-standard","has-post-thumbnail","hentry","category-tuyen-dung"],"_links":{"self":[{"href":"https:\/\/trivita.ai\/en\/wp-json\/wp\/v2\/posts\/3675","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/trivita.ai\/en\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/trivita.ai\/en\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/trivita.ai\/en\/wp-json\/wp\/v2\/users\/1"}],"replies":[{"embeddable":true,"href":"https:\/\/trivita.ai\/en\/wp-json\/wp\/v2\/comments?post=3675"}],"version-history":[{"count":1,"href":"https:\/\/trivita.ai\/en\/wp-json\/wp\/v2\/posts\/3675\/revisions"}],"predecessor-version":[{"id":3677,"href":"https:\/\/trivita.ai\/en\/wp-json\/wp\/v2\/posts\/3675\/revisions\/3677"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/trivita.ai\/en\/wp-json\/wp\/v2\/media\/3811"}],"wp:attachment":[{"href":"https:\/\/trivita.ai\/en\/wp-json\/wp\/v2\/media?parent=3675"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/trivita.ai\/en\/wp-json\/wp\/v2\/categories?post=3675"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/trivita.ai\/en\/wp-json\/wp\/v2\/tags?post=3675"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}