Sponsors
  • Habsida
  • Figma
  • Rezi
  • Notion for Startups
  • Habsida
  • Figma
  • Rezi
  • Notion for Startups
  • Habsida
  • Figma
  • Rezi
  • Notion for Startups

Stay updated with our latest job postings by following us on LinkedIn and join our Discord community for daily notifications.

Dev Korea

AI Researcher - Speech / Audio

Seoul (Hybrid) β€’ Full-time

  • Machine Learning
No longer accepting applications

Insights about this position

  • Visa sponsorship: No

    The company cannot sponsor a Korean work visa for this role.

  • Korean language proficiency: Not required

    No knowledge of Korean is required for this role.

  • Workplace type: Hybrid

    A mix of on-site and remote work.

Are you seeking an opportunity to work on intriguing technical challenges within a well-funded company operating in a rapidly growing industry? Look no further!

Join a leading AI company that empowers individuals to create content using synthetic voices and AI actors. We are currently seeking skilled Speech & Audio Researchers to join this great research team.

This company has gained recognition for its natural-sounding AI voice actors and realistic virtual humans. With over 1.4 million subscriber users on platforms like YouTube and TikTok, they have been revolutionizing the content creator market for the past few years.

As a Speech & Audio Researcher, you will be at the forefront of research on the latest algorithms in deep learning and machine learning. Your primary responsibilities will include topics such as speech synthesis, voice conversion, audio pre/post-processing, signal separation, sound quality enhancement, speaker recognition, emotion recognition, and situation recognition.

You will play a crucial role in improving product quality, particularly in the English language, while expanding our capabilities to support German and Spanish. Additionally, you will experiment with innovative approaches to enhance English synthesis and ensure controllability over naturalness, tone, speed, and other speech characteristics. Your insights and contributions will enable our product to effectively predict context.

Key Responsibilities

  • Conduct research on the latest algorithms in deep learning and machine learning related to speech and audio.
  • Drive speech synthesis, voice conversion, audio pre/post-processing, signal separation, sound quality enhancement, speaker recognition, emotion recognition, and situation recognition.
  • Stay up to date with recent advancements in text-to-speech (TTS) technologies.
  • Implement algorithms effectively into commercial products.
  • Manage and analyze large speech and audio datasets.
  • Collaborate with the team to apply deep learning techniques and methodologies.

Required Skills

  • Strong knowledge and experience in deep learning.
  • Ability to commercialize algorithms and integrate them into commercial products.
  • Proficiency in handling large speech and audio datasets.
  • Up-to-date understanding of the latest trends in deep learning and machine learning for Speech and Audio.
  • Join our brilliant team and be part of a company that is redefining the speech and audio synthesis space!

You must be based in South Korea for this position.