Student Researcher (Speech Foundation Model - Seed) - 2027 Start (PhD)

Beijing ByteDance Technology Co Ltd

  • San Jose, CA
  • 30+ days ago

    Highlights

    Explore methods to advance model capabilities in areas such as speech generation, speech understanding, and multimodal modeling involving speech, language, or vision. The team focuses on the forefront of research and product development in speech and audio, music, natural language understanding, and multimodal deep learning.

    Numbers & Facts

    LocationSan Jose, CA

    Description

    About the team

    The mission of the Seed Speech team is to enrich interactive and creative processes through the application of multimodal speech technologies. The team focuses on the forefront of research and product development in speech and audio, music, natural language understanding, and multimodal deep learning.

    Responsibilities

    • Conduct research on speech foundation models and related systems.
    • Explore methods to advance model capabilities in areas such as speech generation, speech understanding, and multimodal modeling involving speech, language, or vision.
    • Design and prototype algorithms, models, or system components.
    • Collaborate with the team to advance research directions.

    Minimum Qualifications

    • Currently pursuing a PhD in computer science, electrical engineering, mathematics, or a related field.
    • Strong programming skills and solid foundation in algorithms and data structures, proficient in Python or C/C++.
    • Demonstrated research track record with publications in conferences related to speech, audio, multimodal learning, or machine learning.

    Preferred Qualifications

    • Experience related to speech generation, speech understanding, multimodal modeling, or representation learning.
    • Strong problem-solving ability and experience conducting independent research.
    • Ability to collaborate effectively in a research environment.

    Similar Jobs