通义Token Foundary事业部-多模态通用音频大模型算法工程师/专家-未来生活实验室
社招全职3年以上技术类-算法地点:北京 | 杭州状态:招聘♡ 收藏
工作描述
任职要求 1、学历经验:硕士及以上,1年以上音频/语音/音乐生成大模型研发经验。 2、核心技术:精通音频生成算法,对DiT、Flow Matching、MLLM、vocoder等技术有自己的理解和实操经验。 3、底层技术:熟悉音频表征(HuBERT/WavLM等)及高音质编解码方案(EnCodec/DAC等),对音质敏感。 4、多模态能力:有 AudioLD…
登录查看完整工作描述
微信扫码,1秒登录
包括英文材料
学历+
大模型+
https://www.youtube.com/watch?v=xZDB1naRUlk
You will build projects with LLMs that will enable you to create dynamic interfaces, interact with vast amounts of text data, and even empower LLMs with the capability to browse the internet for research papers.
https://www.youtube.com/watch?v=zjkBMFhNj_g
算法+
https://roadmap.sh/datastructures-and-algorithms
Step by step guide to learn Data Structures and Algorithms in 2025
https://www.hellointerview.com/learn/code
A visual guide to the most important patterns and approaches for the coding interview.
https://www.w3schools.com/dsa/
还有更多 •••
相关职位
社招3年以上产品类-用户型
1. 经验背景:3 年以上互联网或 AI 产品经验,具备多模态 AIGC 产品的完整落地经历;有视频生成、多模态模型平台、模型工具链或底层模型相关经验者优先。 2. 技术理解能力:对多模态 AI 技术
更新于 2026-08-13北京|杭州
社招3年以上运营类-内容运营
1、本科及以上学历,5年以上年海外运营经验,有AI 产品或创作者平台经验优先 2、熟悉海外主流社交媒体及社区生态,包括Discord、Instagram、TikTok、X、Reddit、YouTube
更新于 2026-08-27北京|杭州
社招3年以上技术类-算法
Qualifications 1、熟悉主流音频架构(如 Whisper, VITS, AudioLM, Vall-E,CosyVoice); 2、精通音频信号处理及神经编解码器(Neural Code
更新于 2026-09-20北京|杭州