通义Token Foundary事业部-多模态通用音频大模型算法工程师/专家-未来生活实验室
社招全职3年以上技术类-算法地点:北京 | 杭州状态:招聘
工作描述
任职要求 1、学历经验:硕士及以上,1年以上音频/语音/音乐生成大模型研发经验。 2、核心技术:精通音频生成算法,对DiT、Flow Matching、MLLM、vocoder等技术有自己的理解和实操经验。 3、底层技术:熟悉音频表征(HuBERT/WavLM等)及高音质编解码方案(EnCodec/DAC等),对音质敏感。 4、多模态能力:有 AudioLD…
登录查看完整工作描述
微信扫码,1秒登录
包括英文材料
学历+
大模型+
https://www.youtube.com/watch?v=xZDB1naRUlk
You will build projects with LLMs that will enable you to create dynamic interfaces, interact with vast amounts of text data, and even empower LLMs with the capability to browse the internet for research papers.
https://www.youtube.com/watch?v=zjkBMFhNj_g
算法+
https://roadmap.sh/datastructures-and-algorithms
Step by step guide to learn Data Structures and Algorithms in 2025
https://www.hellointerview.com/learn/code
A visual guide to the most important patterns and approaches for the coding interview.
https://www.w3schools.com/dsa/
还有更多 •••
相关职位
社招3年以上产品类-用户型
1. 经验背景:3 年以上互联网或 AI 产品经验,具备多模态 AIGC 产品的完整落地经历;有视频生成、多模态模型平台、模型工具链或底层模型相关经验者优先。 2. 技术理解能力:对多模态 AI 技术
更新于 2026-08-13北京|杭州
社招3年以上运营类-内容运营
1. 3-5年品牌市场、活动策划或整合营销相关经验,有AIGC类产品成功案例者优先。 2. 具备完整的品牌活动从 0 到 1 操盘经验,抗压性强,能在多任务并行环境下精准管控节奏与交付品质。 3. 熟
更新于 2026-08-10北京|杭州
社招3年以上技术类-算法
Qualifications 1、熟悉主流音频架构(如 Whisper, VITS, AudioLM, Vall-E,CosyVoice); 2、精通音频信号处理及神经编解码器(Neural Code
更新于 2026-08-13北京|杭州