千问千问C端事业群-大模型数据研发高级专家-北京/杭州
社招全职3年以上技术类-开发地点:北京 | 杭州状态:招聘
工作描述
任职要求 1. 主导过LLM、VLM、ASR或TTS大模型预训练及微调语料数据建设工作,有丰富的数据交付经验; 2. 精通大规模分布式数据处理技术(如spark/flink/ray等),拥有从0到1搭建全模态数据处理pipeline的丰富实战经验; 3. 深刻理解大模型训练数据的特性与需求,对高质量语料建设有自己的方法论沉淀,具备从无到有构建数据分类、质量体系及数据画像的能力,善于从模型效…
登录查看完整工作描述
微信扫码,1秒登录
包括英文材料
大模型+
https://www.youtube.com/watch?v=xZDB1naRUlk
You will build projects with LLMs that will enable you to create dynamic interfaces, interact with vast amounts of text data, and even empower LLMs with the capability to browse the internet for research papers.
https://www.youtube.com/watch?v=zjkBMFhNj_g
语音识别+
https://developer.nvidia.com/blog/essential-guide-to-automatic-speech-recognition-technology/
Over the past decade, AI-powered speech recognition systems have slowly become part of our everyday lives, from voice search to virtual assistants in contact centers, cars, hospitals, and restaurants.
语音合成+
https://www.ibm.com/think/topics/text-to-speech
Text to speech (TTS) is a type of technology that converts text on a digital interface into natural-sounding audio.
还有更多 •••
相关职位

社招3年以上技术类-开发
1. 主导过LLM、VLM、ASR或TTS大模型预训练及微调语料数据建设工作,有丰富的数据交付经验; 2. 精通大规模分布式数据处理技术(如spark/flink/ray等),拥有从0到1搭建全模态数
更新于 2026-04-06北京|杭州
社招3年以上技术类-开发
1. 主导设计和落地过大规模AI语料数据处理平台或数据生产系统,具备从0到1或大规模演进语料数据体系的经验,能够从工程架构层面构建支撑语料数据全生命周期的系统能力; 2. 具备大规模数据处理经验,熟悉
更新于 2026-07-01北京|杭州
实习阿里巴巴2027
1.计算机科学、人工智能、软件工程、数学、统计学或相关专业本科及以上学历; 2.熟练掌握 Python,具备扎实的编程能力与工程实现经验,熟悉 Linux 开发环境及常用数据处理工具; 3.了解深度学
更新于 2026-05-19北京|杭州