腾讯微信-语音识别算法工程师
社招全职微信AI助手技术地点:北京状态:招聘
工作描述
任职要求 1.计算机、人工智能、语音信号处理等相关专业,硕士及以上学历优先; 2.具备流式语音识别模型研发经验,熟悉模型训练、流式解码、状态缓存和延迟优化; 3.熟悉 CTC、Transducer、Conformer、Transformer、语言模型融合等技术; 4.具备语音理解或上下文建模经验,熟悉热词增强、语义纠错、意图理解、实体识别或多轮对话建模中的一种或多种; 5.熟悉语音数据处理流程,具备数据清洗、伪标注、难例分析和效果归因能力; 6.熟练使用 PyTorch,具备较强的模型训练、实验设计和问题定位能力; 7.具备良好的工程意识,关注识别准确率、首字延迟、尾字延迟、实时率及线上稳定性。 加分项 1.有大规模流式 …
登录查看完整工作描述
微信扫码,1秒登录
包括英文材料
学历+
语音识别+
https://developer.nvidia.com/blog/essential-guide-to-automatic-speech-recognition-technology/
Over the past decade, AI-powered speech recognition systems have slowly become part of our everyday lives, from voice search to virtual assistants in contact centers, cars, hospitals, and restaurants.
缓存+
https://hackernoon.com/the-system-design-cheat-sheet-cache
The cache is a layer that stores a subset of data, typically the most frequently accessed or essential information, in a location quicker to access than its primary storage location.
https://www.youtube.com/watch?v=bP4BeUjNkXc
Caching strategies, Distributed Caching, Eviction Policies, Write-Through Cache and Least Recently Used (LRU) cache are all important terms when it comes to designing an efficient system with a caching layer.
https://www.youtube.com/watch?v=dGAgxozNWFE
Transformer+
https://huggingface.co/learn/llm-course/en/chapter1/4
Breaking down how Large Language Models work, visualizing how data flows through.
https://poloclub.github.io/transformer-explainer/
An interactive visualization tool showing you how transformer models work in large language models (LLM) like GPT.
https://www.youtube.com/watch?v=wjZofJX0v4M
Breaking down how Large Language Models work, visualizing how data flows through.
PyTorch+
https://datawhalechina.github.io/thorough-pytorch/
PyTorch是利用深度学习进行数据科学研究的重要工具,在灵活性、可读性和性能上都具备相当的优势,近年来已成为学术界实现深度学习算法最常用的框架。
https://www.youtube.com/watch?v=V_xro1bcAuA
Learn PyTorch for deep learning in this comprehensive course for beginners. PyTorch is a machine learning framework written in Python.
还有更多 •••
相关职位