阿里云阿里云智能-大模型算法专家-杭州
社招全职5年以上云智能集团地点:杭州状态:招聘
工作描述
任职要求 1. 具备5年以上算法经验,参与过完整的大模型相关项目经历,有CV、视频、多模态相关模型+云计算对口行业实战算法项目经验加分; 2. 扎实的算法基础,熟悉大模型的prompt、微调、RLHF、Agent RL等大模型训练技术, 熟练掌握至少一种深度学习框架,如PyTorch、TensorFlow等; 3. 自驱力强,有良好的协同能力,具备独立承担单一算法…
登录查看完整工作描述
微信扫码,1秒登录
包括英文材料
算法+
https://roadmap.sh/datastructures-and-algorithms
Step by step guide to learn Data Structures and Algorithms in 2025
https://www.hellointerview.com/learn/code
A visual guide to the most important patterns and approaches for the coding interview.
https://www.w3schools.com/dsa/
大模型+
https://www.youtube.com/watch?v=xZDB1naRUlk
You will build projects with LLMs that will enable you to create dynamic interfaces, interact with vast amounts of text data, and even empower LLMs with the capability to browse the internet for research papers.
https://www.youtube.com/watch?v=zjkBMFhNj_g
Prompt+
https://cloud.google.com/vertex-ai/generative-ai/docs/learn/prompts/introduction-prompt-design
A prompt is a natural language request submitted to a language model to receive a response back.
https://learn.microsoft.com/en-us/azure/ai-foundry/openai/concepts/prompt-engineering
These techniques aren't recommended for reasoning models like gpt-5 and o-series models.
https://www.youtube.com/watch?v=LWiMwhDZ9as
Learn and master the fundamentals of Prompt Engineering and LLMs with this 5-HOUR Prompt Engineering Crash Course!
RLHF+
[英文] What is RLHF?
https://aws.amazon.com/what-is/reinforcement-learning-from-human-feedback/
Reinforcement learning from human feedback (RLHF) is a machine learning (ML) technique that uses human feedback to optimize ML models to self-learn more efficiently.
https://www.ibm.com/think/topics/rlhf
Reinforcement learning from human feedback (RLHF) is a machine learning technique in which a “reward model” is trained with direct human feedback, then used to optimize the performance of an artificial intelligence agent through reinforcement learning.
还有更多 •••
相关职位
社招3年以上技术类-算法
1. 人工智能、计算机、数学、统计等相关专业的硕士或以上学历; 2. 扎实的深度学习算法基础,熟悉时间序列信号预测、对抗学习、强化学习中的一种或多种; 3. 具备 LLM 相关基础,有 RL、SFT
更新于 2026-06-17杭州
社招4年以上技术类-算法
1.具有以Agent、Agentic RAG为基础的智能助手类产品的实际研发工作经验,有大规模项目案例落地经验优先。具备大语言模型从数据集构建到产品落地的经验,具备优秀的算法抽象能力,能够将业务场景问
更新于 2026-07-21上海|杭州
社招4年以上
1、计算机或相关专业本科以上学历,硕士博士优先; 2、具备NLP大模型算法知识,有大模型相关使用经验,做过对话机器人的优先; 3、具有良好的英文阅读能力,可以快速理解前沿论文和技术文档并评测效果; 4
更新于 2026-02-05杭州
