阿里巴巴算法工程师-医疗大模型
实习兼职阿里巴巴2027届实习生地点:杭州状态:招聘
工作描述
任职要求 1. 熟悉先进大模型的算法原理,有丰富的大模型训练(pre-train/post-train)或数据样本合成方面的经验或技巧优先; 2. 有大模型强化学习训练上有经验者优先,包括但不限于RFT、RLVR、PPO、GRPO等; 3. 在大模型方向有高质量(ACL、EM…
登录查看完整工作描述
微信扫码,1秒登录
包括英文材料
大模型+
https://www.youtube.com/watch?v=xZDB1naRUlk
You will build projects with LLMs that will enable you to create dynamic interfaces, interact with vast amounts of text data, and even empower LLMs with the capability to browse the internet for research papers.
https://www.youtube.com/watch?v=zjkBMFhNj_g
算法+
https://roadmap.sh/datastructures-and-algorithms
Step by step guide to learn Data Structures and Algorithms in 2025
https://www.hellointerview.com/learn/code
A visual guide to the most important patterns and approaches for the coding interview.
https://www.w3schools.com/dsa/
强化学习+
https://cloud.google.com/discover/what-is-reinforcement-learning?hl=en
Reinforcement learning (RL) is a type of machine learning where an "agent" learns optimal behavior through interaction with its environment.
https://huggingface.co/learn/deep-rl-course/unit0/introduction
This course will teach you about Deep Reinforcement Learning from beginner to expert. It’s completely free and open-source!
https://www.kaggle.com/learn/intro-to-game-ai-and-reinforcement-learning
Build your own video game bots, using classic and cutting-edge algorithms.
还有更多 •••
相关职位
实习阿里巴巴2027
1.熟悉先进多模态大模型的算法原理,有丰富的多模态大模型训练(pre-train/post-train)或数据样本合成方面的经验或技巧优先; 2.有大模型强化学习训练上有经验者优先,包括但不限于RFT
更新于 2026-03-17北京|杭州
社招3年以上技术类-开发
1. 扎实的代码能力、工程能力、数据结构和基础算法功底;熟练Python或JAVA; 2. 熟悉NLP、CV、RL相关的算法和技术,熟悉大模型训练; 3. 有医疗大模型经验优先,有参与过大影响力模型项
更新于 2026-08-13北京|上海|杭州
实习阿里巴巴研究型实
1、硕士研究生及以上在读,计算机、人工智能、软件工程、信息安全、统计、数学等相关专业; 2、具备良好的编程基础和计算机视觉、机器学习等算法方向的研究和实践经验,能熟练运用pytorch等深度学习框架解
更新于 2026-07-02杭州