阿里巴巴研究型实习生-高性能算子生成 Agent 研发
实习兼职阿里巴巴研究型实习生地点:杭州状态:招聘
工作描述
任职要求 我们期望你是: 1. 计算机科学、电子工程、人工智能或相关领域博士研究生。 2. 具备大语言模型及 Agent 系统开发经验,熟悉 Prompt Engineering、模型评测、RAG/向量检索及常见 Agent 框架;有代码生成相关项目经验者优先。 3. 熟悉模型后训练方法,具备 SFT、RLHF/RLAIF、DPO/PPO/GRPO 等一种或多种方法的实践经验,能够围绕业务目标构建训练数据、奖励机制与评测体系。 4. 熟悉 AI 芯片架构(如 GPU/NPU/TPU),掌握 CUDA、Triton、Cute、TilieLang 等至少一种算子编程技术,熟悉GPU常用优化方法。 5. 具备多智能体系统开发经验,熟悉任务规划、工具调用、反思优化、分布式协作等机制;有 AutoGen、LangGraph、…
登录查看完整工作描述
微信扫码,1秒登录
包括英文材料
AI agent+
https://www.ibm.com/think/ai-agents
Your one-stop resource for gaining in-depth knowledge and hands-on applications of AI agents.
Prompt+
https://cloud.google.com/vertex-ai/generative-ai/docs/learn/prompts/introduction-prompt-design
A prompt is a natural language request submitted to a language model to receive a response back.
https://learn.microsoft.com/en-us/azure/ai-foundry/openai/concepts/prompt-engineering
These techniques aren't recommended for reasoning models like gpt-5 and o-series models.
https://www.youtube.com/watch?v=LWiMwhDZ9as
Learn and master the fundamentals of Prompt Engineering and LLMs with this 5-HOUR Prompt Engineering Crash Course!
RAG+
https://www.youtube.com/watch?v=sVcwVQRHIc8
Learn how to implement RAG (Retrieval Augmented Generation) from scratch, straight from a LangChain software engineer.
SFT+
https://cameronrwolfe.substack.com/p/understanding-and-using-supervised
Understanding how SFT works from the idea to a working implementation...
RLHF+
[英文] What is RLHF?
https://aws.amazon.com/what-is/reinforcement-learning-from-human-feedback/
Reinforcement learning from human feedback (RLHF) is a machine learning (ML) technique that uses human feedback to optimize ML models to self-learn more efficiently.
https://www.ibm.com/think/topics/rlhf
Reinforcement learning from human feedback (RLHF) is a machine learning technique in which a “reward model” is trained with direct human feedback, then used to optimize the performance of an artificial intelligence agent through reinforcement learning.
CUDA+
https://developer.nvidia.com/blog/even-easier-introduction-cuda/
This post is a super simple introduction to CUDA, the popular parallel computing platform and programming model from NVIDIA.
https://www.youtube.com/watch?v=86FAWCzIe_4
Lean how to program with Nvidia CUDA and leverage GPUs for high-performance computing and deep learning.
Triton Inference Server+
https://docs.nvidia.com/deeplearning/triton-inference-server/user-guide/docs/index.html
Triton Inference Server is an open source inference serving software that streamlines AI inferencing.
还有更多 •••
相关职位
实习阿里巴巴研究型实
1.信息安全、网络空间安全等相关专业的硕士生/博士生; 2.熟悉隐私保护方向前沿技术,对多方安全计算 (MPC) 、联邦学习、隐私保护机器学习 (PPML)等至少一个领域有深入理解; 3.满足以下一个
更新于 2026-03-20北京
实习阿里巴巴研究型实
1. 扎实的工程能力,优良的编程风格,熟悉C++语言和常用设计模式,具备复杂系统的调试与开发能力; 2. 优秀的沟通表达能力、团队合作意识和经验;具备快速学习的能力,以及深入钻研技术问题的决心; 3.
更新于 2026-03-17深圳
实习阿里巴巴研究型实
1.计算机或相关方向博士、硕士在读; 2.在数据库、人工智能、存储、文件系统、操作系统、网络等领域有顶会论文发表经验; 3.有系统研发能力,并善于与团队交流合作; 4.富有创造力,有面向落地场景解决实
更新于 2026-03-17杭州