阿里巴巴高德大模型算法实习生(LBS智能检索与Agent方向)
实习兼职阿里巴巴研究型实习生地点:北京状态:招聘
工作描述
任职要求 1. 在读本科、硕士或博士,计算机、人工智能、数学、电子信息、统计等相关专业,能够保证一段相对完整的实习投入。 2. 掌握机器学习与深度学习的基础原理,理解 Transformer 结构与注意力机制,了解预训练语言模型的基本范式;对信息检索或自然语言处理中的至少一个子方向有较系统的认识。 3. 熟练使用 Python,能独立完成数据处理、模型训练与实验分析;熟悉 PyTorch,了解 Hugging Face Transformers 生态的基本用法。 4. 具备扎实的工程基本功和调试能力,写得出结构清晰、可复现的实验代码;有能力读懂英文论文与开源项目源码。 5. 有良好的问题拆解与表达能力,遇到效果不达预期时能主动做归因分析,而不是只调参数;愿意在真实数据里反复打磨细节。 加分项:以下经历不是硬性门槛,但会显著加分: ● 有搜索、推荐、广告或问答系统的实习、科研或竞赛经历,理解过完整的召回—排序—重排链路 ● 做过 RAG 或知识库问答的完整项目,踩过文档切分、检索精度、幻觉抑制等实际问题 ● 有大模型微调实战经验,熟悉 TRL / PEFT / LLaMA-Factory / DeepSpeed / Megatron 等训练框架中的一种或多种 ● 熟悉向量数据库与检索引擎,如 FAISS、Milvus、Elasticsearch、Vespa 等 ● 有 vLLM、S…
登录查看完整工作描述
微信扫码,1秒登录
包括英文材料
机器学习+
https://www.youtube.com/watch?v=0oyDqO8PjIg
Learn about machine learning and AI with this comprehensive 11-hour course from @LunarTech_ai.
https://www.youtube.com/watch?v=i_LwzRVP7bg
Learn Machine Learning in a way that is accessible to absolute beginners.
https://www.youtube.com/watch?v=NWONeJKn6kc
Learn the theory and practical application of machine learning concepts in this comprehensive course for beginners.
https://www.youtube.com/watch?v=PcbuKRNtCUc
Learn about all the most important concepts and terms related to machine learning and AI.
深度学习+
https://d2l.ai/
Interactive deep learning book with code, math, and discussions.
Transformer+
https://huggingface.co/learn/llm-course/en/chapter1/4
Breaking down how Large Language Models work, visualizing how data flows through.
https://poloclub.github.io/transformer-explainer/
An interactive visualization tool showing you how transformer models work in large language models (LLM) like GPT.
https://www.youtube.com/watch?v=wjZofJX0v4M
Breaking down how Large Language Models work, visualizing how data flows through.
信息检索+
https://nlp.stanford.edu/IR-book/information-retrieval-book.html
Christopher D. Manning, Prabhakar Raghavan and Hinrich Schütze, Introduction to Information Retrieval, Cambridge University Press. 2008.
NLP+
https://www.youtube.com/watch?v=fNxaJsNG3-s&list=PLQY2H8rRoyvzDbLUZkbudP-MFQZwNmU4S
Welcome to Zero to Hero for Natural Language Processing using TensorFlow!
https://www.youtube.com/watch?v=R-AG4-qZs1A&list=PLeo1K3hjS3uuvuAXhYjV2lMEShq2UYSwX
Natural Language Processing tutorial for beginners series in Python.
https://www.youtube.com/watch?v=rmVRLeJRkl4&list=PLoROMvodv4rMFqRtEuo6SGjY4XbRIVRd4
The foundations of the effective modern methods for deep learning applied to NLP.
Python+
https://liaoxuefeng.com/books/python/introduction/index.html
中文,免费,零起点,完整示例,基于最新的Python 3版本。
https://www.learnpython.org/
a free interactive Python tutorial for people who want to learn Python, fast.
https://www.youtube.com/watch?v=K5KVEU3aaeQ
Master Python from scratch 🚀 No fluff—just clear, practical coding skills to kickstart your journey!
https://www.youtube.com/watch?v=rfscVS0vtbw
This course will give you a full introduction into all of the core concepts in python.
PyTorch+
https://datawhalechina.github.io/thorough-pytorch/
PyTorch是利用深度学习进行数据科学研究的重要工具,在灵活性、可读性和性能上都具备相当的优势,近年来已成为学术界实现深度学习算法最常用的框架。
https://www.youtube.com/watch?v=V_xro1bcAuA
Learn PyTorch for deep learning in this comprehensive course for beginners. PyTorch is a machine learning framework written in Python.
RAG+
https://www.youtube.com/watch?v=sVcwVQRHIc8
Learn how to implement RAG (Retrieval Augmented Generation) from scratch, straight from a LangChain software engineer.
大模型+
https://www.youtube.com/watch?v=xZDB1naRUlk
You will build projects with LLMs that will enable you to create dynamic interfaces, interact with vast amounts of text data, and even empower LLMs with the capability to browse the internet for research papers.
https://www.youtube.com/watch?v=zjkBMFhNj_g
LLaMA-Factory+
https://llamafactory.readthedocs.io/en/latest/
LLaMA Factory is an easy-to-use and efficient platform for training and fine-tuning large language models.
DeepSpeed+
https://www.youtube.com/watch?v=pDGI668pNg0
Megatron+
https://www.youtube.com/watch?v=hc0u4avAkuM
Faiss+
https://faiss.ai/index.html
Faiss is a library for efficient similarity search and clustering of dense vectors.
https://huggingface.co/learn/llm-course/en/chapter5/6
In this section we’ll use this information to build a search engine that can help us find answers to our most pressing questions about the library!
Milvus+
[英文] Tutorials Overview
https://milvus.io/docs/tutorials-overview.md
This page provides a list of tutorials for you to interact with Milvus.
https://www.baeldung.com/milvus-tutorial-intro
In this tutorial, we’ll explore Milvus, a highly scalable open-source vector database.
https://www.youtube.com/watch?v=7ejr_ZzU9jw
Discover the power of Milvus, an open-source vector database revolutionizing AI applications.
https://www.youtube.com/watch?v=Yhv19le0sBw
Vector databases have been trending recently as they power modern search, recommendations, and AI-driven applications.
ElasticSearch+
https://www.youtube.com/watch?v=a4HBKEda_F8
Learn about Elasticsearch with this comprehensive course designed for beginners, featuring both theoretical concepts and hands-on applications using Python (though applicable to any programming language). The course is structured in two parts: first covering essential Elasticsearch fundamentals including index management, document storage, text analysis, pipeline creation, search functionality, and advanced features like semantic search and embeddings; followed by a practical section where you'll build a real-world website using Elasticsearch as a search engine, working with the Astronomy Picture of the Day (APOD) dataset to implement features such as data cleaning pipelines, tokenization, pagination, and aggregations.
还有更多 •••
相关职位
实习阿里巴巴研究型实
学历背景 ꔷ 统招硕士及以上学历,计算机、统计学等相关专业优先。 算法能力 ꔷ 扎实的数据结构、机器学习、深度学习基础 ꔷ 熟悉树模型、深度学习、NLP/CV算法、强化学习、LLM大模型、多模态大模型
更新于 2026-08-13北京
社招3年以上技术类-算法
1、计算机科学、人工智能、机器人学、自动化或相关专业硕士及以上学历,博士优先; 2、在以下至少一个方向有丰富的经验和思考: 视觉-语言-导航大模型(VLN) 视觉-语言-动作大模型(VLA) 大型行为
更新于 2026-03-26北京
社招3年以上技术类-算法
岗位要求: 1、计算机、电子信息工程、自动化控制、数学、信息安全等相关专业背景,硕士及以上学历; 2、在机器学习或深度学习领域有实习或者项目经历,具备以下一个或多个方向的研究和应用经验,如多模态数据处
更新于 2026-08-13北京