蚂蚁金服蚂蚁集团-AI算法专家(搜索推荐方向)-AIRS
社招全职3年以上技术类-算法地点:北京 | 杭州状态:招聘
工作描述
任职要求 1. 深厚的算法积淀:精通搜索或推荐领域的基础理论。在深度学习、大规模分布式训练、特征工程等方面有深厚的工程实践经验,理解复杂链路下的系统性瓶颈; 2. 大模型实战经验:对 LLM / 强化学习有深刻理解,并有成功的业务落地经验。能够熟练进行模型微调(SFT)、强化学习对齐(RLHF)或 RAG 系统的架构设计;有多模态理解与生成(AIGC)经验者加分,能应用于素材理解与生成式创意优化; 3. 卓越的问题解决能力:具备极强的数据洞察力和实验设计能力,能够从复杂的业务现象中抽象出底层算法问题,并给出系统性的解决方案; 4. 行业影响力: (1)在相关领域顶级会议(如 KDD、NeurIPS、SIGIR、WSDM、ICDE 等)有高质量论文发表者优先; (2)在知名开源项目、高水平 K…
登录查看完整工作描述
微信扫码,1秒登录
包括英文材料
算法+
https://roadmap.sh/datastructures-and-algorithms
Step by step guide to learn Data Structures and Algorithms in 2025
https://www.hellointerview.com/learn/code
A visual guide to the most important patterns and approaches for the coding interview.
https://www.w3schools.com/dsa/
深度学习+
https://d2l.ai/
Interactive deep learning book with code, math, and discussions.
特征工程+
https://www.ibm.com/think/topics/feature-engineering
Feature engineering preprocesses raw data into a machine-readable format. It optimizes ML model performance by transforming and selecting relevant features.
https://www.kaggle.com/learn/feature-engineering
Better features make better models. Discover how to get the most out of your data.
大模型+
https://www.youtube.com/watch?v=xZDB1naRUlk
You will build projects with LLMs that will enable you to create dynamic interfaces, interact with vast amounts of text data, and even empower LLMs with the capability to browse the internet for research papers.
https://www.youtube.com/watch?v=zjkBMFhNj_g
强化学习+
https://cloud.google.com/discover/what-is-reinforcement-learning?hl=en
Reinforcement learning (RL) is a type of machine learning where an "agent" learns optimal behavior through interaction with its environment.
https://huggingface.co/learn/deep-rl-course/unit0/introduction
This course will teach you about Deep Reinforcement Learning from beginner to expert. It’s completely free and open-source!
https://www.kaggle.com/learn/intro-to-game-ai-and-reinforcement-learning
Build your own video game bots, using classic and cutting-edge algorithms.
SFT+
https://cameronrwolfe.substack.com/p/understanding-and-using-supervised
Understanding how SFT works from the idea to a working implementation...
RLHF+
[英文] What is RLHF?
https://aws.amazon.com/what-is/reinforcement-learning-from-human-feedback/
Reinforcement learning from human feedback (RLHF) is a machine learning (ML) technique that uses human feedback to optimize ML models to self-learn more efficiently.
https://www.ibm.com/think/topics/rlhf
Reinforcement learning from human feedback (RLHF) is a machine learning technique in which a “reward model” is trained with direct human feedback, then used to optimize the performance of an artificial intelligence agent through reinforcement learning.
还有更多 •••
相关职位
社招3年以上技术类-算法
职位要求 具备扎实的计算机、机器学习和算法基础,能够从问题本质出发分析和解决复杂问题,对搜索与大模型技术保持持续好奇。 具备优秀的编程能力,熟练使用 Python;具备 C/C++、分布式训练或高性能
更新于 2026-08-26北京|上海|杭州
社招2年以上云智能集团
1. 计算机、人工智能等相关专业,2年以上搜索、推荐、NLP、大模型等领域算法经验。 2. 熟练掌握机器学习/深度学习算法,熟悉deepspeed、megatron等大模型训练框架,在搜推/NLP/大
更新于 2026-01-22北京|杭州
社招2年以上技术类-算法
我们正在寻找两类核心人才,只要你符合其中之一,欢迎与我们共同定义未来: 方向一:搜索资深专家(向 AI 场景转型) ● 具备丰富的搜索系统研发经验, 熟悉搜索基础算法与工程技术,对搜索业务有深刻理解。
更新于 2026-09-14杭州|北京