蚂蚁金服蚂蚁国际-大模型算法工程师-金融算法
社招全职3年以上技术类-算法地点:深圳状态:招聘
工作描述
任职要求 1. 熟练掌握大模型的核心算法,深入理解 Qwen、DeepSeek、Llama 等主流开源大模型的架构设计与关键技术细节;具备大模型预训练、后训练(如指令微调、对齐优化)及基于强化学习的训练(如 RLHF)等实战经验者优先; 2. 具备扎实的编程能力,熟练使用 PyTorch、Hugging Face Transformers、Megatron-LM、SG…
登录查看完整工作描述
微信扫码,1秒登录
包括英文材料
大模型+
https://www.youtube.com/watch?v=xZDB1naRUlk
You will build projects with LLMs that will enable you to create dynamic interfaces, interact with vast amounts of text data, and even empower LLMs with the capability to browse the internet for research papers.
https://www.youtube.com/watch?v=zjkBMFhNj_g
算法+
https://roadmap.sh/datastructures-and-algorithms
Step by step guide to learn Data Structures and Algorithms in 2025
https://www.hellointerview.com/learn/code
A visual guide to the most important patterns and approaches for the coding interview.
https://www.w3schools.com/dsa/
Llama+
https://github.com/LlamaFamily/Llama-Chinese
Llama中文社区,实时汇总最新Llama学习资料,构建最好的中文Llama大模型开源生态,完全开源可商用。
https://www.llama.com/docs/overview/
This guide provides information and resources to help you set up Llama including how to access the model, hosting, how-to and integration guides.
系统设计+
https://roadmap.sh/system-design
Everything you need to know about designing large scale systems.
https://www.youtube.com/watch?v=F2FmTdLtb_4
This complete system design tutorial covers scalability, reliability, data handling, and high-level architecture with clear explanations, real-world examples, and practical strategies.
强化学习+
https://cloud.google.com/discover/what-is-reinforcement-learning?hl=en
Reinforcement learning (RL) is a type of machine learning where an "agent" learns optimal behavior through interaction with its environment.
https://huggingface.co/learn/deep-rl-course/unit0/introduction
This course will teach you about Deep Reinforcement Learning from beginner to expert. It’s completely free and open-source!
https://www.kaggle.com/learn/intro-to-game-ai-and-reinforcement-learning
Build your own video game bots, using classic and cutting-edge algorithms.
RLHF+
[英文] What is RLHF?
https://aws.amazon.com/what-is/reinforcement-learning-from-human-feedback/
Reinforcement learning from human feedback (RLHF) is a machine learning (ML) technique that uses human feedback to optimize ML models to self-learn more efficiently.
https://www.ibm.com/think/topics/rlhf
Reinforcement learning from human feedback (RLHF) is a machine learning technique in which a “reward model” is trained with direct human feedback, then used to optimize the performance of an artificial intelligence agent through reinforcement learning.
PyTorch+
https://datawhalechina.github.io/thorough-pytorch/
PyTorch是利用深度学习进行数据科学研究的重要工具,在灵活性、可读性和性能上都具备相当的优势,近年来已成为学术界实现深度学习算法最常用的框架。
https://www.youtube.com/watch?v=V_xro1bcAuA
Learn PyTorch for deep learning in this comprehensive course for beginners. PyTorch is a machine learning framework written in Python.
还有更多 •••
相关职位
校招金融服务平台
【任职资格】 1.优秀的探索与创新能力,在ACL/EMNLP/NAACL/NeurIPS/ICML/ICLR等顶级会议上发表论文者优先; 2.扎实的算法功底,在ACM/ICPC、NOI/IOI、Top
更新于 2026-06-03北京|上海

校招AI 算法类
岗位要求: 1. 2026 届硕士及以上学历,计算机/人工智能/软件工程/信息安全/金融工程等相关专业优先 2. 具备优秀的问题分析与解决能力,能深入解决多模态/智能体在**训练、推理、评测与落地**
杭州
社招3年以上微信支付-基础平
1.熟练掌握PyTorch、Swift、VeRL等一种或多种深度学习框架,具备多模态模型(如VLM)开发与调优经验; 2.熟悉多模态大模型训练技术,包括增量预训练(CPT)、有监督微调(SFT)、强
更新于 2026-06-08深圳