
广发证券AI架构师
社招全职5年以上算法地点:广州状态:招聘
工作描述
任职要求 1、计算机科学、人工智能、数学、统计学或相关领域的硕士或博士学位; 2、5年以上AI/机器学习领域从业经验,其中2年以上AI架构设计或技术负责人经验; 3、深度理解大模型技术栈,熟悉预训练、微调(SFT)、强化学习(RLHF/PPO/DPO)、推理优化等核心技术; 4、熟练掌握PyTorch、TensorFlow等深度学习框架,了解HuggingFace生态; 5、具备大规模行业训练语料库建设经验,熟悉指令集设计、复杂数据标注与合成方法; 6、具备模型工程化落地能力,了解推理加速、模型压缩与部署技术; 7、具备复杂系统架构设计能力,能够平…
登录查看完整工作描述
微信扫码,1秒登录
包括英文材料
学历+
机器学习+
https://www.youtube.com/watch?v=0oyDqO8PjIg
Learn about machine learning and AI with this comprehensive 11-hour course from @LunarTech_ai.
https://www.youtube.com/watch?v=i_LwzRVP7bg
Learn Machine Learning in a way that is accessible to absolute beginners.
https://www.youtube.com/watch?v=NWONeJKn6kc
Learn the theory and practical application of machine learning concepts in this comprehensive course for beginners.
https://www.youtube.com/watch?v=PcbuKRNtCUc
Learn about all the most important concepts and terms related to machine learning and AI.
系统设计+
https://roadmap.sh/system-design
Everything you need to know about designing large scale systems.
https://www.youtube.com/watch?v=F2FmTdLtb_4
This complete system design tutorial covers scalability, reliability, data handling, and high-level architecture with clear explanations, real-world examples, and practical strategies.
大模型+
https://www.youtube.com/watch?v=xZDB1naRUlk
You will build projects with LLMs that will enable you to create dynamic interfaces, interact with vast amounts of text data, and even empower LLMs with the capability to browse the internet for research papers.
https://www.youtube.com/watch?v=zjkBMFhNj_g
SFT+
https://cameronrwolfe.substack.com/p/understanding-and-using-supervised
Understanding how SFT works from the idea to a working implementation...
强化学习+
https://cloud.google.com/discover/what-is-reinforcement-learning?hl=en
Reinforcement learning (RL) is a type of machine learning where an "agent" learns optimal behavior through interaction with its environment.
https://huggingface.co/learn/deep-rl-course/unit0/introduction
This course will teach you about Deep Reinforcement Learning from beginner to expert. It’s completely free and open-source!
https://www.kaggle.com/learn/intro-to-game-ai-and-reinforcement-learning
Build your own video game bots, using classic and cutting-edge algorithms.
RLHF+
[英文] What is RLHF?
https://aws.amazon.com/what-is/reinforcement-learning-from-human-feedback/
Reinforcement learning from human feedback (RLHF) is a machine learning (ML) technique that uses human feedback to optimize ML models to self-learn more efficiently.
https://www.ibm.com/think/topics/rlhf
Reinforcement learning from human feedback (RLHF) is a machine learning technique in which a “reward model” is trained with direct human feedback, then used to optimize the performance of an artificial intelligence agent through reinforcement learning.
还有更多 •••
相关职位

社招1年以上高级技术职位
1. 方案设计: 洞察行业机会点,将公司大模型技术、Agent框架转化为具备竞争力的商业化解决方案。需对方案的商业可行性、ROI及市场竞争力负责; 2. 端到端交付: 负责AI项目从POC到正式交付的
更新于 2023-02-23深圳|上海|北京

实习
1.针对B端客户,负责AIGC模型产品的线上效果评估,协助算法改进业务场景效果; 2.针对B端客户,协助设计prompt模板,协助客户优化提示词,能根据技术迭代与客户需求找到AI产品下的创作的新方法;
更新于 2025-11-18北京|深圳|上海
社招3年以上公司事务平台
法律/法律科技/AI方向5年+经验,有互联网法务场景经验优先 熟悉LLM应用、RAG、Tool Use、多轮任务流等技术 能将法律判断抽象为可执行的系统结构 重视准确性、溯源、权限控制与审计留痕 需提
更新于 2026-04-17北京
