拼多多【商业化】大模型应用开发工程师
社招全职3年以上技术类地点:上海状态:招聘
任职要求
1.计算机、数学等相关专业背景,三年以上互联网相关行业从业经验; 2.具有高并发分布式系统的设计能力,熟悉微服务、缓存、消息队列的常用组件。 3.熟悉机器学习和深度学习的基本原理,做过算法工程相关工作; 4.熟练使用python语言(或java,C++)。有大模型微调、应用服务开发(GPU部署、Langchain、大模型API)经验者优先。 5.有文档问答类项目经验者优先。
工作职责
1.负责拼多多核心电商搜索、推荐、商业化场景大模型Agent的开发与优化工作,支持业务场景(如AI交互式对话搜索,智能导购,图文创意生成、数字人等)高效落地; 2.负责大模型Agent、RAG系统全流程研发工作,结合业务需要,与算法团队搭档,推进 AIGC 项目在各个场景落地以及效果的持续优化。 3.设计高并发分布式架构,优化检索-生成链路性能,解决高并发环境下的延迟问题,保障服务高性能和SLA。 4.探索大模型在电商推荐、搜索、广告投放等场景的落地,推进技术、产品、数据的闭环协同。
包括英文材料
高并发+
https://www.baeldung.com/concurrency-principles-patterns
In this tutorial, we’ll discuss some of the design principles and patterns that have been established over time to build highly concurrent applications.
https://www.baeldung.com/java-concurrency
Handling concurrency in an application can be a tricky process with many potential pitfalls. A solid grasp of the fundamentals will go a long way to help minimize these issues.
https://www.oreilly.com/library/view/concurrency-in-go/9781491941294/
You’ll understand how Go chooses to model concurrency, what issues arise from this model, and how you can compose primitives within this model to solve problems.
https://www.oreilly.com/library/view/modern-concurrency-in/9781098165406/
With this book, you'll explore the transformative world of Java 21's key feature: virtual threads.
https://www.youtube.com/watch?v=qyM8Pi1KiiM
https://www.youtube.com/watch?v=wEsPL50Uiyo
分布式系统+
https://www.distributedsystemscourse.com/
The home page of a free online class in distributed systems.
https://www.youtube.com/watch?v=7VbL89mKK3M&list=PLOE1GTZ5ouRPbpTnrZ3Wqjamfwn_Q5Y9A
微服务+
https://learn.microsoft.com/en-us/training/modules/dotnet-microservices/
Microservice applications are composed of small, independently versioned, and scalable customer-focused services that communicate with each other by using standard protocols and well-defined interfaces.
https://microservices.io/
Microservices - also known as the microservice architecture - is an architectural style that structures an application as a collection of two or more services.
https://spring.io/microservices
Building small, self-contained, ready to run applications can bring great flexibility and added resilience to your code.
https://www.ibm.com/think/topics/microservices
Microservices, or microservices architecture, is a cloud-native architectural approach in which a single application is composed of many loosely coupled and independently deployable smaller components or services.
https://www.youtube.com/watch?v=CqCDOosvZIk
https://www.youtube.com/watch?v=hmkF77F9TLw
Learn about software system design and microservices.
缓存+
https://hackernoon.com/the-system-design-cheat-sheet-cache
The cache is a layer that stores a subset of data, typically the most frequently accessed or essential information, in a location quicker to access than its primary storage location.
https://www.youtube.com/watch?v=bP4BeUjNkXc
Caching strategies, Distributed Caching, Eviction Policies, Write-Through Cache and Least Recently Used (LRU) cache are all important terms when it comes to designing an efficient system with a caching layer.
https://www.youtube.com/watch?v=dGAgxozNWFE
消息队列+
https://www.youtube.com/watch?v=xErwDaOc-Gs
机器学习+
https://www.youtube.com/watch?v=0oyDqO8PjIg
Learn about machine learning and AI with this comprehensive 11-hour course from @LunarTech_ai.
https://www.youtube.com/watch?v=i_LwzRVP7bg
Learn Machine Learning in a way that is accessible to absolute beginners.
https://www.youtube.com/watch?v=NWONeJKn6kc
Learn the theory and practical application of machine learning concepts in this comprehensive course for beginners.
https://www.youtube.com/watch?v=PcbuKRNtCUc
Learn about all the most important concepts and terms related to machine learning and AI.
深度学习+
https://d2l.ai/
Interactive deep learning book with code, math, and discussions.
算法+
https://roadmap.sh/datastructures-and-algorithms
Step by step guide to learn Data Structures and Algorithms in 2025
https://www.hellointerview.com/learn/code
A visual guide to the most important patterns and approaches for the coding interview.
https://www.w3schools.com/dsa/
Python+
https://liaoxuefeng.com/books/python/introduction/index.html
中文,免费,零起点,完整示例,基于最新的Python 3版本。
https://www.learnpython.org/
a free interactive Python tutorial for people who want to learn Python, fast.
https://www.youtube.com/watch?v=K5KVEU3aaeQ
Master Python from scratch 🚀 No fluff—just clear, practical coding skills to kickstart your journey!
https://www.youtube.com/watch?v=rfscVS0vtbw
This course will give you a full introduction into all of the core concepts in python.
Java+
https://www.youtube.com/watch?v=eIrMbAQSU34
Master Java – a must-have language for software development, Android apps, and more! ☕️ This beginner-friendly course takes you from basics to real coding skills.
C+++
https://www.learncpp.com/
LearnCpp.com is a free website devoted to teaching you how to program in modern C++.
https://www.youtube.com/watch?v=ZzaPdXTrSb8
大模型+
https://www.youtube.com/watch?v=xZDB1naRUlk
You will build projects with LLMs that will enable you to create dynamic interfaces, interact with vast amounts of text data, and even empower LLMs with the capability to browse the internet for research papers.
https://www.youtube.com/watch?v=zjkBMFhNj_g
LangChain+
https://python.langchain.com/docs/tutorials/
New to LangChain or LLM app development in general? Read this material to quickly get up and running building your first applications.
https://www.freecodecamp.org/news/beginners-guide-to-langchain/
LangChain is a popular framework for creating LLM-powered apps.
相关职位
社招ACG
-负责百度知识管理平台的基础服务架构、相关组件与模块设计与开发 -提升百度知识管理平台商业化服务稳定性,确保业务高可用 -构建百度知识管理平台私有云/公有云交付能力 -提升交付质量与效率,支持标准化、规模化大客户项目落地 -参与百度知识管理平台商业化业务开放能力建设;满足各类第三方生态接入,满足客户的二次开发需求
更新于 2025-04-10
社招技术类
1)负责拼多多核心电商搜索、推荐、商业化场景大模型AIGC算法的开发与优化,支持业务场景(如AI交互式对话搜索,智能导购,图文创意生成、数字人等)高效落地; 2)负责大模型Agent、RAG系统全流程研发工作,包括样本标注,数据处理,模型训练(PreTrain、SFT、RL等),Prompt Engineer,WorkFlow设计与开发,评价指标设计; 3)负责Diffusion、Flux等算法在电商图像、视频生成领域的算法优化,追踪前沿技术,持续提升大模型内容生成的质量,赋能业务创新。
更新于 2025-09-01
社招3年以上技术-开发
1、深度参与蚂蚁国际化战略,与生态伙伴合作,一带一路技术出海,业务覆盖全球; 2、和海外技术生态对接,主导技术难题攻关,面临跨洲跨国家的技术挑战;持续提升产品的扩展性,降低技术输出的成本; 3、从海外钱包经营者的视角看清楚端增长逻辑,设计数字化运营系统,助力经营决策; 4、设计并建设商业化能力,帮助海外钱包客户增加收入类型与规模; 5、负责基于大语言模型创新应用的技术方案和系统设计评审;把握复杂系统的设计,确保系统的架构质量;对现存或未来系统进行宏观的思考,规划形成统一的框架、平台或组件; 6、负责基于大语言模型创新应用的开发和利用RAG、SFT、RL等技术优化应用效果。
更新于 2025-06-12