北京雁栖湖应用数学研究院 北京雁栖湖应用数学研究院

  • 关于我们
    • 院长致辞
    • 理事会
    • 协作机构
    • 参观来访
  • 人员
    • 管理层
    • 科研人员
    • 博士后
    • 来访学者
    • 行政团队
    • 学术支持
  • 学术研究
    • 研究团队
    • 公开课
    • 讨论班
    • 期刊
  • 招生招聘
    • 教研人员
    • 博士后
    • 学生
  • 会议
    • 学术会议
    • 工作坊
    • 论坛
  • 学院生活
    • 住宿
    • 交通
    • 配套设施
    • 周边旅游
  • 新闻
    • 新闻动态
    • 通知公告
    • 资料下载
关于我们
院长致辞
理事会
协作机构
参观来访
人员
管理层
科研人员
博士后
来访学者
行政团队
学术支持
学术研究
研究团队
公开课
讨论班
期刊
招生招聘
教研人员
博士后
学生
会议
学术会议
工作坊
论坛
学院生活
住宿
交通
配套设施
周边旅游
新闻
新闻动态
通知公告
资料下载
清华大学 "求真书院"
清华大学丘成桐数学科学中心
清华三亚国际数学论坛
上海数学与交叉学科研究院
河套数学与交叉学科研究院
BIMSA > 随机过程
随机过程
This course studies stochastic processes with reinforcement, in which past observations influence future dynamics. The course is organized around three interconnected topics (see Syllabus) and a common question: How does feedback shape the long-term behavior of a stochastic system?

The course emphasizes representative examples, probabilistic ideas, and qualitative understanding rather than technically demanding proofs. Martingales, exchangeability, coupling, stochastic approximation, concentration inequalities, and hidden Markov structures will be introduced through the models in which they arise.

Each meeting consists of three 45-minute periods. Lectures will be combined with in-class problem sessions, during which the instructor and students work through examples, calculations, simulations, and selected proof arguments together.

As this is a non-credit course, there will be no formal examination or assessment. No regular take-home homework will be assigned. Optional reading and computational experiments may be suggested for interested students.
讲师
秦硕
日期
2026年09月04日 至 2027年01月08日
位置
Weekday Time Venue Online ID Password
周五 09:50 - 12:15 RUC - - -
修课要求
A solid undergraduate course in probability is recommended. Familiarity with conditional expectation, Markov chains, and basic convergence concepts is helpful. Martingale and concentration methods needed in the course will be reviewed or introduced as they arise. The course is intended primarily for senior undergraduate students and beginning graduate students interested in probability and stochastic processes. It is also recommended that the audience take the course “Online Learning” by Yuval Peres in parallel with this course
课程大纲
1. Urn Models (Weeks 1–4): Classical Pólya urns, exchangeability and random limits, generalized and nonlinear urns, stochastic approximation, and phase transitions caused by different reinforcement strengths.

2. Reinforced Random Walks (Weeks 5–10): Vertex-reinforced, edge-reinforced, and step-reinforced random walks, with emphasis on one-dimensional models. Topics include local times, strong reinforcement and trapping, localization of vertex-reinforced walks, partial exchangeability and random-environment representations, spatial Markov descriptions of local-time processes, the elephant random walk and its phase transition, and random recursive tree representations. Results on higher-dimensional lattices and general graphs will be discussed at a survey level.

3. Stochastic Multi-Armed Bandits (Weeks 11–15): Stochastic bandit models, regret, failures of naïve reinforcement and greedy policies, explore-then-commit, Upper Confidence Bound algorithms, information-theoretic lower bounds, and Thompson sampling. Particular attention will be paid to the relationship between reinforcement, learning, and the exploration–exploitation tradeoff.

4. Comparative Case Studies (Week 16): A comparison of random limits, localization, lock-in, and controlled exploration across urns, reinforced random walks, and stochastic bandits.
参考资料
R. Pemantle, A Survey of Random Processes with Reinforcement, Probability Surveys 4 (2007), 1–79.
T. Lattimore and C. Szepesvári, Bandit Algorithms, Cambridge University Press, 2020.
A. Slivkins, Introduction to Multi-Armed Bandits, Foundations and Trends in Machine Learning, 2019.

Additional notes and selected research articles will be provided by the instructor.
听众
Advanced Undergraduate , Graduate
视频公开
公开
笔记公开
公开
语言
中文
讲师介绍
Shuo Qin has been the first Chern Instructor at BIMSA. He obtained a Ph.D. in mathematics in 2024 from New York University under the supervision of Prof. Pierre Tarrès. His work is in probability theory, especially in random processes with memory or reinforcement.
北京雁栖湖应用数学研究院
CONTACT

No. 544, Hefangkou Village Huaibei Town, Huairou District Beijing 101408

北京市怀柔区 河防口村544号
北京雁栖湖应用数学研究院 101408

Tel. 010-60661855 Tel. 010-60661855
Email. administration@bimsa.cn

版权所有 © 北京雁栖湖应用数学研究院

京ICP备2022029550号-1

京公网安备11011602001060 京公网安备11011602001060