亲爱的研友该休息了!由于当前在线用户较少,发布求助请尽量完整地填写文献信息,科研通机器人24小时在线,伴您度过漫漫科研夜!身体可是革命的本钱,早点休息,好梦!

Modeling collective motion for fish schooling via multi-agent reinforcement learning

强化学习 运动(物理) 集体运动 基于Agent的模型 计算机科学 人工智能 人工神经网络 过程(计算) 集体行为 钢筋 先验与后验 动力学(音乐) 心理学 社会心理学 社会学 认识论 操作系统 哲学 教育学 人类学
作者
Xin Wang,Shuo Liu,Yifan Yu,Shengzhi Yue,Ying Liu,Fumin Zhang,Yuanshan Lin
出处
期刊:Ecological Modelling [Elsevier BV]
卷期号:477: 110259-110259 被引量:8
标识
DOI:10.1016/j.ecolmodel.2022.110259
摘要

Complex collective motion patterns can emerge from very simple local interactions among individual agents. However, it is still unclear how and why the interactions among individuals lead to the emergence of collective motion. Modeling is an effective way to understand the mechanisms that govern collective animal motions. In this work, to avoid imposing fixed sets of rules on collective motion models a priori as classical approaches do, we propose a new method of modeling collective motion for fish schooling via multi-agent reinforcement learning. We model each fish individual as an artificial learning agent, whose policy is acquired by using mean field Q-learning (MFQ). The observation of each fish agent is represented as a multi-channel image, where each channel describes a different feature, such as an agent's position or an agent's orientation. The policy of an agent is approximated with a neural network trained with the MFQ algorithm, during which, agents are rewarded (or penalized) according to the number of neighbors and consecutive collisions between individuals. We study the dynamics of collective motion that emerge from the learned policy. The experimental results show that the learned policy can produce collective motion in groups of various sizes. In addition, three different collective motion patterns observed in nature emerged during the training process. The learned policy can help us gain new insight into how and why individual interactions lead to collective motion. This study also demonstrates that multi-agent reinforcement learning has great potential to be a new approach for analysis and modeling of collective motion.

科研通智能强力驱动
Strongly Powered by AbleSci AI
科研通是完全免费的文献互助平台,具备全网最快的应助速度,最高的求助完成率。 对每一个文献求助,科研通都将尽心尽力,给求助人一个满意的交代。
实时播报
1秒前
1秒前
学术孤儿应助科研通管家采纳,获得10
1秒前
大模型应助科研通管家采纳,获得10
1秒前
香蕉觅云应助科研通管家采纳,获得10
1秒前
呼呼不爱噜噜应助白落提采纳,获得10
2秒前
Orange应助一颗石头鱼采纳,获得10
3秒前
4秒前
chenchen发布了新的文献求助30
6秒前
完美世界应助Joyi采纳,获得10
7秒前
努力TOP完成签到 ,获得积分10
7秒前
14秒前
卑微学术人完成签到 ,获得积分10
15秒前
17秒前
18秒前
领导范儿应助wenwen采纳,获得20
19秒前
美丽涑发布了新的文献求助200
19秒前
Prof.Z发布了新的文献求助10
19秒前
白落提完成签到,获得积分10
21秒前
欢喜以莲发布了新的文献求助10
21秒前
Joyi发布了新的文献求助10
22秒前
ASAMIMI发布了新的文献求助10
23秒前
大模型应助落寞的笑寒采纳,获得10
25秒前
充电宝应助铠甲勇士采纳,获得10
29秒前
30秒前
高高的夏波完成签到,获得积分10
37秒前
41秒前
外向的问儿完成签到 ,获得积分10
41秒前
田様应助英俊纸飞机采纳,获得10
43秒前
llj完成签到 ,获得积分10
43秒前
斯文的白玉完成签到 ,获得积分0
45秒前
铠甲勇士发布了新的文献求助10
47秒前
欧阳懿完成签到 ,获得积分10
48秒前
ASAMIMI发布了新的文献求助10
49秒前
CAT完成签到,获得积分10
52秒前
等待冰之完成签到 ,获得积分10
52秒前
科研通AI6.3应助Joyi采纳,获得10
55秒前
1分钟前
1分钟前
十亩间发布了新的文献求助10
1分钟前
高分求助中
Markov Chain Monte Carlo 10000
(应助此贴封号)【重要!!请各用户(尤其是新用户)详细阅读】【科研通的精品贴汇总】 10000
Common Foundations of American and East Asian Modernisation: From Alexander Hamilton to Junichero Koizumi 5000
Pediatric Dermoscopy Trichoscopy & Onychoscopy 1000
悉尼大学博士学位论文,题目:Modelling and testing of one-sided stitched laminated composites. 作者:Kristopher P. Plain 700
Matrix Methods in Data Mining and Pattern Recognition Second Edition 610
International Security Studies and Technology :Approaches, Assessments, and Frontiers 500
热门求助领域 (近24小时)
化学 材料科学 医学 生物 纳米技术 工程类 有机化学 化学工程 生物化学 计算机科学 内科学 物理 复合材料 催化作用 细胞生物学 无机化学 光电子学 物理化学 电极 基因
热门帖子
关注 科研通微信公众号,转发送积分 7571452
求助须知:如何正确求助?哪些是违规求助? 9151007
关于积分的说明 19572628
捐赠科研通 7156493
什么是DOI,文献DOI怎么找? 3264048
关于科研通互助平台的介绍 2429357
邀请新用户注册赠送积分活动 2254200