Multi-objective Multiagent Credit Assignment Through Difference Rewards in Reinforcement Learning

Yliniemi, Logan; Tumer, Kagan

doi:10.1007/978-3-319-13563-2_35

Logan Yliniemi²⁷ &
Kagan Tumer²⁷

Part of the book series: Lecture Notes in Computer Science ((LNTCS,volume 8886))

Included in the following conference series:

Asia-Pacific Conference on Simulated Evolution and Learning

2866 Accesses
6 Citations

Abstract

Multiagent systems have had a powerful impact on the real world. Many of the systems it studies (air traffic, satellite coordination, rover exploration) are inherently multi-objective, but they are often treated as single-objective problems within the research. A very important concept within multiagent systems is that of credit assignment: clearly quantifying an individual agent’s impact on the overall system performance. In this work we extend the concept of credit assignment into multi-objective problems, broadening the traditional multiagent learning framework to account for multiple objectives. We show in two domains that by leveraging established credit assignment principles in a multi-objective setting, we can improve performance by (i) increasing learning speed by up to 10x (ii) reducing sensitivity to unmodeled disturbances by up to 98.4% and (iii) producing solutions that dominate all solutions discovered by a traditional team-based credit assignment schema. Our results suggest that in a multiagent multi-objective problem, proper credit assignment is as important to performance as the choice of multi-objective algorithm.

This is a preview of subscription content, log in via an institution to check access.

Access this chapter

Log in via an institution

Chapter: USD 29.95; Price excludes VAT (USA)

eBook: USD 39.99; Price excludes VAT (USA)

Softcover Book: USD 54.99; Price excludes VAT (USA)

Tax calculation will be finalised at checkout

Purchases are for personal use only

Institutional subscriptions

Preview

Unable to display preview. Download preview PDF.

References

Agogino, A.K., Tumer, K.: Analyzing and visualizing multi-agent rewards in dynamic and stochastic domains. JAAMAS 17(2), 320–338 (2008)
Google Scholar
Arthur, W.B.: Inductive reasoning and bounded rationality (the El Farol Problem). American Economic Review 84(406) (1994)
Google Scholar
Damiani, S., Verfaillie, G., Charmeau, M.-C.: An earth watching satellite constellation: How to manage a team of watching agents with limited communications. In: AAMAS (2005)
Google Scholar
Fonseca, C.M., Fleming, P.J.: On the performance assessment and comparison of stochastic multiobjective optimizers. In: Ebeling, W., Rechenberg, I., Voigt, H.-M., Schwefel, H.-P. (eds.) PPSN 1996. LNCS, vol. 1141, pp. 584–593. Springer, Heidelberg (1996)
Chapter Google Scholar
Kaelbling, L.P., Littman, M.L., Moore, A.W.: Reinforcement learning: A survey. Journal of Artificial Intelligence Research (1996)
Google Scholar
Pareto, V.: Manual of Political Economy. MacMillan Press Ltd. (1927)
Google Scholar
Rubenstein, M., Cabrera, A., Werfel, J., Habibi, G., McLurkin, J., Nagpal, R.: Collective transport of complex objects by simple robots: Theory and experiments. In: AAMAS (2013)
Google Scholar
Sherstov, A.A., Stone, P.: Function approximation via tile coding: Automating parameter choice. In: Zucker, J.-D., Saitta, L. (eds.) SARA 2005. LNCS (LNAI), vol. 3607, pp. 194–205. Springer, Heidelberg (2005)
Chapter Google Scholar
Sutton, R., Barto, A.G.: Reinforcement Learning: An Introduction. MIT Press (1998)
Google Scholar
Tomlin, C., Pappas, G.J., Sastry, S.: Conflict resolution for air traffic management: A study in multiagent hybrid systems. IEEE Transactions on Automatic Control 43(4), 509–521 (1998)
Article MATH MathSciNet Google Scholar
Wolpert, D.H., Wheeler, K., Tumer, K.: Collective intelligence for control of distributed dynamical systems. Europhysics Letters 49(6) (2000)
Google Scholar
Wooldridge, M.: An Introduction to MultiAgent Systems. John Wiley and Sons (2002)
Google Scholar

Download references

Author information

Authors and Affiliations

Oregon State University, Corvallis, Oregon, USA
Logan Yliniemi & Kagan Tumer

Authors

Logan Yliniemi
View author publications
You can also search for this author in PubMed Google Scholar
Kagan Tumer
View author publications
You can also search for this author in PubMed Google Scholar

Editor information

Editors and Affiliations

Otago University, Dunedin, New Zealand
Grant Dick
Victoria University of Welling, New Zealand
Will N. Browne
University of Otago, Dunedin, New Zealand
Peter Whigham
Unitec Institute of Technology, Victoria University of Wellington, New Zealand
Mengjie Zhang
Le Quy Don Technical University, Hanoi, Vietnam
Lam Thu Bui
Department of Computer Science and Intelligent Systems, Graduate School of Engineering, Osaka Prefecture University, 1-1 Gakuen-cho, Naka-ku, 599-8531, Sakai, Osaka, Japan
Hisao Ishibuchi
Department of Computing, University of Surrey, GU2 7XH, Guildford, Surrey, UK
Yaochu Jin
RMIT University, Melbourne, Australia
Xiaodong Li
Department of Electrical & Electronic Engineering, Xi’an Jiaotong-Liverpool University, Suzhou, China
Yuhui Shi
Indian Institute of Information Technology and Management, Gwalior, India
Pramod Singh
Department of Electrical and Computer Engineering, National University of Singapore, 4 Engineering Drive 3, 117576, Singapore, Singapore
Kay Chen Tan
USTC-Birmingham Joint Research Institute in Intelligent Computation and Its Applications (UBRI), School of Computer Science and Technology, University of Science and Technology of China, 230027, Hefei, China
Ke Tang

Rights and permissions

Reprints and permissions

Copyright information

About this paper

Cite this paper

Yliniemi, L., Tumer, K. (2014). Multi-objective Multiagent Credit Assignment Through Difference Rewards in Reinforcement Learning. In: Dick, G., et al. Simulated Evolution and Learning. SEAL 2014. Lecture Notes in Computer Science, vol 8886. Springer, Cham. https://doi.org/10.1007/978-3-319-13563-2_35

Download citation

DOI: https://doi.org/10.1007/978-3-319-13563-2_35
Publisher Name: Springer, Cham
Print ISBN: 978-3-319-13562-5
Online ISBN: 978-3-319-13563-2
eBook Packages: Computer ScienceComputer Science (R0)

Publish with us

Policies and ethics