Open Access Open Access  Restricted Access Subscription Access

Spatio-Temporal Learning for a Rescue Robot

Ludmilla Kleinmann, Jonathan Rabe, Bärbel Mertsching

Abstract


The presented work aims at developing an autonomous rescue system with flexible behavior. The embodied cognitive system is based on the interrelated multiple memory structures whose functionality is modeled at a high level of abstraction. The biologically motivated learning processes are incorporated in the system corresponding to the kind of memory they belong to. The system was tested and evaluated in an unknown maze in a simulator by searching and navigating a set of specified objects in a given order.

Full Text:

PDF

References


N. M. White, Scholarpedia, 2007; doi:10.4249/scholarpedia.2663.

M. V. Butz, O. Sigaud., P. Gerard, “Internal models and anticipations in

adaptive learning systems”, in Anticipatory behavior in adaptive learning systems: Foundations, theories, and systems, M. V. Butz, O. Sigaud,, P. Gerard, Eds, Berlin Heidelberg: Springer-Verlag, pp. 86–109, 2003.

J. E. Laird, P. S. Rosenbloom, A. Newell, “Chunking in SOAR: The anatomy of a general learning mechanism”, Machine Learning, vol 1, pp.11-46, 1986.

J.R. Anderson, C. Lebiere, “The atomic components of thought”, ,Lawrence Erlbaum, 1998.

R. O’Reilly, J. Rudy, “Conjunctive representations in learning and memory: principles of cortical and hippocampal function”, PsycholRev pp. 108:311–345, 2001.

E.T. Rolls, R.P. Kesner; “A computational theory of hippocampal function, and empirical tests of the theory”; Progress in Neurobiology vol.79, pp.1–48, 2006.

E.T. Rolls, “Memory, attention, and decision-making: A unifying computational neuroscience approach”, Oxford University Press, Oxford, 2008.

M.E. Hasselmo, “What is the function of hippocampal theta rhythm?—

Linking behavioral data to phasic properties of field”, Hippocampus, Vol. 15 (7), pp.936–49, 2005.

A. Gorchetchnikov, S. Grossberg, “Space, time and learning in the

hippocampus: How fine spatial and temporal scales are expanded into population codes for behavioral control”,. Neural Netw., vol. 20, pp.182–93, 2007.

H. Wagatsuma,, Y. Yamaguchi, “Neural dynamics of the cognitive map

in the hippocampus”, Cogn. Neurodyn.,vol. 1, pp.119–141, 2007.

N. Sato, Y. Yamaguchi, “Simulation of Human Episodic Memory by

Using a Computational Model of the Hippocampus”, Adv. Artificial Intelligence, 2010.

A. M. Nuxoll,, “Enhancing Intelligent Agents with Episodic Memory”,

PhD thesis, University of Michigan, 2007.

C. Brom, K. Peskova,, J. Lukavsky, “What does your actor remembertowards characters with a full episodic memory”, Lecture Notes in Computer Science, Proc. of 4th ICVS, pp.9–101, 2007.

T. Deutsch, A. Gruber, R. Lang, V. Velik, “Episodic memory for

autonomous agents”, Proc. of IEEE HSI Human System Interactions

Conference, Krakow, Poland, May 25-27, 2008.

N.S. Kuppuswami, S. Cho, J. Kim, “A cognitive control architecture for

an artificial creature using episodic memory”, Proc. SICE-ICASE Int.

Joint Conf., pp. 3104–3110, Busan, Korea, October 2006.

D. Tecuci, “A Generic Memory Module for Events”, PhD thesis,

University of Texas in Austin, 2007.

W. C. Ho, K. Dautenhahn, C.L. Nehaniv, “Autobiographic agents in

dynamic virtual environments - performance comparison for different

memory control architectures”, in Proc. of IEEE Congress on Evolutionary Computation, pp. 573–580, 2005.

R. S.Sutton, A. G. Barto, “Reinforcement Learning: An Introduction”,

The MIT Press, Cambridge, MA, 1998.

R. Bellman, “A Markovian Decision Process”,. Journal of Mathematics

and Mechanics, vol. 6, 1957.

R. Bellman, “Dynamic Programming”, Princeton University Press,

Princeton, NJ, 1957.

Y. Niv, “Reinforcement Learning in the Brain”, Journal of

Mathematical Psychology, Vol. 53,(3), pp. 139-154, 2009.

W. Schultz, P. Dayan., R. R. Montague ,”A neural substrate of

prediction and reward”, Science, vol. 275, pp.1593–1599, 1997.

N. Maier, T. Schneirla, “Principles of Animal Psychology”, Dover

Publications, New York, 1964.

G. Wustmann, K. Rein, R. Wolf, M. Heisenberg, “A New Paradigm for

Operant Conditioning of Drosophila Melanogaster”, Journal of Comparative Physiology, vol. 179, pp. 429-436, 1996.

RL-Toolbox, University of Graz : http://www.igi.tugraz.at/riltoolbox/

general/overview.html (seen on 20.06.2012)

E. Tolman,, “Cognitive maps in rats and men”, Psychological Review,

vol. 55, pp. 189-208, 1948.

F.V. Jensen, “An Introduction To Bayesian Networks”, UCL Press,

K. Murphy, “ Dynamic Bayesian Networks: Representation, Inference

and Learning”, Ph.D. thesis, Computer Science Division, University of

California, Berkeley, 2002.

C. Boutilier.,R. Dearden, M. Goldszmidt, “ Exploiting structure in

policy construction”,. In Proc. IJCAI, pp. 1104-1111, 1995.




DOI: http://dx.doi.org/10.21535%2FProICIUS.2012.v8.776

Refbacks

  • There are currently no refbacks.