<?xml version="1.0" encoding="UTF-8"?><xml><records><record><source-app name="Biblio" version="7.x">Drupal-Biblio</source-app><ref-type>5</ref-type><contributors><authors><author><style face="normal" font="default" size="100%">Arnaud J Blanchard</style></author><author><style face="normal" font="default" size="100%">Lola Cañamero</style></author></authors><secondary-authors><author><style face="normal" font="default" size="100%">Martin V Butz</style></author><author><style face="normal" font="default" size="100%">Olivier Sigaud</style></author><author><style face="normal" font="default" size="100%">Giovanni Pezzulo</style></author><author><style face="normal" font="default" size="100%">Gianluca Baldassarre</style></author></secondary-authors></contributors><titles><title><style face="normal" font="default" size="100%">Anticipating Rewards in Continuous Time and Space: A Case Study in Developmental Robotics</style></title><secondary-title><style face="normal" font="default" size="100%">Anticipatory Behavior in Adaptive Learning Systems: From Brains to Individual and Social Behavior</style></secondary-title><tertiary-title><style face="normal" font="default" size="100%">Lecture Notes in Artificial Intelligence</style></tertiary-title></titles><dates><year><style  face="normal" font="default" size="100%">2007</style></year></dates><urls><web-urls><url><style face="normal" font="default" size="100%">https://www.springer.com/gp/book/9783540742616</style></url></web-urls></urls><publisher><style face="normal" font="default" size="100%">Springer</style></publisher><pub-location><style face="normal" font="default" size="100%">Berlin, Heidelberg</style></pub-location><volume><style face="normal" font="default" size="100%">4520</style></volume><pages><style face="normal" font="default" size="100%">267–284</style></pages><isbn><style face="normal" font="default" size="100%">978-3-540-74261-6</style></isbn><language><style face="normal" font="default" size="100%">eng</style></language><abstract><style face="normal" font="default" size="100%">This paper presents the first basic principles, implementation and experimental results of what could be regarded as a new approach to reinforcement learning, where agents—physical robots interacting with objects and other agents in the real world—can learn to anticipate rewards using their sensory inputs. Our approach does not need discretization, notion of events, or classification, and instead of learning rewards for the different possible actions of an agent in all the situations, we propose to make agents learn only the main situations worth avoiding and reaching. However, the main focus of our work is not reinforcement learning as such, but modeling cognitive development on a small autonomous robot interacting with an “adult” caretaker, typically a human, in the real world; the control architecture follows a Perception-Action approach incorporating a basic homeostatic principle. This interaction occurs in very close proximity, uses very coarse and limited sensory-motor capabilities, and affects the “well-being” and affective state of the robot. The type of anticipatory behavior we are concerned with in this context relates to both sensory and reward anticipation. We have applied and tested our model on a real robot.</style></abstract></record><record><source-app name="Biblio" version="7.x">Drupal-Biblio</source-app><ref-type>17</ref-type><contributors><authors><author><style face="normal" font="default" size="100%">Lola Cañamero</style></author><author><style face="normal" font="default" size="100%">Arnaud J Blanchard</style></author><author><style face="normal" font="default" size="100%">Jacqueline Nadel</style></author></authors></contributors><titles><title><style face="normal" font="default" size="100%">Attachment Bonds for Human-Like Robots</style></title><secondary-title><style face="normal" font="default" size="100%">International Journal of Humanoid Robotics</style></secondary-title></titles><dates><year><style  face="normal" font="default" size="100%">2006</style></year></dates><urls><web-urls><url><style face="normal" font="default" size="100%">http://www.worldscientific.com/doi/abs/10.1142/S0219843606000771</style></url></web-urls></urls><volume><style face="normal" font="default" size="100%">3</style></volume><pages><style face="normal" font="default" size="100%">301–320</style></pages><language><style face="normal" font="default" size="100%">eng</style></language><abstract><style face="normal" font="default" size="100%">If robots are to be truly integrated in humans' everyday environment, they cannot be simply (pre-)designed and directly taken &quot;off the shelf&quot; and embedded into a real-life setting. Also, technical excellence and human-like appearance and &quot;superficial&quot; traits of their behavior are not enough to make social robots trusted, believable, and accepted. Fuller and deeper integration into human environments would require that, like children, robots develop embedded in the social environment in which they will fulfill their roles. An important element to bootstrap and guide this integration is the establishment of affective bonds between the &quot;infant&quot; robot and the adults among whom it develops, from whom it learns, and who it will later have to look after. In this paper, we present a Perception–Action architecture and experiments to simulate imprinting — the establishment of strong attachment links with a &quot;caregiver&quot; — in a robot. Following recent theories, we do not consider imprinting as rigidly timed and irreversible, but as a more flexible phenomenon that allows for further adaptation as a result of reward-based learning through experience. After the initial imprinting, adaptation is achieved in the context of a history of &quot;affective&quot; interactions between the robot and a human, driven by &quot;distress&quot; and &quot;comfort&quot; responses in the robot.</style></abstract><issue><style face="normal" font="default" size="100%">3</style></issue></record></records></xml>