Header logo is


2019


no image
How do people learn how to plan?

Jain, Y. R., Gupta, S., Rakesh, V., Dayan, P., Callaway, F., Lieder, F.

Conference on Cognitive Computational Neuroscience, September 2019 (conference)

re

[BibTex]

2019


[BibTex]


no image
An ACT-R approach to investigating mechanisms of performance-related changes in an interrupted learning task

Wirzberger, M., Borst, J. P., Krems, J. F., Rey, G. D.

41st Annual Meeting of the Cognitive Science Society., July 2019 (conference)

re

[BibTex]

[BibTex]


no image
What’s in the Adaptive Toolbox and How Do People Choose From It? Rational Models of Strategy Selection in Risky Choice

Mohnert, F., Pachur, T., Lieder, F.

41st Annual Meeting of the Cognitive Science Society, July 2019 (conference)

re

[BibTex]


no image
Measuring how people learn how to plan

Jain, Y. R., Callaway, F., Lieder, F.

RLDM 2019, July 2019 (conference)

re

[BibTex]

[BibTex]


no image
Measuring how people learn how to plan

Jain, Y. R., Callaway, F., Lieder, F.

41st Annual Meeting of the Cognitive Science Society, July 2019 (conference)

re

[BibTex]

[BibTex]


no image
A model-based explanation of performance related changes in abstract stimulus-response learning

Wirzberger, M., Borst, J. P., Krems, J. F., Rey, G. D.

52nd Annual Meeting of the Society for Mathematical Psychology, July 2019 (conference)

Abstract
Stimulus-response learning constitutes an important part of human experience over the life course. Independent of the domain, it is characterized by changes in performance with increasing task progress. But what cognitive mechanisms are responsible for these changes and how do additional task requirements affect the related dynamics? To inspect that in more detail, we introduce a computational modeling approach that investigates performance-related changes in learning situations with reference to chunk activation patterns. It leverages the cognitive architecture ACT-R to model learner behavior in abstract stimulus-response learning in two conditions of task complexity. Additional situational demands are reflected in embedded secondary tasks that interrupt participants during the learning process. Our models apply an activation equation that also takes into account the association between related nodes of information and the similarity between potential responses. Model comparisons with two human datasets (N = 116 and N = 123 participants) indicate a good fit in terms of both accuracy and reaction times. Based on the existing neurophysiological mapping of ACT-R modules on defined human brain areas, we convolve recorded module activity into simulated BOLD responses to investigate underlying cognitive mechanisms in more detail. The resulting evidence supports the connection of learning effects in both task conditions with activation-related patterns to explain changes in performance.

re

[BibTex]

[BibTex]


no image
A cognitive tutor for helping people overcome present bias

Lieder, F., Callaway, F., Jain, Y., Krueger, P., Das, P., Gul, S., Griffiths, T.

RLDM 2019, July 2019 (conference)

re

[BibTex]

[BibTex]


no image
Introducing the Decision Advisor: A simple online tool that helps people overcome cognitive biases and experience less regret in real-life decisions

Iwama, G., Greenberg, S., Moore, D., Lieder, F.

40th Annual Meeting of the Society for Judgement and Decision Making, June 2019 (conference)

re

[BibTex]

[BibTex]


Thumb xl learning tactile servoing thumbnail
Learning Latent Space Dynamics for Tactile Servoing

Sutanto, G., Ratliff, N., Sundaralingam, B., Chebotar, Y., Su, Z., Handa, A., Fox, D.

In Proceedings of the IEEE International Conference on Robotics and Automation (ICRA) 2019, IEEE, International Conference on Robotics and Automation, May 2019 (inproceedings) Accepted

am

pdf video [BibTex]

pdf video [BibTex]


no image
X-ray Optics Fabrication Using Unorthodox Approaches

Sanli, U., Baluktsian, M., Ceylan, H., Sitti, M., Weigand, M., Schütz, G., Keskinbora, K.

Bulletin of the American Physical Society, APS, 2019 (article)

mms pi

[BibTex]

[BibTex]


no image
Nanoscale detection of spin wave deflection angles in permalloy

Gross, F., Träger, N., Förster, J., Weigand, M., Schütz, G., Gräfe, J.

{Applied Physics Letters}, 114(1), American Institute of Physics, Melville, NY, 2019 (article)

mms

DOI [BibTex]

DOI [BibTex]


no image
Spatial Continuity Effect vs. Spatial Contiguity Failure. Revising the Effects of Spatial Proximity Between Related and Unrelated Representations

Beege, M., Wirzberger, M., Nebel, S., Schneider, S., Schmidt, N., Rey, G. D.

Frontiers in Education, 4:86, 2019 (article)

Abstract
The split-attention effect refers to learning with related representations in multimedia. Spatial proximity and integration of these representations are crucial for learning processes. The influence of varying amounts of proximity between related and unrelated information has not yet been specified. In two experiments (N1 = 98; N2 = 85), spatial proximity between a pictorial presentation and text labels was manipulated (high vs. medium vs. low). Additionally, in experiment 1, a control group with separated picture and text presentation was implemented. The results revealed a significant effect of spatial proximity on learning performance. In contrast to previous studies, the medium condition leads to the highest transfer, and in experiment 2, the highest retention score. These results are interpreted considering cognitive load and instructional efficiency. Findings indicate that transfer efficiency is optimal at a medium distance between representations in experiment 1. Implications regarding the spatial contiguity principle and the spatial contiguity failure are discussed.

re

link (url) DOI [BibTex]


no image
Generation of switchable singular beams with dynamic metasurfaces

Yu, P., Li, J., Li, X., Schütz, G., Hirscher, M., Zhang, S., Liu, N.

{ACS Nano}, 13(6):7100-7106, American Chemical Society, Washington, DC, 2019 (article)

mms

DOI [BibTex]

DOI [BibTex]


no image
Extracting the dynamic magnetic contrast in time-resolved X-ray transmission microscopy

Schaffers, T., Feggeler, T., Pile, S., Meckenstock, R., Buchner, M., Spoddig, D., Ney, V., Farle, M., Wende, H., Wintz, S., Weigand, M., Ohldag, H., Ollefs, K, Ney, A.

{Nanomaterials}, 9(7), MDPI, Basel, Schweiz, 2019 (article)

mms

DOI [BibTex]

DOI [BibTex]


no image
Coherent excitation of heterosymmetric spin waves with ultrashort wavelengths

Dieterle, G., Förster, J., Stoll, H., Semisalova, A. S., Finizio, S., Gangwar, A., Weigand, M., Noske, M., Fähnle, M., Bykova, I., Gräfe, J., Bozhko, D. A., Musiienko-Shmarova, H. Y., Tiberkevich, V., Slavin, A. N., Back, C. H., Raabe, J., Schütz, G., Wintz, S.

{Physical Review Letters}, 122(11), American Physical Society, Woodbury, N.Y., 2019 (article)

mms

DOI [BibTex]

DOI [BibTex]


no image
Reprogrammability and Scalability of Magnonic Fibonacci Quasicrystals

Lisiecki, F., Rychły, J., Kuświk, P., Głowiński, H., Kłos, J. W., Groß, F., Bykova, I., Weigand, M., Zelent, M., Goering, E. J., Schütz, G., Gubbiotti, G., Krawczyk, M., Stobiecki, F., Dubowik, J., Gräfe, J.

Physical Review Applied, 11, pages: 054003, 2019 (article)

Abstract
Magnonic crystals are systems that can be used to design and tune the dynamic properties of magnetization. Here, we focus on one-dimensional Fibonacci magnonic quasicrystals. We confirm the existence of collective spin waves propagating through the structure as well as dispersionless modes; the reprogammability of the resonance frequencies, dependent on the magnetization order; and dynamic spin-wave interactions. With the fundamental understanding of these properties, we lay a foundation for the scalable and advanced design of spin-wave band structures for spintronic, microwave, and magnonic applications.

mms

link (url) DOI [BibTex]

link (url) DOI [BibTex]


no image
Coordinated molecule-modulated magnetic phase with metamagnetism in metal-organic frameworks

Son, K., Kim, J. Y., Schütz, G., Kang, S. G., Moon, H. R., Oh, H.

{Inorganic Chemistry}, 58(14):8895-8899, American Chemical Society, Washington, DC, 2019 (article)

mms

DOI [BibTex]

DOI [BibTex]


no image
A special issue on hydrogen-based Energy storage

Hirscher, M.

{International Journal of Hydrogen Energy}, 44, pages: 7737, Elsevier, Amsterdam, 2019 (misc)

mms

DOI [BibTex]

DOI [BibTex]


no image
Novel X-ray lenses for direct and coherent imaging

Sanli, U. T.

Universität Stuttgart, Stuttgart, 2019 (phdthesis)

mms

link (url) DOI [BibTex]

link (url) DOI [BibTex]


no image
Doing more with less: Meta-reasoning and meta-learning in humans and machines

Griffiths, T., Callaway, F., Chang, M., Grant, E., Krueger, P. M., Lieder, F.

Current Opinion in Behavioral Sciences, 2019 (article)

re

DOI [BibTex]

DOI [BibTex]


no image
Nanoscale X-ray imaging of spin dynamics in Yttrium iron garnet

Förster, J., Wintz, S., Bailey, J., Finizio, S., Josten, E., Meertens, D., Dubs, C., Bozhko, D. A., Stoll, H., Dieterle, G., Traeger, N., Raabe, J., Slavin, A. N., Weigand, M., Gräfe, J., Schütz, G.

2019 (misc)

mms

link (url) [BibTex]

link (url) [BibTex]


no image
Magnons in a Quasicrystal: Propagation, Extinction, and Localization of Spin Waves in Fibonacci Structures

Lisiecki, F., Rychły, J., Kuświk, P., Głowiński, H., Kłos, J. W., Groß, F., Träger, N., Bykova, I., Weigand, M., Zelent, M., Goering, E. J., Schütz, G., Krawczyk, M., Stobiecki, F., Dubowik, J., Gräfe, J.

Physical Review Applied, 11, pages: 054061, 2019 (article)

Abstract
Magnonic quasicrystals exceed the possibilities of spin-wave (SW) manipulation offered by regular magnonic crystals, because of their more complex SW spectra with fractal characteristics. Here, we report the direct x-ray microscopic observation of propagating SWs in a magnonic quasicrystal, consisting of dipolar coupled permalloy nanowires arranged in a one-dimensional Fibonacci sequence. SWs from the first and second band as well as evanescent waves from the band gap between them are imaged. Moreover, additional mini band gaps in the spectrum are demonstrated, directly indicating an influence of the quasiperiodicity of the system. Finally, the localization of SW modes within the Fibonacci crystal is shown. The experimental results are interpreted using numerical calculations and we deduce a simple model to estimate the frequency position of the magnonic gaps in quasiperiodic structures. The demonstrated features of SW spectra in one-dimensional magnonic quasicrystals allow utilizing this class of metamaterials for magnonics and make them an ideal basis for future applications.

mms

link (url) DOI [BibTex]

link (url) DOI [BibTex]


no image
Reconfigurable nanoscale spin wave majority gate with frequency-division multiplexing

Talmelli, G., Devolder, T., Träger, N., Förster, J., Wintz, S., Weigand, M., Stoll, H., Heyns, M., Schütz, G., Radu, I., Gräfe, J., Ciubotaru, F., Adelmann, C.

2019 (misc)

Abstract
Spin waves are excitations in ferromagnetic media that have been proposed as information carriers in spintronic devices with potentially much lower operation power than conventional charge-based electronics. The wave nature of spin waves can be exploited to design majority gates by coding information in their phase and using interference for computation. However, a scalable spin wave majority gate design that can be co-integrated alongside conventional Si-based electronics is still lacking. Here, we demonstrate a reconfigurable nanoscale inline spin wave majority gate with ultrasmall footprint, frequency-division multiplexing, and fan-out. Time-resolved imaging of the magnetisation dynamics by scanning transmission x-ray microscopy reveals the operation mode of the device and validates the full logic majority truth table. All-electrical spin wave spectroscopy further demonstrates spin wave majority gates with sub-micron dimensions, sub-micron spin wave wavelengths, and reconfigurable input and output ports. We also show that interference-based computation allows for frequency-division multiplexing as well as the computation of different logic functions in the same device. Such devices can thus form the foundation of a future spin-wave-based superscalar vector computing platform.

mms

link (url) [BibTex]

link (url) [BibTex]


no image
Prototyping Micro- and Nano-Optics with Focused Ion Beam Lithography

Keskinbora, K.

SL48, pages: 46, SPIE.Spotlight, SPIE Press, Bellingham, WA, 2019 (book)

mms

DOI [BibTex]

DOI [BibTex]


no image
Structural and magnetic properties of FePt-Tb alloy thin films

Schmidt, N. Y., Laureti, S., Radu, F., Ryll, H., Luo, C., d\textquotesingleAcapito, F., Tripathi, S., Goering, E., Weller, D., Albrecht, M.

{Physical Review B}, 100(6), American Physical Society, Woodbury, NY, 2019 (article)

mms

DOI [BibTex]

DOI [BibTex]


Thumb xl screenshot from 2019 03 21 12 11 19
Automated Generation of Reactive Programs from Human Demonstration for Orchestration of Robot Behaviors

Berenz, V., Bjelic, A., Mainprice, J.

ArXiv, 2019 (article)

Abstract
Social robots or collaborative robots that have to interact with people in a reactive way are difficult to program. This difficulty stems from the different skills required by the programmer: to provide an engaging user experience the behavior must include a sense of aesthetics while robustly operating in a continuously changing environment. The Playful framework allows composing such dynamic behaviors using a basic set of action and perception primitives. Within this framework, a behavior is encoded as a list of declarative statements corresponding to high-level sensory-motor couplings. To facilitate non-expert users to program such behaviors, we propose a Learning from Demonstration (LfD) technique that maps motion capture of humans directly to a Playful script. The approach proceeds by identifying the sensory-motor couplings that are active at each step using the Viterbi path in a Hidden Markov Model (HMM). Given these activation patterns, binary classifiers called evaluations are trained to associate activations to sensory data. Modularity is increased by clustering the sensory-motor couplings, leading to a hierarchical tree structure. The novelty of the proposed approach is that the learned behavior is encoded not in terms of trajectories in a task space, but as couplings between sensory information and high-level motor actions. This provides advantages in terms of behavioral generalization and reactivity displayed by the robot.

am

Support Video link (url) [BibTex]


no image
Cognitive Prostheses for Goal Achievement

Lieder, F., Chen, O. X., Krueger, P. M., Griffiths, T.

Nature Human Behavior, 2019 (article)

re

DOI [BibTex]

DOI [BibTex]


no image
Remediating cognitive decline with cognitive tutors

Das, P., Callaway, F., Griffiths, T., Lieder, F.

RLDM 2019, 2019 (conference)

re

[BibTex]

[BibTex]


no image
Interpreting first-order reversal curves beyond the Preisach model: An experimental permalloy microarray investigation

Groß, F., Ilse, S. E., Schütz, G., Gräfe, J., Goering, E.

{Physical Review B}, 99(6), American Physical Society, Woodbury, NY, 2019 (article)

mms

DOI [BibTex]


no image
Effects of system response delays on elderly humans’ cognitive performance in a virtual training scenario

Wirzberger, M., Schmidt, R., Georgi, M., Hardt, W., Brunnett, G., Rey, G. D.

Scientific Reports, 9:8291, 2019 (article)

Abstract
Observed influences of system response delay in spoken human-machine dialogues are rather ambiguous and mainly focus on perceived system quality. Studies that systematically inspect effects on cognitive performance are still lacking, and effects of individual characteristics are also often neglected. Building on benefits of cognitive training for decelerating cognitive decline, this Wizard-of-Oz study addresses both issues by testing 62 elderly participants in a dialogue-based memory training with a virtual agent. Participants acquired the method of loci with fading instructional guidance and applied it afterward to memorizing and recalling lists of German nouns. System response delays were randomly assigned, and training performance was included as potential mediator. Participants’ age, gender, and subscales of affinity for technology (enthusiasm, competence, positive and negative perception of technology) were inspected as potential moderators. The results indicated positive effects on recall performance with higher training performance, female gender, and less negative perception of technology. Additionally, memory retention and facets of affinity for technology moderated increasing system response delays. Participants also provided higher ratings in perceived system quality with higher enthusiasm for technology but reported increasing frustration with a more positive perception of technology. Potential explanations and implications for the design of spoken dialogue systems are discussed.

re

link (url) DOI [BibTex]

link (url) DOI [BibTex]


no image
Bistability of magnetic states in Fe-Pd nanocap arrays

Aravind, P. B., Heigl, M., Fix, M., Groß, F., Gräfe, J., Mary, A., Rajgowrav, C. R., Krupiński, M., Marszałek, M., Thomas, S., Anantharaman, M. R., Albrecht, M.

Nanotechnology, 30, pages: 405705, 2019 (article)

Abstract
Magnetic bistability between vortex and single domain states in nanostructures are of great interest from both fundamental and technological perspectives. In soft magnetic nanostructures, the transition from a uniform collinear magnetic state to a vortex state (or vice versa) induced by a magnetic field involves an energy barrier. If the thermal energy is large enough for overcoming this energy barrier, magnetic bistability with a hysteresis-free switching occurs between the two magnetic states. In this work, we tune this energy barrier by tailoring the composition of FePd alloys, which were deposited onto self-assembled particle arrays forming magnetic vortex structures on top of the particles. The bifurcation temperature, where a hysteresis-free transition occurs, was extracted from the temperature dependence of the annihilation and nucleation field which increases almost linearly with Fe content of the magnetic alloy. This study provides insights into the magnetization reversal process associated with magnetic bistability, which allows adjusting the bifurcation temperature range by the material properties of the nanosystem.

mms

link (url) [BibTex]

link (url) [BibTex]


no image
An international laboratory comparison study of volumetric and gravimetric hydrogen adsorption measurements

Hurst, K. E., Gennett, T., Adams, J., Allendorf, M. D., Balderas-Xicohténcatl, R., Bielewski, M., Edwards, B., Espinal, L., Fultz, B., Hirscher, M., Hudson, M. S. L., Hulvey, Z., Latroche, M., Liu, D., Kapelewski, M., Napolitano, E., Perry, Z. T., Purewal, J., Stavila, V., Veenstra, M., White, J. L., Yuan, Y., Zhou, H., Zlotea, C., Parilla, P.

{ChemPhysChem}, 20(15):1997-2009, Wiley-VCH, Weinheim, Germany, 2019 (article)

mms

DOI [BibTex]

DOI [BibTex]


no image
Hydrogen Energy

Hirscher, M., Autrey, T., Orimo, S.

{ChemPhysChem}, 20, pages: 1153-1411, Wiley-VCH, Weinheim, Germany, 2019 (misc)

mms

link (url) DOI [BibTex]

link (url) DOI [BibTex]


no image
Superior magnetic performance in FePt L10 nanomaterials

Son, K., Ryu, G. H., Jeong, H., Fink, L., Merz, M., Nagel, P., Schuppler, S., Richter, G., Goering, E., Schütz, G.

{Small}, 15(34), Wiley, Weinheim, Germany, 2019 (article)

mms

DOI [BibTex]

DOI [BibTex]


no image
A rational reinterpretation of dual process theories

Milli, S., Lieder, F., Griffiths, T.

2019 (article)

re

DOI [BibTex]

DOI [BibTex]


no image
Systematic experimental study on quantum sieving of hydrogen isotopes in metal-amide-imidazolate frameworks with narrow 1-D channels

Mondal, S. S., Kreuzer, A., Behrens, K., Schütz, G., Holdt, H., Hirscher, M.

{ChemPhysChem}, 20(10):1311-1315, Wiley-VCH, Weinheim, Germany, 2019 (article)

mms

DOI [BibTex]

DOI [BibTex]


no image
The route to supercurrent transparent ferromagnetic barriers in superconducting matrix

Ivanov, Y. P., Soltan, S., Albrecht, J., Goering, E., Schütz, G., Zhang, Z., Chuvilin, A.

{ACS Nano}, 13(5):5655-5661, American Chemical Society, Washington, DC, 2019 (article)

mms

DOI [BibTex]

DOI [BibTex]


no image
Artifacts from manganese reduction in rock samples prepared by focused ion beam (FIB) slicing for X-ray microspectroscopy

Macholdt, D. S., Förster, J., Müller, M., Weber, B., Kappl, M., Kilcoyne, A. L. D., Weigand, M., Leitner, J., Jochum, K. P., Pöhlker, C., Andreae, M. O.

{Geoscientific instrumentation, methods and data systems}, 8(1):97-111, Copernicus Publ., Göttingen, 2019 (article)

mms

DOI [BibTex]

DOI [BibTex]


no image
Mixed-state magnetotransport properties of MgB2 thin film prepared by pulsed laser deposition on an Al2O3 substrate

Alzayed, N. S., Shahabuddin, M., Ramey, S. M., Soltan, S.

{Journal of Materials Science: Materials in Electronics}, 30(2):1547-1552, Springer, Norwell, MA, 2019 (article)

mms

DOI [BibTex]

DOI [BibTex]


no image
Comparison of theories of fast and ultrafast magnetization dynamics

Fähnle, M.

{Journal of Magnetism and Magnetic Materials}, 469, pages: 28-29, NH, Elsevier, Amsterdam, 2019 (article)

mms

DOI [BibTex]

DOI [BibTex]


no image
Concepts for improving hydrogen storage in nanoporous materials

Broom, D. P., Webb, C. J., Fanourgakis, G. S., Froudakis, G. E., Trikalitis, P. N., Hirscher, M.

{International Journal of Hydrogen Energy}, 44(15):7768-7779, Elsevier, Amsterdam, 2019 (article)

mms

DOI [BibTex]

DOI [BibTex]


no image
Controlling dislocation nucleation-mediatd plasticity in nanostructures via surface modification

Shin, J., Chen, L. Y., Sanli, U. T., Richter, G., Labat, S., Richard, M., Cornelius, T., Thomas, O., Gianola, D. S.

{Acta Materialia}, 166, pages: 572-586, Elsevier Science, Kidlington, 2019 (article)

mms

DOI [BibTex]

DOI [BibTex]


no image
Reprogrammability and scalability of magnonic Fibonacci quasicrystals

Lisiecki, F., Rychly, J., Kuswik, P., Glowinski, H., Klos, J. W., Groß, F., Bykova, I., Weigand, M., Zelent, M., Goering, E. J., Schütz, G., Gubbiotti, G., Krawczyk, M., Stobiecki, F., Dubowik, J., Gräfe, J.

{Physical Review Applied}, 11(5), American Physical Society, College Park, Md. [u.a.], 2019 (article)

mms

DOI [BibTex]

DOI [BibTex]

2006


no image
Learning operational space control

Peters, J., Schaal, S.

In Robotics: Science and Systems II (RSS 2006), pages: 255-262, (Editors: Gaurav S. Sukhatme and Stefan Schaal and Wolfram Burgard and Dieter Fox), Cambridge, MA: MIT Press, RSS , 2006, clmc (inproceedings)

Abstract
While operational space control is of essential importance for robotics and well-understood from an analytical point of view, it can be prohibitively hard to achieve accurate control in face of modeling errors, which are inevitable in complex robots, e.g., humanoid robots. In such cases, learning control methods can offer an interesting alternative to analytical control algorithms. However, the resulting learning problem is ill-defined as it requires to learn an inverse mapping of a usually redundant system, which is well known to suffer from the property of non-covexity of the solution space, i.e., the learning system could generate motor commands that try to steer the robot into physically impossible configurations. A first important insight for this paper is that, nevertheless, a physically correct solution to the inverse problem does exits when learning of the inverse map is performed in a suitable piecewise linear way. The second crucial component for our work is based on a recent insight that many operational space controllers can be understood in terms of a constraint optimal control problem. The cost function associated with this optimal control problem allows us to formulate a learning algorithm that automatically synthesizes a globally consistent desired resolution of redundancy while learning the operational space controller. From the view of machine learning, the learning problem corresponds to a reinforcement learning problem that maximizes an immediate reward and that employs an expectation-maximization policy search algorithm. Evaluations on a three degrees of freedom robot arm illustrate the feasability of our suggested approach.

am ei

link (url) [BibTex]

2006


link (url) [BibTex]


no image
Reinforcement Learning for Parameterized Motor Primitives

Peters, J., Schaal, S.

In Proceedings of the 2006 International Joint Conference on Neural Networks, pages: 73-80, IJCNN, 2006, clmc (inproceedings)

Abstract
One of the major challenges in both action generation for robotics and in the understanding of human motor control is to learn the "building blocks of movement generation", called motor primitives. Motor primitives, as used in this paper, are parameterized control policies such as splines or nonlinear differential equations with desired attractor properties. While a lot of progress has been made in teaching parameterized motor primitives using supervised or imitation learning, the self-improvement by interaction of the system with the environment remains a challenging problem. In this paper, we evaluate different reinforcement learning approaches for improving the performance of parameterized motor primitives. For pursuing this goal, we highlight the difficulties with current reinforcement learning methods, and outline both established and novel algorithms for the gradient-based improvement of parameterized policies. We compare these algorithms in the context of motor primitive learning, and show that our most modern algorithm, the Episodic Natural Actor-Critic outperforms previous algorithms by at least an order of magnitude. We demonstrate the efficiency of this reinforcement learning method in the application of learning to hit a baseball with an anthropomorphic robot arm.

am ei

link (url) DOI [BibTex]

link (url) DOI [BibTex]