Zum Hauptinhalt springen

Showing 1–4 of 4 results for author: Schramm, L

Searching in archive cs. Search in all archives.
.
  1. arXiv:2407.12163  [pdf, ps, other

    cs.LG cs.RO

    Bellman Diffusion Models

    Authors: Liam Schramm, Abdeslam Boularias

    Abstract: Diffusion models have seen tremendous success as generative architectures. Recently, they have been shown to be effective at modelling policies for offline reinforcement learning and imitation learning. We explore using diffusion as a model class for the successor state measure (SSM) of a policy. We find that enforcing the Bellman flow constraints leads to a simple Bellman update on the diffusion… ▽ More

    Submitted 16 July, 2024; originally announced July 2024.

  2. arXiv:2407.05511  [pdf, other

    cs.LG cs.RO

    Provably Efficient Long-Horizon Exploration in Monte Carlo Tree Search through State Occupancy Regularization

    Authors: Liam Schramm, Abdeslam Boularias

    Abstract: Monte Carlo tree search (MCTS) has been successful in a variety of domains, but faces challenges with long-horizon exploration when compared to sampling-based motion planning algorithms like Rapidly-Exploring Random Trees. To address these limitations of MCTS, we derive a tree search algorithm based on policy optimization with state occupancy measure regularization, which we call {\it Volume-MCTS}… ▽ More

    Submitted 7 July, 2024; originally announced July 2024.

    Comments: To be published in ICML 2024 Conference Proceedings

  3. arXiv:2207.01115  [pdf, other

    cs.LG cs.AI cs.RO

    USHER: Unbiased Sampling for Hindsight Experience Replay

    Authors: Liam Schramm, Yunfu Deng, Edgar Granados, Abdeslam Boularias

    Abstract: Dealing with sparse rewards is a long-standing challenge in reinforcement learning (RL). Hindsight Experience Replay (HER) addresses this problem by reusing failed trajectories for one goal as successful trajectories for another. This allows for both a minimum density of reward and for generalization across multiple goals. However, this strategy is known to result in a biased value function, as th… ▽ More

    Submitted 3 July, 2022; originally announced July 2022.

  4. arXiv:2005.10418  [pdf, other

    cs.LG eess.SY stat.ML

    Learning to Transfer Dynamic Models of Underactuated Soft Robotic Hands

    Authors: Liam Schramm, Avishai Sintov, Abdeslam Boularias

    Abstract: Transfer learning is a popular approach to bypassing data limitations in one domain by leveraging data from another domain. This is especially useful in robotics, as it allows practitioners to reduce data collection with physical robots, which can be time-consuming and cause wear and tear. The most common way of doing this with neural networks is to take an existing neural network, and simply trai… ▽ More

    Submitted 20 May, 2020; originally announced May 2020.

    Comments: ICRA 2020