<!-- mobian-agent-page publisher="dailydev" canonical="https://daily.dev/posts/in-context-exploration-exploitation-for-reinforcement-learning---spotify-research-ant96hbpg" -->

---
title: In-context Exploration-Exploitation for Reinforcement...
description: A groundbreaking approach called in-context exploration-exploitation (ICEE) is presented in this post, which allows for real-time adaptation to users&#x27;...
canonical: https://daily.dev/posts/in-context-exploration-exploitation-for-reinforcement-learning---spotify-research-ant96hbpg
twitter:card: summary_large_image
twitter:site: @dailydotdev
og:type: website
og:site_name: daily.dev
og:title: In-context Exploration-Exploitation for Reinforcement Learning - Spotify Research | daily.dev
og:description: A groundbreaking approach called in-context exploration-exploitation (ICEE) is presented in this post, which allows for real-time adaptation to users&#x27;...
og:url: https://daily.dev/posts/in-context-exploration-exploitation-for-reinforcement-learning---spotify-research-ant96hbpg
og:image: https://api.daily.dev/og/posts/Ant96hBpg.png
og:image:alt: In-context Exploration-Exploitation for Reinforcement Learning - Spotify Research
og:image:width: 1200
og:image:height: 630
og:locale: en
---

> ## Documentation Index
> Fetch the complete documentation index at: https://daily.dev/llms.txt
> Use this file to discover all available pages before exploring further.

# In-context Exploration-Exploitation for Reinforcement Learning - Spotify Research

**[Spotify Research](https://daily.dev/sources/spotify_research)** · 6 min read · 0 upvotes · 0 comments

## Summary

A groundbreaking approach called in-context exploration-exploitation (ICEE) is presented in this post, which allows for real-time adaptation to users' preferences through neural network inference. ICEE is able to perform Bayesian belief updates without the need for explicit Bayesian inference and can solve new RL tasks within a few episodes.

## Full article

daily.dev links to this article rather than hosting it. Read it at the original source: <https://research.atspotify.com/2024/05/in-context-exploration-exploitation-for-reinforcement-learning/>

---

Tags: [#machine-learning](https://daily.dev/tags/machine-learning), [#reinforcement-learning](https://daily.dev/tags/reinforcement-learning)

[View this post on daily.dev](https://daily.dev/posts/in-context-exploration-exploitation-for-reinforcement-learning---spotify-research-ant96hbpg)

```json
{"@context":"https://schema.org","@graph":[{"@type":"Organization","@id":"https://daily.dev/#organization","name":"daily.dev","url":"https://daily.dev","logo":{"@type":"ImageObject","url":"https://daily.dev/apple-touch-icon.png","width":180,"height":180},"sameAs":["https://twitter.com/dailydotdev","https://github.com/dailydotdev","https://www.linkedin.com/company/daily-dev-ltd"]},{"@type":"WebSite","@id":"https://daily.dev/#website","url":"https://daily.dev","name":"daily.dev","publisher":{"@id":"https://daily.dev/#organization"},"potentialAction":{"@type":"SearchAction","target":{"@type":"EntryPoint","urlTemplate":"https://daily.dev/search?q={search_term_string}"},"query-input":"required name=search_term_string"}}]}
{"@context":"https://schema.org","@type":"TechArticle","headline":"In-context Exploration-Exploitation for Reinforcement Learning - Spotify Research","url":"https://daily.dev/posts/in-context-exploration-exploitation-for-reinforcement-learning---spotify-research-ant96hbpg","mainEntityOfPage":{"@type":"WebPage","@id":"https://daily.dev/posts/in-context-exploration-exploitation-for-reinforcement-learning---spotify-research-ant96hbpg"},"datePublished":"2024-05-07T09:41:36.678Z","dateModified":"2024-05-09T08:43:33.910Z","description":"A groundbreaking approach called in-context exploration-exploitation (ICEE) is presented in this post, which allows for real-time adaptation to users'...","image":"https://media.daily.dev/image/upload/f_auto,q_auto/v1/posts/78984d9ecad14ab51e7b53df6a7316f9?_a=AQAEuiZ","thumbnailUrl":"https://media.daily.dev/image/upload/f_auto,q_auto/v1/posts/78984d9ecad14ab51e7b53df6a7316f9?_a=AQAEuiZ","isAccessibleForFree":true,"articleSection":"Spotify Research","inLanguage":"en","publisher":{"@type":"Organization","name":"daily.dev","url":"https://daily.dev","logo":{"@type":"ImageObject","url":"https://daily.dev/apple-touch-icon.png","width":180,"height":180}},"author":{"@type":"Organization","name":"Spotify Research","logo":"https://media.daily.dev/image/upload/t_logo,f_auto/v1/logos/a020ee10033c40eaa05cc420c7fb45af","url":"https://daily.dev/sources/spotify_research"},"commentCount":0,"discussionUrl":"https://daily.dev/posts/in-context-exploration-exploitation-for-reinforcement-learning---spotify-research-ant96hbpg","interactionStatistic":[{"@type":"InteractionCounter","interactionType":{"@type":"LikeAction"},"userInteractionCount":0},{"@type":"InteractionCounter","interactionType":{"@type":"CommentAction"},"userInteractionCount":0}],"keywords":"machine-learning,reinforcement-learning","timeRequired":"PT6M"}
{"@context":"https://schema.org","@type":"BreadcrumbList","itemListElement":[{"@type":"ListItem","position":1,"name":"Home","item":"https://daily.dev"},{"@type":"ListItem","position":2,"name":"Spotify Research","item":"https://daily.dev/sources/spotify_research"},{"@type":"ListItem","position":3,"name":"In-context Exploration-Exploitation for Reinforcement Learning - Spotify Research"}]}
```

