Sfoglia per Correlatore RESTELLI, MARCELLO
Integrating behavioral cloning into a reinforcement learning pipeline
2022/2023 D'Silva, Andrea
Learning from logged bandit feedback techniques for targeting optimization of online advertising
2016/2017 GASPARINI, MARGHERITA
Machine learning techniques for evaluating nasal airflow : preliminary results
2017/2018 ROMANI, GIANLUCA
On the sample complexity of inverse reinforcement learning
2022/2023 Lazzati, Filippo
Online gradient descent for online portfolio optimization with transaction costs
2018/2019 BERNASCONI de LUCA, MARTINO
Optimization of digital advertising campaigns in non-stationary environments through a reinforcement learning algorithm
2015/2016 ITALIA, ELENA MARIA
Policy gradient algorithms for the asset allocation problem
2015/2016 NECCHI, PIERPAOLO GIORGIO
Progettazione e sviluppo di un sistema odometrico basato su mouse ottici
2010/2011 CADARIO, STEFANO
REC-NS-MAB : an algorithm for recurrent concepts in non-stationary multi-armed bandits
2020/2021 RE, GERLANDO
Reinforcement learning control for functional electrical stimulation of the upper limb
2016/2017 DI FEBBO, DAVIDE
Reinforcement learning: from theory to algorithms
PIROTTA, MATTEO
Ricerca di equilibri approssimati per i giochi in forma estesa basati su simulazione
2010/2011 BELLEN, MARCO
Stochastic linear bandits with global-local structure
2021/2022 Gonzales, Francesco Fulco
Legenda icone accesso al fulltext
- File accessibili da tutti
- File accessibili dagli utenti autorizzati
- File accessibili da tutti o solo dagli utenti autorizzati, a partire dalla la data indicata nella scheda
- File non accessibili