Publications
International Conferences
2026
- Worst-Case Regret Bounds for Combinatorial Bandits with Ranking Feedback NeurIPS 2026 To appear Link Paper Poster
- Online Compatible Reward Identification from Preference Feedback ICML 2026 Link Paper Poster
- Fast Mixing Steady-State Control in Markov Decision Processes ICML 2026 Link Paper Poster
- Learning to Rank from Incomplete Rankings ICML 2026 Link Paper Poster
- Reusing Trajectories in Policy Gradients Enables Fast Convergence ICML 2026 Link Paper arXiv Poster
2025
- Tightening Regret Lower and Upper Bounds in Restless Rising Bandits NeurIPS 2025 Link Paper Poster Slides
- Sleeping Reinforcement Learning ICML 2025 Link Paper Poster
- Towards Theoretical Understanding of Sequential Decision Making with Preference Feedback ICML 2025 Link Paper Poster
- Convergence Analysis of Policy Gradient Methods with Dynamic Stochasticity ICML 2025 Link Paper Poster
- Position: Constants are Critical in Regret Bounds for Reinforcement Learning ICML 2025 Link Paper Poster
2024
- Last-Iterate Global Convergence of Policy Gradients for Constrained Reinforcement Learning NeurIPS 2024 Link Paper arXiv Poster Slides
- Factored-Reward Bandits with Intermediate Observations ICML 2024 Link Paper Poster Slides
- Best Arm Identification for Stochastic Rising Bandits ICML 2024 Spotlight Link Paper arXiv Poster
- Learning Optimal Deterministic Policies with Stochastic Policy Gradients ICML 2024 Spotlight Link Paper arXiv Poster
- Graph-Triggered Rising Bandits ICML 2024 Link Paper Poster
- Autoregressive Bandits AISTATS 2024 Link Paper arXiv Poster Slides
2023
- Dynamical Linear Bandits ICML 2023 Link Paper arXiv Poster Slides
- Dynamic Pricing with Volume Discounts in Online Settings IAAI 2023 Innovative Application of AI Award Link Paper arXiv Poster Slides Award
2022
Journals
2026
- Bridging Rested and Restless Bandits with Graph-Triggering: Rising and Rotting Journal of Machine Learning Research, 2026 Link Paper arXiv
- Trading-off Statistical and Computational Efficiency via W-step Markov Decision Processes: A Policy Gradient Approach Machine Learning, 2026 Link Paper Poster Slides
2025
- Human-AI Interaction in Safety-Critical Network Infrastructures iScience, 2025 Link Paper
- Factored-Reward Bandits with Intermediate Observations: Regret Minimization and Best Arm Identification Artificial Intelligence, 2025 Link Paper
- Generalizing the Regret: an Analysis of Lower and Upper Bounds Journal of Artificial Intelligence Research, 2025 Link Paper
2024
- A Reinforcement Learning Controller Optimizing Costs and Battery State of Health in Smart Grids Journal of Energy Storage, 2024 Link Paper
2023
- ARLO: A Framework for Automated Reinforcement Learning Expert Systems with Applications, 2023 Link Paper arXiv
2022
- An Online State of Health Estimation Method for Lithium-Ion Batteries based on Time Partitioning and Data-Driven Model Identification Journal of Energy Storage, 2022 Link Paper
2021
Workshops
2026
- Combinatorial Bandits with Plackett-Luce Feedback: A Worst-Case Analysis EWRL 2026 Link Paper Poster
- Variance-Aware Optimal Ranking in Log-Concave Random Utility Models EWRL 2026 Link Paper Poster
- Modeling Incomparability: A New Rationality Paradigm for Preference-Based Reinforcement Learning ICML 2026 EIML Workshop Paper Poster
- Robust Learning to Rank from Incomplete Rankings under Positional Censoring ICML 2026 EIML Workshop Spotlight Paper Poster Slides
2025
- Trading-off Reward Maximization and Stability in Sequential Decision Making EWRL 2025 Link Paper Poster
- A Theoretical Perspective on Sequential Decision Making with Preference Feedback EWRL 2025 Link Paper Poster
- A Novel Self-Normalized Bernstein-Like Dimension-Free Inequality and Regret Bounds for Generalized Kernelized Bandits EWRL 2025 Link Paper Poster
- Gym4ReaL: A Benchmark Suite for Evaluating Reinforcement Learning in Realistic Domains EWRL 2025 Link Paper Poster
- Power Grid Control with Graph-Based Distributed Reinforcement Learning ECML 2025 MLSPS Workshop Paper arXiv Poster
2024
- State and Action Factorization in Power Grids ECML 2024 MLSPS Workshop Paper arXiv Poster Slides
- Open Problem: Tight Bounds for Bernoulli Rewards in Kernelized Multi-Armed Bandits ICML 2024 ARLET Workshop Link Paper Poster
- Intermediate Observations in Factored-Reward Bandits AAMAS 2024 ALA Workshop Link Paper Slides
2023
- Online Learning in Autoregressive Dynamics EWRL 2023 Link Paper Poster
- Stochastic Rising Bandits: A Best Arm Identification Approach EWRL 2023 Link Paper Poster
- A Best Arm Identification Approach for Stochastic Rising Bandits ICML 2023 F4LCD Workshop Link Paper Poster
2022
Book Chapters
2026
Preprints
2026
- Learning a Ranking from Human Feedback in Log-Concave Random Utility Models arXiv:2610.07973 Paper arXiv
- Reusing Past Samples in Proximal Policy Optimization: When and How Does It Help? arXiv:2610.01399 Paper arXiv
2025
- Online Dynamic Pricing of Complementary Products arXiv:2511.22291 Paper arXiv
- Generalized Kernelized Bandits: A Novel Self-Normalized Bernstein-Like Dimension-Free Inequality and Regret Bounds arXiv:2508.01681 Paper arXiv
- Gym4ReaL: A Suite for Benchmarking Real-World Reinforcement Learning arXiv:2507.00257 Paper arXiv
- Learning Deterministic Policies with Policy Gradients in Constrained Markov Decision Processes arXiv:2506.05953 Paper arXiv
- A refined Analysis of UCBVI arXiv:2502.17370 Paper arXiv
2024
Technical Reports
2024
* Equal Contribution.
Experience
- Jan 2026 - now Assistant Professor Politecnico di Milano Dipartimento di Elettronica, Informazione e Bioingegneria
- Jun 2024 - Jan 2026 Postdoctoral Researcher Politecnico di Milano Dipartimento di Elettronica, Informazione e Bioingegneria
- Nov 2020 - Jun 2024 Research Scientist ML cube During Ph.D. in Information Technology
- Jan 2020 - Oct 2020 Research Assistant Politecnico di Milano Dipartimento di Elettronica, Informazione e Bioingegneria
Education
- Nov 2020 - Jun 2024 Ph.D. in Information Technology Politecnico di Milano Focus on Reinforcement Learning and Online Learning Advisor: Prof. Marcello Restelli Link Thesis Slides
- Sep 2017 - Dec 2019 M.Sc. in Computer Science and Engineering Politecnico di Milano
- Sep 2014 - Jul 2017 B.Sc. in Engineering of Computing Systems Politecnico di Milano
- Sep 2008 - Jul 2014 High School Diploma in Computer Science IIS Galileo Galilei Crema Main focus: C, Java, HTML, CSS, Javascript