Shie Mannor
Researcher Next ID · RN-042065
Researcher · Computer Science
Reading, Israel
- Works count
- 634
- Citation count
- 19,016
- H-index
- 62
- i10-index
- 249
Research interests
Publications
Explainable Artificial Intelligence (XAI) techniques for energy and power systems: Review, challenges and opportunities
Energy and AI · 2022 · https://doi.org/10.1016/j.egyai.2022.100169
Deep Learning Reconstruction of Ultrashort Pulses
Conference on Lasers and Electro-Optics · 2018 · 10.1364/cleo_si.2018.sth4n.1
Reward Constrained Policy Optimization
arXiv (Cornell University) · 2018 · https://doi.org/10.48550/arxiv.1805.11074
Encouraging Physical Activity in Patients With Diabetes: Intervention Using a Reinforcement Learning System
Journal of Medical Internet Research · 2017 · 10.2196/jmir.7994
A Deep Hierarchical Approach to Lifelong Learning in Minecraft
Proceedings of the AAAI Conference on Artificial Intelligence · 2017 · https://doi.org/10.1609/aaai.v31i1.10744
Bayesian Reinforcement Learning: A Survey
Foundations and Trends® in Machine Learning · 2015 · https://doi.org/10.1561/2200000049
Robustness and generalization
Machine Learning · 2011 · https://doi.org/10.1007/s10994-011-5268-1
Sparse Algorithms Are Not Stable: A No-Free-Lunch Theorem
IEEE Transactions on Pattern Analysis and Machine Intelligence · 2011 · https://doi.org/10.1109/tpami.2011.177
Robust Regression and Lasso
IEEE Transactions on Information Theory · 2010 · 10.1109/tit.2010.2048503
Percentile Optimization for Markov Decision Processes with Parameter Uncertainty
Operations Research · 2009 · https://doi.org/10.1287/opre.1080.0685
Fully Parallel Stochastic LDPC Decoders
IEEE Transactions on Signal Processing · 2008 · https://doi.org/10.1109/tsp.2008.929671
Robustness and Regularization of Support Vector Machines
arXiv (Cornell University) · 2008 · https://doi.org/10.48550/arxiv.0803.3490
Action Elimination and Stopping Conditions for the Multi-Armed Bandit and Reinforcement Learning Problems
Journal of Machine Learning Research · 2006
Stochastic decoding of LDPC codes
IEEE Communications Letters · 2006 · https://doi.org/10.1109/lcomm.2006.060570
Automatic basis function construction for approximate dynamic programming and reinforcement learning
Journal · 2006 · https://doi.org/10.1145/1143844.1143901
The cross entropy method for classification
Journal · 2005 · https://doi.org/10.1145/1102351.1102422
A Tutorial on the Cross-Entropy Method
Annals of Operations Research · 2005 · https://doi.org/10.1007/s10479-005-5724-z
Reinforcement learning with Gaussian processes
Journal · 2005 · https://doi.org/10.1145/1102351.1102377
Basis Function Adaptation in Temporal Difference Reinforcement Learning
Annals of Operations Research · 2005 · https://doi.org/10.1007/s10479-005-5732-z
The Kernel Recursive Least-Squares Algorithm
IEEE Transactions on Signal Processing · 2004 · https://doi.org/10.1109/tsp.2004.830985
Dynamic abstraction in reinforcement learning via clustering
Journal · 2004 · https://doi.org/10.1145/1015330.1015355
The Sample Complexity of Exploration in the Multi-Armed Bandit Problem
Journal · 2004 · https://doi.org/10.5555/1005332.1005355
Bayes meets bellman: the Gaussian process approach to temporal difference learning
Journal · 2003
PAC Bounds for Multi-armed Bandit and Markov Decision Processes
Lecture notes in computer science · 2002 · https://doi.org/10.1007/3-540-45435-7_18
Q-Cut—Dynamic Discovery of Sub-goals in Reinforcement Learning
Lecture notes in computer science · 2002 · https://doi.org/10.1007/3-540-36755-1_25
Current projects
No projects listed.