Multi-armed Bandits with Compensation
First TextWorld Problems—Microsoft Research Montreal’s latest AI competition is really cooking
This week, Microsoft Research threw down the gauntlet with the launch of a competition challenging researchers around the world to develop AI agents that can solve text-based games. Conceived by the Machine Reading Comprehension team…
A Deep Learning Theory: Global minima and over-parameterization
One empirical finding in deep learning is that simple methods such as stochastic gradient descent (SGD) have a remarkable ability to fit training data. From a capacity perspective, this may not be surprising— modern neural…
Fast, accurate, stable and tiny – Breathing life into IoT devices with an innovative algorithmic approach
In the larger quest to make the Internet of Things (IoT) a reality for people everywhere, building devices that can be both ultrafunctional and beneficent isn’t a simple matter. Particularly in the arena of resource-constrained,…
Learning to teach: Mutually enhanced learning and teaching for artificial intelligence
Teaching is super important. From an individual perspective, a student learning on his or her own is never ideal; a student needs a teacher’s guidance and perspective to be more effectively educated. Taking the societal…