关于Cesa-Bianchi和Lugosi《预测、学习与博弈》中定理2.3的注记
机器学习
2010-11-29 v1
摘要
本文给出了时变势函数指数加权平均预测器损失界的一个修正证明。该算法的遗憾项被上界为 sqrt{n ln(N)}(关于 n 一致),其中 N 是专家数量,n 是步数。
引用
@article{arxiv.1011.5668,
title = {On Theorem 2.3 in "Prediction, Learning, and Games" by Cesa-Bianchi and Lugosi},
author = {Alexey Chernov},
journal= {arXiv preprint arXiv:1011.5668},
year = {2010}
}
备注
3 pages; excerpt from arXiv:1005.1918, simplified and rewritten using the notation of the monograph by Cesa-Bianchi and Lugosi