课题基金 / 基金详情

RI: Small: Random Perturbation Methods in Sequential Learning

RI: Small: Random Perturbation Methods in Sequential Learning
RI:小:顺序学习中的随机扰动方法
批准号:
2007055
负责人:
Ambuj Tewari
金额:
$45.0万
依托单位国家:
美国
项目类别:
Continuing Grant
财政年份:
2020
资助国家:
美国
项目状态:
已结题
起止时间:
2020-10-01 至 2024-09-30

项目摘要

项目成果

Ambuj Tewari的其他基金

相似基金

相关文献

中文摘要
翻译
点击翻译按钮获取中文摘要
英文摘要
Neither babies nor machines begin learning from a blank slate. Just like a baby comes into the world with brain structures that predispose her to learn motor and language skills, a machine has to be given enough prior structure to help it learn. This prior structure is called inductive bias in the field of machine learning. Inductive bias can take many forms, which is why there are many different sorts of machine learning algorithms. For example, the machine could be told that similar inputs should produce similar outputs, or that it should prefer simpler models over complex ones. Recently a class of methods has emerged that uses randomness to inject inductive bias into machine learning algorithms. However, researchers do not fully understand the power and limitations of these methods. For example, what is the relationship between injecting randomness and having a preference for simpler models? This project studies such fundamental questions about the power of randomness in designing machine learning algorithms. The algorithms developed in this project can be applied to many problems of practical interest including the discovery of cheap renewable energy sources.The technical goals of this project are divided into three categories according to the underlying sequential learning problem: online learning, bandit problems, and reinforcement learning. In online learning, the project examines the universality of perturbations. That is, are perturbation-based algorithms powerful enough to realize optimal performance guarantees in any online convex optimization problem? This work also aims to discover universal perturbation-based online learning algorithms that succeed in learning a problem as soon as the problem is online learnable. In bandit problems, random perturbations are used to design algorithms that are robust to non-stationarity and corruptions in the observed rewards. In reinforcement learning, exploration strategies based on random perturbations are designed that are both computationally tractable and sample efficient.This award reflects NSF's statutory mission and has been deemed worthy of support through evaluation using the Foundation's intellectual merit and broader impacts review criteria.
期刊论文(17)
专著(0)
科研奖励(0)
会议论文
Online Agnostic Multiclass Boosting
在线不可知多类提升
DOI: --
发表时间: 2022
期刊: Advances in Neural Information Processing Systems 35}
影响因子: --
作者: [Raman, Vinod, Tewari, Ambuj]
通讯作者: Tewari, Ambuj
DOI: --
发表时间: 2021-11
期刊:
影响因子: --
作者: [Ziping Xu;Ambuj Tewari]
通讯作者: Ziping Xu;Ambuj Tewari
DOI: --
发表时间: 2021-07
期刊: ArXiv
影响因子: --
作者: [Yuntian Deng-;Xingyu Zhou;Baekjin Kim;Ambuj Tewari;Abhishek Gupta;N. Shroff]
通讯作者: Yuntian Deng-;Xingyu Zhou;Baekjin Kim;Ambuj Tewari;Abhishek Gupta;N. Shroff
DOI: 10.48550/arxiv.2211.09403
发表时间: 2022-11
期刊:
影响因子: --
作者: [Chinmaya Kausik;Kevin Tan;Ambuj Tewari]
通讯作者: Chinmaya Kausik;Kevin Tan;Ambuj Tewari
15
    Efficient Algorithms with Statistical Guarantees for High Dimensional Time Series
    CAREER: New Frontiers in Sequential Decision Making with a View Towards Mobile Health Applications
    RI: Small: Collaborative Research: Statistical ranking theory without a canonical loss
    国内基金
    海外基金
    昼夜节律性small RNA在血斑形成时间推断中的法医学应用研究
    • 批准号:
    • 项目类别:
      省市级项目
    • 资助金额:
      --
    • 批准年份:
      2024
    • 负责人:
    • 依托单位:
    tRNA-derived small RNA上调YBX1/CCL5通路参与硼替佐米诱导慢性疼痛的机制研究
    • 批准号:
    • 项目类别:
      省市级项目
    • 资助金额:
      10.0万元
    • 批准年份:
      2022
    • 负责人:
      张祥忠
    • 依托单位:
    Small RNA调控I-F型CRISPR-Cas适应性免疫性的应答及分子机制
    Small RNAs调控解淀粉芽胞杆菌FZB42生防功能的机制研究
    • 批准号:
      31972324
    • 项目类别:
      面上项目
    • 资助金额:
      58.0万元
    • 批准年份:
      2019
    • 负责人:
      高学文
    • 依托单位: