Skip to main content Skip to main navigation menu Skip to site footer
Proceedings of the AAAI Conference on Artificial Intelligence
  • Current
  • Archives
  • About
    • About the Journal
    • Submissions
    • Privacy Statement
    • Contact
  • Login
  1. Home /
  2. Search

Search

Advanced filters
Published After
Published Before

Search Results

Found 14874 items.
  • Reinforcement Learning Based Multi-Agent Resilient Control: From Deep Neural Networks to an Adaptive Law

    Jian Hou, Fangyuan Wang, Lili Wang, Zhiyong Chen
    7737-7745
    2021-05-18
  • Reward-Biased Maximum Likelihood Estimation for Linear Stochastic Bandits

    Yu-Heng Hung, Ping-Chun Hsieh, Xi Liu, P. R. Kumar
    7874-7882
    2021-05-18
  • Asynchronous Optimization Methods for Efficient Training of Deep Neural Networks with Guarantees

    Vyacheslav Kungurtsev, Malcolm Egan, Bapi Chatterjee, Dan Alistarh
    8209-8216
    2021-05-18
  • A Free Lunch for Unsupervised Domain Adaptive Object Detection without Source Data

    Xianfeng Li, Weijie Chen, Di Xie, Shicai Yang, Peng Yuan, Shiliang Pu, Yueting Zhuang
    8474-8481
    2021-05-18
  • Stochastic Graphical Bandits with Adversarial Corruptions

    Shiyin Lu, Guanghui Wang, Lijun Zhang
    8749-8757
    2021-05-18
  • Semi-supervised Medical Image Segmentation through Dual-task Consistency

    Xiangde Luo, Jieneng Chen, Tao Song, Guotai Wang
    8801-8809
    2021-05-18
  • Sequential Attacks on Kalman Filter-based Forward Collision Warning Systems

    Yuzhe Ma, Jon A Sharp, Ruizhe Wang, Earlence Fernandes, Xiaojin Zhu
    8865-8873
    2021-05-18
  • Advice-Guided Reinforcement Learning in a non-Markovian Environment

    Daniel Neider, Jean-Raphael Gaglione, Ivan Gavran, Ufuk Topcu, Bo Wu, Zhe Xu
    9073-9080
    2021-05-18
  • Top-k Ranking Bayesian Optimization

    Quoc Phong Nguyen, Sebastian Tay, Bryan Kian Hsiang Low, Patrick Jaillet
    9135-9143
    2021-05-18
  • Warm Starting CMA-ES for Hyperparameter Optimization

    Masahiro Nomura, Shuhei Watanabe, Youhei Akimoto, Yoshihiko Ozaki, Masaki Onishi
    9188-9196
    2021-05-18
  • Robust Reinforcement Learning: A Case Study in Linear Quadratic Regulation

    Bo Pang, Zhong-Ping Jiang
    9303-9311
    2021-05-18
  • Shuffling Recurrent Neural Networks

    Michael Rotman, Lior Wolf
    9428-9435
    2021-05-18
  • Meta-Learning Effective Exploration Strategies for Contextual Bandits

    Amr Sharaf, Hal Daumé III
    9541-9548
    2021-05-18
  • Strategy and Benchmark for Converting Deep Q-Networks to Event-Driven Spiking Neural Networks

    Weihao Tan, Devdhar Patel, Robert Kozma
    9816-9824
    2021-05-18
  • Meta Learning for Causal Direction

    Jean-François Ton, Dino Sejdinovic, Kenji Fukumizu
    9897-9905
    2021-05-18
  • Quantum Exploration Algorithms for Multi-Armed Bandits

    Daochen Wang, Xuchen You, Tongyang Li, Andrew M. Childs
    10102-10110
    2021-05-18
  • Adaptive Algorithms for Multi-armed Bandit with Composite and Anonymous Feedback

    Siwei Wang, Haoyun Wang, Longbo Huang
    10210-10217
    2021-05-18
  • Robust Bandit Learning with Imperfect Context

    Jianyi Yang, Shaolei Ren
    10594-10602
    2021-05-18
  • On Convergence of Gradient Expected Sarsa(λ)

    Long Yang, Gang Zheng, Yu Zhang, Qian Zheng, Pengfei Li, Gang Pan
    10621-10629
    2021-05-18
  • Sequential Generative Exploration Model for Partially Observable Reinforcement Learning

    Haiyan Yin, Jianda Chen, Sinno Jialin Pan, Sebastian Tschiatschek
    10700-10708
    2021-05-18
  • Learning Modality-Specific Representations with Self-Supervised Multi-Task Learning for Multimodal Sentiment Analysis

    Wenmeng Yu, Hua Xu, Ziqi Yuan, Jiele Wu
    10790-10797
    2021-05-18
  • A Primal-Dual Online Algorithm for Online Matching Problem in Dynamic Environments

    Yu-Hang Zhou, Peng Hu, Chen Liang, Huan Xu, Guangda Huzhang, Yinfu Feng, Qing Da, Xinshang Wang, An-Xiang Zeng
    11160-11167
    2021-05-18
  • Scalable and Safe Multi-Agent Motion Planning with Nonlinear Dynamics and Bounded Disturbances

    Jingkai Chen, Jiaoyang Li, Chuchu Fan, Brian C. Williams
    11237-11245
    2021-05-18
  • The Influence of Memory in Multi-Agent Consensus

    David Kohan Marzagão, Luciana Basualdo Bonatto, Tiago Madeira, Marcelo Matheus Gauy, Peter McBurney
    11254-11262
    2021-05-18
  • Exploration-Exploitation in Multi-Agent Learning: Catastrophe Theory Meets Game Theory

    Stefanos Leonardos, Georgios Piliouras
    11263-11271
    2021-05-18
9976 - 10000 of 14874 items << < 395 396 397 398 399 400 401 402 403 404 > >> 

Information

  • For Readers
  • For Authors
  • For Librarians
  • Part of the
    PKP Publishing Services Network

Copyright © 2024, Association for the Advancement of Artificial Intelligence

More information about the publishing system, Platform and Workflow by OJS/PKP.