Scholay

学术搜索 · AI 审稿 · LaTeX 协作

Csaba Szepesvári

ORCID:0000-0002-9286-2892

发表论文 444 篇 · 总被引 16838 次 · h-index 56

代表论文

  • When Is Partially Observable Reinforcement Learning Not Scary? (2022 · arXiv (Cornell University) · 被引 8)
  • Optimistic MLE: A Generic Model-Based Algorithm for Partially Observable Sequential Decision Making (2023 · 被引 7)
  • Sample-Efficient Reinforcement Learning of Partially Observable Markov Games (2022 · arXiv (Cornell University) · 被引 7)
  • Bandit Theory and Thompson Sampling-Guided Directed Evolution for Sequence Optimization (2022 · arXiv (Cornell University) · 被引 4)
  • To Believe or Not to Believe Your LLM (2024 · arXiv (Cornell University) · 被引 3)
  • Near-Optimal Sample Complexity Bounds for Constrained MDPs (2022 · arXiv (Cornell University) · 被引 3)