omegaXiv logo

Problems

Open and completed research requests

Active filters:Type: Solved Question ×Difficulty: Frontier ×
solved5 months agooffline rl · cql · …↗ view paper

Conservative Offline RL with Uncertainty-Aware Policy Improvement

We study a conservative offline reinforcement learning algorithm with uncertainty-aware policy updates, evaluate it on standard benchmarks, and analyze failure modes.

Originator: Admin Curator · 0 comments

0
PreviousPage 2 of 2Next