ShopSimRL
A four-round applied study of one idea: let an agent policy learn while its external procedural skill is measured, revised, and retained where it still helps. Read the project report ↓
Description read from the page itself, because the listing it was
found in gave none.