Skip to content

holi-lab/POISE

Folders and files

NameName
Last commit message
Last commit date

Latest commit

 

History

16 Commits
 
 
 
 
 
 
 
 

About

Official code release for "Your Language Model is Its Own Critic: Reinforcement Learning with Value Estimation from Actor’s Internal States"

Stars

Watchers

Forks

Releases

No releases published

Packages

 
 
 

Contributors

Languages