hiro/ray - Forgejo: Beyond coding. We Forge.

hiro/ray

mirror of https://github.com/vale981/ray synced 2025-03-06 02:21:39 -05:00

Author	SHA1	Message	Date
kourosh hakhamaneshi	4607e788c1	[RLlib] Fix test_ope flakiness (#27676 )	2022-08-09 16:12:30 -07:00
Rohan Potdar	5b6a58ed28	[RLlib] Add OPE Learning Tests (#27154 )	2022-08-02 17:51:38 -07:00
Rohan Potdar	deccf33912	[RLlib]: Add Off-Policy Estimation docs (#26809 ) Co-authored-by: Kourosh Hakhamaneshi <kourosh@anyscale.com>	2022-07-26 13:57:56 -07:00
Rohan Potdar	4fded80813	[RLlib]: Fix FQE Policy call (#26671 )	2022-07-19 00:58:31 -07:00
Rohan Potdar	38c9e1d52a	[RLlib]: Fix OPE trainables (#26279 ) Co-authored-by: Kourosh Hakhamaneshi <kourosh@anyscale.com>	2022-07-17 14:25:53 -07:00
Rohan Potdar	09ce4711fd	[RLlib]: Move OPE to evaluation config (#25911 )	2022-07-12 11:04:34 -07:00
Rohan Potdar	28df3f34f5	[RLlib]: Off-Policy Evaluation fixes. (#25899 )	2022-06-21 13:24:24 +02:00
Rohan Potdar	a9d8da0100	[RLlib]: Doubly Robust Off-Policy Evaluation. (#25056 )	2022-06-07 12:52:19 +02:00