hiro/ray - Forgejo: Beyond coding. We Forge.

hiro/ray

mirror of https://github.com/vale981/ray synced 2025-03-06 10:31:39 -05:00

Author	SHA1	Message	Date
matthewdeng	138b273136	[rllib] Add tests for examples using ray client (#16271 ) * [rllib] add tests for examples using ray client * rename test_client to test_ray_client	2021-06-09 10:39:14 -07:00
Eric Liang	810f5c803a	Disable flaky object spilling test on OSX & adjust test timeouts (#15986 ) * blacklist * move it * adjust according to bazel timeouts * fix build * move to large * Update BUILD	2021-05-24 09:49:59 -07:00
Sven Mika	2d34216660	[RLlib] APEX-DQN: Bug fix for torch and add learning test. (#15762 )	2021-05-20 09:27:03 +02:00
Sven Mika	d2c755ccef	[RLlib] Examples scripts add argparse help and replace `--torch` with `--framework`. (#15832 )	2021-05-18 13:18:12 +02:00
Sven Mika	308ea62430	[RLlib] Fix "seed" setting to work in all frameworks and w/ all CUDA versions. (#15682 )	2021-05-18 11:00:24 +02:00
Sven Mika	f25d58492d	[Testing] Dependabot for RLlib. (#15812 )	2021-05-17 18:24:13 +02:00
Sven Mika	d89fb82bfb	[RLlib] Add simple curriculum learning API and example script. (#15740 )	2021-05-16 17:35:10 +02:00
Sven Mika	bc09e75b78	[RLlib] Fix 3 flakey test cases. (#15785 )	2021-05-16 12:20:33 +02:00
Ian Rodney	00c913cbc6	[Flaky] Mark `test_nested_observation_spaces` as Flaky (#15794 )	2021-05-14 12:08:52 -07:00
Sven Mika	c4a3e1589b	[RLlib] CQL: Bug fixes and OPE example added to test and offline_rl.py example. (#15761 )	2021-05-13 09:17:23 +02:00
Sven Mika	16ddab49f5	[RLlib] Trainer._evaluate -> Trainer.evaluate; Also make evaluation possible w/o evaluation worker set. (#15591 )	2021-05-12 12:16:00 +02:00
Sven Mika	46f6fa2361	[RLlib] Example script for restoring 1 agent (out of n) from a checkpoint (multi-agent). (#15540 )	2021-05-10 16:09:05 +02:00
Eric Liang	ff36ae594b	Remove flaky tag from newly unflaky tests (#15639 )	2021-05-05 12:15:46 -07:00
Kai Fricke	1d52ab819f	[release] release 1.3.0 results and test updates (#15366 ) Convert a number of release tests and add logs for release 1.3.0	2021-05-04 22:10:04 +01:00
Sven Mika	e973b726c2	[RLlib] Support native tf.keras.Models (part 2) - Default keras models for Vision/RNN/Attention. (#15273 )	2021-04-30 19:26:30 +02:00
Sven Mika	fc3a65f9d4	[RLlib] Split test_checkpoint_restore tests into 3 and make each "large" (from "enormous"). (#15499 )	2021-04-30 12:33:12 +02:00
Sven Mika	bb8a286cbc	[RLlib] Support native tf.keras.Model (milestone toward obsoleting ModelV2 class). (#14684 )	2021-04-27 10:44:54 +02:00
Eric Liang	af01a47d59	Add support for tune,serve,rllib tests to flaky builder (#15447 )	2021-04-22 15:03:29 -07:00
Sven Mika	41968512ca	[RLlib] Partial GPU examples (for learner and workers). (#15334 )	2021-04-20 08:46:05 +02:00
SangBin Cho	1d87e4447d	[Test] increase the test size of test io that consistenly times out (#15341 )	2021-04-15 14:02:41 -07:00
Sven Mika	45d6560759	[RLlib] Fix flakey custom_fast_model_torch/tf tests. (#15330 )	2021-04-15 16:10:29 +02:00
SangBin Cho	27ab0c7633	[Test] Skip the failing rllib example test. (#15321 )	2021-04-14 20:19:44 -07:00
Sven Mika	5254d2fb36	[RLlib] Support parallelizing evaluation and training (optional). (#15040 )	2021-04-13 09:53:35 +02:00
Eric Liang	b90cc51c27	[RLlib] Attempt splitting rollout test to avoid initial timeout (#14999 )	2021-03-30 19:20:02 +02:00
Eric Liang	af8a93f2a4	Deflake some RLlib tests (#14947 ) * fix * update * 100 * flake	2021-03-26 11:45:17 -07:00
Raoul Khouri	c37fa3f389	[RLlib] Example and test for custom Trainer wrapper experiments (#14652 )	2021-03-24 16:22:46 +01:00
Sven Mika	78c64ca151	[RLlib] Attention net example script: Clarifications on how to use with Trainer.compute_action. (#14864 )	2021-03-23 19:33:01 +01:00
Sven Mika	69202c6a7d	[RLlib] Obsolete usage tracking dict via sample batch. (#13065 )	2021-03-17 08:18:15 +01:00
Michael Luo	020c9439dd	[RLlib] CQL Documentation + Tests (#14531 )	2021-03-11 18:51:39 +01:00
Sven Mika	732197e23a	[RLlib] Multi-GPU for tf-DQN/PG/A2C. (#13393 )	2021-03-08 15:41:27 +01:00
Sven Mika	ef944bc5f0	[RLlib] Re-enable placement group support for RLlib. (#14384 )	2021-03-05 08:16:24 +01:00
Sven Mika	5637d89ecc	[RLlib] Serve + RLlib example script. (#14416 )	2021-03-03 14:33:03 +01:00
Richard Liaw	a2d2275ee1	Revert "[RLlib + Tune] Add placement group support to RLlib. (#14289 )" (#14360 ) This reverts commit `6cd0cd3bd9`.	2021-02-25 14:27:35 -08:00
Sven Mika	4cd5c1da2c	[RLlib] Remove flaky test case for mixed (tf+torch) policies trainer. (#14357 )	2021-02-25 14:07:05 -08:00
Sven Mika	6cd0cd3bd9	[RLlib + Tune] Add placement group support to RLlib. (#14289 )	2021-02-25 16:01:31 +01:00
Sven Mika	8000258333	[RLlib] R2D2 Implementation. (#13933 )	2021-02-25 12:18:11 +01:00
Sven Mika	3d20d58c90	[RLlib] Tune trial + checkpoint selection example. (#14209 )	2021-02-22 12:52:37 +01:00
Sven Mika	4db86404ad	[RLlib] Issue #13507 : Fix MB-MPO CartPole Env's reward function as well as MB-MPO running into a traj. view API related issue. (#14037 )	2021-02-11 18:58:46 +01:00
Sven Mika	eb0038612f	[RLlib] Extend on_learn_on_batch callback to allow for custom metrics to be added. (#13584 )	2021-02-08 15:02:19 +01:00
Sven Mika	d001af3e59	[RLlib] Allow `rllib rollout` to run distributed via evaluation workers. (#13718 )	2021-02-08 12:05:16 +01:00
Sven Mika	0a0d9183fe	[RLlib] Trajectory view API example script (enhancements and tf2 support). (#13786 )	2021-02-02 18:42:18 +01:00
Raoul Khouri	714c367b9d	[RLlib] Trainer._validate_config idempotentcy correction (issue 13427) (#13556 )	2021-02-02 13:11:57 +01:00
Yuri Rocha	b01b0f80aa	[RLlib] Fix multiple Unity3DEnvs trying to connect to the same custom port (#13519 )	2021-01-28 13:28:08 +01:00
Sven Mika	daf0bef285	[RLlib] Dreamer: Fix broken import and add compilation test case. (#13553 )	2021-01-21 16:30:26 +01:00
Sven Mika	e74947cc94	[RLlib] Env directory cleanup and tests. (#13082 )	2021-01-19 10:09:39 +01:00
Sven Mika	56878221ed	[RLlib] Redo: Make TFModelV2 fully modular like TorchModelV2 (soft-deprecate register_variables, unify var names wrt torch). (#13363 )	2021-01-14 14:44:33 +01:00
Kai Fricke	25f10a947a	Revert "[RLlib] Make TFModelV2 behave more like TorchModelV2: Obsolete register_variables. Unify variable dicts. (#13339 )" (#13361 ) This reverts commit `e2b2abb88b`.	2021-01-12 12:33:57 +01:00
Sven Mika	e2b2abb88b	[RLlib] Make TFModelV2 behave more like TorchModelV2: Obsolete register_variables. Unify variable dicts. (#13339 )	2021-01-11 22:42:30 +01:00
Sven Mika	6f342a2221	[RLlib] Preparatory PR for: Documentation on Model Building. (#13260 )	2021-01-08 10:56:09 +01:00
Sven Mika	28ac4243f4	[RLlib] Deflake test case: 2-step game MADDPG. (#13121 )	2020-12-30 18:37:37 -05:00

1 2 3 4

159 commits