hiro/ray - Forgejo: Beyond coding. We Forge.

hiro/ray

mirror of https://github.com/vale981/ray synced 2025-03-06 18:41:40 -05:00

Author	SHA1	Message	Date
Will Drevo	fa878e2d4d	Added example to user guide for cloud checkpointing (#20045 ) Co-authored-by: will <will@anyscale.com> Co-authored-by: Antoni Baum <antoni.baum@protonmail.com> Co-authored-by: Kai Fricke <kai@anyscale.com>	2021-11-15 15:43:06 +00:00
Amog Kamsetty	a74cf7ff1c	[Train] Torch Prepare utilities (#20254 ) * update * formatting * fix failures * fix session tests * address comments * add to api docs * package refactor * wip * wip * wip * finish * finish * fix * comment * fix * install horovod for docs * address comment * Update python/ray/train/session.py Co-authored-by: matthewdeng <matthew.j.deng@gmail.com> * Update python/ray/train/torch.py Co-authored-by: matthewdeng <matthew.j.deng@gmail.com> * address comments * try fix docs * fix doc build failure * fix * fix * fix * try fix doc highlighting * fix docs Co-authored-by: matthewdeng <matthew.j.deng@gmail.com>	2021-11-15 07:34:17 -08:00
Qing Wang	1172195571	[Java] Remove global named actor and global pg (#20135 ) This PR removes global named actor and global PGs. I believe these APIs are not used widely in OSS. CPP part is not included in this PR. @kfstorm @clay4444 @raulchen Please take a look if this change is reasonable. IMPORTANT NOTE: This is a Java API change and will lead backward incompatibility in Java global named actor and global PG usage. CPP part is not included in this PR. INCLUDES: Remove setGlobalName() and getGlobalActor() APIs. Remove getGlobalPlacementGroup() and setGlobalPG Add getActor(name, namespace) API Add getPlacementGroup(name, namespace) API Update doc pages.	2021-11-15 16:28:53 +08:00
matthewdeng	e22632dabc	[train] wrap BackendExecutor in ray.remote() (#20123 ) * [train] wrap BackendExecutor in ray.remote() * wip * fix trainer tests * move CheckpointManager to Trainer * [tune] move force_on_current_node to ml_utils * fix import * force on head node * init ray * split test files * update example * move tests to ray client * address comments * move comment * address comments	2021-11-13 15:30:44 -08:00
Sven Mika	e5ead6a4b0	[RLlib; Documentation] Minor fixes "rllib in 60s" and per-feature sigils. (#20248 )	2021-11-13 22:10:47 +01:00
Amog Kamsetty	65a17da2ec	[Train] Refactor Backends (#20312 ) * wip * finish * comment * fix * install horovod for docs * address comment * fix doc build failure	2021-11-13 11:05:53 -08:00
Antoni Baum	1b867520e6	[docs]Add pyarrow as a dependency (#20320 )	2021-11-13 16:00:58 +00:00
matthewdeng	e77cc926be	[train] minor doc updates (#20271 )	2021-11-12 17:20:23 -08:00
Tricia Fu	e59c14117f	[Doc] [Serve] Add summary sub header to each page (#20231 )	2021-11-12 14:18:42 -08:00
xwjiang2010	cdf70c2900	[Tune] Remove legacy resources implementations in Runner and Executor. (#19773 )	2021-11-12 12:33:39 -08:00
Siyuan (Ryans) Zhuang	3b62388a9a	[Workflow] Workflow tail recursion optimization (#19928 ) * tail recursion optimization	2021-11-12 09:13:40 -08:00
Kai Fricke	246787cdd9	Revert "[RLlib] POC: `PGTrainer` class that works by sub-classing, not `trainer_template.py`. (#20055 )" (#20284 ) This reverts commit `6f85af435f`.	2021-11-12 13:09:43 +00:00
Kai Fricke	d88fdd6e38	[tune] refactor SyncConfig (#20155 )	2021-11-12 09:36:15 +00:00
Michael Galarnyk	dbeb2e2f73	Add Ray Serve Blogs to Doc(#19846 ) The Serving ML Models in Production blog links is inline with the latest Ray Summit talk on Ray Serve.	2021-11-11 15:10:36 -08:00
Edward Oakes	59698aa89c	[Serve] add survey link (#20230 )	2021-11-11 15:10:10 -08:00
Jules S. Damji	71a162d8ab	Fixed code snippet to include config parameter and a minor typo (#20193 ) Signed-off-by: Jules S.Damji <jules@anyscale.com> Co-authored-by: Jules S.Damji <jules@anyscale.com>	2021-11-11 18:37:03 +00:00
Dmitri Gekhtman	8971422d8f	[autoscaler] Use drain node api in autoscaler before terminating nodes (#20013 ) * wip * Draft * Use bytest for node id * remove stray helm change * fix autoscaler init arg * don't forget to instantiate new load metrics dict * remove extraneous diff * Timeout, comments, function signature. * typo * another comment * tweak * docstring * shorter timeout * Use a better error code * missing self * Dedent example * Add drain node prometheus metric. * comment * Update tests part 1: test_autoscaler.py * Update tests part 2: test_resource_demand_scheduler * lint * Update tests part 3: test_autoscaling_policy * Unit tests for new Prometheus metric and DrainNode error handling. * comment * removed unused function * Try adding ability to mock out process termination to fake node provider * Add integration test. * fix * fix * lint * Improve log message * fix * Simplify test * Fix doc example * remove unused dict * Mock out process termination in a subclass * Add add doc string and comment explaining prune active ips. * Comment: wtf is use_node_id_as_ip * one more comment * more explanation * period * tweak	2021-11-11 08:31:40 -08:00
Sven Mika	6f85af435f	[RLlib] POC: `PGTrainer` class that works by sub-classing, not `trainer_template.py`. (#20055 )	2021-11-11 12:16:20 +01:00
Will Drevo	2fdb1c46c7	[RLlib; Documentation] Added atari pip installs to Pong-v0 example. (#20225 ) * Added imports to Pongv0 example * Added comment * Apply suggestions from code review Co-authored-by: will <will@anyscale.com> Co-authored-by: Sven Mika <sven@anyscale.io>	2021-11-11 09:08:02 +01:00
Tobias Kaymak	893f57591d	[serve] Add Google Cloud Storage as a backend (#20104 )	2021-11-10 19:45:19 -08:00
Edward Oakes	082a4af3e6	[serve] Remove lingering backend/endpoint wording in docs (#20229 )	2021-11-10 16:49:29 -08:00
Sven Mika	ebd56b57db	[RLlib; documentation] "RLlib in 60sec" overhaul. (#20215 )	2021-11-10 22:20:06 +01:00
matthewdeng	790e22f9ad	[tune] move force_on_current_node to ml_utils (#20211 )	2021-11-10 10:21:24 -08:00
Sven Mika	143d23a278	[RLlib] Issue 20062: Action inference examples missing (#20144 )	2021-11-10 18:49:06 +01:00
Kim Pevey	82a5bf68fa	[Docs] Add note for multi-node on Windows (#20184 ) * add note for multi-node on Windows * update message Co-authored-by: Philipp Moritz <pcmoritz@gmail.com>	2021-11-09 16:02:01 -08:00
Kim Pevey	bbacb6d828	[Docs] Streaming MapReduce: Remove breaking city (#20187 )	2021-11-09 11:15:57 -08:00
Kai Fricke	9c2b8c8501	[tune] Deprecate DurableTrainable (#19880 )	2021-11-08 20:56:07 +00:00
Amog Kamsetty	b1f24768a1	[Tune] More fixes to PTL Tutorial (#20065 ) * ptl-fix-2 * improve * fix	2021-11-08 09:13:44 -08:00
Jules S. Damji	e6343f0e69	Fixed a broken code snippet with a missing method (#20130 ) Signed-off-by: Jules S.Damji <jules@anyscale.com> Co-authored-by: Jules S.Damji <jules@anyscale.com>	2021-11-08 07:56:32 +09:00
Alex Wu	81194f5660	[workflow][docs] Fix api comparison formatting (#20069 ) ## Why are these changes needed? The API comparison formatting uses \`code\` which is rendered as italicization not code. This PR puts the code in code blocks instead of italics. ## Related issue number ## Checks	2021-11-05 17:05:35 -07:00
matthewdeng	78e9ff7c91	[train][datasets] add example for big data training (#20042 ) * [train][datasets] add example for big data training * add title docstring * lint and dependencies * add dask_ml requirement	2021-11-05 09:28:48 -07:00
Chen Shen	320f9dc234	[Core][CoreWorker] increase the default port range (#19541 ) * increase the port range * Update doc/source/configure.rst Co-authored-by: Edward Oakes <ed.nmi.oakes@gmail.com> Co-authored-by: Edward Oakes <ed.nmi.oakes@gmail.com>	2021-11-05 09:25:44 -07:00
Eric Liang	6102912494	Dataset doc updates (#19815 )	2021-11-04 18:13:40 -07:00
javi-redondo	11371768c1	Update Ray client docs with working_dir explanation (#18294 )	2021-11-04 14:52:28 -07:00
javi-redondo	5781f44cc9	Update config.rst to reflect min nodes on autoscaling up (#15589 )	2021-11-04 13:40:50 -07:00
Philipp Moritz	a64e32c53b	[docs] Fix broken links in documentation and add linkcheck to documentation (#20030 ) Co-authored-by: Richard Liaw <rliaw@berkeley.edu>	2021-11-04 13:19:43 -07:00
Alex Wu	36a214386f	[docs] PyData ray dataset talk (#20038 )	2021-11-04 10:33:18 -07:00
Amog Kamsetty	f67b526b7a	[Tune] Fix PTL tutorial docs (#19999 )	2021-11-04 09:21:28 -07:00
Richard Hamnett	f4256a4ddc	[Doc] Update installation.rst for nightly build (#20034 ) Ensure clean removal of previous ray nightly before updating.	2021-11-03 12:05:07 -07:00
Avnish Narayan	026bf01071	[RLlib] Upgrade gym version to 0.21 and deprecate pendulum-v0. (#19535 ) * Fix QMix, SAC, and MADDPA too. * Unpin gym and deprecate pendulum v0 Many tests in rllib depended on pendulum v0, however in gym 0.21, pendulum v0 was deprecated in favor of pendulum v1. This may change reward thresholds, so will have to potentially rerun all of the pendulum v1 benchmarks, or use another environment in favor. The same applies to frozen lake v0 and frozen lake v1 Lastly, all of the RLlib tests and have been moved to python 3.7 * Add gym installation based on python version. Pin python<= 3.6 to gym 0.19 due to install issues with atari roms in gym 0.20 * Reformatting * Fixing tests * Move atari-py install conditional to req.txt * migrate to new ale install method * Fix QMix, SAC, and MADDPA too. * Unpin gym and deprecate pendulum v0 Many tests in rllib depended on pendulum v0, however in gym 0.21, pendulum v0 was deprecated in favor of pendulum v1. This may change reward thresholds, so will have to potentially rerun all of the pendulum v1 benchmarks, or use another environment in favor. The same applies to frozen lake v0 and frozen lake v1 Lastly, all of the RLlib tests and have been moved to python 3.7 * Add gym installation based on python version. Pin python<= 3.6 to gym 0.19 due to install issues with atari roms in gym 0.20 Move atari-py install conditional to req.txt migrate to new ale install method Make parametric_actions_cartpole return float32 actions/obs Adding type conversions if obs/actions don't match space Add utils to make elements match gym space dtypes Co-authored-by: Jun Gong <jungong@anyscale.com> Co-authored-by: sven1977 <svenmika1977@gmail.com>	2021-11-03 16:24:00 +01:00
Will Drevo	f359b21541	[RLlib; Docs] Updated RLlib training example page (#19932 )	2021-11-03 12:34:18 +01:00
Sven Mika	2d24ef0d32	[RLlib] Add all simple learning tests as `framework=tf2`. (#19273 ) * Unpin gym and deprecate pendulum v0 Many tests in rllib depended on pendulum v0, however in gym 0.21, pendulum v0 was deprecated in favor of pendulum v1. This may change reward thresholds, so will have to potentially rerun all of the pendulum v1 benchmarks, or use another environment in favor. The same applies to frozen lake v0 and frozen lake v1 Lastly, all of the RLlib tests and Tune tests have been moved to python 3.7 * fix tune test_sampler::testSampleBoundsAx * fix re-install ray for py3.7 tests Co-authored-by: avnishn <avnishn@uw.edu>	2021-11-02 12:10:17 +01:00
Will Drevo	97f04b118d	[RLlib; Docs] Added fixes to CartPole example. (#19908 ) * Added fixes to CartPole example * Apply suggestions from code review Co-authored-by: will <will@anyscale.com> Co-authored-by: Sven Mika <sven@anyscale.io>	2021-11-02 10:06:39 +01:00
xwjiang2010	c48d86e469	[CI] change git protocol to use https. (#19964 )	2021-11-01 19:38:58 -07:00
Kim Pevey	3ff4fde0f5	[Doc] Update newsreader example (#19893 )	2021-10-29 22:25:40 -07:00
Kim Pevey	8aa61566fa	[Doc] Example docs minor wording fixes (#19890 )	2021-10-29 22:15:35 -07:00
Kim Pevey	96480d97d6	[DOC] Minor typos/fixes to Tips for First Timers (#19887 ) * fix typos * some more fixes Co-authored-by: Philipp Moritz <pcmoritz@gmail.com>	2021-10-29 22:13:15 -07:00
Philipp Moritz	0a5942d8b0	[Documentation] Fix quotes for windows installations (#19859 ) * [Documentation] Fix quotes for windows installations * update * formatting	2021-10-29 10:54:38 -07:00
architkulkarni	fdefd875c3	[Doc] [runtime env] Move runtime env section up one level, add inbound links (#19863 )	2021-10-29 12:02:39 -05:00
Antoni Baum	f2773267c7	[docs] Tune doc fixes (#19791 )	2021-10-29 11:45:29 +02:00

... 2 3 4 5 6 ...

1871 commits