hiro/ray - Forgejo: Beyond coding. We Forge.

hiro/ray

mirror of https://github.com/vale981/ray synced 2025-03-07 02:51:39 -05:00

Author	SHA1	Message	Date
Richard Liaw	cf357f6bce	[docs] Add a talks section for ray.data (#20444 )	2021-11-16 14:30:08 -08:00
Antoni Baum	3f9ded55f7	[tune] Merge `Analysis` into `ExperimentAnalysis` (#20197 ) Co-authored-by: Kai Fricke <kai@anyscale.com>	2021-11-16 16:47:12 +00:00
Amog Kamsetty	4f88796d5a	[Train] Move to beta (#20378 )	2021-11-16 08:19:30 -08:00
Kai Fricke	3e6ba5d6d2	Revert "Revert [RLlib] POC: `PGTrainer` class that works by sub-classing, not `trainer_template.py`." (#20285 ) * Revert "Revert "[RLlib] POC: `PGTrainer` class that works by sub-classing, not `trainer_template.py`. (#20055)" (#20284)" This reverts commit `246787cdd9`. Co-authored-by: sven1977 <svenmika1977@gmail.com>	2021-11-16 12:26:47 +01:00
Eric Liang	460cf86858	Split blocks automatically into 500MB chunks on file read and transformation (#20235 ) This PR adds support for automatic block splitting on read and map transforms, to keep block size bounded to ~500MiB. This avoids potential OOM situations where a map task may consume too much intermediate Python heap memory, or too much object store shared memory for one block.	2021-11-15 22:25:11 -08:00
Antoni Baum	ec81f52061	[Docs] Fix typo in C++ Placement Group example (#20386 )	2021-11-16 08:19:09 +09:00
Will Drevo	fa878e2d4d	Added example to user guide for cloud checkpointing (#20045 ) Co-authored-by: will <will@anyscale.com> Co-authored-by: Antoni Baum <antoni.baum@protonmail.com> Co-authored-by: Kai Fricke <kai@anyscale.com>	2021-11-15 15:43:06 +00:00
Amog Kamsetty	a74cf7ff1c	[Train] Torch Prepare utilities (#20254 ) * update * formatting * fix failures * fix session tests * address comments * add to api docs * package refactor * wip * wip * wip * finish * finish * fix * comment * fix * install horovod for docs * address comment * Update python/ray/train/session.py Co-authored-by: matthewdeng <matthew.j.deng@gmail.com> * Update python/ray/train/torch.py Co-authored-by: matthewdeng <matthew.j.deng@gmail.com> * address comments * try fix docs * fix doc build failure * fix * fix * fix * try fix doc highlighting * fix docs Co-authored-by: matthewdeng <matthew.j.deng@gmail.com>	2021-11-15 07:34:17 -08:00
Qing Wang	1172195571	[Java] Remove global named actor and global pg (#20135 ) This PR removes global named actor and global PGs. I believe these APIs are not used widely in OSS. CPP part is not included in this PR. @kfstorm @clay4444 @raulchen Please take a look if this change is reasonable. IMPORTANT NOTE: This is a Java API change and will lead backward incompatibility in Java global named actor and global PG usage. CPP part is not included in this PR. INCLUDES: Remove setGlobalName() and getGlobalActor() APIs. Remove getGlobalPlacementGroup() and setGlobalPG Add getActor(name, namespace) API Add getPlacementGroup(name, namespace) API Update doc pages.	2021-11-15 16:28:53 +08:00
Sven Mika	e5ead6a4b0	[RLlib; Documentation] Minor fixes "rllib in 60s" and per-feature sigils. (#20248 )	2021-11-13 22:10:47 +01:00
Amog Kamsetty	65a17da2ec	[Train] Refactor Backends (#20312 ) * wip * finish * comment * fix * install horovod for docs * address comment * fix doc build failure	2021-11-13 11:05:53 -08:00
matthewdeng	e77cc926be	[train] minor doc updates (#20271 )	2021-11-12 17:20:23 -08:00
Tricia Fu	e59c14117f	[Doc] [Serve] Add summary sub header to each page (#20231 )	2021-11-12 14:18:42 -08:00
xwjiang2010	cdf70c2900	[Tune] Remove legacy resources implementations in Runner and Executor. (#19773 )	2021-11-12 12:33:39 -08:00
Siyuan (Ryans) Zhuang	3b62388a9a	[Workflow] Workflow tail recursion optimization (#19928 ) * tail recursion optimization	2021-11-12 09:13:40 -08:00
Kai Fricke	246787cdd9	Revert "[RLlib] POC: `PGTrainer` class that works by sub-classing, not `trainer_template.py`. (#20055 )" (#20284 ) This reverts commit `6f85af435f`.	2021-11-12 13:09:43 +00:00
Kai Fricke	d88fdd6e38	[tune] refactor SyncConfig (#20155 )	2021-11-12 09:36:15 +00:00
Michael Galarnyk	dbeb2e2f73	Add Ray Serve Blogs to Doc(#19846 ) The Serving ML Models in Production blog links is inline with the latest Ray Summit talk on Ray Serve.	2021-11-11 15:10:36 -08:00
Edward Oakes	59698aa89c	[Serve] add survey link (#20230 )	2021-11-11 15:10:10 -08:00
Jules S. Damji	71a162d8ab	Fixed code snippet to include config parameter and a minor typo (#20193 ) Signed-off-by: Jules S.Damji <jules@anyscale.com> Co-authored-by: Jules S.Damji <jules@anyscale.com>	2021-11-11 18:37:03 +00:00
Dmitri Gekhtman	8971422d8f	[autoscaler] Use drain node api in autoscaler before terminating nodes (#20013 ) * wip * Draft * Use bytest for node id * remove stray helm change * fix autoscaler init arg * don't forget to instantiate new load metrics dict * remove extraneous diff * Timeout, comments, function signature. * typo * another comment * tweak * docstring * shorter timeout * Use a better error code * missing self * Dedent example * Add drain node prometheus metric. * comment * Update tests part 1: test_autoscaler.py * Update tests part 2: test_resource_demand_scheduler * lint * Update tests part 3: test_autoscaling_policy * Unit tests for new Prometheus metric and DrainNode error handling. * comment * removed unused function * Try adding ability to mock out process termination to fake node provider * Add integration test. * fix * fix * lint * Improve log message * fix * Simplify test * Fix doc example * remove unused dict * Mock out process termination in a subclass * Add add doc string and comment explaining prune active ips. * Comment: wtf is use_node_id_as_ip * one more comment * more explanation * period * tweak	2021-11-11 08:31:40 -08:00
Sven Mika	6f85af435f	[RLlib] POC: `PGTrainer` class that works by sub-classing, not `trainer_template.py`. (#20055 )	2021-11-11 12:16:20 +01:00
Will Drevo	2fdb1c46c7	[RLlib; Documentation] Added atari pip installs to Pong-v0 example. (#20225 ) * Added imports to Pongv0 example * Added comment * Apply suggestions from code review Co-authored-by: will <will@anyscale.com> Co-authored-by: Sven Mika <sven@anyscale.io>	2021-11-11 09:08:02 +01:00
Tobias Kaymak	893f57591d	[serve] Add Google Cloud Storage as a backend (#20104 )	2021-11-10 19:45:19 -08:00
Edward Oakes	082a4af3e6	[serve] Remove lingering backend/endpoint wording in docs (#20229 )	2021-11-10 16:49:29 -08:00
Sven Mika	ebd56b57db	[RLlib; documentation] "RLlib in 60sec" overhaul. (#20215 )	2021-11-10 22:20:06 +01:00
matthewdeng	790e22f9ad	[tune] move force_on_current_node to ml_utils (#20211 )	2021-11-10 10:21:24 -08:00
Sven Mika	143d23a278	[RLlib] Issue 20062: Action inference examples missing (#20144 )	2021-11-10 18:49:06 +01:00
Kim Pevey	82a5bf68fa	[Docs] Add note for multi-node on Windows (#20184 ) * add note for multi-node on Windows * update message Co-authored-by: Philipp Moritz <pcmoritz@gmail.com>	2021-11-09 16:02:01 -08:00
Kai Fricke	9c2b8c8501	[tune] Deprecate DurableTrainable (#19880 )	2021-11-08 20:56:07 +00:00
Amog Kamsetty	b1f24768a1	[Tune] More fixes to PTL Tutorial (#20065 ) * ptl-fix-2 * improve * fix	2021-11-08 09:13:44 -08:00
Jules S. Damji	e6343f0e69	Fixed a broken code snippet with a missing method (#20130 ) Signed-off-by: Jules S.Damji <jules@anyscale.com> Co-authored-by: Jules S.Damji <jules@anyscale.com>	2021-11-08 07:56:32 +09:00
Alex Wu	81194f5660	[workflow][docs] Fix api comparison formatting (#20069 ) ## Why are these changes needed? The API comparison formatting uses \`code\` which is rendered as italicization not code. This PR puts the code in code blocks instead of italics. ## Related issue number ## Checks	2021-11-05 17:05:35 -07:00
Chen Shen	320f9dc234	[Core][CoreWorker] increase the default port range (#19541 ) * increase the port range * Update doc/source/configure.rst Co-authored-by: Edward Oakes <ed.nmi.oakes@gmail.com> Co-authored-by: Edward Oakes <ed.nmi.oakes@gmail.com>	2021-11-05 09:25:44 -07:00
Eric Liang	6102912494	Dataset doc updates (#19815 )	2021-11-04 18:13:40 -07:00
javi-redondo	11371768c1	Update Ray client docs with working_dir explanation (#18294 )	2021-11-04 14:52:28 -07:00
javi-redondo	5781f44cc9	Update config.rst to reflect min nodes on autoscaling up (#15589 )	2021-11-04 13:40:50 -07:00
Philipp Moritz	a64e32c53b	[docs] Fix broken links in documentation and add linkcheck to documentation (#20030 ) Co-authored-by: Richard Liaw <rliaw@berkeley.edu>	2021-11-04 13:19:43 -07:00
Alex Wu	36a214386f	[docs] PyData ray dataset talk (#20038 )	2021-11-04 10:33:18 -07:00
Amog Kamsetty	f67b526b7a	[Tune] Fix PTL tutorial docs (#19999 )	2021-11-04 09:21:28 -07:00
Richard Hamnett	f4256a4ddc	[Doc] Update installation.rst for nightly build (#20034 ) Ensure clean removal of previous ray nightly before updating.	2021-11-03 12:05:07 -07:00
Avnish Narayan	026bf01071	[RLlib] Upgrade gym version to 0.21 and deprecate pendulum-v0. (#19535 ) * Fix QMix, SAC, and MADDPA too. * Unpin gym and deprecate pendulum v0 Many tests in rllib depended on pendulum v0, however in gym 0.21, pendulum v0 was deprecated in favor of pendulum v1. This may change reward thresholds, so will have to potentially rerun all of the pendulum v1 benchmarks, or use another environment in favor. The same applies to frozen lake v0 and frozen lake v1 Lastly, all of the RLlib tests and have been moved to python 3.7 * Add gym installation based on python version. Pin python<= 3.6 to gym 0.19 due to install issues with atari roms in gym 0.20 * Reformatting * Fixing tests * Move atari-py install conditional to req.txt * migrate to new ale install method * Fix QMix, SAC, and MADDPA too. * Unpin gym and deprecate pendulum v0 Many tests in rllib depended on pendulum v0, however in gym 0.21, pendulum v0 was deprecated in favor of pendulum v1. This may change reward thresholds, so will have to potentially rerun all of the pendulum v1 benchmarks, or use another environment in favor. The same applies to frozen lake v0 and frozen lake v1 Lastly, all of the RLlib tests and have been moved to python 3.7 * Add gym installation based on python version. Pin python<= 3.6 to gym 0.19 due to install issues with atari roms in gym 0.20 Move atari-py install conditional to req.txt migrate to new ale install method Make parametric_actions_cartpole return float32 actions/obs Adding type conversions if obs/actions don't match space Add utils to make elements match gym space dtypes Co-authored-by: Jun Gong <jungong@anyscale.com> Co-authored-by: sven1977 <svenmika1977@gmail.com>	2021-11-03 16:24:00 +01:00
Will Drevo	f359b21541	[RLlib; Docs] Updated RLlib training example page (#19932 )	2021-11-03 12:34:18 +01:00
Sven Mika	2d24ef0d32	[RLlib] Add all simple learning tests as `framework=tf2`. (#19273 ) * Unpin gym and deprecate pendulum v0 Many tests in rllib depended on pendulum v0, however in gym 0.21, pendulum v0 was deprecated in favor of pendulum v1. This may change reward thresholds, so will have to potentially rerun all of the pendulum v1 benchmarks, or use another environment in favor. The same applies to frozen lake v0 and frozen lake v1 Lastly, all of the RLlib tests and Tune tests have been moved to python 3.7 * fix tune test_sampler::testSampleBoundsAx * fix re-install ray for py3.7 tests Co-authored-by: avnishn <avnishn@uw.edu>	2021-11-02 12:10:17 +01:00
Will Drevo	97f04b118d	[RLlib; Docs] Added fixes to CartPole example. (#19908 ) * Added fixes to CartPole example * Apply suggestions from code review Co-authored-by: will <will@anyscale.com> Co-authored-by: Sven Mika <sven@anyscale.io>	2021-11-02 10:06:39 +01:00
Philipp Moritz	0a5942d8b0	[Documentation] Fix quotes for windows installations (#19859 ) * [Documentation] Fix quotes for windows installations * update * formatting	2021-10-29 10:54:38 -07:00
architkulkarni	fdefd875c3	[Doc] [runtime env] Move runtime env section up one level, add inbound links (#19863 )	2021-10-29 12:02:39 -05:00
Antoni Baum	f2773267c7	[docs] Tune doc fixes (#19791 )	2021-10-29 11:45:29 +02:00
Rohan138	b9c9cc5946	[RLlib] Updated PettingZoo+RLlib tutorial; Removed pettingzoo example script (#19069 ) * Updated PettingZoo+RLlib tutorial Updated the tutorial and added link to the blog post by the PettingZoo team. * Ran linting * Converted link to tinyurl for linting * fixed line lengths * Decrease num_workers to 1 * Added comments * Decreased num_workers * Decreased timesteps * Increased num_workers * Update links and remove pettingzoo_env.py * remove pettingzoo.py script from tests Co-authored-by: sven1977 <svenmika1977@gmail.com>	2021-10-29 10:57:10 +02:00
Yi Cheng	68ec652be7	[gcs] New option to increase gcs grpc client threads and fix issues in hybrid scheduling (#19663 ) ## Why are these changes needed? - Since broadcasting is moving to grpc, introducing the option to increase the client side thread number - For hybrid schedule, ignore the threshold if gcs based actor scheduler is enabled With these fixing, actor creation rate > 600actor/s vs ~ 140 actor/s ## Related issue number	2021-10-28 22:40:18 -07:00

1 2 3 4 5 ...

1546 commits