hiro/ray - Forgejo: Beyond coding. We Forge.

hiro/ray

mirror of https://github.com/vale981/ray synced 2025-03-06 10:31:39 -05:00

Author	SHA1	Message	Date
Antoni Baum	7ff914b06e	[AIR][Docs] Set `logging_strategy="epoch"` for HF (#27917 )	2022-08-16 16:45:46 -07:00
Richard Liaw	759fbd9502	[air][minor] Use drop_columns in docs (#27852 )	2022-08-16 14:01:25 -07:00
Simon Mo	b9a2fb79b6	[AIR][Docs] Remove the excessive printing from Torch examples (#27903 )	2022-08-16 09:09:54 -07:00
matthewdeng	58495fe594	[data][docs] fix broken links (#27818 )	2022-08-12 11:17:34 -07:00
Eric Liang	52f7b89865	[docs] Editing pass on clusters docs, removing legacy material and fixing style issues (#27816 )	2022-08-12 00:15:03 -07:00
Simon Mo	2fbfc87f5c	[Serve] Update AIR Examples to use new API, add linked guide (#27733 )	2022-08-11 13:01:17 -07:00
Cheng Su	bc5d8d9176	[AIR] Replace references of `to_tf` with `iter_tf_batches` (#27672 )	2022-08-09 16:00:02 -07:00
Richard Liaw	fb43bd5baf	[air/docs] Update train gettingstarted (#27655 ) Co-authored-by: matthewdeng <matthew.j.deng@gmail.com>	2022-08-08 15:45:00 -07:00
Cheng Su	aeb2346804	[AIR] Replace references of `to_torch` with `iter_torch_batches` (#27574 )	2022-08-07 20:14:12 -07:00
Eric Liang	9b467e3954	[docs] Improve the "Why Ray" and "Why AIR" sections of the docs (#27480 )	2022-08-05 18:42:45 -07:00
Richard Liaw	4629a3a649	[air/docs] Update Trainer documentation (#27481 ) Co-authored-by: xwjiang2010 <xwjiang2010@gmail.com> Co-authored-by: Kai Fricke <kai@anyscale.com> Co-authored-by: xwjiang2010 <87673679+xwjiang2010@users.noreply.github.com> Co-authored-by: Eric Liang <ekhliang@gmail.com>	2022-08-05 11:21:19 -07:00
xwjiang2010	ff2b728e9a	[air] add tuner user guide (#26837 ) Co-authored-by: Kai Fricke <kai@anyscale.com> Co-authored-by: Richard Liaw <rliaw@berkeley.edu>	2022-08-03 09:43:42 -07:00
Clark Zinzow	df124d0ad5	[AIR - Datasets] Hide tensor extension from UDFs. (#27019 ) We previously added automatic tensor extension casting on Datasets transformation outputs to allow the user to not have to worry about tensor column casting; however, this current state creates several issues: 1. Not all tensors are supported, which means that we’ll need to have an opaque object dtype (i.e. ndarray of ndarray pointers) fallback for the Pandas-only case. Known unsupported tensor use cases: a. Heterogeneous-shaped (i.e. ragged) tensors b. Struct arrays 2. UDFs will expect a NumPy column and won’t know what to do with our TensorArray type. E.g., torchvision transforms don’t respect the array protocol (which they should), and instead only support Torch tensors and NumPy ndarrays; passing a TensorArray column or a TensorArrayElement (a single item in the TensorArray column) fails. Implicit casting with object dtype fallback on UDF outputs can make the input type to downstream UDFs nondeterministic, where the user won’t know if they’ll get a TensorArray column or an object dtype column. 3. The tensor extension cast fallback warning spams the logs. This PR: 1. Adds automatic casting of tensor extension columns to NumPy ndarray columns for Datasets UDF inputs, meaning the UDFs will never have to see tensor extensions and that the UDF input column types will be consistent and deterministic; this fixes both (2) and (3). 2. No longer implicitly falls back to an opaque object dtype when TensorArray casting fails (e.g. for ragged tensors), and instead raises an error; this fixes (4) but removes our support for (1). 3. Adds a global enable_tensor_extension_casting config flag, which is True by default, that controls whether we perform this automatic casting. Turning off the implicit casting provides a path for (1), where the tensor extension can be avoided if working with ragged tensors in Pandas land. Turning off this flag also allows the user to explicitly control their tensor extension casting, if they want to work with it in their UDFs in order to reap the benefits of less data copies, more efficient slicing, stronger column typing, etc.	2022-07-28 10:37:45 -07:00
Kai Fricke	3924a4b7cc	[air/train] Rename BaseWorkerMixin, only log info torch loop for rank 0 (#27098 ) This PR - only prints train_loop info strings (e.g. `train_loop_utils.py:298 -- Moving model to device: cpu`) for rank 0 workers for torch - renames `BaseWorkerMixin` to `RayTrainWorker` as the name comes up often in output and is more meaningful Signed-off-by: Kai Fricke <kai@anyscale.com>	2022-07-27 20:11:59 +01:00
matthewdeng	113c4d7fab	[air][data] move train_test_split to ray.data.Dataset (#27065 )	2022-07-27 09:53:37 -07:00
Balaji Veeramani	89f7f2a567	[Datasets] Add `size` parameter to `ImageFolderDatasource` (#26975 ) If you read a folder with differently-sized images, `ImageFolderDatasource` errors. This PR fixes the issue by resizing images to a user-specified size.	2022-07-26 14:57:38 -07:00
Balaji Veeramani	55988992b9	[AIR] Rename `limit` parameter as `max_categories` (#26977 )	2022-07-26 10:10:40 -07:00
Jiao	5315f1e643	[AIR] Enable other notebooks previously marked with # REGRESSION (#26896 ) Co-authored-by: Richard Liaw <rliaw@berkeley.edu>	2022-07-25 13:40:21 -07:00
matthewdeng	df638b3f0f	[Datasets] Automatically cast tensor columns when building Pandas blocks. (#26924 ) This PR just applies the changes from the following PRs: [Datasets] Automatically cast tensor columns when building Pandas blocks. #26684 reverted by Revert "[Datasets] Automatically cast tensor columns when building Pandas blocks." #26921 [AIR - Datasets] Fix TensorDtype construction from string and fix example. #26904 This fixes the test failures introduced in the originally reverted PRs.	2022-07-25 12:12:10 -07:00
Eric Liang	36c46e9686	[docs] Improve AIR table of contents titles (#26858 )	2022-07-22 17:17:49 -07:00
Jiao	db027d86af	[P0][AIR] Fix train to serve notebooks (#26821 ) Co-authored-by: Simon Mo <simon.mo@hey.com>	2022-07-21 18:04:13 -07:00
Balaji Veeramani	ac1d21027d	[AIR] Add framework-specific checkpoints (#26777 )	2022-07-20 19:33:27 -07:00
xwjiang2010	e7957f4a3e	[air] update offline/online rl example and enable them. (#26786 )	2022-07-20 14:06:03 -07:00
Sumanth Ratna	759966781f	[air] Allow users to use instances of `ScalingConfig` (#25712 ) Co-authored-by: Xiaowei Jiang <xwjiang2010@gmail.com> Co-authored-by: matthewdeng <matthew.j.deng@gmail.com> Co-authored-by: Kai Fricke <krfricke@users.noreply.github.com>	2022-07-18 15:46:58 -07:00
Amog Kamsetty	3a345a470c	[AIR/Docs] Add Predictor Docs (#25833 )	2022-07-16 21:14:21 -07:00
Richard Liaw	799311b2f7	[air/docs] update examples to remove pandas again (#26598 )	2022-07-16 08:40:44 -07:00
Jiao	647e12b6c7	[AIR] Fix convert_existing_pytorch_code_to_ray_air notebook (#26523 )	2022-07-14 14:30:55 -07:00
Amog Kamsetty	6595bd6e2d	[AIR] Introduce better scoring API for `BatchPredictor` (#26451 ) Signed-off-by: Amog Kamsetty <amogkamsetty@yahoo.com> As discussed offline, allow configurability for feature columns and keep columns in BatchPredictor for better scoring UX on test datasets.	2022-07-14 11:26:12 -07:00
Richard Liaw	a0ce3c111b	[air/data] Concatenator preprocessor (#26526 )	2022-07-14 10:26:14 -07:00
Jiao	15dbc0362a	[AIR][Docs] Fix torch_image_example (#26453 )	2022-07-13 21:59:24 -07:00
Antoni Baum	5ed10ef921	[AIR/CI] Fix Hugging Face notebook example (#26475 )	2022-07-13 09:16:42 -07:00
Eric Liang	4c04c8d92c	[doc] Rename toc entry for libraries back to "Ray Libraries" (#26485 )	2022-07-12 14:23:36 -07:00
Richard Liaw	1abe908c22	[air/docs] improve consistency of getting started (#26247 )	2022-07-11 20:16:37 -07:00
Richard Liaw	5892a76a44	[air/tune] Documentation testing fixes (#26409 )	2022-07-09 19:47:21 -07:00
Amog Kamsetty	cc43bcccb4	[AIR] Update TensorflowPredictor to new API (#26215 ) Updates TensorflowPredictor to use the new _predict_pandas API. Also as agreed upon offline, removes the extra configurations from TensorflowPredictor (column selection, concatenation) in favor of having this be done via a Preprocessor.	2022-07-08 13:04:49 -07:00
Antoni Baum	ea94cda1f3	[AIR] Replace `train.` with `session.` (#26303 ) This PR replaces legacy API calls to `train.` with AIR `session.` in Train code, examples and docs. Depends on https://github.com/ray-project/ray/pull/25735	2022-07-07 16:29:04 -07:00
Antoni Baum	d1966899bb	[Docs] Small fix to AIR examples descriptions (#26227 )	2022-07-05 17:16:56 -07:00
Simon Mo	88a219c7f2	Revert "Revert "[AIR][Serve] Rename ModelWrapperDeployment -> PredictorDeployment"" (#26231 )	2022-07-05 13:26:49 -07:00
Antoni Baum	128f9e5664	[AIR] Move integration logging callbacks to AIR (#26126 ) As the integration logging callbacks are commonly used with AIR Trainers, they should be moved from the tune package to the air package. The old imports will still work, but raise a deprecation warning.	2022-06-28 17:25:19 -07:00
Stephanie Wang	c9be251b7a	Revert "[AIR][Serve] Rename ModelWrapperDeployment -> PredictorDeployment (#25962 )" (#26176 ) This reverts commit `68692b3464`.	2022-06-28 17:07:07 -07:00
Simon Mo	68692b3464	[AIR][Serve] Rename ModelWrapperDeployment -> PredictorDeployment (#25962 )	2022-06-28 10:26:10 -07:00
Antoni Baum	0ec198acc2	[AIR] Remove unnecessary pandas from examples (#26009 ) Removes unnecessary pandas usage from AIR examples. Helps ensure users do not follow bad practices.	2022-06-24 14:38:23 -07:00
Antoni Baum	94492c2b49	[AIR/Docs] Improve user guide gallery (#26016 ) Moves logging to examples, reorders to match TOC, adds missing entry.	2022-06-23 17:51:01 -07:00
Antoni Baum	91dd360f9d	[AIR/train] Move predictors to `ray.train` (#25769 )	2022-06-15 17:02:15 -07:00
Kai Fricke	fdf85ea403	[air] Add tutorial to convert existing pytorch code to Ray AIR (#25723 )	2022-06-14 18:11:32 -07:00
Antoni Baum	5e9a8eb5f6	[AIR/data] Move preprocessors to `ray.data` (#25599 ) Moves ray.air.Preprocessor and ray.air.preprocessors to ray.data to converge on the agreed upon package structure discussed internally.	2022-06-13 12:57:59 -07:00
Amog Kamsetty	1316a2d05e	[AIR/Train] Move `ray.air.train` to `ray.train` (#25570 )	2022-06-08 21:34:18 -07:00
Antoni Baum	7616435ed0	[Docs] Capitalize Ray AIR (#25597 )	2022-06-08 14:37:53 -07:00
xwjiang2010	29a063afdf	[air] add feast example (#25417 )	2022-06-07 14:55:42 -07:00
Simon Mo	7471b1fa41	[Serve] [AIR] ModelWrapper improvements and docs (#25003 ) * batching collation code and tests * wip notebook for np and dataframe * finish content * reset ray-more-libs changes * add comments * run through * Apply suggestions from code review Co-authored-by: shrekris-anyscale <92341594+shrekris-anyscale@users.noreply.github.com> * rename package * lint * richard's comment Co-authored-by: shrekris-anyscale <92341594+shrekris-anyscale@users.noreply.github.com>	2022-06-07 08:53:10 -07:00

1 2

63 commits