Comments (6)
It would be nice to update the Readme if the following is not true anymore: "We are committed to bring this library to stable release, ..."
from data.
Hi folks - I'll be updating the README shortly and also adding a pinned issue here to collect feedback.
As of July 2023, we have paused active development on TorchData and will pause new releases. We have learnt a lot from building it and hearing from users, but also believe we need to re-evaluate the technical design and approach given how much the industry has changed since we began the project.
During the rest of 2023 we will be re-evaluating our plans in this space. Data loading is really important and we want to make sure we're able to solve the problem effectively. Please reach out if you have any suggestions or comments.
from data.
but also believe we need to re-evaluate the technical design and approach given how much the industry has changed since we began the project.
Would be interested to know what areas you think aren't working well or aren't suitable for building upon for the future? Are there alternative approaches/examples to TorchData you would recommend using in the meantime? Appreciate it might be too soon to say. Thanks.
from data.
Please see #1196 for a full update!
from data.
cc: @laurencer
from data.
Sorry for the ping, but I'm very curious to hear updates on this! My colleague has developed a lib on top of torchdata for streaming and chipping satellite imagery: https://zen3geo.readthedocs.io/en/latest/walkthrough.html I'm curious if we can expect torchdata to be carried through to a stable release and be actively developed beyond that.
from data.
Related Issues (20)
- MultiplexerLongest example snippet isn't very useful
- Passing dict in datapipe/dataset will have memory leak problem HOT 3
- Roadmap for mixed chain of multithread and multiprocessing pipelines? HOT 2
- DataLoader2 Memory Behavior is very strange on Epoch Resets HOT 9
- FileExistsError when using `on_disk_cache` and multiple workers HOT 1
- Dataloader2 with FullSyncIterDataPipe throws error during initilization HOT 3
- Make archive datapipes faster HOT 1
- An iterator that can stream over stdin
- torchdata has a very low accuracy
- Future of torchdata and dataloading HOT 54
- Calling __iter__ twice on DataLoader2 causes hang with MPRS HOT 2
- Loading `.tfrecords` files that require a deserialization method
- S3FileLoaderIterDataPipe buffer_size
- Iterating a data pipe, created with random split, ends in error as the code tries to iterate past the data pipe lenght
- `v2.1.2+cu118` and `v2.1.1+cu118` run into torchdata `ImportError: libssl.so.3: cannot open shared object file: No such file or directory`, that `v2.1.0+cu118` doesn't have an issue with HOT 1
- PyTorch 2.2: import torchdata fails on ubuntu-20.04 github runners HOT 3
- Dataloader is slow with iterdatapipes and shuffle that has large in-memory fields (because traverse_dps is slow) HOT 3
- DataLoader2 with multiprocess raise exception: Can not request next item while we are still waiting response for previous request HOT 1
- Move to removesuffix string method after python 3.8 support is dropped
Recommend Projects
-
React
A declarative, efficient, and flexible JavaScript library for building user interfaces.
-
Vue.js
🖖 Vue.js is a progressive, incrementally-adoptable JavaScript framework for building UI on the web.
-
Typescript
TypeScript is a superset of JavaScript that compiles to clean JavaScript output.
-
TensorFlow
An Open Source Machine Learning Framework for Everyone
-
Django
The Web framework for perfectionists with deadlines.
-
Laravel
A PHP framework for web artisans
-
D3
Bring data to life with SVG, Canvas and HTML. 📊📈🎉
-
Recommend Topics
-
javascript
JavaScript (JS) is a lightweight interpreted programming language with first-class functions.
-
web
Some thing interesting about web. New door for the world.
-
server
A server is a program made to process requests and deliver data to clients.
-
Machine learning
Machine learning is a way of modeling and interpreting data that allows a piece of software to respond intelligently.
-
Visualization
Some thing interesting about visualization, use data art
-
Game
Some thing interesting about game, make everyone happy.
Recommend Org
-
Facebook
We are working to build community through open source technology. NB: members must have two-factor auth.
-
Microsoft
Open source projects and samples from Microsoft.
-
Google
Google ❤️ Open Source for everyone.
-
Alibaba
Alibaba Open Source for everyone
-
D3
Data-Driven Documents codes.
-
Tencent
China tencent open source team.
from data.