
LakeFS
DVC
Monte Carlo Data
Git Large File Storage
ArtiVC
Tonic AI
AWS Lake Formation
Dagster
Cachely.dev
nxCloud
Cachely is the managed self-hosted remote cache for Nx and Turborepo - the cache backend you'd otherwise build and run yourself, hosted for you on Cloudflare's edge (R2). It's a drop-in replacement for a DIY @nx/s3-cache / S3 bucket setup: point your build tool at Cachely with a token and two environment variables, and share build cache across CI and every developer's laptop.
Unlike a self-hosted cache, Cachely enforces read-only tokens at the API, so pull-request and fork builds can read but never write - closing the Nx cache-poisoning attack (CVE-2025-36852). It adds ROI reporting (the real build minutes and dollars the cache saved), per-tool insights, and build-optimization suggestions on top.
Pricing is a flat per-workspace subscription with no per-seat fees - add every developer, bot, and CI actor without watching the bill. Cachely never stores your source code; it caches only task outputs and their content hashes. Nx and Turborepo today; Bazel on the roadmap.
Cachely.devNo features have been listed yet.
No Cachely.dev videos yet. You could help us improve this page by suggesting one.
Based on our record, LakeFS seems to be more popular. It has been mentiond 6 times since March 2021. We are tracking product recommendations and mentions on various public social media platforms and blogs. They can help you identify which product is more popular and what people think of it.
I would add https://github.com/gaul/s3proxy to your list. - Source: Hacker News / over 2 years ago
* data state - this is contents of both your data and metadata at a given point in time. if your data doesn't fit into a single database, this can be difficult to manage. We use this technology to help us: https://lakefs.io/. - Source: Hacker News / almost 3 years ago
Saltcured, find these comments super insightful! > Yeah, there's a lot of hidden magic/assumptions in having a "writable snapshot of a specific version" of production data. That's absolutely a huge assumption. This technology has been a game changer for us: https://lakefs.io/ > It becomes a headache when there is too much contention to use these sandboxes, or too much manual effort to reset them to a desired... - Source: Hacker News / almost 3 years ago
You should not store your data in git itself, but rather use git to version your data sets. The currently best option for that is (IMHO) https://lakefs.io though there are a few others in various states of usability/maturity. Source: over 3 years ago
I mean if you're ready to adopt a new framework into your ecosystem this is one of the major usecases for LakeFS. Source: over 3 years ago
DVC - Diablo Valley College consists of two campuses serving more than 22,000 students in Contra Costa County each semester with a wide variety of program options.
nxCloud - nxCloud is a commercial OwnCloud provider
Monte Carlo Data - Monte Carloโs Data Observability platform increases trust in data by eliminating data downtime, so engineers innovate more and fix less.
Git Large File Storage - Git Large File Storage (LFS) replaces large files such as audio samples, videos, datasets, and graphics with text pointers.
ArtiVC - ArtiVC (Artifact Version Control) is a version control system for large files.
Tonic AI - The fake data company