One private library for every local AI model
All your models.
Organized. Ready. Far less disk.
Tensor Archive finds the models scattered across LM Studio, Ollama, ComfyUI and your folders, then groups related versions into one private library. See what fits, compare models on your own hardware and keep every checkpoint, quantization, LoRA and adapter losslessly compact—ready when an app asks.
The first scan is read-only. Nothing is moved, uploaded or deleted.
Powered by U.S. Patent-Pending Technology
Your models, already understood.
LM Studio connected
Ollama connected
ComfyUI connected
What happens after you install it
Install it once. Tensor Archive takes it from there.
Discovery is read-only. Reclaiming space happens only after exact recovery has been proved and you choose to proceed.
It maps the library you already own.
Known LM Studio, Ollama and ComfyUI locations appear automatically. Full Scan can include other folders and attached drives.
Discovery reads model format headers and changes nothing.It understands what belongs together.
Related quantizations, checkpoints, LoRAs and adapters become one understandable model family instead of unrelated files.
You see every usable version—and where the repeated storage is.It proves recovery before reclaiming space.
Tensor Archive stores shared data once, rebuilds a test copy and verifies it byte for byte before the original working copy can be retired.
You decide when to proceed. No destructive quantization and no “close enough.”The model appears when an app needs it.
Tensor Archive prepares the exact working files for the requesting app. After unload, the temporary copy can be reclaimed again.
The compact archive stays ready for the next request.Model Discovery
Know what fits before a 30 GB download.
Model Discovery checks your operating system, memory and available acceleration locally, then shows which model sizes make sense for text, image, video, voice, music and audio—before you download any weights.
Choose what you want to make.
Model Lab
Benchmarks narrow the field. Your machine makes the final call.
Model Lab is a local comparison workspace. Ask a normal question, run a controlled prompt or compare two answers side by side—with the exact model identity, runtime, prompt and measured timing kept visible.
- LOCAL Nothing in the prompt or result needs to leave your computer.
- REAL Runs through Tensor Runtime or a connected local runtime such as Ollama.
- REPEATABLE Save the prompt, environment and result as a local receipt.
The archive identifies repeated tensor relationships while preserving a byte-exact reconstruction path…
- Ready
- 1.8 s
- Decode
- 42.1 tok/s
- Memory
- 8.4 GB
The system stores shared model data once and retains the metadata needed to rebuild each original version…
- Ready
- 2.4 s
- Decode
- 37.8 tok/s
- Memory
- 8.7 GB
Available now · one library behind your model apps
One library. Every model app gets the exact version it needs.
Your runtime or creative app still performs the inference and generation. Tensor Archive handles discovery, exact files and storage behind it.




The storage engine underneath everything
Keep every version. Reclaim the repeated space.
In a measured base-plus-adapter family, five standalone deployments occupied 907.3 MB. Tensor Archive kept the same five deployable versions in 261.6 MB. Every restored byte matched.
This is the proof behind “far less disk.” The full benchmark suite includes storage comparisons, RAM usage, packing speed, baselines, methodology and exact-restore evidence.
Measured model-and-adapter family versus five standalone deployments.
Peak packaging memory versus ZipLLM on the same TA-Bench v1 source data.
Differed after restoring all 50 files in the sequential-checkpoint test.
T2T Network · delivery · elastic storage · optional compute
Your AI library can live beyond this machine. So can the work.
T2T turns a local-first library into an elastic one. Receive verified public models, move complete protected model families off this disk to eligible T2T capacity or Tensor Archive Cloud, and recover the exact bytes automatically when an app asks for them.
Optional Mutual Compute extends the same verified path to approved model preparation and reconstruction work. The result comes back verified; your local catalog remains the authority.
Verified objects arrive through T2T and can remain compact until you need a working version.
Protected network or Cloud capacity keeps it recoverable; Tensor Archive restores what is missing on demand.
Approved model preparation and exact reconstruction can run on selected capacity; a verified result returns ready to use.
One install · one local model system
Keep the models. Get your disk space back.
Discover, compare, compact and use every local model from one private library. Free download for macOS, Windows and Linux.