AI & ML interests

Bit-exact lossless model compression

Recent Activity

Organization Card

ISIRO Runtime

Lossless inference layer for self-hosted AI

Smaller weights · More KV cache · No quantization

Baseline model weights compressed to .tic with bit-exact decode via ISIRO Runtime

ISIRO compresses AI model weights losslessly.
Weights are bit-exact, with a hash manifest for verification, and are served directly in the compressed state.
No quantization, no retraining, no model changes.

Lower cost to self-host models in cloud and on-prem.

Download sample .tic models or compile your own with ISIRO compiler.
What is TIC?

Install: curl -fsSL https://isiro.ai/install.sh | sh

Documentation · GitHub · isiro.ai

datasets 0

None public yet