Market Opportunity
Storage pain for LoRA models — aggressive size reduction via targeted compression targets a $4.8B = 1,000,000 AI teams/orgs x $4.8K/year average spend on model artifact tooling, storage optimization, and related workflows total addressable market with low saturation and a year-over-year growth rate of 30-45% -- driven by rapid adoption of fine-tuning and model customization across industries.
Key trends driving demand: LoRA & adapter growth -- LoRA has become the dominant cheap fine-tuning approach, multiplying the number of adapter artifacts that need storage and versioning.; On-device and edge inference -- demand for compact model artifacts to run locally on consumer devices and low-cost servers increases need for compression.; Open-source model proliferation -- many small teams publish dozens of adapters for models, creating an explosion of discrete artifacts rather than few monolithic models..
Key competitors include AutoGPTQ (community projects), bitsandbytes, Hugging Face (Optimum / Model Hub), NVIDIA TensorRT / TensorRT-LLM.