#compression
Every summary, chronological. Filter by category, tag, or source from the rail.
Tag · #compression
PrismML's Ternary Compression for On-Device LLMs
PrismML is shrinking high-performance LLMs to fit on consumer hardware by using 'ternary' weight compression, achieving 98% benchmark parity with original models.
TechCrunch — AI
Showing 1 of 1