Get the latest tech news

Intel squeezed a 1.58-bit LLM down to 1.485 bits without changing a single weight


Intel's BITCOS format compresses ternary model weights below 1.58 bits by exploiting zero-heavy distributions, boosting decoding speed up to 27% on GPUs.

None

Get the Android app

Or read this on r/technology

Read more on:

Photo of Intel

Intel

Photo of bits

bits

Photo of bit LLM

bit LLM

Related news:

News photo

Details about Intel's next-gen Nova Lake CPUs keep leaking — an attempt to establish a timeline based on what we know so far

News photo

Linux Ready With Fix For Intel's FRED Crashing Some Games Under Wine / Steam Play

News photo

AI developer vibe codes DLSS 5 onto Intel CPU's integrated graphics — Intel Arc 140T runs neural rendering in 360p at 10 frames per second