PrismML이 27B 모델을 5.9GB로 압축하며 성능을 98.2%로 유지하는 기술을 발표했습니다.
PrismML은 Qwen3.8 27B 기반의 3진 가중치 모델을 공개했습니다. 이 모델은 전체 정밀도 모델보다 크기를 9배 이상 줄이며, 성능은 98.2%를 유지합니다. 가중치를 -1, 0, +1 세 값으로 표현해 실효 비트 수를 1.76비트로 낮춘 것이 특징입니다.
PrismML unveils a 27B model that maintains 98.2% performance while compressing to 5.9GB.
PrismML has introduced a ternary weight model based on Qwen3.8 27B. This model reduces size by over nine times compared to full precision models while maintaining 98.2% performance. It represents weights with three values (-1, 0, +1) and reduces the effective bit count per weight to 1.76 bits.