Performance diference versus unquantized 8b param version?

#1
by ashercn97 - opened

Hey! I was just wondering how much the performance degraded with the quantization? I didn't see anything on the page (correct me if I'm wrong!).

Thanks!

Hi :) on average we see ~85% the performance of the full precision model (in zero-shot). For many tasks it performs close to parity if they aren't very complicated or heavily visual. We will probably add more details to the card at some point.

Awesome!

ashercn97 changed discussion status to closed

Sign up or log in to comment