Mastodon Feed: Post

Mastodon Feed

dysfun@treehouse.systems ("gaytabase") wrote:

LOL, apparently AI boosters are making a fuss about GLM 5.2, XAI's new open weights model

here's the amount of GPU RAM you'd need to make this fast for various levels of quantisation (where less quantisation = smaller, but shitter)

truly democratision of technology in action!

a quantisation table ranging from 217GB at 1 bit to 1.51TB at 16-bit