Mira Murati's Thinking Machines just released Inkling-Small, an open source model that's 1/4 the size of their original Inkling but matches (and sometimes beats) its performance.
This is the direction AI is heading - smaller, faster, more efficient models that don't need hyperscale infrastructure.
At BHK Cloud, we've been ready for this.
RTX 3090s at $0.15/hr, perfectly suited for running these efficient open-source models.
Buy Now Pay Later - zero upfront, pay after 1 month.
Spin up an RTX 3090 in 60 seconds. Storage at $2.49/TB. Zero egress between GPU and storage.
Try BHK Cloud free
Originally posted on reddit