Colibrì proof-of-concept gains frontier-level 1.5-TB AI model — novel approach runs on only 25GB of RAM and shows promise for local AI setups

Let’s get the elephant out of the way: Colibrì’s speed on Vincenzo’s setup is only about 0.05 to 0.1 tokens per second on average, a measure that’s unusable for practical conversation — imagine just one question taking h…

SK hynix and TetraMem collaborate on experimental chip to bolster energy efficiency for edge AI devices — memristor-based in-memory SoC research leaves performa

Conceptually, the approach is somewhat analogous to Nvidia’s NVFP4 philosophy, in that both seek to achieve higher effective precision from low-precision hardware. However, the implementations are fundamentally different…