1 article tagged with this topic
Liquid AI ships a 2.6B parameter model hitting 260 tokens/sec on an RTX 3090. Limited capability, but enough for daily chores — and it lowers the bar