I recently come across this project, which seems very interesting because it let you run literally huge models on consumer hardware.
I am fancying yo try run Qwen3.8 flash next with this on my setup. I don’t expect anything usable, but just for the fun of it. What stopping me is the download size of the model.
Has anybody tried colibri? Just to have an idea of speed to expect. 1t/s? 0.1t/s?




Try subwave. You can have AI djs choose music from your navidrone library based on shows you create by defining prompts, moods and styles.
Pretty neat…
Fast developing and pretty solid already.