Nothing comparable but inspired from DwarfStar I wrote a little inference engine for Intel Xe-LP (no XMX) 32GB laptops. The only model supported right now is a quantized Gemma-4, but I don't exclude in the future to support other MoE of similar size. Too bad we have no Qwen 3.8 35B-A3B yet.
I'm also looking into expanding the protocol and the engine to support various steering techniques.
Cannot read the vibe coded website because it hangs and lags.
So recently opposition to the local LLM narrative pops up here and there and we get reassurance immediately. Is this automated by AI now? Make a sentiment analysis and produce slop articles that suppress last week's opposition?
Nothing comparable but inspired from DwarfStar I wrote a little inference engine for Intel Xe-LP (no XMX) 32GB laptops. The only model supported right now is a quantized Gemma-4, but I don't exclude in the future to support other MoE of similar size. Too bad we have no Qwen 3.8 35B-A3B yet.
I'm also looking into expanding the protocol and the engine to support various steering techniques.
https://github.com/simoneiacomino/xenolith
the problem is the dsv4 checkpoint so quantized isn't very good
Cannot read the vibe coded website because it hangs and lags.
So recently opposition to the local LLM narrative pops up here and there and we get reassurance immediately. Is this automated by AI now? Make a sentiment analysis and produce slop articles that suppress last week's opposition?