I've been quite intrigued by a very "close to metal" assistant LLM like this, especially because I've seen folks working at it from the other direction of getting an assistant to be specialized with skills and tools for a specific CLI or code repo. I dug around a bit and I saw some of the execution was for remote models like voice. Since it's built to work well with something like qemu, have you experimented with pointing it at a local model for some of the work? Eg, I could see a reasonably quantized qwen 3.8 flash doing quite well on launching apps, shuttling around system calls, etc.
Very cool! I'll keep an eye out. I'm sure it wouldn't be too hard to point some of the abstractions at a local hosted service layer. Take your time on it and good luck on the experiments!
I've been quite intrigued by a very "close to metal" assistant LLM like this, especially because I've seen folks working at it from the other direction of getting an assistant to be specialized with skills and tools for a specific CLI or code repo. I dug around a bit and I saw some of the execution was for remote models like voice. Since it's built to work well with something like qemu, have you experimented with pointing it at a local model for some of the work? Eg, I could see a reasonably quantized qwen 3.8 flash doing quite well on launching apps, shuttling around system calls, etc.
That’s a very valid point. However, I’ve only just started, and I genuinely believe there is significant potential for growth and success.
Very cool! I'll keep an eye out. I'm sure it wouldn't be too hard to point some of the abstractions at a local hosted service layer. Take your time on it and good luck on the experiments!
Holy sloperoni