Qwen 3.6 27b GLM 5.2 fine-tune?
Hi everyone, Since both models are open weights and GLM seems to find that secret to frontier model reasoning, why don't we see any Qwen GLM…
Hi everyone, Since both models are open weights and GLM seems to find that secret to frontier model reasoning, why don't we see any Qwen GLM…
This is my third post about designing an orchestration library for agents. I want to share the architecture decisions as I go and to put a…
I think of purchasing 2 DGX Sparks for my office (because a 700+W workstation would be intolerable) for LLM-centric work (inference only, no fine-tuning). I know…
A few months ago, I bought a RX 7900 XTX 24g to start toying with local LLM, at 900€ new. Little I knew that now I…
I was about to say how much better 595.71.05 drivers were at reverting my dual 3090s to a lower power state when idle, with my 3090s…
Everything runs locally in your browser using custom WebGPU kernels written by Fable 5 (before it was shut down) and Opus 4.8. The video was recorded…
submitted by /u/fallingdowndizzyvr [link] [comments]
I created a simple RAG API using medical Wikipedia articles that you can point your agent to and use freely. It may be useful in allowing…
HF blocks download threads I want to signal this early to the community since this seems to be a very recent change. Upon starting a KoboldCpp…
submitted by /u/PAiERAlabs [link] [comments]