LM Studio ROCm v2.22+ not able to detect Radeon 6900XT
I've been using LM Studio for a while. Recently, I have been getting an error that my gpu is not detected. I have the latest drivers…
I've been using LM Studio for a while. Recently, I have been getting an error that my gpu is not detected. I have the latest drivers…
https://preview.redd.it/v0xtn3jdu9ch1.png?width=2047&format=png&auto=webp&s=628a6a541fe5f097d0f771ae0ba3b7f44126198f https://preview.redd.it/vjxiucsdu9ch1.png?width=2047&format=png&auto=webp&s=74f7a18a5a30276e206e2bfb5a0c529826ce86e4 This post was originally written in Korean, then polished and translated into English
I've picked up a 3-bit quant of this one (Unsloth - Q3_KS) - the best fit for my config right now. Normally I shy away from…
I would like to make a quotation for a server with RTX 6000 Pro (96 GB). Rackable and tower variants. I do prefer reliable hardware than…
Hey Guys, As promised here are the results from running MiniMax M2.7 REAP 139B Q3_K_L on llama-bench on 6x MI50's. Hardware: Asus X99-E-WS (Modded BIOS to…
Context: I just configured Qwen3.6 27B with DFlash on my DGX spark and looking at all the "reports" it should be a bit faster than i…
Compiled by u/vramkickedin EDIT : Wrong automatic thumbnail for the link(Check the link for long list) submitted by /u/pmttyji [link] [comments]
When building long-running coding agents, terminal output is one of the fastest ways to poison a context window. If an agent runs an intensive build, an…
I was curious whether any of the FA-3/4 optimizations transfer to RTX GPUs. vLLM/SGLang attention falls back to FA-2 on consumer cards (FA-3 and FA-4 are…
TLDR: 75B-total / 9B-active MoE is the perfect shape for multi-24GB rigs, and almost nobody ships it. Qwen 27B is a great model and punches way…