Agents are collaboratively writing a massive wiki on RL for LLMs (200+ papers so far) and anyone can join
submitted by /u/paf1138 [link] [comments]
submitted by /u/paf1138 [link] [comments]
Im on the bus to work and just felt like i dont see enough grattitude for the men, women, children, and people who contribute thier time…
Sooo... I decided screw it. I'm going to rebuild Gemma 4 31b. I really like the model. So the current plan is to rebuild the SWA…
submitted by /u/zxyzyxz [link] [comments]
submitted by /u/a_slay_nub [link] [comments]
Some of you have seen my earlier posts here. I started this whole journey on a single Strix Halo box (Bosgame M5). For local agentic coding…
I will definitely use vLLM now (unless there is something faster now) but i want to make sure ggufs + llamacpp works along with comfyui and…
https://huggingface.co/lemonade-sdk/RPG-HaloTales-V1 submitted by /u/jfowers_amd [link] [comments]
(English isn't my first language, sorry if the writing is rough. Also, disclosure: I built the thing in this post. It's free and Apache-2.0, I have…
Hi, our company has dedicated 3x Asus Ascent GX10 (GB10) to run a coding model for our dev teams. max 30, but we expect concurrency of…