[Discussion] non developers: what are some use cases for your local models?
i get curious about this a lot, and majority of the times i get a response which is similar to mine that is : it fulfills…
i get curious about this a lot, and majority of the times i get a response which is similar to mine that is : it fulfills…
submitted by /u/arsenyinfo [link] [comments]
My weekend sideproject was implementing the paper 'Approximating Softmax for FPGAs with Taylor Series and Pade Approximants' The paper’s motivation is the hardware constraints limiting exponential…
submitted by /u/pmigdal [link] [comments]
XYZ-Aquila is a family of open-weight Deep Search agents developed by XYZ AI Lab. XYZ-Aquila-mini is post-trained from Qwen3.6-35B-A3B through a bounded-exploration AI4AI pipeline: humans define…
Jensen Huang on 𝕏: https://x.com/JensenHuang/status/2081698060330250294 submitted by /u/Nunki08 [link] [comments]
submitted by /u/celsowm [link] [comments]
submitted by /u/InternationalGap3698 [link] [comments]
I finished the speed leg of my spec-decode benchmarking for Qwen3.6-27B, main algorithms across quants. Overall: the heavier the quant, the more spec-decode buys you (10…
Forked SGLang, wrote TeilLang FlashAttention for V100, used open-source marlin-v100, ungated flashinfer for sm70, made Dflash work for Qwen3.5/3.6 models, added Laguna S2.1 support, tried to…