Qwen3.8-Max Aquarium break simulator
Qwen3.8-Max was one of the few (alongside opus4.8 and opus 5) that was able to nail aquarium break simulation in oneshot!! submitted by /u/kms_dev [link] [comments]
Qwen3.8-Max was one of the few (alongside opus4.8 and opus 5) that was able to nail aquarium break simulation in oneshot!! submitted by /u/kms_dev [link] [comments]
Qwen3.8-Max oneshots across 35 prompts https://oneshotlm.com/model/qwen-qwen3-8-max/ submitted by /u/kms_dev [link] [comments]
I am a college student, and I am currently experimenting with making smaller models more efficient and useful, but I am currently using a slow custom…
Has anyone followed nousresearch work on Hermes? I mean we are Q3 2026. We have some crazy models trickling down from HGX territory to multi gpu…
submitted by /u/adefa [link] [comments]
submitted by /u/Fun_Librarian_7699 [link] [comments]
Hugging Face Artificial Analysis Should be a sweet spot for general work. Seems like coding is the only part that is inferior to Qwen. submitted by…
I tested Ling-3.0-flash with hard bugs and it fixed bugs that qwen3.6-27b could not. This models speed faster than deepseek v4 flash but almost the same…
Ran DeepSeek V4-Flash-0731 — the full official checkpoint, not a re-quant — on commodity used hardware. Sharing because I couldn't find anyone else publishing Ampere results…
In this paper, the authors tackle continued pretraining without the risk of catastrophic forgetting, by identifying parameters which can safely be changed without risking identified concepts,…