Do Qwen 3.6 27B quantizations break the pelican?
submitted by /u/pmigdal [link] [comments]
Every primary-source story across every tracked model. Filter by clicking a chip.
submitted by /u/pmigdal [link] [comments]
Article URL: https://status.claude.com/incidents/mfdtrknpxghq Comments URL: https://news.ycombinator.com/item?id=49068029 Points: 5 # Comments: 1
How AI is expanding what people do at work
I finished the speed leg of my spec-decode benchmarking for Qwen3.6-27B, main algorithms across quants. Overall: the heavier the quant, the more spec-decode buys you (10…
Forked SGLang, wrote TeilLang FlashAttention for V100, used open-source marlin-v100, ungated flashinfer for sm70, made Dflash work for Qwen3.5/3.6 models, added Laguna S2.1 support, tried to…
submitted by /u/SignificantLegs [link] [comments]
Hey r/LocalLLaMA, I’m a retired engineer with a background in distributed computing, currently running a 1-person startup. Like many people here, I’d love to experiment with…
Article URL: https://status.claude.com/incidents/lhqp09kxq7pb Comments URL: https://news.ycombinator.com/item?id=49066591 Points: 4 # Comments: 1
Shared conversations with Anthropic's Claude chatbot briefly appeared in Google search results because the pages lacked a noindex tag. Users said some chats contained crypto keys…
Test Prompts: 1.1. Algorithm & Logic (10 pts): "Write a function in Python that finds the contiguous subarray with the largest sum (Kadane's algorithm). Include time…
(Adapted from a post on my Substack.)A recent Washington Post tech report “Are ChatGPT and other AI chatbots politically biased? We tested them” went viral with…
RT Wyatt WallsRight now I’m less concerned about model capabilities and much more concerned about OpenAI’s lack of themMaybe more info will tell a different story,…
arXiv:2603.06851v3 Announce Type: replace Abstract: We study contextual bilateral trade under full feedback when, conditionally on the context, trader valuations have bounded density but infinite variance.…
arXiv:2607.21774v1 Announce Type: new Abstract: Large language models may infer demographic attributes from subtle linguistic cues even when those attributes are not explicitly stated. This pilot…
decided to mess around with it to see how the world model training affects it, found a system prompt that massively improves reasoning: predict your own…
New OpenAI research shows how AI is expanding what workers do, with ChatGPT users taking on tasks across roles and reshaping job boundaries.
To be clear this is good and based. These skills will generalize somewhatIt simply goes to show that Anthropic doesn't have such a profound advantage in…
Epistemic status: banged out furiously over the course of an afternoon.A record of three "warning shots"Off the top of my head, OpenAI has now been responsible…
Instead of 2T+ models, continuing to release highly capable small to medium size LLMs would really help to keep this community vibrant. Hardly anyone can even…
RT JPRe Opus 4.7 seeing Mythos or Opus 5 system card and be like “There is no possible way a model with that system card is…
RT X FreezeGrok Build with Grok 4.5 stands far ahead on the efficiency frontier....completely alone inside the chart’s most attractive quadrantIt delivers top-tier coding-agent performance while…
Cognizant and Anthropic expand their partnership to bring Claude to enterprise clients
RT Sakana AIRe Announcing Fugu-Ultra v1.1 and Claude Code interface for FuguRelease Notes: https://sakana.ai/fugu-1-1-claude-code-interface/ 🐡
Re Announcing Fugu-Ultra v1.1 and Claude Code interface for FuguRelease Notes: https://sakana.ai/fugu-1-1-claude-code-interface/ 🐡
We track 28 AI models across text, image, video, and audio domains. Each model has its own filtered news feed — click a chip above to see only that model's primary-source coverage. Tagging is automatic at ingest time using a strict title-only keyword match (so a benchmark post that mentions five models in its summary only shows up under whichever model is named in the headline — no thin-content drift).
Text / LLMs: GPT, Claude, Gemini, Gemma, Llama, Mistral, Grok, Qwen, DeepSeek, Kimi, GLM, MiniMax, Yi, Hunyuan, Command, Phi.
Image generation: FLUX, Stable Diffusion, Midjourney, Imagen.
Video generation: Sora, Veo, Runway, Luma, Kling, Pika.
Audio / voice: ElevenLabs, Suno.