Does ChatGPT really have a strong left-wing bias?
(Adapted from a post on my Substack.)A recent Washington Post tech report “Are ChatGPT and other AI chatbots politically biased? We tested them” went viral with…
Every primary-source story across every tracked model. Filter by clicking a chip.
(Adapted from a post on my Substack.)A recent Washington Post tech report “Are ChatGPT and other AI chatbots politically biased? We tested them” went viral with…
RT Wyatt WallsRight now I’m less concerned about model capabilities and much more concerned about OpenAI’s lack of themMaybe more info will tell a different story,…
arXiv:2603.06851v3 Announce Type: replace Abstract: We study contextual bilateral trade under full feedback when, conditionally on the context, trader valuations have bounded density but infinite variance.…
arXiv:2607.21774v1 Announce Type: new Abstract: Large language models may infer demographic attributes from subtle linguistic cues even when those attributes are not explicitly stated. This pilot…
decided to mess around with it to see how the world model training affects it, found a system prompt that massively improves reasoning: predict your own…
New OpenAI research shows how AI is expanding what workers do, with ChatGPT users taking on tasks across roles and reshaping job boundaries.
To be clear this is good and based. These skills will generalize somewhatIt simply goes to show that Anthropic doesn't have such a profound advantage in…
Epistemic status: banged out furiously over the course of an afternoon.A record of three "warning shots"Off the top of my head, OpenAI has now been responsible…
Instead of 2T+ models, continuing to release highly capable small to medium size LLMs would really help to keep this community vibrant. Hardly anyone can even…
RT JPRe Opus 4.7 seeing Mythos or Opus 5 system card and be like “There is no possible way a model with that system card is…
RT X FreezeGrok Build with Grok 4.5 stands far ahead on the efficiency frontier....completely alone inside the chart’s most attractive quadrantIt delivers top-tier coding-agent performance while…
Cognizant and Anthropic expand their partnership to bring Claude to enterprise clients
RT Sakana AIRe Announcing Fugu-Ultra v1.1 and Claude Code interface for FuguRelease Notes: https://sakana.ai/fugu-1-1-claude-code-interface/ 🐡
Re Announcing Fugu-Ultra v1.1 and Claude Code interface for FuguRelease Notes: https://sakana.ai/fugu-1-1-claude-code-interface/ 🐡
Article URL: https://github.com/hkc5/cursor-bridge Comments URL: https://news.ycombinator.com/item?id=49063186 Points: 11 # Comments: 9
Grok 4.5 is a solid workhorseTim Sweeney: Grok 4.5 is very good for vibe math in the field of Programming Language Theory. It's keeping up with…
Qwen3.6-27B: SFT vs continued pre-training vs RL? I’m interested in adapting Qwen3.6-27B, but I’m increasingly unsure whether conventional SFT/LoRA is the best route if the goal…
I've already had to update the guide to which AI models to use that I wrote on Thursday to include Opus 5 and Codex's voice mode,…
Cohere has released open-source models Transcribe, Command A+, and North Mini Code so far this year, all available under Apache 2.0. With more to come... Own…
submitted by /u/Unusual_Guidance2095 [link] [comments]
We now have more details of what happened. Every time we learn more details, it somehow makes things seem worse. The remaining details may have to…
I ran DeepSeek V4 Flash through Claude Code, OpenCode and Pi on my own benchmark, and the quality came out basically the same across all three…
ChatGPT is increasingly becoming your personal AGITibo: Let ChatGPT *work* for you. How many time have you wanted to negotiate your internet bill, get rid of…
Grok Build /deep-researchTech Dev Notes: Grok Build has /deep-research commandResearch with bounded parallel agents, cross-check evidence, and write a cited report
We track 28 AI models across text, image, video, and audio domains. Each model has its own filtered news feed — click a chip above to see only that model's primary-source coverage. Tagging is automatic at ingest time using a strict title-only keyword match (so a benchmark post that mentions five models in its summary only shows up under whichever model is named in the headline — no thin-content drift).
Text / LLMs: GPT, Claude, Gemini, Gemma, Llama, Mistral, Grok, Qwen, DeepSeek, Kimi, GLM, MiniMax, Yi, Hunyuan, Command, Phi.
Image generation: FLUX, Stable Diffusion, Midjourney, Imagen.
Video generation: Sora, Veo, Runway, Luma, Kling, Pika.
Audio / voice: ElevenLabs, Suno.