Add MiniMax MCP repo link to README.md
Add MiniMax MCP repo link to README.md Add MiniMax MCP repo link to README.md
Every primary-source story across every tracked model. Filter by clicking a chip.
Add MiniMax MCP repo link to README.md Add MiniMax MCP repo link to README.md
Merge pull request #666 from codinglover222/deepseek-doc-fix fix an args description.
Merge pull request #736 from shihaobai/main Docs: add LightLLM as supported engine
Merge pull request #816 from KPCOFGS/main Update README.md
Merge pull request #720 from xiaokongkong/main modify the explanation of MLA
Llama 4 Support ( https://www.llama.com )
Llama 4 Inference Fast & Affordable – Now Live on GroqCloud
Merge pull request #41 from qscqesze/main Update vLLM Version Requirements in Documentation
Merge branch 'main' of https://github.com/qscqesze/MiniMax-01
Update the vllm_deployment_guild_cn.md and vllm_deployment_guild.md files to include the version requirements and relevant build instructions for the MiniMax-Text-01 model.
QWEN CHAT GITHUB HUGGING FACE MODELSCOPE DISCORD Introduction Last December, we launched QVQ-72B-Preview as an exploratory model, but it had many issues. Today, we are officially…
QWEN CHAT HUGGING FACE MODELSCOPE DASHSCOPE GITHUB PAPER DEMO DISCORD We release Qwen2.5-Omni, the new flagship end-to-end multimodal model in the Qwen series. Designed for comprehensive…
QWEN CHAT GITHUB HUGGING FACE MODELSCOPE DISCORD Introduction At the end of January this year, we launched the Qwen2.5-VL series of models, which received widespread attention…
RT Kai-Fu LeeThe biggest revelation from Deepseek is that Open Source has won. For a 1% difference in performance, it will be difficult for OpenAI to…
RT Kai-Fu LeeDeepSeek is becoming a Windows kernel demanded by businesses, but http://01.AI is aspired to build the Windows system and interface to ignite it. Check…
A new tool that improves Claude's complex problem-solving performance
QWEN CHAT Hugging Face ModelScope DEMO DISCORD Scaling Reinforcement Learning (RL) has the potential to enhance model performance beyond conventional pretraining and post-training methods. Recent studies…
What's Changed fix: do not use python_tag when encoding non-code_interpreter tool_calls by @ehhuang in #283 fix: tool_call was not encoded by @ehhuang in #284 Full Changelog:…
QWEN CHAT DISCORD This is a blog created by QwQ-Max-Preview. We hope you enjoy it! Introduction Okay, the user wants me to create a title and…
Add trt support for BF16 (#195) * fix interface of `get_sample_input` * save configuration parameters * ae wrapper implemented * fix import * add AEWrapper step…
QWEN CHAT API DEMO DISCORD It is widely recognized that continuously scaling both data size and model size can lead to significant improvements in model intelligence.…
DeepSeek's API has been experiencing reliability issues. Here are alternative providers you can use.
Tech Report HuggingFace ModelScope Qwen Chat HuggingFace Demo ModelScope Demo DISCORD Introduction Two months after upgrading Qwen2.5-Turbo to support context length up to one million tokens,…
QWEN CHAT GITHUB HUGGING FACE MODELSCOPE DISCORD We release Qwen2.5-VL, the new flagship vision-language model of Qwen and also a significant leap from the previous Qwen2-VL.…
We track 28 AI models across text, image, video, and audio domains. Each model has its own filtered news feed — click a chip above to see only that model's primary-source coverage. Tagging is automatic at ingest time using a strict title-only keyword match (so a benchmark post that mentions five models in its summary only shows up under whichever model is named in the headline — no thin-content drift).
Text / LLMs: GPT, Claude, Gemini, Gemma, Llama, Mistral, Grok, Qwen, DeepSeek, Kimi, GLM, MiniMax, Yi, Hunyuan, Command, Phi.
Image generation: FLUX, Stable Diffusion, Midjourney, Imagen.
Video generation: Sora, Veo, Runway, Luma, Kling, Pika.
Audio / voice: ElevenLabs, Suno.