3. 模型与基准 (Models & Benchmarks)
- 2026年AI大模型综合排行榜 - AnyRank - AnyRank
2 minutes ago · 趣事与总结 本月榜单显示, Anthropic 和 Google 在第一梯队竞争极其激烈,二者在编程能力(Code Arena / SWE-bench)与综合指标上交替领先;同时, 国产大模型 (智谱、月之暗面、阿里巴巴)表现异常强劲,稳居全球前20,并在中文特定语境与性价比上确立了护城河。
- AI 编程工具—Cursor进阶使用deepseek V3 模型(deepseek + cursor...
今日模型配置页面,这里我们勾掉之前已经激活的模型.创建deepseek-chat 的模型后,选中这个模型,然后在下面的配置框中输入我们刚才生成的API keys.
- LLM News - 最新AI/ML开源仓库、研究论文动态
2 minutes ago · About LLM News LLM News is an automated tracker for AI/ML developments, focusing on Large Language Models, AGI, and related technologies.
- AI Benchmarks 2026 - MMLU, GPQA, SWE-bench | LM Market Cap
1 day ago · Compare AI models across 17 benchmarks including MMLU, GPQA Diamond, MATH-500, HumanEval, SWE-bench, and Arena Elo. See current leaders, score history, and interactive charts for 350+ models.
- Open-Source LLM Leaderboard 2026 | LM Market Cap
1 day ago · Open LLM leaderboard ranking the best open-source language models by benchmarks, pricing, and capabilities. Compare Llama, DeepSeek, Qwen, Mistral, and Gemma with live scores.
- Best Open Source AI Models & LLM Leaderboard (2026)
1 day ago · Open source AI models ranked: Llama 4, DeepSeek, Qwen, Mistral, and Gemma compared by score, pricing, and capabilities. The open source LLM leaderboard updated hourly.