3. 模型与基准 (Models & Benchmarks)
- 2026年AI大模型综合排行榜 - AnyRank - AnyRank
2 minutes ago · 截至 2026 年 3 月,目前的 LLM 综合排名,主要参考 Code Arena 和各项基准测试得分。 | 本月榜单显示,Anthropic 和 Google 在第一梯队竞争极其激烈,二者在编程能力(Code Arena / S | 前三名:Claude Opus 4.6、Gemini 3.1 Pro、GPT-5.
- LLM News - 最新AI/ML开源仓库、研究论文动态
2 minutes ago · LLM News - Stay updated with the latest AI/ML developments, repositories, and research papers
- DeepSeek Harness—— 编程 AI Agent 的三代进化:从"大而全"到"小而美"...
1 day ago · 2026 年 8 月 13 日,DeepSeek 开源了 DeepSeek Harness(dsh)——不是新模型,而是一套 Agent 的运行时底座。 发布 24 小时内收获约 4.6 万星。 很多人把它当成"又一个对标 Claude Code 的 CLI",但把三代编程 AI Agent 放在一起看,你会发现它站在一条清晰的进化线上。
- Open-Source LLM Leaderboard 2026 | LM Market Cap
1 day ago · Open LLM leaderboard ranking the best open-source language models by benchmarks, pricing, and capabilities. Compare Llama, DeepSeek, Qwen, Mistral, and Gemma with live scores.
- Best Open Source AI Models & LLM Leaderboard (2026)
1 day ago · Open source AI models ranked: Llama 4, DeepSeek, Qwen, Mistral, and Gemma compared by score, pricing, and capabilities. The open source LLM leaderboard updated hourly.
- GitHub Trending 每日榜:热门开源项目 Top 7 | 2026年08月17日 | Ray...
1 day ago · 与 Ollama 等以推理部署为主的工具不同,Unsloth 的差异化在于将 fine-tuning 能力集成到图形化界面中,用户无需编写代码即可完成模型微调。 项目覆盖的模型类型广泛,包括 Qwen3.8、Kimi K3、Gemma 4、DeepSeek-V4 等主流 LLM,以及 FLUX 扩散模型、embedding 模型和音频模型。