3. 模型与基准 (Models & Benchmarks)
- LLM Comparison: Sourced Pricing, Context & Tool Calling ...
8 hours ago · Compare Large Language Models Compare current text-generation listings by API (Application Programming Interface) price, context limit, image input and tool calling. Source: OpenRouter's public catalog, not a benchmark leaderboard. Source checked: 2026-09-19 15:09 U
- AI & LLM Benchmarks 2026: Rankings, Scores & Results
49 minutes ago · AI & LLM Benchmarks 2026 Explore live AI and LLM benchmark rankings across reasoning, coding, math, vision, agents, and tool use. Compare composite indexes first, then open individual evaluations for score provenance, coverage, and methodology.
- AI 编程工具—Cursor进阶使用deepseek V3 模型(deepseek + cursor...
今日模型配置页面,这里我们勾掉之前已经激活的模型.创建deepseek-chat 的模型后,选中这个模型,然后在下面的配置框中输入我们刚才生成的API keys.
- GitHub 开源项目 30 日趋势榜 - kaiyuanbang.cn
2 minutes ago · 榜单分类 GitHub 开源项目 30 日趋势榜 榜单按增长热度排序;已生成当前语言解读的项目进入站内详情,正在生成的同步项目跳转 GitHub 原站。
- AI Benchmarks 2026 - MMLU, GPQA, SWE-bench | LM Market Cap
1 day ago · Compare AI models across 17 benchmarks including MMLU, GPQA Diamond, MATH-500, HumanEval, SWE-bench, and Arena Elo. See current leaders, score history, and interactive charts for 350+ models.
- Open-Source LLM Leaderboard 2026 | LM Market Cap
1 day ago · Open LLM leaderboard ranking the best open-source language models by benchmarks, pricing, and capabilities. Compare Llama, DeepSeek, Qwen, Mistral, and Gemma with live scores.