
DeepSeek-V3
A strong Mixture-of-Experts (MoE) language model with 671B total parameters with 37B activated for each token. - スマートな AI ツールで生産性を向上。


DeepSeek's first-generation reasoning models. DeepSeek-R1-Zero, a model trained via large-scale reinforcement learning without supervised fine-tuning, demonstrated remarkable performance on reasoning. - スマートな AI ツールで生産性を向上。
57
Views
0
Likes
Jan 2026
Added
github.com
Website
Editorial Review
DeepSeek's first-generation reasoning models. DeepSeek-R1-Zero, a model trained via large-scale reinforcement learning without supervised fine-tuning, demonstrated remarkable performance on reasoning.
DeepSeek-R1 は open-source-llm カテゴリーの優れたツールで、AI 支援を必要とするすべてのユーザーに適しています。
Visit the official website to get started
Have an AI tool to share?
Get your product in front of people actively exploring AI tools.
Submit Your Tool
A strong Mixture-of-Experts (MoE) language model with 671B total parameters with 37B activated for each token. - スマートな AI ツールで生産性を向上。

Qwen3 is the large language model series developed by Qwen team, Alibaba Cloud. - スマートな AI ツールで生産性を向上。

Llama3 is a large language model developed by Meta AI. It is the successor to Meta's Llama2 language model. - スマートな AI ツールで生産性を向上。

Mixtral は、Mistral AI の Apache-2.0 の疎な専門家混合モデル ファミリであり、Mixtral 8x7B および 8x22B ベースおよび命令バリアントが含まれます。この独立したガイドでは、ルーティング、アクティブ パラメーターと合計パラメーター、メモリとサービスのコスト、量子化、評価、安全性、最新の代替手段について説明します。