
DeepSeek-R1
DeepSeek's first-generation reasoning models. DeepSeek-R1-Zero, a model trained via large-scale reinforcement learning without supervised fine-tuning, demonstrated remarkable performance on reasoning. - 스마트 AI 도구로 생산성 향상.


A strong Mixture-of-Experts (MoE) language model with 671B total parameters with 37B activated for each token. - 스마트 AI 도구로 생산성 향상.
1
Views
0
Likes
Jan 2026
Added
github.com
Website
A strong Mixture-of-Experts (MoE) language model with 671B total parameters with 37B activated for each token.
DeepSeek-V3은(는) open-source-llm 카테고리의 우수한 도구로, AI 지원이 필요한 모든 사용자에게 적합합니다.
Visit the official website to get started
Have an AI tool to share?
Submit Your Tool
DeepSeek's first-generation reasoning models. DeepSeek-R1-Zero, a model trained via large-scale reinforcement learning without supervised fine-tuning, demonstrated remarkable performance on reasoning. - 스마트 AI 도구로 생산성 향상.

Qwen3 is the large language model series developed by Qwen team, Alibaba Cloud. - 스마트 AI 도구로 생산성 향상.

Llama3 is a large language model developed by Meta AI. It is the successor to Meta's Llama2 language model. - 스마트 AI 도구로 생산성 향상.

Mixtral 8x7B, a high-quality sparse mixture of experts model (SMoE) with open weights. Mixtral outperforms Llama 2 70B on most benchmarks with 6x faster inference. - 스마트 AI 도구로 생산성 향상.