LOADING DATASHEET
LOADING DATASHEET
by DeepSeek
RANKED #227 OF 307 ASSISTANTS · OVERALL #359 OF 6,901 · VIBE SCORE 5.5 · 25 VOICES
DeepSeek-R1-Zero on Hugging Face (text generation). 7,215 downloads. Open weights for local or hosted use.
7 mentions
1 mentions
26 weeks · 8 voices
TOO QUIET
Not enough weekly voices to call a trend yet.
Open weights. You can use a hosted service, or download it and run it yourself, free.
The internet is split on DeepSeek R1 Zero. Biggest praise: getting facts right. Biggest gripe: price.
22 POSTS · 3 COMMENTS · REDDIT · X · HACKER NEWS · BLOG · GITHUB · STACK OVERFLOW · LEMMY
6 thumbs up · 2 thumbs down
1 thumbs up · 3 thumbs down
1 thumbs up · 2 thumbs down
Other models the crowd has fully reviewed, starting with text models like this one.
How was DeepSeek-R1 built; For dummies
Deepseek-r1-Zero is the most uncensored model
[D] r/MachineLearning - a year in review
DeepSeek-R1: How Did They Make an OpenAI-Level Reasoning Model So Damn Efficient?
DeepSeek R1-Zero Removes the Human Bottleneck
(from [8, 12, 14]) Reinforcement learning (RL) has played a crucial role throughout the history of research on large language models (LLMs). Through RL, we created early versions of instruction following models , made important advancements in alignment and…
It has been almost two years since OpenAI released o1, a model that popularized the idea of LLM-based reasoning models. DeepSeek-R1 followed about four months later, together with details of a reinforcement learning with verifiable rewards (RLVR) recipe to…
(from [1, 2, 3]) Scaling is one of the most impactful concepts in the history of AI research. For large language models (LLMs), scaling has mostly been studied in the context of pretraining, where rigorous scaling laws have allowed us to clearly define the…
(from [1, 3, 4]) Recent research on large language models (LLMs) has been heavily focused on reasoning and reinforcement learning (RL). At the center of this research lies Group Relative Policy Optimization (GRPO) [13], the RL optimizer used to train most…
(from [1, 19]) Reinforcement learning (RL) has always played a pivotal role in research on large language models (LLMs), beginning with its use for aligning LLMs to human preferences. More recently, researchers have heavily focused on using RL training to…
AI research team claims to reproduce DeepSeek core technologies for $30 — relatively small R1-Zero model has remarkable problem-solving abilities
Are LLMs Limited by Human Language?