LOADING DATASHEET
LOADING DATASHEET
by Alibaba
RANKED #155 OF 236 ASSISTANTS · OVERALL #254 OF 6,566 · VIBE SCORE 5.6 · 48 VOICES
qwen3-asr-1.7b-ja-anime-GGUF on Hugging Face (automatic speech recognition). 3,217 downloads. Open weights for local or hosted use.
4 mentions
5 mentions
26 weeks · 53 voices
HONEYMOON FADING
Early praise is cooling off in recent weeks.
Open weights. You can use a hosted service, or download it and run it yourself, free.
Open weights: download Qwen3 1.7B once and it's yours. No subscription, no rate limits, works offline.
1.7B parameters, a Q4 file is about 1.1 GB, full FP16 is about 3.4 GB
Raspberry Pi 5 (8GB)
8 GB RAM
~$80 NEWQ4_K_M, ~12 tok/s, 4k context
Mac mini M2 (8GB)
8 GB UNIFIED MEMORY
~$450 USEDQ8_0 or FP16, ~65 tok/s, 32k context
Nvidia GeForce RTX 3060
12 GB VRAM
~$260 USEDFP16 full precision, ~110 tok/s, 32k full context
The three cards above are researched picks. Search the machine you actually have — we only say yes if it should reply at a usable speed, not just load the weights and crawl.
HARDWARE PICKS & PRICES ARE RESEARCHED FROM THE LIVE WEB AND REFRESHED AUTOMATICALLY. TREAT THEM AS BALLPARK, NOT GOSPEL.
Everyday jobs on each of those rigs: how long Qwen3 1.7B takes, and how good the crowd says it is at that kind of work.
| THE JOB | BARE MINIMUMRaspberry Pi 5 (8GB) | SWEET SPOTMac mini M2 (8GB) | FULL POWERNvidia GeForce RTX 3060 |
|---|---|---|---|
| Write an emaila paragraph or two, drafted from a one-line brief | ~42 S | ~8 S | ~5 S |
| Summarize a documenta long report boiled down to the points that matter | ~2.1 MIN★★★★★ | ~23 S★★★★★ | ~14 S★★★★★ |
| Build a websitea small landing page, markup and styles together | ~11 MIN★★★★★ | ~2.1 MIN★★★★★ | ~73 S★★★★★ |
| Build a backendan API with routes, storage and tests | ~35 MIN★★★★★ | ~6.4 MIN★★★★★ | ~3.8 MIN★★★★★ |
| Build a gamea playable browser game in one file | ~21 MIN★★★★★ | ~3.8 MIN★★★★★ | ~2.3 MIN★★★★★ |
STARS ARE THE COMMUNITY SCORE FOR THAT KIND OF WORK, DOCKED FOR HOW SQUASHED THE WEIGHTS GET AT EACH TIER. TIMES ARE COMPUTED FROM RESEARCHED SPEEDS. BALLPARK, NOT GOSPEL.
Not enough discussion about Qwen3 1.7B yet to call it. We found 48 posts but people did not say much either way.
44 POSTS · 4 COMMENTS · DEV FORUMS · REDDIT · GITHUB · LEMMY
Other models the crowd has fully reviewed, starting with text models like this one.
I tested 10 LLMs locally on my MacBook Air M1 (8GB RAM!) – Here's what actually works-
I tested local models on 100+ real RAG tasks. Here are the best 1B model picks
LiquidAI released LFM2.5 Thinking: Runs entirely on-device (phone) with 900MB of memory
Tome (open source local LLM + MCP client) now has Windows support!
Released a Claude Code skill that fine-tunes a small model from your agent's production traces, end-to-end in one conversation
"RLP: Reinforcement as a Pretraining Objective"
MOE Pipeline
New functiongemma model: not worth downloading
I Made LLMs Play Texas Hold’em. The Smallest Model Beat a ~1T Model by Being Too Dumb to Fold
ollama support for qwen3 for tab completion in Continue
Need help with chunking and embedding strategies for my rag model
Genuine question : NEWBIE ALERT