LOADING DATASHEET
LOADING DATASHEET
by OVHcloud AI Endpoints
RANKED #230 OF 271 ASSISTANTS · OVERALL #365 OF 6,567 · VIBE SCORE 5.3 · 29 VOICES
Mistral model for multilingual chat, reasoning, and tool-assisted workflows
29 mentions
0 mentions
26 weeks · 11 voices
TOO QUIET
Not enough weekly voices to call a trend yet.
Open weights: download OVHcloud AI Endpoints 7B Instruct v0.3 once and it's yours. No subscription, no rate limits, works offline.
7.25B parameters, a Q4_K_M GGUF file is about 4.4 GB (FP16 is ~14.5 GB)
NVIDIA GeForce RTX 3050 (6 GB)
6 GB VRAM
~$180 NEWQ4_K_M quantization, ~35 tok/s, 8k context
NVIDIA GeForce RTX 4060 (8 GB)
8 GB VRAM
~$299 NEWQ5_K_M / Q8_0 quantization, ~65 tok/s, 16k context
NVIDIA GeForce RTX 4070 Ti SUPER (16 GB)
16 GB VRAM
~$799 NEWFP16 uncompressed, ~90 tok/s, full 32k context
The three cards above are researched picks. Search the machine you actually have — we only say yes if it should reply at a usable speed, not just load the weights and crawl.
HARDWARE PICKS & PRICES ARE RESEARCHED FROM THE LIVE WEB AND REFRESHED AUTOMATICALLY. TREAT THEM AS BALLPARK, NOT GOSPEL.
Everyday jobs on each of those rigs: how long OVHcloud AI Endpoints 7B Instruct v0.3 takes, and how good the crowd says it is at that kind of work.
| THE JOB | BARE MINIMUMNVIDIA GeForce RTX 3050 (6 GB) | SWEET SPOTNVIDIA GeForce RTX 4060 (8 GB) | FULL POWERNVIDIA GeForce RTX 4070 Ti SUPER (16 GB) |
|---|---|---|---|
| Write an emaila paragraph or two, drafted from a one-line brief | ~14 S★★★★★ | ~8 S★★★★★ | ~6 S★★★★★ |
| Summarize a documenta long report boiled down to the points that matter | ~43 S★★★★★ | ~23 S★★★★★ | ~17 S★★★★★ |
| Build a websitea small landing page, markup and styles together | ~3.8 MIN★★★★★ | ~2.1 MIN★★★★★ | ~89 S★★★★★ |
| Build a backendan API with routes, storage and tests | ~12 MIN★★★★★ | ~6.4 MIN★★★★★ | ~4.6 MIN★★★★★ |
| Build a gamea playable browser game in one file | ~7.1 MIN★★★★★ | ~3.8 MIN★★★★★ | ~2.8 MIN★★★★★ |
STARS ARE THE COMMUNITY SCORE FOR THAT KIND OF WORK, DOCKED FOR HOW SQUASHED THE WEIGHTS GET AT EACH TIER. TIMES ARE COMPUTED FROM RESEARCHED SPEEDS. BALLPARK, NOT GOSPEL.
Not enough discussion about OVHcloud AI Endpoints 7B Instruct v0.3 yet to call it. We found 29 posts but people did not say much either way.
22 POSTS · 7 COMMENTS · REDDIT · STACK OVERFLOW · DEV FORUMS · LEMMY · GITHUB
Other models the crowd has fully reviewed, starting with text models like this one.
Summary: The big AI events of the great 2024
Simple RAG + Web Search + PDF Chat
[P] New collection of Llama, Mistral, Phi, Qwen, and Gemma models for function/tool calling
AI boyfriend users radicalise against OpenAI — and self-host their chatbot companions
D2 sidecars for Mistral-7B-Instruct-v0.3: near-Q8 128K retrieval at lower KV memory
This week in AI - all the Major AI developments in a nutshell
[Help] How to fine-tune Mistral-7B-Instruct-v0.3 to become a DevOps & Cloud tutor AI?
This week in AI - all the Major AI developments in a nutshell
Simple RAG + Web Search + PDF Chat
Let's test different models on counting e in Deepseek.
Comment in r/singularity
Presenting TIS (Token Importance Scoring) - A new way to compress KV cache