Ornith 1.0 35B
DeepReinforce · Ornith
Scores 75.6% on SWE-bench Verified — one of the highest results for any open-weight coding model and above most paid assistants. From the lab DeepReinforce, trained with RL self-improvement for terminal control and tool calling. MIT licensed, 262K context. There is no Ollama library tag yet — pull the GGUF directly with ollama run hf.co/deepreinforce-ai/Ornith-1.0-35B-GGUF.
Params
35B
License
MIT
Context
262K
Min VRAM
20 GB
Min RAM
32 GB
Tier
Heavy
Run locally with Ollama
ollama run hf.co/deepreinforce-ai/Ornith-1.0-35B-GGUFOnce running, Ollama exposes an OpenAI-compatible endpoint at localhost:11434/v1.
Will it run on my hardware?
Ornith 1.0 35B needs about 20 GB of fast memory at Q4. Here is what in our hardware database clears that bar.
Runs at full speed · 6
Runs, but slower on shared memory · 3
6 other tracked configurations do not have enough memory.
Hardware links are affiliate links. We earn a small commission if you buy through them — at no extra cost to you. Disclosure
Specifications
| Parameters | 35B |
|---|---|
| Context window | 262K |
| License | MIT |
| Min VRAM (Q4) | 20 GB |
| Min RAM | 32 GB |
| Recommended tier | Heavy |
| Best for | codingdebuggingcode reviewagentic tasks |
| Reported benchmark | 75.6% on SWE-bench Verified (as stated by the model developer) |
Variants on Ollama
This model has no entry in the Ollama library yet, so it is pulled directly from its HuggingFace GGUF repo.
Ornith 1.0 35B
ollama run hf.co/deepreinforce-ai/Ornith-1.0-35B-GGUFPopularity trend (illustrative)
Illustrative adoption trend, not actual download counts. We do not publish a download number for any model because we have no honest source for it.
Similar models
Mixtral 8x7B
Mistral AI · Apache 2.0
A sparse mixture-of-experts model that activates only 12B parameters per token while carrying 47B total, delivering 70B-class results at a fraction of the compute cost. One of the best open-weight models available under a fully permissive license.
Min VRAM
26 GB
Context
32K
Run with Ollama
ollama run mixtral:8x7bGemma 3 27B
Google · Gemma Terms of Use
The largest Gemma 3 model and one of the best open-weight models in the 20–30B range. Fits in a single RTX 4090 (24 GB) with room to spare at Q4 quantization, and delivers instruction-following quality that rivals models twice its size.
Min VRAM
15 GB
Context
128K
Run with Ollama
ollama run gemma3:27bQwen 2.5 Coder 7B
Alibaba · Apache 2.0
A coding-specialist model fine-tuned on a massive corpus of source code across 40+ programming languages, delivering autocomplete and generation quality that rivals dedicated IDE tools. At 7B it is fast enough for real-time code suggestions on a single consumer GPU.
Min VRAM
4 GB
Context
128K
Run with Ollama
ollama run qwen2.5-coder:7b