Devstral

Agentic coding model from Mistral AI and All Hands AI that runs on a single high-end consumer GPU.

Open SourceSelf HostedOffline CapableGPU Required (24GB+ VRAM)
0.0 (0)

About

Devstral targets agentic coding: Mistral AI built it with All Hands AI, the team behind OpenHands, to explore codebases, edit multiple files, and drive software engineering agent scaffolds rather than autocomplete snippets. Devstral Small, a 24B-parameter model with a 128K context window, scores 53.6 percent on SWE-Bench Verified, which made it the top open-source model on that benchmark at release, ahead of much larger models such as DeepSeek-V3-0324 under the same scaffold, and it is light enough to run on a single RTX 4090 or a Mac with 32 GB of RAM. Weights are Apache-2.0 on Hugging Face, so commercial use and fine-tuning are unrestricted, and deployment is supported through vLLM (recommended), Ollama, llama.cpp, LM Studio, mistral-inference, and Transformers, with a published system prompt file tuned for agent-based interactions. Paired with OpenHands or similar agent frameworks it handles automated issue resolution, multi-file edits, and repository-scale code exploration, making it one of the most practical models to run locally for software engineering agents rather than chat-style code assistance.

Reviews (0)

Leave a Review

No reviews yet. Be the first to review!

Details

Price
Free
Platform
Local/Desktop
Difficulty
Easy (2/5)
License
Apache-2.0
Minimum VRAM
24 GB
Added
Jul 29, 2026

Related Tools

Featured

Lightweight open-weight LLM by Google available in 1B to 27B sizes.

Open SourceSelf HostedOfflineGPU 4GB+
Easy
0.0 (0)

Open-source code LLM family by IBM for enterprise code generation.

Open SourceSelf HostedOfflineGPU 8GB+
Intermediate
0.0 (0)

Open-weight LLM by Meta in 8B and 70B sizes with strong general capabilities.

Open SourceSelf HostedOfflineGPU 8GB+
Intermediate
0.0 (0)
Featured

High-performance open-weight MoE LLM with 671B total parameters.

Open SourceSelf HostedOfflineGPU 24GB+
Advanced
0.0 (0)

Hybrid SSM-Transformer model by AI21 Labs combining Mamba with attention layers.

Open SourceSelf HostedOfflineGPU 12GB+
Intermediate
0.0 (0)

Open-weight code LLM trained on 2 trillion tokens of code and natural language.

Open SourceSelf HostedOfflineGPU 8GB+
Intermediate
0.0 (0)
Browse all Large Language Models (LLMs) tools