LMmy choosing

Model Catalog

Discover models to run locally or use with LM Studio Cloud.

Show

DeepSeek V4 Flash
Available in LM Studio Cloud
Available to download
284B
DeepSeek V4 Flash 0731 is a 284B Mixture-of-Experts model built for coding, tool use, and agentic workflows, with 13B active parameters and a 1M-token context window.
19.1K
3
Updated 5 hours ago
Kimi K3
Available in LM Studio Cloud
Kimi K3 is Moonshot AI’s 2.8T-parameter flagship model for long-horizon agentic work, with native vision, always-on reasoning, and a 1M-token context window.
Updated 6 hours ago
Laguna S 2.1
Available to download
118B
Laguna S 2.1 is Poolside’s 118B Mixture-of-Experts model for agentic coding and long-horizon tool use, with 8B active parameters and a context window of up to 1M tokens.
23.3K
1
Updated 6 hours ago
Bonsai 27B
Available to download
27B
Bonsai 27B is PrismML’s family of 1-bit and ternary Qwen3.6 27B models, with vision, reasoning, tool use, and a 262K-token context window.
95K
12
Updated 6 hours ago
GLM-5.2
Available in LM Studio Cloud
GLM-5.2 is Z.ai’s 753B Mixture-of-Experts flagship model for long-horizon coding and agentic work, with 40B active parameters and flexible reasoning effort.
Updated 6 hours ago
Granite 4.1
Available to download
3B
8B
30B
Granite 4.1 models are new and improved granite models which have gone through an improved post-training pipeline, including supervised finetuning and reinforcement learning alignment, resulting in enhanced tool calling, instruction following, and chat capabilities.
17.3K
56
3
Updated 2 months ago
Nemotron 3 Omni
Available to download
30B
NVIDIA Nemotron 3 Nano Omni is an open multimodal model with highest efficiency that powers sub-agents to complete tasks faster across vision, audio, and language
318.8K
64
Updated 3 months ago
Qwen3.6
Available to download
27B
35B
Qwen3.6 prioritizes stability and real-world utility, offering developers a more intuitive, responsive, and genuinely productive coding experience.
2.5M
222
2
Updated 3 months ago
Gemma 4
Available to download
5.1B
5.1B
7.9B
7.9B
12B
12B
26B
26B
31B
31B
Gemma 4 is Google's most capable family of open models, built from Gemini 3 research. Supports vision input and available in multiple sizes for on-device deployment.
7.9M
1.1K
10
Updated 2 months ago
Nemotron 3 Super
Available to download
120B
NVIDIA Nemotron 3 Super, a 120B open hybrid MoE model (12B active), supporting up to 1M tokens context window
106K
53
Updated 4 months ago
Qwen3.5
Available to download
2B
4B
9B
27B
35B
Qwen3.5 represents a significant leap forward, integrating breakthroughs in multimodal learning, architectural efficiency, reinforcement learning scale, and global accessibility to empower developers and enterprises with unprecedented capability and efficiency
2.3M
414
5
Updated 4 months ago
LFM2-24B-A2B
Available to download
24B
LFM2 is a family of hybrid models designed for on-device deployment. LFM2-24B-A2B is the largest model in the family, scaling the architecture to 24 billion parameters while keeping inference efficient.
81.8K
24
Updated 5 months ago
Qwen3-Coder-Next
Available to download
80B
Qwen3 Coder Next is an 80B MoE with 3B active parameters designed for coding agents and local development. Excels at long-horizon reasoning, complex tool usage, and recovery from execution failures.
288.9K
99
Updated 5 months ago
GLM-4.7
Available to download
30B
Open source coding models by Z.ai, based on a new base model and specializing in coding and tool calling.
331.1K
130
Updated 6 months ago
FunctionGemma
Available to download
270M
FunctionGemma is a lightweight, open model from Google, built as a foundation for creating your own specialized function calling models.
3.1K
55
Updated 7 months ago
Nemotron 3
Available to download
30B
General purpose reasoning and chat model trained from scratch by NVIDIA. Contains 30B total parameters with only 3.5B active at a time for low-latency MoE inference
132.2K
65
Updated 7 months ago
GLM-4.6V-Flash
Available to download
9B
GLM 4.6V Flash is a 9B vision-language model optimized for local deployment and low-latency applications.
290K
75
Updated 7 months ago
Devstral 2
Available to download
24B
123B
Second-generation Devstral for agentic coding. Built for tool use to explore codebases, edit multiple files, and power software engineering agents with newly added vision support.
199.9K
73
2
Updated 7 months ago
Rnj-1
Available to download
8B
Rnj-1 is a family of 8B parameter open-weight, dense models trained from scratch by Essential AI.
74.5K
24
Updated 7 months ago
Ministral 3
Available to download
3B
3B
8B
8B
14B
14B
Ministral 3 series, available in three model sizes: 3B, 8B, and 14B parameters. Provides best of class cost-to-performance ratio.
582.7K
165
6
Updated 8 months ago
Qwen3 Next
Available to download
80B
Hybrid attention architecture, high-sparsity Mixture-of-Experts 80B model (active 3B).
58.3K
33
Updated 8 months ago
Olmo 3
Available to download
7B
7B
32B
Olmo 3 is a family of Open language models designed to enable the science of language models.
39.9K
36
3
Updated 8 months ago
olmOCR 2
Available to download
7B
The olmOCR 2 model is a Vision Language Model (VLM) from Allen AI.
81.6K
21
Updated 8 months ago
minimax-m2
Available to download
230B
MiniMax M2 is a 230B MoE (10B active) model built for coding and agentic workflows
17.6K
45
Updated 8 months ago
gpt-oss-safeguard
Available to download
20B
120B
gpt-oss-safeguard-20b and gpt-oss-safeguard-120b are open safety models from OpenAI, building on gpt-oss. Trained to help classify text content based on customizable policies.
7.2K
40
2
Updated 9 months ago
Qwen3-VL
Available to download
2B
4B
8B
30B
32B
Qwen's latest vision-language model. Includes comprehensive upgrades to visual perception, spatial reasoning, and image understanding.
859.4K
156
5
Updated 9 months ago
Granite 4.0
Available to download
3B
3B
7B
32B
Granite 4.0 language models are lightweight, state-of-the-art open models that natively support multilingual capabilities, coding tasks, RAG, tool use, and JSON output.
92.3K
62
4
Updated 9 months ago
seed-oss
Available to download
36B
Advanced reasoning model from ByteDance with flexible "thinking budget" control and ability to reflect on the length of its own reasoning
59.1K
23
Updated 9 months ago
Qwen3
Available to download
4B
4B
30B
30B
235B
235B
The latest version of the Qwen3 model family, featuring 4B, 30B, and 235B dense and MoE models, both thinking and non-thinking variants.
546.2K
195
6
Updated 9 months ago
gpt-oss
Available to download
20B
120B
OpenAI's first open source LLM. Comes in 2 sizes: 20B and 120B. Supports configurable reasoning effort (low, medium, high). Trained for tool use. Apache 2.0 licensed.
1.8M
397
2
Updated 9 months ago
Qwen3-Coder
Available to download
30B
480B
State-of-the-art, Mixture-of-Experts local coding model with native support for 256K context length. Available in 30B (3B active) and 480B (35B active) sizes.
498.4K
179
2
Updated 9 months ago
Ernie-4.5
Available to download
21B
Medium-size Mixture-of-Experts model from Baidu's new Ernie 4.5 line of foundation models.
27.8K
15
Updated 9 months ago
LFM2
Available to download
350M
700M
1.2B
LFM2 is a new generation of hybrid models developed by Liquid AI, specifically designed for edge AI and on-device deployment. It sets a new standard in terms of quality, speed, and memory efficiency.
78.8K
59
3
Updated 9 months ago
Devstral
Available to download
23.6B
24B
Devstral is a coding model from Mistral AI. It excels at using tools to explore codebases, editing multiple files and power software engineering agents.
85.6K
43
2
Updated 7 months ago
gemma-3n
Available to download
4.5B
6.9B
Gemma 3n is a generative AI model optimized for use in everyday devices, such as phones, laptops, and tablets.
238.9K
103
2
Updated 9 months ago
Mistral Small
Available to download
24B
Mistrall Small is a 'knowledge-dense' 24B multi-modal (image input) local model that supports up to 128 token context length.
86.9K
21
Updated 9 months ago
Magistral
Available to download
23.6B
24B
MistralAI's open-weight reasoning model. 24B dense transformer model supporting up to 128K token context window. The model is capable of long chains of reasoning traces before providing answers.
132.2K
56
2
Updated 9 months ago
mistral-nemo
Available to download
12B
General purpose dense transformer designed for multilingual use cases. Built in collaboration between MistralAI and NVIDIA.
64.6K
15
Updated 9 months ago
qwen2.5-vl
Available to download
3B
7B
32B
72B
Qwen2.5-VL is a performant vision-language model, capable of recognizing common objects and text. Supports context length of 128k tokens in a variety of human languages.
297.7K
35
4
Updated 9 months ago
gemma-3
Available to download
270M
1B
4B
12B
27B
State-of-the-art image + text input models from Google, built from the same research and tech used to create the Gemini models
1.5M
238
5
Updated 9 months ago
phi-4-reasoning
Available to download
3.8B
14.7B
14.7B
Phi-4-mini-reasoning is a lightweight open model built upon synthetic data with a focus on high-quality, reasoning dense data.
219.1K
64
3
Updated 9 months ago
phi-4
Available to download
3B
14B
phi-4 is a state-of-the-art open model built upon a blend of synthetic datasets, data from filtered public domain websites, and acquired academic books and Q&A datasets.
47K
22
2
Updated 9 months ago
Codestral
Available to download
22B
Mistral AI's latest coding model, Codestral can handle both instructions and code completions with ease in over 80 programming languages.
68.8K
36
Updated 9 months ago
Mistral
Available to download
7B
One of the most popular open-source LLMs, Mistral's 7B Instruct model's balance of speed, size, and performance makes it a great general-purpose daily driver.
154.3K
52
Updated 9 months ago
Qwen3 (1st Generation)
Available to download
4B
8B
14B
30B
32B
235B
The first batch of Qwen3 models (Qwen3-2504), a collection of dense and MoE models ranging from 4B to 235B. These are general purpose models that score highly on benchmarks.
674.9K
64
6
Updated 9 months ago
deepseek-r1
Available to download
7B
8B
8B
14B
32B
70B
Distilled version of the DeepSeek-R1-0528 model, created by continuing the post-training process on the Qwen3 8B Base model using Chain-of-Thought (CoT) from DeepSeek-R1-0528.
1M
307
6
Updated 9 months ago