@arminretro
Joined September 2026
Projects
DeepSeek V4 Flash is a Mixture-of-Experts model designed for efficient reasoning, coding, and agentic workflows
MODEL
@arminretro
Joined September 2026
Projects
DeepSeek V4 Flash is a Mixture-of-Experts model designed for efficient reasoning, coding, and agentic workflows
MODEL
Qwen Coder Next is an 80B MoE with 3B active parameters designed for coding agents and local development. Excels at long-horizon reasoning, complex tool usage, and recovery from execution failures.
MODEL
GLM 4.7 Flash is a 30B A3B MoE model form Z.ai. It supports a context length of 128k tokens and achieves strong performance on coding benchmarks among models of similar scale.
MODEL
Qwen Coder Next is an 80B MoE with 3B active parameters designed for coding agents and local development. Excels at long-horizon reasoning, complex tool usage, and recovery from execution failures.
MODEL
GLM 4.7 Flash is a 30B A3B MoE model form Z.ai. It supports a context length of 128k tokens and achieves strong performance on coding benchmarks among models of similar scale.
MODEL