Skip to main content

Model Families Overview

AJ STUDIOZ Cloud Infra hosts models from six major AI research labs, each with distinct strengths.

Gemma (Google)

Lightweight, efficient models from Google DeepMind. Great for production workloads.

Qwen (Alibaba)

High-capability coding and multilingual models with massive scale.

Kimi (Moonshot)

Ultra-large models excelling at agentic and reasoning tasks.

DeepSeek

Top-tier open-source models for coding and analysis.

GLM (Zhipu AI)

Multilingual long-context models with strong Chinese language support.

Mistral & Others

Best-in-class European AI models with multilingual excellence.

Gemma Family (Google DeepMind)

Gemma models are open-weight, production-ready models from Google. They offer excellent quality-per-compute-cost ratio. Best for: General assistant tasks, fast inference, cost-sensitive workloads.

Qwen Family (Alibaba Cloud)

Qwen models offer state-of-the-art performance especially in coding and mathematical reasoning. Best for: Code generation, debugging, mathematical problems, vision tasks.

Kimi Family (Moonshot AI)

Kimi models are among the largest available, optimized for complex multi-step reasoning and agentic workflows. Best for: Complex reasoning, research, multi-step agentic tasks, long context.

DeepSeek Family

DeepSeek produces highly capable open-source models that compete with GPT-4 class models. Best for: Coding, data analysis, structured outputs, STEM tasks.

GLM Family (Zhipu AI)

GLM models excel at long-context understanding and multilingual tasks, especially Chinese. Best for: Long documents, Chinese language, multilingual tasks.

MiniMax Family

MiniMax models offer strong multimodal capabilities. Best for: Vision + language tasks, multimodal reasoning.

Other Models


Choosing the Right Model

Start with: gemma3:27bGreat balance of quality and speed. If you need more capability, try deepseek-v3.2 or glm-5.
Start with: qwen3-coder:480b for best quality, or devstral-2:123b for a balance of speed and quality.For small, fast coding tasks: devstral-small-2:24b or gemma3:12b.
Use: kimi-k2:1t or kimi-k2-thinking for complex multi-step reasoning.Alternatively, cogito-2.1:671b for scientific domains.
Use: gemma3:4b, ministral-3:3b, or gemini-3-flash-preview.These models prioritize speed and cost over raw capability.
Use: qwen3-vl:235b-instruct or minimax-m2.5.