Model Families Overview
AJ STUDIOZ Cloud Infra hosts models from six major AI research labs, each with distinct strengths.Gemma (Google)
Lightweight, efficient models from Google DeepMind. Great for production workloads.
Qwen (Alibaba)
High-capability coding and multilingual models with massive scale.
Kimi (Moonshot)
Ultra-large models excelling at agentic and reasoning tasks.
DeepSeek
Top-tier open-source models for coding and analysis.
GLM (Zhipu AI)
Multilingual long-context models with strong Chinese language support.
Mistral & Others
Best-in-class European AI models with multilingual excellence.
Gemma Family (Google DeepMind)
Gemma models are open-weight, production-ready models from Google. They offer excellent quality-per-compute-cost ratio.
Best for: General assistant tasks, fast inference, cost-sensitive workloads.
Qwen Family (Alibaba Cloud)
Qwen models offer state-of-the-art performance especially in coding and mathematical reasoning.
Best for: Code generation, debugging, mathematical problems, vision tasks.
Kimi Family (Moonshot AI)
Kimi models are among the largest available, optimized for complex multi-step reasoning and agentic workflows.
Best for: Complex reasoning, research, multi-step agentic tasks, long context.
DeepSeek Family
DeepSeek produces highly capable open-source models that compete with GPT-4 class models.
Best for: Coding, data analysis, structured outputs, STEM tasks.
GLM Family (Zhipu AI)
GLM models excel at long-context understanding and multilingual tasks, especially Chinese.
Best for: Long documents, Chinese language, multilingual tasks.
MiniMax Family
MiniMax models offer strong multimodal capabilities.
Best for: Vision + language tasks, multimodal reasoning.
Other Models
Choosing the Right Model
For chat and general tasks
For chat and general tasks
Start with:
gemma3:27bGreat balance of quality and speed. If you need more capability, try deepseek-v3.2 or glm-5.For coding tasks
For coding tasks
Start with:
qwen3-coder:480b for best quality, or devstral-2:123b for a balance of speed and quality.For small, fast coding tasks: devstral-small-2:24b or gemma3:12b.For agentic / research tasks
For agentic / research tasks
Use:
kimi-k2:1t or kimi-k2-thinking for complex multi-step reasoning.Alternatively, cogito-2.1:671b for scientific domains.For fast/cheap inference
For fast/cheap inference
Use:
gemma3:4b, ministral-3:3b, or gemini-3-flash-preview.These models prioritize speed and cost over raw capability.For vision tasks
For vision tasks
Use:
qwen3-vl:235b-instruct or minimax-m2.5.