
GLM-5.2
glm-5-2Parameters: 1T total (32B activated)Quantization: INT4 weight-only (native quantization-aware release from Moonshot AI)Context: 256K tokensStrengths: Long-horizon coding, image and video understanding, generates code and interfaces from visual inputs, large-scale agent orchestration, strong tool callingStructured Outputs: Structured response formatting supportBest for: Agentic coding, design-to-code workflows, multimodal applications, and long-running tool-based tasks that benefit from strong reasoningModel weights: moonshotai/Kimi-K2.6Configuration repo: tinfoilsh/confidential-kimi-k2-6
Vision + Language: Supports text, image, and video inputs with native reasoning and tool calling for agentic workflows.

Gemma 4 31B
gemma4-31bVision + Language: Processes text and image inputs. Features step-by-step reasoning with configurable thinking mode.
Parameters: 117B (5.1B active)Quantization: MXFP4 (native MXFP4 MoE weights, as released by OpenAI)Context: 131K tokensStrengths: Configurable reasoning effort levels, full chain-of-thought access, built-in capabilities including function calling, web browsing, and Python code executionStructured Outputs: Structured response formatting supportBest for: Production use cases requiring configurable reasoning and tool useModel weights: openai/gpt-oss-120bConfiguration repo: tinfoilsh/confidential-gpt-oss-120b

Llama 3.3 70B
llama3-3-70b




