Models
AI models available through the SCX.ai API. From language understanding to audio processing.
Available in Australia
11Models processed in Australian data centers. Models not listed here are processed outside Australia — see each model's Regions column.
| Model | Type | Provider | Regions |
|---|---|---|---|
| Qwen3-32B | language | Alibaba | AU |
| gemma-4-31B-it | language | USAU | |
| Llama-4-Maverick-17B-128E-Instruct | language | Meta | AU |
| MiniMax-M2.5 | language | MiniMax | DEAU |
| MiniMax-M2.7 | language | MiniMax | DEUSAU |
| gpt-oss-120b | language | OpenAI | DEUSAU |
| coder | language | scx.ai | AU |
| MAGPiE | language | scx.ai | USAU |
| E5-Mistral-7B-Instruct | embedding | Mistral | AU |
| scx-stt | audio | SCX | AU |
| Whisper-Large-v3 | audio | OpenAI | AU |
Language Models
23Alibaba
Qwen3-32B
33k ctx32B dense model matching Qwen2.5-72B performance, supporting 119 languages with thinking mode toggle
ReasoningTools
Qwen3.8
1M ctx2.4T sparse MoE (~95B active) flagship from Alibaba — text + image in, toggleable thinking with a 262K reasoning budget, agentic tool use, 1M context
ReasoningVisionTools
Qwen3.8-Max
1M ctx2.4T sparse MoE (~95B active) flagship from Alibaba — text + image in, toggleable thinking with a 262K reasoning budget, agentic tool use, 1M context
ReasoningVisionTools
DeepSeek
DeepSeek-V3.1
131k ctx671B MoE model (37B active) with hybrid thinking/non-thinking modes, tool use, and 128k context
ReasoningTools
DeepSeek-V4-flash
1M ctx284B MoE (13B active) fast tier of DeepSeek V4 with hybrid CSA/HCA sparse attention, toggleable deep reasoning, tool use, and 1M-token context
ReasoningTools
DeepSeek-V4-pro
1M ctx1.6T MoE model (49B active) with hybrid CSA/HCA sparse attention, toggleable thinking modes, tool use, and 1M context
ReasoningTools
DeepSeek-V4.1-flash
1M ctx552B multimodal MoE (8B active on prefill, 16B on decode) — DeepSeek's first Causal Encoder-Decoder model, pairing compressed sparse attention with FP4 KV caching for input-heavy agentic work: native text + image in, continuously controllable reasoning effort, tool use, and 1M context
ReasoningVisionTools
gemma-4-31B-it
131k ctxGoogle Gemma 4 31B dense instruction-tuned multimodal model (text + image in) with a toggleable thinking mode, native tool use, 128k context, and 140+ language coverage
ReasoningVisionTools
Meta
Llama-4-Maverick-17B-128E-Instruct
131k ctx400B MoE (17B active, 128 experts) with native multimodal early fusion supporting 12 languages and images
VisionTools
Meta-Llama-3.3-70B-Instruct
131k ctx70B instruction-tuned model delivering 405B-class text performance with 128k context and GQA
Tools
MiniMax
MiniMax-M2.5
197k ctx456B MoE model (27B active, 256 experts) from MiniMax with SOTA coding (80.2% SWE-Bench) and agentic tool use
ReasoningTools
MiniMax-M2.7
192k ctx230B sparse MoE (10B active, 256 experts, 8 per token) from MiniMax — agentic workflows, 200K context
ReasoningTools
MiniMax-M3
1M ctx428B MoE (23B active, 128 experts, 4 per token) from MiniMax — MiniMax Sparse Attention for 1M context, long-horizon agentic coding, and thinking that can be turned off; text + image in
ReasoningVisionTools
Moonshot AI
Kimi-K2.7-Code
262k ctxCoding-focused agentic model from Moonshot AI built on Kimi K2.6 (DeepSeek-V3-style MLA, 384 routed experts plus one always-active expert) — always-on reasoning tuned for long-horizon software engineering, tool use, and 256K context
ReasoningVisionTools
Kimi-K3
1M ctx2.8T-parameter Stable LatentMoE (16 of 896 experts active) from Moonshot AI with Kimi Delta Attention — native text/image/video understanding, toggleable reasoning, tool use, and 1M-token context
ReasoningVisionTools
OpenAI
gpt-oss-120b
131k ctx117B open-weight MoE (5.1B active) from OpenAI achieving near o4-mini reasoning
ReasoningTools
scx.ai
coder
197k ctxHigh performance coding assistant with reasoning, optimized for algorithms, debugging, and code review
ReasoningTools
MAGPiE
131k ctx117B MoE model from scx.ai achieving near o4-mini reasoning
ReasoningTools
Z.ai
GLM-5.2
1M ctx753B sparse MoE (~40B active) from Z.ai with IndexShare sparse attention — long-horizon agentic coding, toggleable High/Max thinking modes, 1M context
ReasoningTools
GLM-5.2-Fast
1M ctxLatency-optimised preview of GLM-5.2 (753B sparse MoE, ~40B active) from Z.ai — same IndexShare sparse attention and toggleable thinking modes, tuned for faster time-to-first-token
ReasoningTools
GLM-5.3
1M ctx743B sparse MoE from Z.ai on the GLM-5.2 base with scaled post-training — long-horizon agentic coding, always-on thinking with low/high/max effort, 1M context
ReasoningTools
GLM-5.3-Fast
1M ctxFast variant of Z.ai's GLM-5.3 with a 1M context window, 128K maximum output, always-on reasoning, and agentic tool use
ReasoningTools
GLM-5.3-Flash
1M ctxNative multimodal model (text + image in) from Z.ai for efficient coding and long-horizon agent tasks — hybrid sparse and linear attention for accurate long-context behaviour, always-on thinking, 1M context
ReasoningVisionTools
Embedding Models
1| Model | Provider | Dimensions | Context |
|---|---|---|---|
| E5-Mistral-7B-Instruct | Mistral | 4,096 | 33k |
Audio Models
2| Model | Provider | Capabilities |
|---|---|---|
| scx-stt | SCX | Transcription |
| Whisper-Large-v3 | OpenAI | Transcription |