AX Qwen3 Coder Next MLX 4-bit
0
5 commits
1 linked in READMEs
updated Jul 20, 2026
Parameter count: approximately 79.67B logical parameters (80B total, 3B active per token).
4-bitis the quantization precision, not a 4B model-size claim.
This is a revision-pinned, transparent mirror of
mlx-community/Qwen3-Coder-Next-4bit
at commit 7b9321eabb85ce79625cac3f61ea691e4ea984b5.
The model weights, index, configuration, tokenizer, chat template, generation configuration, and tool-parser files are byte-identical to that upstream revision. AutomatosX did not fine-tune, merge, re-quantize, or otherwise alter the model artifacts. We add only mirror documentation, a copy of the declared Apache 2.0 license, and machine-readable provenance.
hf download AutomatosX/AX-Qwen3-Coder-Next-MLX-4bit \
--local-dir ./AX-Qwen3-Coder-Next-MLX-4bit
pip install mlx-lm
mlx_lm.generate \
--model AutomatosX/AX-Qwen3-Coder-Next-MLX-4bit \
--prompt "Write a Python function that merges two sorted lists."
Applications should apply the included chat template for conversational or tool-using prompts.
You can also serve the downloaded model through the OpenAI-compatible API in AX Engine:
ax-engine serve ./AX-Qwen3-Coder-Next-MLX-4bit --port 31418
This release is meant to group a required upstream artifact under the
AutomatosX catalog, not to claim a new conversion. UPSTREAM_README.md
preserves the original upstream model card. ax_provenance.json pins the
source commit and records SHA-256 values and sizes for every mirrored artifact.
The local Hugging Face cache contained an AX-generated model-manifest.json;
it is not present in the pinned upstream repository and was deliberately not
published here.
This is a standard direct-decoding model. It does not contain an MTP head and no MTP conversion was applied.
The upstream model metadata declares Apache License 2.0. See LICENSE, the
upstream model card, and the base-model card for limitations and responsible-use
guidance.
5 commits
AX Qwen3 Coder Next MLX 4-bit
0
5 commits
1 linked in READMEs
updated Jul 20, 2026
Parameter count: approximately 79.67B logical parameters (80B total, 3B active per token).
4-bitis the quantization precision, not a 4B model-size claim.
This is a revision-pinned, transparent mirror of
mlx-community/Qwen3-Coder-Next-4bit
at commit 7b9321eabb85ce79625cac3f61ea691e4ea984b5.
The model weights, index, configuration, tokenizer, chat template, generation configuration, and tool-parser files are byte-identical to that upstream revision. AutomatosX did not fine-tune, merge, re-quantize, or otherwise alter the model artifacts. We add only mirror documentation, a copy of the declared Apache 2.0 license, and machine-readable provenance.
hf download AutomatosX/AX-Qwen3-Coder-Next-MLX-4bit \
--local-dir ./AX-Qwen3-Coder-Next-MLX-4bit
pip install mlx-lm
mlx_lm.generate \
--model AutomatosX/AX-Qwen3-Coder-Next-MLX-4bit \
--prompt "Write a Python function that merges two sorted lists."
Applications should apply the included chat template for conversational or tool-using prompts.
You can also serve the downloaded model through the OpenAI-compatible API in AX Engine:
ax-engine serve ./AX-Qwen3-Coder-Next-MLX-4bit --port 31418
This release is meant to group a required upstream artifact under the
AutomatosX catalog, not to claim a new conversion. UPSTREAM_README.md
preserves the original upstream model card. ax_provenance.json pins the
source commit and records SHA-256 values and sizes for every mirrored artifact.
The local Hugging Face cache contained an AX-generated model-manifest.json;
it is not present in the pinned upstream repository and was deliberately not
published here.
This is a standard direct-decoding model. It does not contain an MTP head and no MTP conversion was applied.
The upstream model metadata declares Apache License 2.0. See LICENSE, the
upstream model card, and the base-model card for limitations and responsible-use
guidance.
5 commits