How to use from
OpenClaw
Start the MLX server
# Install MLX LM:
uv tool install mlx-lm
# Start a local OpenAI-compatible server:
mlx_lm.server --model "dawncr0w/Qwen3.6-27B-uncensored-heretic-v2-Text-Only-oQ8-MLX"
Configure OpenClaw
# Install OpenClaw:
npm install -g openclaw@latest
# Register the local server and set it as the default model:
openclaw onboard --non-interactive --mode local \
  --auth-choice custom-api-key \
  --custom-base-url http://127.0.0.1:8080/v1 \
  --custom-model-id "dawncr0w/Qwen3.6-27B-uncensored-heretic-v2-Text-Only-oQ8-MLX" \
  --custom-provider-id mlx-lm \
  --custom-compatibility openai \
  --custom-text-input \
  --accept-risk \
  --skip-health
Run OpenClaw
openclaw agent --local --agent main --message "Hello from Hugging Face"
Quick Links

Qwen3.6-27B-uncensored-heretic-v2-Text-Only-oQ8-MLX

This repository contains an oMLX oQ8 mixed-precision MLX quantization of llmfan46/Qwen3.6-27B-uncensored-heretic-v2-Native-MTP-Preserved.

This build tracks the BF16 artifact lineage from llmfan46/Qwen3.6-27B-uncensored-heretic-v2-Native-MTP-Preserved-GGUF. oMLX oQ quantization operates on MLX/safetensors checkpoints rather than GGUF files, so this build uses the corresponding BF16 safetensors checkpoint from llmfan46/Qwen3.6-27B-uncensored-heretic-v2-Native-MTP-Preserved and excludes the existing GGUF quantizations.

Variant

  • Quantization: oQ8
  • Variant: Text Only
  • MTP tensors: stripped
  • Text-only: true
  • Approximate target density: 8.63 bpw.
  • The vision/audio-side weights are removed from the output config and shards.

Usage

omlx serve dawncr0w/Qwen3.6-27B-uncensored-heretic-v2-Text-Only-oQ8-MLX

or with mlx-lm:

from mlx_lm import generate, load

model, tokenizer = load("dawncr0w/Qwen3.6-27B-uncensored-heretic-v2-Text-Only-oQ8-MLX")
print(generate(model, tokenizer, "Hello", max_tokens=64))

Validation

Local validation completed with the bundled oMLX runtime:

loader: mlx_lm.load
generation smoke test: passed
prompt: Hello
max tokens: 4
peak memory: 26.798 GB

Source

Upstream model card license tag: apache-2.0.

Downloads last month
52
Safetensors
Model size
27B params
Tensor type
U32
·
BF16
·
MLX
Hardware compatibility
Log In to add your hardware

8-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for dawncr0w/Qwen3.6-27B-uncensored-heretic-v2-Text-Only-oQ8-MLX

Collection including dawncr0w/Qwen3.6-27B-uncensored-heretic-v2-Text-Only-oQ8-MLX