Jasai/qwen36-cpu-inference
0
Qwen3.6-35B-A3B OpenAI-Compatible API
llama.cpp server hosting Qwen3.6-35B-A3B-UD-Q2_K_XL.gguf with OpenAI-compatible /v1/chat/completions endpoint.
Usage
curl -X POST https://YOUR_SPACE.hf.space/v1/chat/completions \
-H "Content-Type: application/json" \
-d '{
"model": "qwen3.6-35b",
"messages": [{"role": "user", "content": "Hello!"}],
"temperature": 0.7,
"max_tokens": 256
}'