qwen36
Here are 12 public repositories matching this topic...
Enables small-to-large self-hosted ai models to use local source code when running tool-calling agentic workloads. We actively data mine 20,900+ (2+ TB) popular github repos using large and small ai models to create reuseable: json, markdown and parquet files for local-first tool-calling models.
-
Updated
May 25, 2026 - Python
Les 5 cercles + 18 stratégies d'optimisation tokens pour Qwen3.6+ (OpenRouter). Promptor v3 Council Edition : audit multi-perspective optionnel (5 advisors + peer review). Reverse Prompt Engineering. Python/PowerShell/OpenCode.
-
Updated
May 29, 2026 - Python
A local LLM inference framework, hand-tuned for the Ryzen AI Max APU. 3467 tok/s prefill, 76.7 tok/s decode on Qwen3.6-35B-A3B-FP8.
-
Updated
Aug 7, 2026 - Python
Run Qwen3.6-35B-A3B full native 258K context on 12GB VRAM (llama.cpp): ncmoe cliff rule, q4_0 KV prefill fix, 4 tuned profiles with scripts
-
Updated
Aug 11, 2026 - PowerShell
Android port of Sonario for quickly summarizing YouTube videos, websites, documents, epubs & pasted text. Use qwen3.6-27b through Groq for fast free summaries, or run local LLMs for greater privacy.
-
Updated
Jul 19, 2026 - Kotlin
Qwen3.6 27B MTP on Modal H100 with llama.cpp and an OpenAI-compatible API
-
Updated
May 11, 2026 - Python
Qwen3.6-27B at 111.6 tok/s (c4) on one DGX Spark—with FR-Spec, an x16 INT8 LM head, and 4.2× faster repeated-prefix TTFT.
-
Updated
Aug 11, 2026 - Python
Optimized Qwen3.6-35B-A3B recipe for DGX Spark — MTP Q4_0, 75-82 t/s
-
Updated
Jul 9, 2026 - Shell
Improve this page
Add a description, image, and links to the qwen36 topic page so that developers can more easily learn about it.
Add this topic to your repo
To associate your repository with the qwen36 topic, visit your repo's landing page and select "manage topics."