Popular repositories Loading
-
qwen-mtp-tuning-guide
qwen-mtp-tuning-guide PublicMeasured MTP (speculative decoding) tuning for Qwen 27B models on llama.cpp — per-GPU recommended n-max, p-min, KV cache, with full raw data
Shell 1
-
openai-edge-tts
openai-edge-tts PublicForked from travisvn/openai-edge-tts
Free, high-quality text-to-speech API endpoint to replace OpenAI, Azure, or ElevenLabs
Python
-
-
chatterbox-tts-api
chatterbox-tts-api PublicForked from travisvn/chatterbox-tts-api
Local, OpenAI-compatible text-to-speech (TTS) API using Chatterbox, enabling users to generate voice cloned speech anywhere the OpenAI API is used (e.g. Open WebUI, AnythingLLM, etc.)
Python
-
searxng
searxng PublicForked from searxng/searxng
SearXNG is a free internet metasearch engine which aggregates results from various search services and databases. Users are neither tracked nor profiled.
Python
-
qwen38-mtp
qwen38-mtp PublicForked from sudoingX/qwen38-mtp
One llama.cpp flag unlocks +33-39% decode speed for Qwen3.8-27B on consumer GPUs. The MTP head already ships inside your GGUF. Recipe, paired benchmarks, probe tool.
Python
If the problem persists, check the GitHub status page or contact support.