jjang-ai / vmlx Star 886 Code Issues Pull requests vMLX - Use MLX models easily - JANGQ (GGUF for MLX) - Not dependant on mlx_vlm macbook persistent-memory mlx openai-api llm lmstudio anthropic-api mcp-server kvcache-optimization kvcache-compression openclaw kvcache-reuse openclaw-agent prefix-cache mlxllm mlxstudio vmlx omlx omlx-alternative Updated Oct 7, 2026 Python
jjang-ai / exploitbot Star 17 Code Issues Pull requests No bs theatricals. Real automated pentesting. Mac only. api caching engine hacking openai pentesting vulnerability-detection vulnerability-scanners mlx uncensored hacking-tools pentesting-tools llm llms anthropic automated-pentesting mlxllm mlxstudio vmlx turboquant Updated Jul 15, 2026 Python
BTankut / glm-5.2-4x-dgx-spark Star 1 Code Issues Pull requests GLM-5.2 744B on 4x NVIDIA DGX Spark (GB10): measured vLLM TP=4 serving recipe, GB10 operational notes, and a single-API tool-plane backend runbook mlx apple-silicon local-llm vmlx glm-5-2 Updated Aug 8, 2026
abedshaaban / hiring-agent Star 0 Code Issues Pull requests AI resume evaluation pipeline with local LLM providers, GitHub enrichment, batch scoring, and explainable outputs. python ai gemini hiring resume-parser llm ollama vmlx Updated Jul 5, 2026 Python
gajsanders / local-llm-backend-bench Star 0 Code Issues Pull requests Reproducible local LLM inference backend benchmarks for Apple Silicon macos benchmarking inference mlx on-device-ai apple-silicon local-inference local-llm qwen lm-studio qwen3-coder llm-benchmark vmlx Updated Oct 2, 2026 Shell