DeepSeek-R1-Distill-Llama-8B
Exceeds Qwen 3.6 27B performance, uncensored and NEO-Di-Matrix quants to bring all that power in quant form. Q4/IQ4s clock in at 94% of full precision (BF16), with Q6 at just under 98%. Even IQ2_M : 83% of BF16. 5 Metrics per quant, plus benchmarks.
This model is natively multimodal. The mmproj file is the vision encoder — you need it alongside the main GGUF to use image/video inputs. Load both files in llama.cpp, LM Studio, or any compatible runtime.
Qwen25-7B-Instruct
This repository contains model weights and configuration files for the post-trained model in the Hugging Face Transformers format. These artifacts are compatible with Hugging Face Transformers, vLLM, SGLang, KTransformers, etc.
openPangu-Ultra-MoE-718B-V1.1 是基于昇腾 NPU 训练的大规模混合专家语言模型,总参数量为718B,激活参数量为39B,同一个模型具备快思考和慢思考两种能力。 相较 [openPangu-Ultra-MoE-718B-V1.0] 版本,V1.1版本主要提升了Agent工具调用能力,降低了幻觉率,其他综合能力也进一步增强。
项目展示
查看全部项目 >- Qwen3-1.7Blike
- Drop-in Jinja templates that fix rendering errors, token waste, and missing features in the official Qwen chat templates. Works in LM Studio, llama.cpp, vLLM, MLX, oMLX, and any engine that supports HuggingFace Jinja templates.like
- This repository contains model weights and configuration files for the post-trained model in the Hugging Face Transformers format. These artifacts are compatible with Hugging Face Transformers, vLLM, SGLang, KTransformers, etc.like
- Qwen3-4Blike
- Qwen25-7B-Instructlike
- DeepSeek-R1-Distill-Qwen-7Blike
- Qwen3-8Blike
- Qwen2.5-Coder-7B-Instructlike
- Multi-angle camera control LoRA for Qwen-Image-Edit-2511 96 camera positions • Trained on 3000+ Gaussian Splatting renders • Built with fal.ailike
- bge-m3 embedding modellike