Custom GGUF quants of Meta’s Llama-3.2-Instruct's finetunes, where the Output Tensors are quantized to Q8_0 or F32 and the Embeddings are kept @F32
Joseph
Joseph717171
AI & ML interests
None yet
Recent Activity
liked a model about 9 hours ago
mlx-community/gemma-4-E4B-it-qat-4bit liked a model 5 days ago
LiquidAI/LFM2.5-8B-A1B