Custom GGUF quants of Metaβs Llama-3.2-Instruct's finetunes, where the Output Tensors are quantized to Q8_0 or F32 and the Embeddings are kept @F32
Joseph
Joseph717171
AI & ML interests
None yet
Recent Activity
liked a model about 1 month ago
google/gemma-4-12B-it-qat-q4_0-unquantized liked a model about 1 month ago
unsloth/Qwen3.8-27B-GGUF liked a model about 1 month ago
google/embeddinggemma-300m-qat-q4_0-unquantized