kogai/laneformer-tp-small-2B-nemov2phase3-smoke-llama3ext-swachange-step-160 Text Generation • 2B • Updated 13 days ago • 552
kogai/laneformer-tp-small-2B-nemov2phase3-smoke-llama3ext-swachange-step-160 Text Generation • 2B • Updated 13 days ago • 552
kogai/laneformer-tp-small-2B-nemov2phase3-smoke-llama3ext-step-160 Text Generation • 2B • Updated 15 days ago • 351
kogai/laneformer-tp-small-2B-nemov2phase3-smoke-llama3ext-step-160 Text Generation • 2B • Updated 15 days ago • 351
view reply Hi @fayismahmood , Short answer: no.As our model requires a custom architecture, the GGUF file would not be enough to run our model in llama-cli. You can try quantization with bitsandbytes on the HF model though.
view article Article Kog Laneformer 2B: The Latency-First Model Behind Kog Inference Engine kogai • 27 days ago • 32
view article Article Kog Laneformer 2B: The Latency-First Model Behind Kog Inference Engine kogai • 27 days ago • 32