Owen Reed
owenreed
AI & ML interests
Efficient LLM inference, KV cache optimization, quantization, speculative decoding, model pruning
Recent Activity
liked a model about 9 hours ago
saarvajanik/facebook-opt-6.7b-gqa-ub-16-best-for-KV-cache liked a model about 9 hours ago
position-specialist-speculative-decoding/llama3-8b-instruct-poss3 liked a model about 9 hours ago
KVCache-ai/DeepSeek-R1-GGML-FP8-HybridOrganizations
None yet