Owen Reed
owenreed
AI & ML interests
Efficient LLM inference, KV cache optimization, quantization, speculative decoding, model pruning
Recent Activity
liked a model 1 day ago
SpeculativeDecoding/doctorboom-qwen2.5-coder-7b-lora liked a model 1 day ago
CedricHwang/qwen2.5-0.5b-modelopt-pruning-gradnas liked a dataset 2 days ago
steven0226/speculative-decoding-bench-rtx4090Organizations
None yet