Youssef Amrani
youssefam
AI & ML interests
Efficient LLM inference, KV cache optimization, quantization, speculative decoding, model pruning
Recent Activity
liked a model about 19 hours ago
KVCache-ai/Qwen3-30BA3B-GGUF upvoted a paper about 19 hours ago
Fathom: Per-Query Read Depth for Sparse Decoding over Offloaded KV Caches upvoted a paper about 19 hours ago
ScienceIDE: Turning World's Scientific Codebase into Agent Learnable EnvironmentsOrganizations
None yet