Zhi Wei Tan
zwtan-cs
AI & ML interests
Efficient LLM inference, KV cache optimization, quantization, speculative decoding, model pruning
Recent Activity
liked a dataset about 22 hours ago
Boxoffice1280/Neurips2026_evaluating_accuracy_KV-cache_reuse_techniques liked a dataset about 22 hours ago
nishant-k/speculative-decoding-benchmark-results liked a dataset about 22 hours ago
weifanjiang/speculative_decoding_benchmarksOrganizations
None yet