Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Docs
  • Enterprise
  • Pricing

  • Log In
  • Sign Up
Zhenhua Han's picture
1 2

Zhenhua Han

hzhua
·
  • hzhua

AI & ML interests

None yet

Organizations

None yet

authored a paper about 1 year ago

RetrievalAttention: Accelerating Long-Context LLM Inference via Vector Retrieval

Paper • 2409.10516 • Published Sep 16, 2024 • 43
authored 2 papers over 1 year ago

MInference 1.0: Accelerating Pre-filling for Long-Context LLMs via Dynamic Sparse Attention

Paper • 2407.02490 • Published Jul 2, 2024 • 27

Parrot: Efficient Serving of LLM-based Applications with Semantic Variable

Paper • 2405.19888 • Published May 30, 2024 • 7
Company
TOS Privacy About Jobs
Website
Models Datasets Spaces Pricing Docs