Hacker Newsnew | past | comments | ask | show | jobs | submit | fromlogin
Medusa: Framework for Accelerating LLM Generation with Multiple Decoding Heads (github.com/fasterdecoding)
5 points by PaulHoule on Dec 23, 2023 | past
Medusa: Simple Framework for Accelerating LLM Generation (github.com/fasterdecoding)
1 point by cmitsakis on Sept 11, 2023 | past

Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: