Faster prompt lookup drafting in llama.cpp
58 points by pptadversary 2 days ago | 9 comments
jadidbourbaki 22 hours ago
Fun update to this: Daniel Lemire added another optimization to make this even faster. https://github.com/jadidbourbaki/llama.cpp/pull/12
replyI’ll benchmark his change and add it to the article, crediting him for this improvement.
PicardsFlute 3 hours ago
So, instead of resolving it privately like you were asked, you come here and complain? Was the bot posts on Reddit not enough for you? Pro tip: There is a right way and a wrong way to approach these things and you are MOST certainly approaching this the wrong way. And before you go accusing me of anything: I am NOT associated with that project, but I DID see what happened. You are acting like a damned child and should be ashamed of yourself.
replyricardobeat 2 hours ago
Is there some missing context here? Where was the author asked to resolve issues in private? And what 'private' channels would be available to them?
replyfrazar0 56 minutes ago
I think they meant to reply to this comment
replyhttps://news.ycombinator.com/item?id=49859982#49863097
but commented on the post.
https://www.reddit.com/r/LocalLLaMA/comments/1wr5ylm/comment...
Any advice for what I can do? Due to this, I cannot create a PR or issue in the llama.cpp repository. However, I am worried about bothering the maintainers on other channels in case it aggravates them further. Thank you for your help!
Email that author (email can usually be found via GitHub, or contact via other private way they've shared somewhere) and explain the situation, don't lambast them publicly on social media or similar ways. If that doesn't work, do the same but for another maintainer. Don't spam all of them straight up, wait a week or something before contacting someone else.