- 211comments
- 61comments
- 6comments
- 216comments
- 79comments
- 14comments
- 80comments
- —discuss
- 27comments
- 26comments
- 26comments
- 23comments
- 100comments
- 594comments
- 41comments
- 10comments
- —discuss
- 91comments
- 1comments
- 30comments
- 54comments
- 11comments
- 171comments
- 25comments
- 264comments
- 18comments
- 12comments
- 74comments
- 13comments
- 11comments
Fun update to this: Daniel Lemire added another optimization to make this even faster. https://github.com/jadidbourbaki/llama.cpp/pull/12
I’ll benchmark his change and add it to the article, crediting him for this improvement.
Btw, if anyone has experience with the open source community in general and llama.cpp in specific, I would greatly appreciate some advice. I’m facing a bit of an interpersonal issue that I really hope is resolved without any ill will. Here is the context:
https://www.reddit.com/r/LocalLLaMA/comments/1wr5ylm/comment...
Any advice for what I can do? Due to this, I cannot create a PR or issue in the llama.cpp repository. However, I am worried about bothering the maintainers on other channels in case it aggravates them further. Thank you for your help!