back
user profile
theanonymousone
7,171karma·1,094submissions·September 27, 2021
recent activity (1,094 total)
6 pts
comment
Thanks a lot. How about Q8 vs FP16/BF16? Have you checked them too?
comment
Have you seen the 8bit quantisation matter a lot? The "consensus" in r/LocalLlama is that up to 4 bits the loss is tolerable.
comment
My question as well. Isn't Tencent a very well-known company? Maybe the mystery is in the model itself?
2 pts
comment
This is a big deal when/if it's working, to me at least. Where can I contribute?
comment
Isn't this link a duplicate? Or I have déjà vu?
1 pts
comment
Hi. Is there a significance to that date or that commit? The commit doesn't look very special, and Git was apparently being used from early April already: https://en.wikipedia.org/…
comment
> Is it worth getting worse results for that reason? > accuracy ranging from 80.8% for Very Polite prompts to 84.8% for Very Rude prompts I can live with that, for now at least.
comment
I have always said please and thank you to LLMs, not to increase accuracy or because I'm stupid. I believe it is more about me than about the LLM, and this is anyway a habit I don't want to …
63 pts
comment
Yes, of course you can destroy it. But how far can you "improve", beyond decent "common sense" behaviour.
comment
Isn't caching a server-side thing? How does the agent affect it, significantly at least?