back
user profile
anotherCodder
4karma·22submissions·September 6, 2024
recent activity (22 total)
comment
vllm on 4x 5090 is getting ~20 tok/s with mtp on (their own thread on the hf card). i had qwen3.8-27b up the day after release, one rtx pro 6000, 140 tok/s spec, 0.156s first token, full 262…
2 pts
comment
Hey HN - I built agnix because I kept losing time to the same class of bug: AI tool configs that are almost right but silently wrong. The trigger: I had a Claude Code skill named `Review-Code`. It nev…
comment
For users of Node.js, Java, Python and very soon Go -
Valkey-Glide:
https://github.com/valkey-io/valkey-glide
It will stay out of the hand of Redis, can promise that, and if you …
comment
Want to understand what dev's will appreciate and will lead decision makers to choose our client.
What will make you choose a client library over the other, and what will make consider refactorin…