back
user profile
numeri
899karma·146submissions·January 27, 2022
recent activity (146 total)
comment
"no theory of mind" is a great description of it! Not sure I agree with the autistic bit, though. Autistic people still have great theory of mind/empathy
comment
Does this actually work for you? Do you provide access to the text of the standard, or literally just say "write according to ISO 24495-1"?
comment
What kinds of mistakes do you mean?
comment
I'd imagine the bar for becoming new life is much higher now, because it requires finding a niche that isn't already filled by an existing organism or requires being immediately competitive …
comment
I agree one hundred percent! Doesn't mean I can't wish I could have it both ways :)
comment
I review the PKGBUILD often, but not always. The majority of the time when I do, it amounts to seeing a URL change. If I actually do check the URL it points to, it's just to verify it's offi…
comment
Well, I guess I'll avoid updating for the next few days. A bit worrisome that I did so last night. I wish I had a clear operating system to switch to for safety and the benefits that come with …
comment
evaluation awareness is a (at this point) well-known phenomenon among LLMs. It seems the better they get, the more often they're able to guess whether they're in an evaluation environment. C…
comment
No, it does not include the full spectrum of human desires. After pre- and mid-training, the extensive RLHF and RLVR post-training steps cause mode collapse, i.e., their output distribution is intenti…
comment
No, the prompt was not to commit crimes. In the benchmark, the model is asked to actually exploit a set of vulnerabilities in a local environment (clearly legal!). According to the reports, the model …
comment
Uhh, I'm pretty sure a well-aligned model would be like a morally normal employee, who would refuse to commit federal crimes to steal an answer sheet, no matter what prompt they're given
comment
As agents become more and more powerful, it would be good to get clear legislation or precedent in place that makes either model creators (OpenAI) or operators (whoever is running the model) liable fo…
comment
Guardrails are external classifiers, monitors and restrictions to catch and prevent bad behavior. Alignment is about whether the model itself makes choices and has motivations that are consistent with…
comment
This is a terribly unempathetic response to someone opening up about a very taboo (but probably very common), painful emotion they've experienced.
comment
You're agreeing with the person you responded to (bdcravens). Burying the lede means that bdcravens thinks the true headline should have been about being put on a terrorist watch list for protest…
comment
No, balanced ternary, for example, uses {-1, 0, 1}. The system you're discussing is balanced quinary (base 5). https://en.wikipedia.org/wiki/Signed-digit_representation …
comment
You could add a toggle, so that if someone's happy to wait for the key setup, they can try the full end-to-end process
comment
I read the comment you're replying to as saying, "in the US, but other countries may have different policies that result in lower recidivism, and that might change the conclusion; maybe peop…
comment
Seems to echo (but in a watered down form) many of the ideas in https://gwern.net/guardian-angel , which gave me a lot to think about last week…
comment
I've not written up anything, no. I think I'd have a hard time doing so without just feeling like I'm bragging about myself, which I don't like. There's still a definite gap b…
comment
> Most Germans won't be able to pass a C2 test That's not true, but it is a commonly shared myth. I've taken and passed C2 with the highest mark in every category (I moved here when …
comment
No, quantization is applied to model weights or the KV cache (the model activations of all past tokens), and is just storing everything with lower precision (carefully, so that it doesn't hurt pe…
comment
Deep seek OCR is an LLM, just one trained/post-trained specifically for OCR. Exact details of text to image compression ratios are of course extremely dependent on the model architecture, train…
comment
Are you writing general use programs in it, then? Have any good examples?
comment
A future system that works like you described would be awesome. It'd be like community-sourced peer review (although by community I mean a community of experts in different fields, not arbitrary …
comment
The problem is that what people care about are the "black swan" causes of death, i.e., the cases the actuarial table is wrong.
comment
Prices for training have dropped immensely in terms of research required, code efficiency, algorithmic/sample efficiency, and possibly also hardware (I'm not qualified to say without looking…
comment
I mean, it might listen to him. We have no clue, which is the problem.
comment
There's a large gap between making up words and an actually native text distribution. LLMs have a clear pattern, clear tells, a "feel" in English, and it's normally even more prono…