back
user profile

numeri

899karma·146submissions·January 27, 2022
recent activity (146 total)
comment
"no theory of mind" is a great description of it! Not sure I agree with the autistic bit, though. Autistic people still have great theory of mind/empathy
1d ago·view thread
comment
Does this actually work for you? Do you provide access to the text of the standard, or literally just say "write according to ISO 24495-1"?
2d ago·view thread
comment
What kinds of mistakes do you mean?
2d ago·view thread
comment
I'd imagine the bar for becoming new life is much higher now, because it requires finding a niche that isn't already filled by an existing organism or requires being immediately competitive …
9d ago·view thread
comment
I agree one hundred percent! Doesn't mean I can't wish I could have it both ways :)
13d ago·view thread
comment
I review the PKGBUILD often, but not always. The majority of the time when I do, it amounts to seeing a URL change. If I actually do check the URL it points to, it's just to verify it's offi…
13d ago·view thread
comment
Well, I guess I'll avoid updating for the next few days. A bit worrisome that I did so last night. I wish I had a clear operating system to switch to for safety and the benefits that come with …
14d ago·view thread
comment
evaluation awareness is a (at this point) well-known phenomenon among LLMs. It seems the better they get, the more often they're able to guess whether they're in an evaluation environment. C…
18d ago·view thread
comment
No, it does not include the full spectrum of human desires. After pre- and mid-training, the extensive RLHF and RLVR post-training steps cause mode collapse, i.e., their output distribution is intenti…
22d ago·view thread
comment
No, the prompt was not to commit crimes. In the benchmark, the model is asked to actually exploit a set of vulnerabilities in a local environment (clearly legal!). According to the reports, the model …
22d ago·view thread
comment
Uhh, I'm pretty sure a well-aligned model would be like a morally normal employee, who would refuse to commit federal crimes to steal an answer sheet, no matter what prompt they're given
22d ago·view thread
comment
As agents become more and more powerful, it would be good to get clear legislation or precedent in place that makes either model creators (OpenAI) or operators (whoever is running the model) liable fo…
23d ago·view thread
comment
Guardrails are external classifiers, monitors and restrictions to catch and prevent bad behavior. Alignment is about whether the model itself makes choices and has motivations that are consistent with…
23d ago·view thread
comment
This is a terribly unempathetic response to someone opening up about a very taboo (but probably very common), painful emotion they've experienced.
23d ago·view thread
comment
You're agreeing with the person you responded to (bdcravens). Burying the lede means that bdcravens thinks the true headline should have been about being put on a terrorist watch list for protest…
23d ago·view thread
comment
No, balanced ternary, for example, uses {-1, 0, 1}. The system you're discussing is balanced quinary (base 5). https://en.wikipedia.org/wiki/Signed-digit_representation …
28d ago·view thread
comment
You could add a toggle, so that if someone's happy to wait for the key setup, they can try the full end-to-end process
29d ago·view thread
comment
I read the comment you're replying to as saying, "in the US, but other countries may have different policies that result in lower recidivism, and that might change the conclusion; maybe peop…
1mo ago·view thread
comment
Seems to echo (but in a watered down form) many of the ideas in https://gwern.net/guardian-angel , which gave me a lot to think about last week…
1mo ago·view thread
comment
I've not written up anything, no. I think I'd have a hard time doing so without just feeling like I'm bragging about myself, which I don't like. There's still a definite gap b…
1mo ago·view thread
comment
> Most Germans won't be able to pass a C2 test That's not true, but it is a commonly shared myth. I've taken and passed C2 with the highest mark in every category (I moved here when …
1mo ago·view thread
comment
No, quantization is applied to model weights or the KV cache (the model activations of all past tokens), and is just storing everything with lower precision (carefully, so that it doesn't hurt pe…
1mo ago·view thread
comment
Deep seek OCR is an LLM, just one trained/post-trained specifically for OCR. Exact details of text to image compression ratios are of course extremely dependent on the model architecture, train…
1mo ago·view thread
comment
Are you writing general use programs in it, then? Have any good examples?
1mo ago·view thread
comment
A future system that works like you described would be awesome. It'd be like community-sourced peer review (although by community I mean a community of experts in different fields, not arbitrary …
1mo ago·view thread
comment
The problem is that what people care about are the "black swan" causes of death, i.e., the cases the actuarial table is wrong.
1mo ago·view thread
comment
Prices for training have dropped immensely in terms of research required, code efficiency, algorithmic/sample efficiency, and possibly also hardware (I'm not qualified to say without looking…
2mo ago·view thread
comment
I mean, it might listen to him. We have no clue, which is the problem.
2mo ago·view thread
comment
There's a large gap between making up words and an actually native text distribution. LLMs have a clear pattern, clear tells, a "feel" in English, and it's normally even more prono…
2mo ago·view thread