back
user profile
helloplanets
2,648karma·541submissions·September 15, 2020
recent activity (541 total)
comment
Not at all. This looks just like someone trying to make a quick buck, hyping their product up with bad benchmarks.
comment
> Both conditions used GitHub Copilot (Claude Sonnet 4.5 or Haiku 4.5, depending on study) running in VS Code within isolated Docker containers. The only difference was Mouse tool availability. ( h…
comment
> These are my personal beliefs, not those of Nym. Why are you posting this on your company's site, littered with ads for the company's product? Post it on a personal blog, or just say th…
comment
Doesn't make sense to fixate on LLMs and not the actual Transformer/attention foundation. The Transformer/attention architecture is the breakthrough, not LLMs. Especially the RLHF chat …
comment
Tangential, but I'm pretty sad about EU having absolutely nothing in the actual SotA LLM market. Especially given the recent events of US completely restricting the actual SotA models. Has this b…
comment
If that ain't getting steganographically tagged...
comment
Dario's been openly talking how worried he is about China and labs getting synthetic training data off their models, for years. Most recently in relation to "Mythos level" capabilities…
comment
The issue is that using Claude Code is an easy compromise for most to make, when you get to use the models 10x cheaper than through API pricing with a custom harness. The cheap tokens are the product.…
comment
Slide number 55 is a beauty.
comment
Which great writers are you thinking of here? True outsider art is very rare afaik.
comment
I actually thought about that while writing the original comment as well. For Emma, Forever Ago is one of my all time favorite albums, good example of raw emotion with no need for any bells or whistle…
comment
I don't believe this is how great music usually comes about, not even Techno. It's missing the other essential piece. Being influenced by and completely immersed in a niche of other brillian…
comment
Deep Research has been using the Orchestrator -> Subagents -> Synthesizer loop since the beginning. It's just strange that they'd put a loop benchmark next to actual model benchmarks. …
comment
OpenAI also announced two days ago that they're starting to make Cerebras style chips themselves [0], will be interesting to see how fast SotA model inference will be by the end of the year. [0]:…
comment
Less so in EU than in US.
comment
I'm not saying your actual point couldn't be valid or fully defensible, just to be clear. My view is that there are people capable of vetting LLM generated code, and people who are not capab…
comment
> Slop means anything produced en masse with complete disregard for truth, accuracy, or usefulness. This doesn't match at all with what the author described in the article. > Anyone trying …
comment
That is not what slop means, though. You're redefining the meaning of the word to suit your view. Why do that? You can just say that LLM generated content is not up to par, or acceptable, ever.
comment
Maxed out 2019 Mac Pro was $50k+. The wheels on that thing were $400. This is a bargain compared to that.
comment
Anthropic already explicitly communicated that they'll store and check all the data from Bedrock or any platform, even if you've selected zero data retention, if using Mythos class models.…
comment
The trope of "ABC players" is tired at this point. To be honest, even talking about "players" is kind of square in my opinion. In a similar way as the saying "don't hate …