back
user profile

helloplanets

2,648karma·541submissions·September 15, 2020
recent activity (541 total)
comment
Not at all. This looks just like someone trying to make a quick buck, hyping their product up with bad benchmarks.
1mo ago·view thread
comment
> Both conditions used GitHub Copilot (Claude Sonnet 4.5 or Haiku 4.5, depending on study) running in VS Code within isolated Docker containers. The only difference was Mouse tool availability. ( h…
1mo ago·view thread
comment
> These are my personal beliefs, not those of Nym. Why are you posting this on your company's site, littered with ads for the company's product? Post it on a personal blog, or just say th…
1mo ago·view thread
comment
Doesn't make sense to fixate on LLMs and not the actual Transformer/attention foundation. The Transformer/attention architecture is the breakthrough, not LLMs. Especially the RLHF chat …
1mo ago·view thread
comment
Tangential, but I'm pretty sad about EU having absolutely nothing in the actual SotA LLM market. Especially given the recent events of US completely restricting the actual SotA models. Has this b…
1mo ago·view thread
comment
If that ain't getting steganographically tagged...
1mo ago·view thread
comment
Dario's been openly talking how worried he is about China and labs getting synthetic training data off their models, for years. Most recently in relation to "Mythos level" capabilities…
1mo ago·view thread
comment
The issue is that using Claude Code is an easy compromise for most to make, when you get to use the models 10x cheaper than through API pricing with a custom harness. The cheap tokens are the product.…
1mo ago·view thread
comment
Slide number 55 is a beauty.
1mo ago·view thread
comment
Which great writers are you thinking of here? True outsider art is very rare afaik.
1mo ago·view thread
comment
I actually thought about that while writing the original comment as well. For Emma, Forever Ago is one of my all time favorite albums, good example of raw emotion with no need for any bells or whistle…
1mo ago·view thread
comment
I don't believe this is how great music usually comes about, not even Techno. It's missing the other essential piece. Being influenced by and completely immersed in a niche of other brillian…
1mo ago·view thread
comment
Deep Research has been using the Orchestrator -> Subagents -> Synthesizer loop since the beginning. It's just strange that they'd put a loop benchmark next to actual model benchmarks. …
1mo ago·view thread
comment
OpenAI also announced two days ago that they're starting to make Cerebras style chips themselves [0], will be interesting to see how fast SotA model inference will be by the end of the year. [0]:…
1mo ago·view thread
comment
Less so in EU than in US.
1mo ago·view thread
comment
I'm not saying your actual point couldn't be valid or fully defensible, just to be clear. My view is that there are people capable of vetting LLM generated code, and people who are not capab…
1mo ago·view thread
comment
> Slop means anything produced en masse with complete disregard for truth, accuracy, or usefulness. This doesn't match at all with what the author described in the article. > Anyone trying …
1mo ago·view thread
comment
That is not what slop means, though. You're redefining the meaning of the word to suit your view. Why do that? You can just say that LLM generated content is not up to par, or acceptable, ever.
1mo ago·view thread
comment
Maxed out 2019 Mac Pro was $50k+. The wheels on that thing were $400. This is a bargain compared to that.
1mo ago·view thread
comment
Anthropic already explicitly communicated that they'll store and check all the data from Bedrock or any platform, even if you've selected zero data retention, if using Mythos class models.…
1mo ago·view thread
comment
The trope of "ABC players" is tired at this point. To be honest, even talking about "players" is kind of square in my opinion. In a similar way as the saying "don't hate …
1mo ago·view thread