back
user profile

schmorptron

1,232karma·577submissions·August 3, 2020
recent activity (577 total)
comment
This is what the author showed off on Twitter at least, with Desktop chromium I think.
10h ago·view thread
comment
The past DeepSeek models and now these new checkpoints score very badly on the ArtificialAnalysis AA-Omniscience and hallucination rate benchmarks. I wonder where that's from? Maybe they're …
3d ago·view thread
comment
You make a good point about cultural things that would make one excited to learn a language. But I really struggle to imagine what could be done to fix that? Injecting more money into the film industr…
13d ago·view thread
comment
What a nice and heartwarming read! I'm german, and really hope we can pivot away from the rise of the far right. How that can be done, I don't know. I just hope we can do so before we start …
13d ago·view thread
comment
App tells customer to review it. Customer reviews the app honestly (the nag is annoying and deserves 1 star). Seems like the nag screen worked? :) (Although in seriousness, I get this point they make …
18d ago·view thread
comment
it has to be so bursty for realtime usecases like chat, which is what most people are using it for today. of course, once (if) stuff like software dark factories start working out for the average pers…
20d ago·view thread
comment
LLM inference unfortunately also seems to be a task that's poorly formed for moderate consumer hardware,as a single user. For a single user use case, the load is bursty but requires the weights t…
20d ago·view thread
comment
It's impossible to keep ads "clearly labelled" and seperate long term. Doesn't work. the incentive structures just don't work. serving answers becomes a cost center, and ads a…
24d ago·view thread
comment
Yeah, this has been my progression as well. Maybe next it'll be just using plain pi when you figured out exactly what you want from omp and what you don't
27d ago·view thread
comment
Are thinking models only the reasonable tradeoff vs using much larger non thinking ones because the cost of output tokens is below that of input tokens?
1mo ago·view thread
comment
That's a more than 2x jump in parameter count. I know it's not a measure of quality by itself, but it will be interesting how it "scales". Bust it looks like they're gonna be …
1mo ago·view thread
comment
yeah, but that's due to enterprise commitments that MS won't train on the user interactions
1mo ago·view thread
comment
I don't think that's it for console manufacturers. They make the majority of their money on game sales, so they want the console itself to be used for as long as possible.
1mo ago·view thread
comment
We're moving towards total surveillance slowly but surely. Age verification. Chat control. To an extent also the digital euro. It all seems hopeless, they're pushing this through despite wha…
1mo ago·view thread
comment
i love that word, and now it's genuinely (hehe) ruined. thanks, claude
1mo ago·view thread
comment
I'm giving them the benefit of the doubt and interpreting it in a charitable way because they sound earnest about it, this is incredibly ambitious and cool-sounding, and I wish them all the best.…
1mo ago·view thread
comment
you can build the datacenter right next to the tank and use the now-warm cooling water to pump into the tanks!
1mo ago·view thread
comment
Cursor's composer models are finetuned kimi
2mo ago·view thread
comment
It's "just" an opencode fork but it adds some nice features to try out while not being a full orchestrator metapackage like oh-my-opencode. Quite nice! Though it would be even nicer if …
2mo ago·view thread
comment
got one answer by reading the rest of the comments, makes sense that the diffusion process is inherently reasoning-like: https://www.inceptionlabs.ai/blog/introducing-mercury-2 …
2mo ago·view thread
comment
What would a diffusing reasoning model look like? have a pre-defined length [thinking] block that gets diffused over a long time, and then the final output block uses what is in that thinking block as…
2mo ago·view thread
comment
The irony of "we train on all of humanity's collective output, but god forbid anyone trains on ours" is still incredible
2mo ago·view thread
comment
Cool project! I'll be trying it out. I've been a big fan of throwing whatever sources I have on a new topic i'm trying to get into into a llm "project" and then asking it to t…
2mo ago·view thread
comment
Maybe this will replace raptor-mini as the "free" model on copilot plans? (but I don't see it at all yet on the student plan, in vscode or the cli)
2mo ago·view thread
comment
the new intel ultra whatevername 3 series seems to come a bit closer there, so the framework pro with its explicit linux support might be an option
2mo ago·view thread
comment
It's interesting that (for example for the explore agent https://github.com/Piebald-AI/claude-code-system-prompts/blo... ) they use a personality "you are a file s…
2mo ago·view thread
comment
i see the reasoning traces in opencode (cli). maybe it's a setting?
2mo ago·view thread
comment
I think part of it is also that we're able to still LARP as full developers of complex systems while vibe coding by seeing an interface that makes us look like l33t h4xx0rs even though we're…
3mo ago·view thread
comment
Aaand critique https://www.lesswrong.com/posts/veFMEzDDyWaer2Sms/sanity-che... …
3mo ago·view thread