back
user profile
schmorptron
1,232karma·577submissions·August 3, 2020
recent activity (577 total)
comment
This is what the author showed off on Twitter at least, with Desktop chromium I think.
comment
The past DeepSeek models and now these new checkpoints score very badly on the ArtificialAnalysis AA-Omniscience and hallucination rate benchmarks. I wonder where that's from? Maybe they're …
comment
You make a good point about cultural things that would make one excited to learn a language. But I really struggle to imagine what could be done to fix that? Injecting more money into the film industr…
comment
What a nice and heartwarming read! I'm german, and really hope we can pivot away from the rise of the far right. How that can be done, I don't know. I just hope we can do so before we start …
comment
App tells customer to review it. Customer reviews the app honestly (the nag is annoying and deserves 1 star). Seems like the nag screen worked? :) (Although in seriousness, I get this point they make …
comment
it has to be so bursty for realtime usecases like chat, which is what most people are using it for today. of course, once (if) stuff like software dark factories start working out for the average pers…
comment
LLM inference unfortunately also seems to be a task that's poorly formed for moderate consumer hardware,as a single user. For a single user use case, the load is bursty but requires the weights t…
comment
It's impossible to keep ads "clearly labelled" and seperate long term. Doesn't work. the incentive structures just don't work. serving answers becomes a cost center, and ads a…
comment
Yeah, this has been my progression as well. Maybe next it'll be just using plain pi when you figured out exactly what you want from omp and what you don't
comment
Are thinking models only the reasonable tradeoff vs using much larger non thinking ones because the cost of output tokens is below that of input tokens?
comment
That's a more than 2x jump in parameter count. I know it's not a measure of quality by itself, but it will be interesting how it "scales". Bust it looks like they're gonna be …
comment
yeah, but that's due to enterprise commitments that MS won't train on the user interactions
comment
I don't think that's it for console manufacturers. They make the majority of their money on game sales, so they want the console itself to be used for as long as possible.
comment
We're moving towards total surveillance slowly but surely. Age verification. Chat control. To an extent also the digital euro. It all seems hopeless, they're pushing this through despite wha…
comment
i love that word, and now it's genuinely (hehe) ruined. thanks, claude
comment
I'm giving them the benefit of the doubt and interpreting it in a charitable way because they sound earnest about it, this is incredibly ambitious and cool-sounding, and I wish them all the best.…
comment
you can build the datacenter right next to the tank and use the now-warm cooling water to pump into the tanks!
comment
Cursor's composer models are finetuned kimi
comment
It's "just" an opencode fork but it adds some nice features to try out while not being a full orchestrator metapackage like oh-my-opencode. Quite nice! Though it would be even nicer if …
comment
got one answer by reading the rest of the comments, makes sense that the diffusion process is inherently reasoning-like: https://www.inceptionlabs.ai/blog/introducing-mercury-2 …
comment
What would a diffusing reasoning model look like? have a pre-defined length [thinking] block that gets diffused over a long time, and then the final output block uses what is in that thinking block as…
comment
The irony of "we train on all of humanity's collective output, but god forbid anyone trains on ours" is still incredible
comment
Cool project! I'll be trying it out. I've been a big fan of throwing whatever sources I have on a new topic i'm trying to get into into a llm "project" and then asking it to t…
comment
Maybe this will replace raptor-mini as the "free" model on copilot plans?
(but I don't see it at all yet on the student plan, in vscode or the cli)
comment
the new intel ultra whatevername 3 series seems to come a bit closer there, so the framework pro with its explicit linux support might be an option
comment
It's interesting that (for example for the explore agent https://github.com/Piebald-AI/claude-code-system-prompts/blo... ) they use a personality "you are a file s…
comment
i see the reasoning traces in opencode (cli). maybe it's a setting?
comment
I think part of it is also that we're able to still LARP as full developers of complex systems while vibe coding by seeing an interface that makes us look like l33t h4xx0rs even though we're…
comment
Aaand critique https://www.lesswrong.com/posts/veFMEzDDyWaer2Sms/sanity-che... …