back
user profile
helloplanets
2,648karma·541submissions·September 15, 2020
recent activity (541 total)
comment
A model being a distilled version of another specific model is a different thing from using synthetic data off of another model. Anthropic goes to insane lengths to block other labs from training off …
comment
In the same vein, Jonathan Blow's modestly named "Preventing the Collapse of Civilization" talk: https://www.youtube.com/watch?v=ZSRHeXYDLko Crazy how time flies, given…
comment
Pretty sure Mythos and Fable have way more params, but they've just been able to use the synthetic data off of them to get the leap in quality from Opus. So, not a distilled version of Mythos or …
comment
With the new Ultracode modes in Claude Code and Codex it's been taken to the next level. I mean, multi agent systems have been available for a long time, but the newest models seem to be much mor…
comment
People don't realize the exponential diffictlty curve with juggling. The highest amount of balls ever juggled is 11.
comment
The actual part on fine-tuning seems very short in the article. Did I miss a page where they have examples of fine-tuning it for different niche use cases? Optimizing models to be fine-tuned is an ama…
comment
Shouldn't the valuation be in Bs instead of Ms?
comment
Yes, it is a 10x markup on the API prices. Depending on whether you factor in cooling costs, data center staff, etc. Or GPU costs and the electricity the GPUs are using only. Either way, inference is …
comment
True, it'd be a whole other situation if the tokens limits were cumulative. I guess it would all come down to whether their Claude Code subscription plans are turning in a profit or not. At least…
comment
The subscription based plans are heavily subsidized, but the direct API inference pricing (which larger companies need to pay) is profitable. Using a full Claude Max 20x plan to 100% of weekly usage w…
comment
I just did this on one .claude directory and >20% of the answers there included some variation of "real", "actual", "exact", "honest", "genuine", &…
comment
It's kind of offputting how much Anthropic models these days keep repeating "real", "genuine" and "honest". They've RL'd that way over the top.
comment
Definitely not just a classifier layered on top, although there is one of those as well. Pretty sure it's different post-training / finetune run and the model weights are different between t…
comment
But this guy's been all over the news? He's not some made up fantasy person with absolutely no real world footprint. Even if his website is sloppy. https://en.wikipedia.org/w…