back
user profile

helloplanets

2,648karma·541submissions·September 15, 2020
recent activity (541 total)
comment
A model being a distilled version of another specific model is a different thing from using synthetic data off of another model. Anthropic goes to insane lengths to block other labs from training off …
22d ago·view thread
comment
In the same vein, Jonathan Blow's modestly named "Preventing the Collapse of Civilization" talk: https://www.youtube.com/watch?v=ZSRHeXYDLko Crazy how time flies, given…
22d ago·view thread
comment
Pretty sure Mythos and Fable have way more params, but they've just been able to use the synthetic data off of them to get the leap in quality from Opus. So, not a distilled version of Mythos or …
22d ago·view thread
comment
With the new Ultracode modes in Claude Code and Codex it's been taken to the next level. I mean, multi agent systems have been available for a long time, but the newest models seem to be much mor…
24d ago·view thread
comment
People don't realize the exponential diffictlty curve with juggling. The highest amount of balls ever juggled is 11.
25d ago·view thread
comment
The actual part on fine-tuning seems very short in the article. Did I miss a page where they have examples of fine-tuning it for different niche use cases? Optimizing models to be fine-tuned is an ama…
1mo ago·view thread
comment
Shouldn't the valuation be in Bs instead of Ms?
1mo ago·view thread
comment
Yes, it is a 10x markup on the API prices. Depending on whether you factor in cooling costs, data center staff, etc. Or GPU costs and the electricity the GPUs are using only. Either way, inference is …
1mo ago·view thread
comment
True, it'd be a whole other situation if the tokens limits were cumulative. I guess it would all come down to whether their Claude Code subscription plans are turning in a profit or not. At least…
1mo ago·view thread
comment
The subscription based plans are heavily subsidized, but the direct API inference pricing (which larger companies need to pay) is profitable. Using a full Claude Max 20x plan to 100% of weekly usage w…
1mo ago·view thread
comment
I just did this on one .claude directory and >20% of the answers there included some variation of "real", "actual", "exact", "honest", "genuine", &…
1mo ago·view thread
comment
It's kind of offputting how much Anthropic models these days keep repeating "real", "genuine" and "honest". They've RL'd that way over the top.
1mo ago·view thread
comment
Definitely not just a classifier layered on top, although there is one of those as well. Pretty sure it's different post-training / finetune run and the model weights are different between t…
1mo ago·view thread
comment
But this guy's been all over the news? He's not some made up fantasy person with absolutely no real world footprint. Even if his website is sloppy. https://en.wikipedia.org/w…
1mo ago·view thread