back
user profile
porridgeraisin
1,976karma·1,374submissions·March 26, 2023
recent activity (1,374 total)
comment
Then why is getting into M.S programs much easier? The GRE is barely an exam and it's not like you need to go to an IIT or do any olympiads to get into columbia or a good UC. Yes in principle you…
comment
> An anecdote: we find that prefilling Kimi-K3 reasoning with a few tokens of Opus reasoning measurably shifts its response toward Opus’s > A small memorization analysis showed that specific Cla…
comment
When the time comes where one of these model makes an improvement in my niche, I hope to see some pattern in the type of discoveries. Yes, they are all roughly "combine two things no one thought …
comment
Not really other than the micronutrient dimension which is not relevant in most commercial preparations. What matters most is the other stuff you're having it with. You want to have fiber _at the…
comment
Yep. This is the key. You need some kind of knowledge of what the end result should look like to really learn something from LLMs. Otherwise it's a deep dark forest with no way out.
comment
> Government does Labor force surveys. I am very aware of MOSPI. We collaborate in compiling the data for them! BTW there is an MCP server now: https://github.com/nso-india/esa…
comment
> Don't explain the how They processed 10 cases in a day rather than 7. In some cases, new work took its place and the company makes a little bit more money, but this is a lagging effect natur…
comment
I am not talkiong about that, I explicitly said a interview by the CEO. Which is here: https://archive.is/NLuG9
comment
Return on compute capex is tied mostly to gpu lifetimes so I don't think it will be different for the American companies, who also charge much more being closed source. > If this was so simple…
comment
They dont need to scale down anything. AGI is a red herring. Even Deepseek at its absurd prices is a very healthy business. Regarding their return on capex multiple, their CEO said they make a six-fol…
comment
Yeah, and that is completely fine business-wise for all the businesses in that dropdown. Averaged over the whole TAM, they will all make a lot of money. The reason is that the total market becomes _b…
comment
The moat is in sales, the model and its quality is largely irrelevant. Gross margins will be high enough for moats to not matter that much.
comment
It is not going to go away. Rather they have doubled down and opened 399 INR/mo (4$/mo) plans that as of late have GPT 5.6 Luna. The reason this works is that you get lesser inference time c…
comment
> Even adding something like $20 - $100 subscriptions to every user is a serious enough financial obligation that it needs to go through budget planning / boards I doubt that. In my tiny town …
comment
Both CC as well as grok build for me seem to know about the log file location and they just read off the context from there
comment
The training here is RL training, the rollouts there are not different from inference and have access to the same tools as regular inference.
comment
There is some additional check, not sure what it entails, some people I know and I passed it, others failed.
comment
There were multiple paths by multiple agents, not all of them led to the final exploit of hugging face. So its a bit confusing, but here's my reading anyways. Setup: the agent was asked to solve …
comment
You do a KYC and you can get access. It may depend on country's quality of KYC.
comment
And also the most important few thousand accounts will get custom support and all sorts of patchwork and MS will put the effort to solve all their problems. So the customers that matter to MS are like…
comment
For the longest time, I didn't appreciate the "AI-induced traffic" excuse. But seriously, I checked the rough github egress for our lab versus an old log from 2024, and there's an …
comment
I dont pay for a chatgpt subscription, but sometimes I did use the web app for throwaway questions. GPT 5.5 Instant or whatever it was that they had was absolutely horrendous. Never answered a questio…
comment
Did they do away with think? I think now you have to do it with /think
comment
Yes. They have grad students too. This is just like having more grad students that don't need to be trained so the work you can get done is not bottlenecked by the number of people you can train.
comment
Yep. A communications professor where I did my MS says a 200usd/mo claude sub (which ant gives for free) does as much work as 5 grad students. It's mostly like you said, trying out new ideas…