back
user profile

andy99

8,223karma·2,382submissions·May 21, 2020
recent activity (2,382 total)
comment
Yeah this makes it unreadable for me, immediately distracted by how stupid it is and lose interest in the article. As someone else noted, it’s people who think they are writing for an “executive” aud…
24d ago·view thread
comment
Rumored M5 ultra bandwidth is apparently 1100 GB/S (M3 Ultra is something like 880). This is 276 GB/S so they’re not really in the same league unfortunately. This will be considerably cheape…
24d ago·view thread
comment
> RAM prices weren't so f*cked I think we'd be seeing 256GB and even 512GB unified memory systems becoming quite common I agree but unfortunately that speaks to the depth of demand right …
24d ago·view thread
comment
I use it exclusively as a server for LLM inference - first more for research but have been using it for coding now that there are sufficiently capable models that run on. For that I’m very happy.
24d ago·view thread
comment
Try poolside that came out yesterday https://news.ycombinator.com/item?id=49004937 or Qwen 3.5 122B A10B, both use more memory and still have experts sized for decent speed at the 395…
24d ago·view thread
comment
I have a framework desktop w/ 128GB that I bought last Christmas and if I’m looking at it right it costs $2000 (CAD) more now because of the RAM shortage (and in any event is apparently out of st…
24d ago·view thread
comment
I don’t understand any of them. Normally even when I’m not really familiar with a tool, I have enough background knowledge to understand why it’s funny e.g. Scheme and Haskell jokes or something. I d…
24d ago·view thread
comment
If an AI researcher was going to pelicanmaxx, they would almost certainly apply the augmentations mentioned in the article during training, e.g. randomly selecting animals and conveyances. You’d want …
25d ago·view thread
comment
This is more a statement of how awful Canadian banks are than anything else. For anyone unaware we have an oligopoly of five identical banks all of which treat their customers like shit and effectivel…
25d ago·view thread
comment
This AI written article seems to be substituting ethics for “taste” and making similar arguments to those from the past. Choosing what to do is more important than doing it is a taste problem, of whic…
25d ago·view thread
comment
Real morality doesn’t pay, shallow “ethics” for business, tech, etc. has a whole industry that pays very all.
25d ago·view thread
comment
To go off topic a bit further, I recently went into Walgreens to buy some bottled water, and other than Evian, it was all advertised as alkaline. Is that just a trend, was it already alkaline and now …
25d ago·view thread
comment
It’s the Strix halo (AMD) with 128 GB shared memory. The 4bit quant is ~75GB. Unfortunately I don’t know about the best way of running on an Nvidia gpu, you could try llama.cpp and offloading as many …
25d ago·view thread
comment
Replying to myself, seems this PR was merged into main and it the model does work with a Vulkan backend on my Framework desktop, I’m getting about 220 tok/s prompt processing and 21 tok/s ou…
25d ago·view thread
comment
These articles are propaganda, it’s not an independent journalist writing it, they’re doing it at the behest of Anthropic or someone like them that’s pushing for regulatory capture.
25d ago·view thread
comment
This appears to be mostly due to fringe / activist views about copyright, rather than anything to do with quality or principle. If it was the latter I could get on board, as in instituting some s…
25d ago·view thread
comment
Seems it’s not fully supported in mainline llama.cpp yet https://github.com/ggml-org/llama.cpp/pull/25165 In the huggingface link they mention building for CPU and for …
25d ago·view thread
comment
This was posted earlier but didn’t get traction, and I made the following comment: Id want to know if “AI” makes a material difference vs just having access to the wrong answer. Like someone could be …
28d ago·view thread
comment
Right, and there are two parallel tracks. First is the “every crack and crevice” part - “ summarize with AI”, “re write with AI”, “help me write”, “analyze with AI”, basically useless features being s…
28d ago·view thread
comment
Id want to know if “AI” makes a material difference vs just having access to the wrong answer. Like someone could be given search access that successfully retrieved wrong answers to questions, would t…
28d ago·view thread
comment
Also, eating say a Big Mac and fries isn’t that unhealthy when done occasionally, there’s a lot of salt and fat but nothing horrendous. Compared to the load on your liver and pancreas etc of consuming…
28d ago·view thread
comment
I get about 12 tok/s with 27B 8 bit, 50 with 35B A3B 8 bit, and 12 with 3.5 122B A10B 4 bit. The latter is about 80 GB iirc. it feels like the best balance between using as much memory as I can a…
28d ago·view thread
comment
Qwen 3.5 to 3.6 was a big jump for the same size, e.g. 29 to 32 on artificial analysis intelligence for the 35BA3B models. Although I don’t think anyone has released a better model of that size since.…
28d ago·view thread