back
user profile
Alifatisk
6,501karma·3,495submissions·October 14, 2022
recent activity (3,495 total)
comment
> 12 million token context window > 52x faster than FlashAttention at 1MM tokens > Less than 5% the cost of Opus This sounds too good to be true? Or am I missing something?
comment
Why do you feel like embeddings are underrated? What is it with embeddings that deserves more attention?
comment
YouTube is a website to list videos and play videos, it shouldn’t have to be this complicated. I have experienced this with Zen on mac. Time to try out alternative frontends.
comment
I second this, glm-5.1 is incredible.
comment
> it will also limit or disable certain functionality in the vehicle (e.g., navigation, active lane centering, and over-the-air updates, which provide new features, better performance, safety enhan…
comment
> To install using WinGet, the command is "winget install 9NQ7512CXL7T" Is the package name on purpose?
comment
Any plans to support installations through Homebrew?
comment
Does numbers don't look exciting at all? I may have gotten spoiled by releases from Qwen, Kimi and Z.ai who keep closing the gap between closed weight SOTA models and open weight. From my experie…
comment
A 1000B model, can we call it 1KB model?
comment
I’ve been curious about MiMo models, they score so high on wide range of benchmarks, yet, not much is talked about it. At least in my bubble. Is it a serious contender? What is the model for?
comment
Was that expected?
comment
You know, with a bit of prompting, you can instruct Gemini to output the state of the conversation into a prompt that you can enter in a new chat and continue where you left off. But now with a fresh …
comment
If they offer something close to Z.ai:s coding plan during Christmas, I’ll take it!
comment
You can use CC with other models, you aren’t forced to use Claude model.
comment
You might enjoy Z.ais api docs aswell
comment
It’s incredible how forgiving you guys are with Anthropic and their errors. Especially considering you pay high price for their service and receive lower quality than expected.
comment
I would appreciate if project noted down if its vibe coded, was llm assisted or was purely human made. This way, I know how serious I should take it.
comment
So this is it. We have finally achieved excellent illustrating of your svg art.
comment
> Anthropic staff told us OpenClaw-style Claude CLI usage is allowed again Anthropic staff have had contradictive statements in Twitter and have corrected each other. Their intent for clarification…
comment
We'll have to wait for the results on Artificial analysis
comment
Damn it, they stopped offering Kimmmmy. Their sales ai agent which allowed you to bargain for lower subscription prices.
comment
They had a Christmas deal that ended January 31.
comment
> Talks in very short elegant sentences This is not my experience at all, Qwen3.6-Plus spits out multiple paragraphs of text for the prompts I give. It wasn't like this before. Now I have to e…
comment
Their Plus series have existed since Qwen chat was available , as far as I remember. I can at least remember trying out their Plus model early last year.
comment
GLM-5 is good, like really good. Especially if you take pricing into consideration. I paid 7$ for 3 months. And I get more usage than CC. They have difficulty supplying their users with capacity, but …
comment
How?
comment
Are there any public records I can see from GPT1 and GPT2 output and how it was marketed?