back
user profile
HenryNdubuaku
946karma·141submissions·October 24, 2021
recent activity (141 total)
comment
Although GPT OSS 20B has 1.7B activated parameters which will be fast, 20B weights is a lot for developers to bundle or consumers to download. That’s the actual problem.
comment
Real-time video and audio inference.
comment
The license change doesn’t affect you based on your explanation actually, the licence has been updated with clearer words. We really appreciate you as a user, please share any more feedback you have, …
comment
Done, thanks, let us know anything else.
comment
This was one of the issues we set out to solve, so not as much as you’d expect.
comment
It’s absolutely fine to share your thoughts, that’s the point of this post, we want to understand where people’s heads are at, it’s what determines our next decisions. What do you really think? I’m ge…
comment
Thanks for sharing your thoughts. Honestly, I’d be annoyed too and it might sound like an excuse, but our circumstance was quite unique, it was a difficult decision at that time being an open-source c…
comment
Understandable, though to explain, Cactus is still free for personal & small projects if you fall into that category. We’re early and would definitely consider your concerns on license in our next…
comment
It can incorporate any tool you want at all. This company’s app use exactly that feature, you can download and get a sense of it before digging in. https://anythingllm.com/mobile …
comment
Thanks for noticing! The app is just a demo for the framework, so devs can compare the open-source models against frontier Cloud models and make a decision. We removed the comparison now so those scre…
comment
Thanks for the kind words, we’ve improved performance now actually, follow the instructions on the core repo. Same model should run 3x faster on the same phone. These improvements are still being push…
comment
Cactus is free for hobbyists and personal projects, but we charge a tiny fee for commercial use which comes with more features that are relevant for enterprises.
comment
400mb if you ship the model as an asset. However, you can also build the app to download the model post-install, Cactus SDKs support this, as well as agentic workflows you’d need.
comment
Thanks!
comment
[flagged]
comment
We are writing our own backend, but tflite (now called LiteRT) was not faster than GGML when we tested and GGML is already well supported. But we are moving away completely anyway.
comment
We are following Ollama's design, but not verbatim due to apps being sandboxed. Phones are resource-constrained, we saw significant battery overhead with in-process HTTP listeners so we stuck wit…
comment
Thanks for the comment, but: 1) The commit history goes back to April. 2) LlaMa.cpp licence is included in the Repo where necessary like Ollama, until it is deprecated. 3) Flutter isolates behave like…
comment
Please feel free to join our Discord: https://discord.com/invite/bNurx3AXTJ