It would be beneficial to have a hardware-optimized Llama lineup with a clearer naming scheme and distinct performance tiers, for example:
- Llama 4.0 Phone (Lite / Standard / Max) – For mobile devices.
- Llama 4.0 Workstation (Lite / Standard / Max) – For PCs and laptops.
- Llama 4.0 Server (Lite / Standard / Max) – For high-performance computing.
This approach would enable developers to select the appropriate model based on both device type and performance needs.
What do you think? For example now I feel like 3.3 70B is more for laptops/PCs, and the previous 3.2 3B for phones, is a bit confusing to me.