If part of the marketing talk is about the ability to run a local AI (via tools like Ollama, vLLM, LLM Studio, ...) Qualcomm seems to be leaving late. For more than a year and a half since the first Gen Snapdragon X is marketed, I still find it difficult to find inference engines that rely on the NPU Hexagon capabilities while the CUDA at Nvidia, rockM AMD or MLX at Apple are supported by the main engines of inference It's really a pity because …
This story is only covered by news sources that have yet to be evaluated by the independent media monitoring agencies we use to assess the quality and reliability of news outlets on our platform. Learn more here.
If part of the marketing talk is about the ability to run a local AI (via tools like Ollama, vLLM, LLM Studio, ...) Qualcomm seems to be leaving late. For more than a year and a half since the first Gen Snapdragon X is marketed, I still find it difficult to find inference engines that rely on the NPU Hexagon capabilities while the CUDA at Nvidia, rockM AMD or MLX at Apple are supported by the main engines of inference It's really a pity because …