FP8 is the format modern quantised LLM inference increasingly targets, and INT2 suits aggressively compressed always-on models. Since Qualcomm publishes no throughput figure, supported precision is the one axis on which its generations can honestly be compared.
Snapdragon 8 Elite Gen 5
The real change is the precision set: INT2 and FP8 arrive, which matters more for modern quantised inference than the headline percentage.
