Measure LLM inference performance with KleidiAI and SME2 on Android
Introduction
Understand how SME2 and KleidiAI accelerate LLM inference in llama.cpp
Trace how KleidiAI and SME2 accelerate llama.cpp from model load to token decode
Build llama.cpp with KleidiAI and SME2 enabled
Measure SME2 acceleration in llama.cpp on Android
Next Steps
Measure LLM inference performance with KleidiAI and SME2 on Android
Share
Bring your insights to the conversation.
Give Feedback
How would you rate this Learning Path?
What is the primary reason for your feedback ?
Thank you! We're grateful for your feedback.
- Have more feedback? Log an issue on GitHub.
- Want to collaborate? Join our Discord server.
Continue Learning
Read related resources
Find more information about the topics in this Learning Path:
Join the Arm Developer Program
Connect, upskill, and build with the Arm Developer Community. Join today for hands-on technical resources and education materials, along with the support of Arm engineers and the broader ecosystem.