Deploy a RAG-based Chatbot with llama-cpp-python using KleidiAI on Google Axion processors
Introduction
Demo
Set up a RAG based LLM Chatbot
Deploy a RAG-based LLM backend server
Deploy RAG-based LLM frontend server
The RAG Chatbot and its Performance
Next Steps
Deploy a RAG-based Chatbot with llama-cpp-python using KleidiAI on Google Axion processors
Share
Bring your insights to the conversation.
Give Feedback
How would you rate this Learning Path?
What is the primary reason for your feedback ?
Thank you! We're grateful for your feedback.
- Have more feedback? Log an issue on GitHub.
- Want to collaborate? Join our Discord server.
Continue Learning
Read related resources
Find more information about the topics in this Learning Path:
Join the Arm Developer Program
Connect, upskill, and build with the Arm Developer Community. Join today for hands-on technical resources and education materials, along with the support of Arm engineers and the broader ecosystem.