About this Course
Run large language models locally using GGUF and llama.cpp. Optimize inference speed and manage hardware constraints. This AI/ML curriculum is designed to give you hands-on experience and deep conceptual understanding.
Across 4 intensive modules, you'll tackle real-world challenges and build practical projects that reinforce your learning. By the end of this journey, you'll have the skills and proof of work to demonstrate your expertise.
What you'll learn
Master the core concepts of introduction to local inference.
Gain hands-on experience with running models with llama.cpp.
Understand the architecture behind advanced inference techniques.
Implement production-grade optimization and deployment.
W1
Introduction to Local Inference
Master the core concepts of introduction to local inference.
4 videos•63m
3 readings
4 topics
1 homework
W2
Running Models with llama.cpp
Gain hands-on experience with running models with llama.cpp.
4 videos•45m
3 readings
4 topics
1 homework
W3
Advanced Inference Techniques
Understand the architecture behind advanced inference techniques.
4 videos•97m
3 readings
4 topics
1 homework
W4
Optimization and Deployment
Implement production-grade optimization and deployment.
4 videos•57m
3 readings
4 topics
1 homework
01
Learn
Watch curated videos and read study resources
02
Practice
Practice what you learned
03
Build Projects
Build projects using your new gained knowledge
04
Submit & Verify
Submit your project and get verified by our system
References
Rate this course
Help the community find verified technical paths.
Community Insights
0Join the discussion
Sign in to share your thoughts and technical insights.
Loading insights...