Author: Prompt Engineering - Bewertung: 93x - Views:1659
Colibri: Run GLM 5.2 on 25GB of RAM on consumer hardware!
A 744B Mixture-of-Experts model activates only ~40B parameters per token — and only ~11 GB of those change from token to token (the routed experts).
LINKS:
https://github.com/JustVugg/colibri
https://z.ai/blog/glm-5.2
DwarfStar-4 Video: https://youtu.be/9gHcmhUDJfw
DSpark video: https://youtu.be/eFgknPFK-g0
My voice to text App: whryte.com
Website: https://engineerprompt.ai/
RAG Beyond Basics Course:
https://prompt-s-site.thinkific.com/courses/rag
Signup for Newsletter, localgpt:
https://tally.so/r/3y9bb0
Let's Connect:
🦾 Discord: https://discord.com/invite/t4eYQRUcXB
☕ Buy me a Coffee: https://ko-fi.com/promptengineering
|🔴 Patreon: https://www.patreon.com/PromptEngineering
💼Consulting: https://calendly.com/engineerprompt/consulting-call
📧 Business Contact: [email protected]
Become Member: http://tinyurl.com/y5h28s6h
💻 Pre-configured localGPT VM: https://bit.ly/localGPT (use Code: PromptEngineering for 50% off).
Signup for Newsletter, localgpt:
https://tally.so/r/3y9bb0
SOCIAL SHARE CARD GENERATOR