I havent slept since yesterday, my eyes are about to die on me but i reached a breakthroughhh!!!! finally made this CPU Native and might be able to achieve one more thing which i will disclose next devlog. This will be huge for Local LLM users.
This is possible using something called ThinkingCaps and ContextFabric. These methods will allow for actually usable fast 9B CPU native model inferencing with like enough context length and Symbolic Reasoning and Coding capabilities and verification stuff
Also made a Web UI for Investigating and testing out smaller quantities of test runs which is supposed to run on Local.
I will probably send the first milestone report for next devlog and also start working on a Demo URL that will be hosted on Modal
Comments 0
No comments yet. Be the first!
Sign in to join the conversation.