| Hi stardance! Please make sure to try both demos out at https://tinyapper.craisin.tech and https://yapperpedia.craisin.tech. Additionally, tinyapper can be kinda slow (no GPU in my server unfortunately) and dumb (it is severely undertrained), please be patient and understand that these are silly proof of concept models!
Tinyapper 2.0.0
More Tinyapper, the LLM trained from scratch! This time the model is trained to be fully conversational with user, assistant, and eos tokens, along with a session attached KV cache.
Tech Changes
- New conversational demo at tinyapper.craisin.tech
- Updated attention conditional for multiple infills for chat
- Updated KV cache so chats can be cached
- Reiterated how datasets are made
- Retrained model for yapperpedia so the model is slightly more coherent
Future Development
- Standardizing datasets to use memmap for larger datasets
- Retrain conversational model
- Fine tune the model
While this will probably be the last of me working on this project for Stardance, I’ll definitely work on this more in my own time. Keep an eye on the Github for updates if you are interested, stars and follows there are noticed and appreciated!
- 4 devlogs
- 14h
- 16.51x multiplier
- 225 Stardust



