DEVLOG #35 - A better retry is still not model learning
SL-LLM-R is meant to change future model behavior, not just get a better answer on the second attempt. Many AI systems can retry with more context and look like they learned even when the model itself stayed the same. I only care about a correction as learning if it affects later behavior, does not break other abilities, and can still be undone.
Comments 0
No comments yet. Be the first!
Sign in to join the conversation.