Results from my experiment – with fine-tuning, diffusion models can beat pre-trained models almost 9 times larger, and autoregressive fine-tuned models 1.5 times larger
Results from my experiment – with fine-tuning, diffusion models can beat pre-trained models almost 9 times larger, and autoregressive fine-tuned models 1.5 times larger
I am training an Agent using Soft Actor Critic, and this is the performance after 12k steps of training (yes it’s a little violent rn but I will either train for longer, increase frequency or implement imitation learning)