i'm honestly pretty shocked it got this good too considering what i was working with a few months ago, i expect that i can make it even better with a bit more training too
Профиль
Pangaea
Профиль Vivelyhi, i made those ArcSys script modding tools, and the Xrd SAMMI/WebSockets integration. posting stuff i make here. @TopTwentyNotes on twitter Support my work: https://ko-fi.com/pangaea__
as a comparison, this is what the model was doing a couple months ago bsky.app/profile/topt...
Pangaeanew model, this time it's trained to predict further ahead in the future, hoping that if i tweak this right it will learn to plan ahead better and maybe execute real combos, there's already some improvement to sols short strings but it doesn't seem to be enough yet
different type of training, it's used to make models that predict stuff better at actually accomplishing a specific task instead of just predicting stuff bsky.app/profile/topt...
Pangaeafor those who dont know what "RL" means: its reinforcement learning. the models i was showing before were trained to mimic player behavior as accurately as possible, this one has a second phase of training which tells it to estimate which decisions were "good" and focus on those actions
no actually! thats why i'm so surprised by it bsky.app/profile/topt...
Pangaeain this case i actually expected zato to be unusually low, as eddies attacks wouldn't be counted because theyre a separate entity changing state, not zato. there must be some other thing boosting him unusually high, that's what happens when you have few samples of some characters
this isn't a perfect metric, it over-focuses on attacks, probably has some noise because a button press doesnt necessarily cause any given state change, and the average probably gets raised by stuff like mashing FD, i'm surprised at how clean it seems to be for characters with a lot of match samples