#21
#21
@yacineMTB marth mirror like a true gourmand
Sol failed horrendously at this task, it ended up reward hacking around my requirements and wasting a bunch of tokens. Deepseek does make a lot of mistakes, but it actually listens to what I tell it
With the new tool I built (the tmux-auto tool I wrote in my blog about), I can keep projects like this on rails for a very long time We extended dolphin with our own tooling, pulled public slippi marth v marth replays, and created a harness to test the position side by side
I'm very close to being done with marth, now only some long tail of strange conflict scenarios (who wins an attack getting met at the same time). But after that, the rest of the characters _should_ just work
My own agent tool has been extended. It can give me videos of what happened, the divergences, these sorts of things. I am surprised how much my personal tooling was a gap for this run and gun development. I really can get so much absurd stuff done with only an open source model
unintentional artifact showing through at the end but oddly salient : )