Beau
Okay, Jo. So, we've talked about all these incredible, like, theoretical ways to make VLIW processors scream. We've got our fancy compiler doing the scheduling, we've unrolled our loops, we've optimized memory access... but how do we actually know if it's working? You know? How do we prove it's fast and not just... theoretically fast?