Side by side

Compare two models on the same machine

Useful when two models look interchangeable on paper. Grouped-query attention, expert routing and cache compression mean models of the same size can behave nothing alike once they are loaded, and the gap usually shows up in the context column rather than the weights.