Screenshot from movi2 paper. Size-wise ropebwt is already competitive, but I don't really get why it's _this_ much slower? Traversing a Btree ought to only be a bit slower than reading the relevant cacheline directly. Looks like the diff is more than just batching.
1 likes 3 replies
?