To the second footnote: you could utilize Polar's lazyframe API to do that cosin...

minimaxir · 2025-02-24T21:04:55 1740431095

That would get around memory limitations but I still think that would be slow.

kipukun · 2025-02-24T21:26:58 1740432418

You'd be surprised. As long as your query is using Polars natives and not a UDF (which drops it down to Python), you may get good results.

jononor · 2025-02-25T09:14:48 1740474888

A (simple) benchmark would be great to figure out where the practical limits of such an approach are. Runtime is expected to grow with O(n*2) which will get painful at some point.