Given an arbitrary full license (plain) text, it would be cool if we could do a similarity search and get those entries listed whose texts are similar. The metric of "similar" would need to be defined, plus maybe some sort of threshold would need to be configurable; but maybe that's already going too far, and just some simple similarity search with sane defaults would suffice.
Given an arbitrary full license (plain) text, it would be cool if we could do a similarity search and get those entries listed whose texts are similar. The metric of "similar" would need to be defined, plus maybe some sort of threshold would need to be configurable; but maybe that's already going too far, and just some simple similarity search with sane defaults would suffice.