20M+ Indian legal documents with citation graphs and vector embeddings – potential uses for legal NLP? [D]
A r/MachineLearning discussion thread exploring the potential NLP applications of a large-scale dataset comprising over 20 million Indian legal documents, enriched with citation graphs and pre-compute