Constarium
← Search

Data · dataset · 2024

Legal Case Retrieval Large Language Model Evaluation Datasets

Listed in ScienceDB

We collected and organized a 11.3MB legal model evaluation dataset for case retrieval, which is stored in JSON format (*. json).

Description

The dataset consists of two files: Case_Setroval_Satasets.json and Case_Setroval_Satasets_llm.json. Among them, the Case_Setroval_Satasets.json file contains preprocessed datasets for various tasks, specifically designed for evaluating and testing large language models, with a data volume of 3.73MB.

On the other hand, the Case_Setroval_Satasets_llm.json file adds a comprehensive dataset of large language model prediction results based on the former, with a data size of 7.66MB. 

Links

Where it is published

Catalogue records · 1

Topics

Inferred from text
Legal systems 69%
Provenance · 1 source records, 10 field assertions
SourceKeyLast seenRaw
ScienceDB10.57760/sciencedb.175299 d agoJSON v1
FieldAssertionExtractorEvidence
concepts[field].anzsrc:group:4805enrichment · scidb cntaxonomy-embedding@1.0.0title+keywords+description (69%)
concepts[field].local:field:earth-environmentalmapping · scidb cnconnector:scidb_cn@1.0.0
concepts[field].local:field:engineeringmapping · scidb cnconnector:scidb_cn@1.0.0
concepts[field].local:field:humanitiesmapping · scidb cnconnector:scidb_cn@1.0.0
concepts[field].local:field:life-sciencesmapping · scidb cnconnector:scidb_cn@1.0.0
concepts[field].local:field:social-sciencemapping · scidb cnconnector:scidb_cn@1.0.0
descriptionsource · scidb cnconnector:scidb_cn@1.0.0/metadata/dc/description
license_textsource · scidb cnconnector:scidb_cn@1.0.0
publication_datesource · scidb cnconnector:scidb_cn@1.0.0
titlesource · scidb cnconnector:scidb_cn@1.0.0/metadata/dc/title