SQ
ENT-00001093 · Dataset

SQuAD

SQuAD is a dataset associated with Stanford NLP, classified in LXKeys.world as Language / Training Dataset.

Level IV ConnectedDocumentedDocumented
Human record

Overview

SQuAD is a documented stanford question answering dataset used in artificial-intelligence research, evaluation or dataset construction. Its record captures the task context, public source, evaluation role and relationships to models, agents or research systems that rely on comparable measurements.

Timeline

  1. 2016Initial dataset release

    SQuAD entered the documented public record in 2016. This event is retained at the precision supported by the Entry’s reviewed source history.

Capabilities

Evaluation referenceComparable measurementResearch reproducibilityModel documentation support

Known limitations

Benchmark interpretation depends on task design, data quality and contamination controls.Scores should not be treated as a complete measure of intelligence.
Technical and structured detail

Technical description

Structured reference record for SQuAD. The technical layer captures source URL, creator, first-public year, evaluation or dataset role, measurable task family and relationships to model testing, retrieval, safety, coding, reasoning or multimodal assessment.

Controlled tags

DatasetEvaluationDocumentationEstablishedBatch 03 Candidate
Relationship graph

SQuAD

Open in full graph
SQuAD
Created or maintained byOutgoing
Stanford NLP

SQuAD is associated with Stanford NLP.

Machine readable layer

Structured entry data for human tools, AI systems and machine clients.

{
    "@context": [
        "https://schema.org",
        {
            "lxw": "https://lxkeys.world/schema/"
        }
    ],
    "@type": "Thing",
    "identifier": "ENT-00001093",
    "name": "SQuAD",
    "alternateName": [],
    "additionalType": {
        "category": "Data and Evaluation",
        "type": "Dataset",
        "subtype": "",
        "lxkeysEntity": false
    },
    "description": "SQuAD is a dataset associated with Stanford NLP, classified in LXKeys.world as Language / Training Dataset.",
    "creator": "Stanford NLP",
    "url": "https://lxkeys.world/entry.php?id=ENT-00001093&lang=en",
    "sameAs": "https://rajpurkar.github.io/SQuAD-explorer/",
    "image": "",
    "lxkeysWorld": {
        "worldId": "ENT-00001093",
        "kind": "Data and Evaluation",
        "type": "Dataset",
        "subtype": "",
        "classification": "Language / Training Dataset",
        "organization": "Stanford NLP",
        "originContext": "AI evaluation and dataset record",
        "firstPublicAppearance": "2016",
        "currentStatus": "Documented",
        "documentationStatus": "Documented",
        "documentationIndex": {
            "total": 74,
            "documentation": 25,
            "evidence": 13,
            "structure": 25,
            "relationships": 11,
            "level": "Level IV Connected"
        },
        "lxCalendarium": {
            "start_date_utc": "2023-04-01",
            "created_utc": "2026-06-16T23:59:14+00:00",
            "created_dypclt": "D-0 Y-2 P-3 C-3 L-21 T-3",
            "updated_utc": "2026-06-16T23:59:14+00:00",
            "updated_dypclt": "D-0 Y-2 P-3 C-3 L-21 T-3"
        },
        "facts": [],
        "capabilities": [
            "Evaluation reference",
            "Comparable measurement",
            "Research reproducibility",
            "Model documentation support"
        ],
        "limitations": [
            "Benchmark interpretation depends on task design, data quality and contamination controls.",
            "Scores should not be treated as a complete measure of intelligence."
        ],
        "tags": [
            "Dataset",
            "Evaluation",
            "Documentation",
            "Established",
            "Batch 03 Candidate"
        ],
        "timeline": [
            {
                "date": "2016",
                "title": "Initial dataset release",
                "description": "SQuAD entered the documented public record in 2016. This event is retained at the precision supported by the Entry’s reviewed source history.",
                "source_url": "https://rajpurkar.github.io/SQuAD-explorer/",
                "verification_status": "source_backed_official"
            }
        ],
        "relationships": [
            {
                "target": "Stanford NLP",
                "type": "Created or maintained by",
                "description": "SQuAD is associated with Stanford NLP.",
                "evidence_level": "documentary"
            }
        ],
        "sources": [
            {
                "label": "SQuAD public reference",
                "url": "https://rajpurkar.github.io/SQuAD-explorer/",
                "source_type": "Primary or Research Source",
                "verification_status": "verified",
                "authority": "OFFICIAL",
                "role": "primary",
                "update_enabled": true,
                "authority_basis": "curated-corpus-refresh-2026-09-08"
            }
        ],
        "canonical": [],
        "imageMeta": []
    }
}
Proofs

Contribute to this record

Submit a source, correction or comment. Public changes remain moderated.

Submit proof or comment

Approved comments

No approved public comment yet.