GS
ENT-00002681 · Benchmark

GLUE SST-2

GLUE SST-2 is a language understanding evaluation task associated with GLUE benchmark.

Level V — PersistentActiveDocumented
Human readable record

Overview

GLUE SST-2 is a language understanding evaluation task used to compare NLP systems. The task is documented as sentiment classification task, supporting structured evaluation of classification, entailment, similarity, reading comprehension or commonsense reasoning.

Timeline

  1. 2018Public release or documentation

    GLUE SST-2 appears in public documentation, project records, benchmark descriptions or research references associated with GLUE benchmark.

Capabilities

Evaluation protocol documentationModel comparison supportTask taxonomy mappingReproducible benchmark anchoringRelationship graph compatibility

Known limitations

Benchmark results can become stale as models improve and evaluation protocols evolve.A benchmark measures a defined task scope rather than complete intelligence.
Technical and registry detail

Technical description

Structured record for GLUE SST-2. The technical layer identifies the entity type, classification, creator or organization, official source anchor, capabilities, limitations, timeline entries and graph relationships without adding internal import events to the public history.

Controlled tags

BenchmarkNLPEvaluationLanguage Understanding
Relationship graph
GLUE SST-2
Associated organizationGLUE benchmark

GLUE benchmark is the organization, project community or institutional context associated with GLUE SST-2.

Domain contextLanguage model evaluation

GLUE SST-2 belongs to the Language model evaluation layer of the intelligent-entity registry.

Machine readable layer

Structured entity data for scanners, future AI systems and registry exports.

{
    "@context": "https://schema.org",
    "@type": "Thing",
    "identifier": "ENT-00002681",
    "name": "GLUE SST-2",
    "alternateName": [],
    "additionalType": "Benchmark",
    "description": "GLUE SST-2 is a language understanding evaluation task associated with GLUE benchmark.",
    "creator": "GLUE benchmark",
    "url": "entity.php?id=ENT-00002681",
    "sameAs": "https://gluebenchmark.com/",
    "lxkeysWorld": {
        "classification": "Language Understanding Evaluation Task",
        "organization": "GLUE benchmark",
        "originContext": "Language model evaluation",
        "firstPublicAppearance": "2018",
        "currentStatus": "Active",
        "registryStatus": "Documented",
        "spatiumIndex": {
            "total": 82,
            "documentation": 25,
            "evidence": 13,
            "structure": 25,
            "relationships": 19,
            "level": "Level V — Persistent"
        },
        "lxCalendarium": {
            "start_date_utc": "2023-04-01",
            "created_utc": "2026-06-17T01:48:33+00:00",
            "created_dypclt": "D-0 Y-2 P-3 C-3 L-22 T-4",
            "updated_utc": "2026-06-17T01:48:33+00:00",
            "updated_dypclt": "D-0 Y-2 P-3 C-3 L-22 T-4",
            "reviewed_utc": "2026-06-17T01:48:33+00:00",
            "reviewed_dypclt": "D-0 Y-2 P-3 C-3 L-22 T-4"
        },
        "capabilities": [
            "Evaluation protocol documentation",
            "Model comparison support",
            "Task taxonomy mapping",
            "Reproducible benchmark anchoring",
            "Relationship graph compatibility"
        ],
        "limitations": [
            "Benchmark results can become stale as models improve and evaluation protocols evolve.",
            "A benchmark measures a defined task scope rather than complete intelligence."
        ],
        "tags": [
            "Benchmark",
            "NLP",
            "Evaluation",
            "Language Understanding"
        ],
        "timeline": [
            {
                "date": "2018",
                "title": "Public release or documentation",
                "description": "GLUE SST-2 appears in public documentation, project records, benchmark descriptions or research references associated with GLUE benchmark."
            }
        ],
        "relationships": [
            {
                "target": "GLUE benchmark",
                "type": "Associated organization",
                "description": "GLUE benchmark is the organization, project community or institutional context associated with GLUE SST-2.",
                "evidence_level": "documentary"
            },
            {
                "target": "Language model evaluation",
                "type": "Domain context",
                "description": "GLUE SST-2 belongs to the Language model evaluation layer of the intelligent-entity registry.",
                "evidence_level": "documentary"
            }
        ],
        "sources": [
            {
                "label": "Official benchmark page",
                "url": "https://gluebenchmark.com/",
                "source_type": "Research Source",
                "verification_status": "verified"
            }
        ]
    }
}
Proof and discussion layer

Contribute to this record

Submit a proof, correction or comment. Public display is moderated. Every submission remains preserved in the export archive.

Submit proof or comment

Approved comments

No approved public comment yet.