GL
ENT-00001091 · Benchmark

GLUE

GLUE is a AI benchmark associated with NYU / DeepMind / University of Washington, classified in LXKeys.world as AI Benchmark / Evaluation.

Niveau IV ConnectéDocumentéDocumenté
Fiche publique

Vue d’ensemble

GLUE is a documented general language understanding evaluation used in artificial-intelligence research, evaluation or dataset construction. Its record captures the task context, public source, evaluation role and relationships to models, agents or research systems that rely on comparable measurements.

Chronologie

  1. 2018Initial benchmark release

    GLUE entered the documented public record in 2018. This event is retained at the precision supported by the Entry’s reviewed source history.

Capacités

Evaluation referenceComparable measurementResearch reproducibilityModel documentation support

Limites connues

Benchmark interpretation depends on task design, data quality and contamination controls.Scores should not be treated as a complete measure of intelligence.The registry keeps GLUE as the benchmark suite; its constituent tasks are represented as benchmark components.
Détail technique et structuré

Description technique

Structured reference record for GLUE. The technical layer captures source URL, creator, first-public year, evaluation or dataset role, measurable task family and relationships to model testing, retrieval, safety, coding, reasoning or multimodal assessment.

Tags contrôlés

DatasetEvaluationDocumentationEstablishedBatch 03 CandidateCurated Parent Entity
Graphe relationnel

GLUE

Ouvrir dans le graphe complet
GLUE
Créé ou maintenu parSortante
NYU / DeepMind / University of Washington

GLUE is associated with NYU / DeepMind / University of Washington.

Couche lisible par machine

Données structurées de l’entrée pour les outils humains, les systèmes IA et les clients machine.

{
    "@context": [
        "https://schema.org",
        {
            "lxw": "https://lxkeys.world/schema/"
        }
    ],
    "@type": "Thing",
    "identifier": "ENT-00001091",
    "name": "GLUE",
    "alternateName": [],
    "additionalType": {
        "category": "Data and Evaluation",
        "type": "Benchmark",
        "subtype": "",
        "lxkeysEntity": false
    },
    "description": "GLUE is a AI benchmark associated with NYU / DeepMind / University of Washington, classified in LXKeys.world as AI Benchmark / Evaluation.",
    "creator": "NYU / DeepMind / University of Washington",
    "url": "https://lxkeys.world/entry.php?id=ENT-00001091&lang=fr",
    "sameAs": "https://gluebenchmark.com/",
    "image": "",
    "lxkeysWorld": {
        "worldId": "ENT-00001091",
        "kind": "Data and Evaluation",
        "type": "Benchmark",
        "subtype": "",
        "classification": "AI Benchmark / Evaluation",
        "organization": "NYU / DeepMind / University of Washington",
        "originContext": "AI evaluation and dataset record",
        "firstPublicAppearance": "2018",
        "currentStatus": "Documented",
        "documentationStatus": "Documented",
        "documentationIndex": {
            "total": 74,
            "documentation": 25,
            "evidence": 13,
            "structure": 25,
            "relationships": 11,
            "level": "Level IV Connected"
        },
        "lxCalendarium": {
            "start_date_utc": "2023-04-01",
            "created_utc": "2026-06-16T23:59:14+00:00",
            "created_dypclt": "D-0 Y-2 P-3 C-3 L-21 T-3",
            "updated_utc": "2026-06-16T23:59:14+00:00",
            "updated_dypclt": "D-0 Y-2 P-3 C-3 L-21 T-3"
        },
        "facts": [],
        "capabilities": [
            "Evaluation reference",
            "Comparable measurement",
            "Research reproducibility",
            "Model documentation support"
        ],
        "limitations": [
            "Benchmark interpretation depends on task design, data quality and contamination controls.",
            "Scores should not be treated as a complete measure of intelligence.",
            "The registry keeps GLUE as the benchmark suite; its constituent tasks are represented as benchmark components."
        ],
        "tags": [
            "Dataset",
            "Evaluation",
            "Documentation",
            "Established",
            "Batch 03 Candidate",
            "Curated Parent Entity"
        ],
        "timeline": [
            {
                "date": "2018",
                "title": "Initial benchmark release",
                "description": "GLUE entered the documented public record in 2018. This event is retained at the precision supported by the Entry’s reviewed source history.",
                "source_url": "https://gluebenchmark.com/",
                "verification_status": "source_backed_official"
            }
        ],
        "relationships": [
            {
                "target": "NYU / DeepMind / University of Washington",
                "type": "Created or maintained by",
                "description": "GLUE is associated with NYU / DeepMind / University of Washington.",
                "evidence_level": "documentary"
            }
        ],
        "sources": [
            {
                "label": "GLUE public reference",
                "url": "https://gluebenchmark.com/",
                "source_type": "Primary or Research Source",
                "verification_status": "verified",
                "authority": "OFFICIAL",
                "role": "primary",
                "update_enabled": true,
                "authority_basis": "curated-corpus-refresh-2026-09-08"
            }
        ],
        "canonical": [],
        "imageMeta": []
    }
}
Preuves

Contribuer à cette entrée

Soumettez une source, une correction ou un commentaire. Les modifications publiques restent modérées.

Soumettre une preuve ou un commentaire

Commentaires approuvés

Aucun commentaire public approuvé pour le moment.