HL
ENT-00001078 · Benchmark

Humanity's Last Exam

Humanity's Last Exam is a AI benchmark associated with Center for AI Safety, classified in LXKeys.world as AI Benchmark / Evaluation.

Niveau IV ConnectéActifDocumenté
Fiche publique

Vue d’ensemble

Humanity's Last Exam is a documented expert-level multimodal benchmark used in artificial-intelligence research, evaluation or dataset construction. Its record captures the task context, public source, evaluation role and relationships to models, agents or research systems that rely on comparable measurements.

Chronologie

  1. 2025Initial benchmark release

    Humanity's Last Exam entered the documented public record in 2025. This event is retained at the precision supported by the Entry’s reviewed source history.

Capacités

Evaluation referenceComparable measurementResearch reproducibilityModel documentation support

Limites connues

Benchmark interpretation depends on task design, data quality and contamination controls.Scores should not be treated as a complete measure of intelligence.
Détail technique et structuré

Description technique

Structured reference record for Humanity's Last Exam. The technical layer captures source URL, creator, first-public year, evaluation or dataset role, measurable task family and relationships to model testing, retrieval, safety, coding, reasoning or multimodal assessment.

Tags contrôlés

BenchmarkEvaluationDocumentationBatch 03 Candidate
Graphe relationnel

Humanity's Last Exam

Ouvrir dans le graphe complet
Humanity's Last Exam
Créé ou maintenu parSortante
Center for AI Safety

Humanity's Last Exam is associated with Center for AI Safety.

Couche lisible par machine

Données structurées de l’entrée pour les outils humains, les systèmes IA et les clients machine.

{
    "@context": [
        "https://schema.org",
        {
            "lxw": "https://lxkeys.world/schema/"
        }
    ],
    "@type": "Thing",
    "identifier": "ENT-00001078",
    "name": "Humanity's Last Exam",
    "alternateName": [],
    "additionalType": {
        "category": "Data and Evaluation",
        "type": "Benchmark",
        "subtype": "",
        "lxkeysEntity": false
    },
    "description": "Humanity's Last Exam is a AI benchmark associated with Center for AI Safety, classified in LXKeys.world as AI Benchmark / Evaluation.",
    "creator": "Center for AI Safety",
    "url": "https://lxkeys.world/entry.php?id=ENT-00001078&lang=fr",
    "sameAs": "https://lastexam.ai/",
    "image": "",
    "lxkeysWorld": {
        "worldId": "ENT-00001078",
        "kind": "Data and Evaluation",
        "type": "Benchmark",
        "subtype": "",
        "classification": "AI Benchmark / Evaluation",
        "organization": "Center for AI Safety",
        "originContext": "AI evaluation and dataset record",
        "firstPublicAppearance": "2025",
        "currentStatus": "Active",
        "documentationStatus": "Documented",
        "documentationIndex": {
            "total": 74,
            "documentation": 25,
            "evidence": 13,
            "structure": 25,
            "relationships": 11,
            "level": "Level IV Connected"
        },
        "lxCalendarium": {
            "start_date_utc": "2023-04-01",
            "created_utc": "2026-06-16T23:59:14+00:00",
            "created_dypclt": "D-0 Y-2 P-3 C-3 L-21 T-3",
            "updated_utc": "2026-06-16T23:59:14+00:00",
            "updated_dypclt": "D-0 Y-2 P-3 C-3 L-21 T-3"
        },
        "facts": [],
        "capabilities": [
            "Evaluation reference",
            "Comparable measurement",
            "Research reproducibility",
            "Model documentation support"
        ],
        "limitations": [
            "Benchmark interpretation depends on task design, data quality and contamination controls.",
            "Scores should not be treated as a complete measure of intelligence."
        ],
        "tags": [
            "Benchmark",
            "Evaluation",
            "Documentation",
            "Batch 03 Candidate"
        ],
        "timeline": [
            {
                "date": "2025",
                "title": "Initial benchmark release",
                "description": "Humanity's Last Exam entered the documented public record in 2025. This event is retained at the precision supported by the Entry’s reviewed source history.",
                "source_url": "https://lastexam.ai/",
                "verification_status": "source_backed_official"
            }
        ],
        "relationships": [
            {
                "target": "Center for AI Safety",
                "type": "Created or maintained by",
                "description": "Humanity's Last Exam is associated with Center for AI Safety.",
                "evidence_level": "documentary"
            }
        ],
        "sources": [
            {
                "label": "Humanity's Last Exam public reference",
                "url": "https://lastexam.ai/",
                "source_type": "Primary or Research Source",
                "verification_status": "verified",
                "authority": "OFFICIAL",
                "role": "primary",
                "update_enabled": true,
                "authority_basis": "curated-corpus-refresh-2026-09-08"
            }
        ],
        "canonical": [],
        "imageMeta": []
    }
}
Preuves

Contribuer à cette entrée

Soumettez une source, une correction ou un commentaire. Les modifications publiques restent modérées.

Soumettre une preuve ou un commentaire

Commentaires approuvés

Aucun commentaire public approuvé pour le moment.