HL
ENT-00001078 · Evaluation System

Humanity's Last Exam

Humanity's Last Exam is a expert-level multimodal benchmark associated with Center for AI Safety.

Level IV — ConnectedActiveConnected
Human readable record

Overview

Humanity's Last Exam is a documented expert-level multimodal benchmark used in artificial-intelligence research, evaluation or dataset construction. Its record captures the task context, public source, evaluation role and relationships to models, agents or research systems that rely on comparable measurements.

Timeline

  1. 2025Humanity's Last Exam public release or documentation

    Humanity's Last Exam appears in public documentation as a expert-level multimodal benchmark associated with Center for AI Safety.

Capabilities

Evaluation referenceComparable measurementResearch reproducibilityModel documentation support

Known limitations

Benchmark interpretation depends on task design, data quality and contamination controls.Scores should not be treated as a complete measure of intelligence.
Technical and registry detail

Technical description

Structured reference record for Humanity's Last Exam. The technical layer captures source URL, creator, first-public year, evaluation or dataset role, measurable task family and relationships to model testing, retrieval, safety, coding, reasoning or multimodal assessment.

Controlled tags

BenchmarkEvaluationDocumentationBatch 03 Candidate
Relationship graph
Humanity's Last Exam
Created or maintained byCenter for AI Safety

Humanity's Last Exam is associated with Center for AI Safety.

Machine readable layer

Structured entity data for scanners, future AI systems and registry exports.

{
    "@context": "https://schema.org",
    "@type": "Thing",
    "identifier": "ENT-00001078",
    "name": "Humanity's Last Exam",
    "alternateName": [],
    "additionalType": "Evaluation System",
    "description": "Humanity's Last Exam is a expert-level multimodal benchmark associated with Center for AI Safety.",
    "creator": "Center for AI Safety",
    "url": "entity.php?id=ENT-00001078",
    "sameAs": "https://lastexam.ai/",
    "lxkeysWorld": {
        "classification": "Expert-level multimodal benchmark",
        "organization": "Center for AI Safety",
        "originContext": "AI evaluation and dataset record",
        "firstPublicAppearance": "2025",
        "currentStatus": "Active",
        "registryStatus": "Connected",
        "spatiumIndex": {
            "total": 74,
            "documentation": 25,
            "evidence": 13,
            "structure": 25,
            "relationships": 11,
            "level": "Level IV — Connected"
        },
        "lxCalendarium": {
            "start_date_utc": "2023-04-01",
            "created_utc": "2026-06-16T23:59:14+00:00",
            "created_dypclt": "D-0 Y-2 P-3 C-3 L-21 T-3",
            "updated_utc": "2026-06-16T23:59:14+00:00",
            "updated_dypclt": "D-0 Y-2 P-3 C-3 L-21 T-3",
            "reviewed_utc": "2026-06-16T23:59:14+00:00",
            "reviewed_dypclt": "D-0 Y-2 P-3 C-3 L-21 T-3"
        },
        "capabilities": [
            "Evaluation reference",
            "Comparable measurement",
            "Research reproducibility",
            "Model documentation support"
        ],
        "limitations": [
            "Benchmark interpretation depends on task design, data quality and contamination controls.",
            "Scores should not be treated as a complete measure of intelligence."
        ],
        "tags": [
            "Benchmark",
            "Evaluation",
            "Documentation",
            "Batch 03 Candidate"
        ],
        "timeline": [
            {
                "date": "2025",
                "title": "Humanity's Last Exam public release or documentation",
                "description": "Humanity's Last Exam appears in public documentation as a expert-level multimodal benchmark associated with Center for AI Safety."
            }
        ],
        "relationships": [
            {
                "target": "Center for AI Safety",
                "type": "Created or maintained by",
                "description": "Humanity's Last Exam is associated with Center for AI Safety.",
                "evidence_level": "documentary"
            }
        ],
        "sources": [
            {
                "label": "Humanity's Last Exam public reference",
                "url": "https://lastexam.ai/",
                "source_type": "Primary or Research Source",
                "verification_status": "verified"
            }
        ]
    }
}
Proof and discussion layer

Contribute to this record

Submit a proof, correction or comment. Public display is moderated. Every submission remains preserved in the export archive.

Submit proof or comment

Approved comments

No approved public comment yet.