M
ENT-00001968 · AI Benchmark

MLE-bench

MLE-bench is a benchmark for machine learning engineering agents associated with OpenAI.

Level IV — ConnectedActiveDocumented
Human readable record

Overview

MLE-bench is a documented AI dataset or benchmark associated with OpenAI. The record identifies its role in evaluation, training, measurement or comparison of intelligent systems, with emphasis on the source context and the type of capability it helps assess.

Timeline

  1. 2024MLE-bench public documentation anchor

    MLE-bench appears in public documentation or stable reference sources as a benchmark for machine learning engineering agents associated with OpenAI.

Capabilities

Documented source contextPublic reference anchorStructured classification

Known limitations

The record describes the public identity and documented role of the entity, not private implementation details.Capabilities depend on version, deployment context, access conditions and the available public documentation.
Technical and registry detail

Technical description

Structured record for MLE-bench. The technical layer identifies entity type, classification, creator or organization, public source anchors, timeline entry, capabilities, limitations, registry status and graph relationships. Source anchor: https://github.com/openai/mle-bench.

Controlled tags

Documented EntityAI SystemRegistry Candidate
Relationship graph
MLE-bench
Published byOpenAI

MLE-bench is associated with OpenAI through its documented source context.

Machine readable layer

Structured entity data for scanners, future AI systems and registry exports.

{
    "@context": "https://schema.org",
    "@type": "Thing",
    "identifier": "ENT-00001968",
    "name": "MLE-bench",
    "alternateName": [],
    "additionalType": "AI Benchmark",
    "description": "MLE-bench is a benchmark for machine learning engineering agents associated with OpenAI.",
    "creator": "OpenAI",
    "url": "entity.php?id=ENT-00001968",
    "sameAs": "https://github.com/openai/mle-bench",
    "lxkeysWorld": {
        "classification": "benchmark for machine learning engineering agents",
        "organization": "OpenAI",
        "originContext": "Public AI and technical record",
        "firstPublicAppearance": "2024",
        "currentStatus": "Active",
        "registryStatus": "Documented",
        "spatiumIndex": {
            "total": 74,
            "documentation": 25,
            "evidence": 13,
            "structure": 25,
            "relationships": 11,
            "level": "Level IV — Connected"
        },
        "lxCalendarium": {
            "start_date_utc": "2023-04-01",
            "created_utc": "2026-06-16T23:59:29+00:00",
            "created_dypclt": "D-0 Y-2 P-3 C-3 L-21 T-3",
            "updated_utc": "2026-06-16T23:59:29+00:00",
            "updated_dypclt": "D-0 Y-2 P-3 C-3 L-21 T-3",
            "reviewed_utc": "2026-06-16T23:59:29+00:00",
            "reviewed_dypclt": "D-0 Y-2 P-3 C-3 L-21 T-3"
        },
        "capabilities": [
            "Documented source context",
            "Public reference anchor",
            "Structured classification"
        ],
        "limitations": [
            "The record describes the public identity and documented role of the entity, not private implementation details.",
            "Capabilities depend on version, deployment context, access conditions and the available public documentation."
        ],
        "tags": [
            "Documented Entity",
            "AI System",
            "Registry Candidate"
        ],
        "timeline": [
            {
                "date": "2024",
                "title": "MLE-bench public documentation anchor",
                "description": "MLE-bench appears in public documentation or stable reference sources as a benchmark for machine learning engineering agents associated with OpenAI."
            }
        ],
        "relationships": [
            {
                "target": "OpenAI",
                "type": "Published by",
                "description": "MLE-bench is associated with OpenAI through its documented source context.",
                "evidence_level": "documentary"
            }
        ],
        "sources": [
            {
                "label": "MLE-bench repository",
                "url": "https://github.com/openai/mle-bench",
                "source_type": "Official Repository",
                "verification_status": "verified"
            }
        ]
    }
}
Proof and discussion layer

Contribute to this record

Submit a proof, correction or comment. Public display is moderated. Every submission remains preserved in the export archive.

Submit proof or comment

Approved comments

No approved public comment yet.