QU
ENT-00002826 · Benchmark

QuixBugs

QuixBugs is a AI benchmark associated with QuixBugs contributors, classified in LXKeys.world as AI Benchmark / Evaluation.

Level V PersistentActiveDocumented
Human record

Overview

QuixBugs is a documented benchmark or dataset used to evaluate code intelligence, program synthesis, debugging, semantic parsing or data reasoning. It is recorded as program repair benchmark.

Timeline

  1. 2017Initial benchmark release

    QuixBugs entered the documented public record in 2017. This event is retained at the precision supported by the Entry’s reviewed source history.

Capabilities

Evaluation protocol documentationModel comparison supportTask taxonomy mappingReproducible benchmark anchoringRelationship graph compatibility

Known limitations

Benchmark results can become stale as models improve and evaluation protocols evolve.A benchmark measures a defined task scope rather than complete intelligence.
Technical and structured detail

Technical description

Structured LXKeys.world registry record for QuixBugs. Entity type: Benchmark; classification: AI Benchmark / Evaluation; creator/organization context: QuixBugs contributors. Canonical source anchor: https://github.com/jkoppel/QuixBugs. The record tracks source authority, first public appearance, timeline, capabilities, limitations, registry status and graph relationships. Automatic refresh is limited to sources explicitly classified as OFFICIAL and enabled for updates; documentary and research references remain non-authoritative unless reviewed.

Controlled tags

BenchmarkCode IntelligenceSoftware EngineeringEvaluation
Relationship graph

QuixBugs

Open in full graph
QuixBugs
Associated organizationOutgoing
QuixBugs contributors

QuixBugs contributors is the organization, project community or institutional context associated with QuixBugs.

Domain contextOutgoing
Code and data reasoning evaluation

QuixBugs belongs to the Code and data reasoning evaluation layer of the intelligent-entity registry.

Machine readable layer

Structured entry data for human tools, AI systems and machine clients.

{
    "@context": [
        "https://schema.org",
        {
            "lxw": "https://lxkeys.world/schema/"
        }
    ],
    "@type": "Thing",
    "identifier": "ENT-00002826",
    "name": "QuixBugs",
    "alternateName": [],
    "additionalType": {
        "category": "Data and Evaluation",
        "type": "Benchmark",
        "subtype": "",
        "lxkeysEntity": false
    },
    "description": "QuixBugs is a AI benchmark associated with QuixBugs contributors, classified in LXKeys.world as AI Benchmark / Evaluation.",
    "creator": "QuixBugs contributors",
    "url": "https://lxkeys.world/entry.php?id=ENT-00002826&lang=en",
    "sameAs": "https://github.com/jkoppel/QuixBugs",
    "image": "",
    "lxkeysWorld": {
        "worldId": "ENT-00002826",
        "kind": "Data and Evaluation",
        "type": "Benchmark",
        "subtype": "",
        "classification": "AI Benchmark / Evaluation",
        "organization": "QuixBugs contributors",
        "originContext": "Code and data reasoning evaluation",
        "firstPublicAppearance": "2017",
        "currentStatus": "Active",
        "documentationStatus": "Documented",
        "documentationIndex": {
            "total": 82,
            "documentation": 25,
            "evidence": 13,
            "structure": 25,
            "relationships": 19,
            "level": "Level V Persistent"
        },
        "lxCalendarium": {
            "start_date_utc": "2023-04-01",
            "created_utc": "2026-06-17T01:48:33+00:00",
            "created_dypclt": "D-0 Y-2 P-3 C-3 L-22 T-4",
            "updated_utc": "2026-06-17T01:48:33+00:00",
            "updated_dypclt": "D-0 Y-2 P-3 C-3 L-22 T-4"
        },
        "facts": [],
        "capabilities": [
            "Evaluation protocol documentation",
            "Model comparison support",
            "Task taxonomy mapping",
            "Reproducible benchmark anchoring",
            "Relationship graph compatibility"
        ],
        "limitations": [
            "Benchmark results can become stale as models improve and evaluation protocols evolve.",
            "A benchmark measures a defined task scope rather than complete intelligence."
        ],
        "tags": [
            "Benchmark",
            "Code Intelligence",
            "Software Engineering",
            "Evaluation"
        ],
        "timeline": [
            {
                "date": "2017",
                "title": "Initial benchmark release",
                "description": "QuixBugs entered the documented public record in 2017. This event is retained at the precision supported by the Entry’s reviewed source history.",
                "source_url": "https://github.com/jkoppel/QuixBugs",
                "verification_status": "source_backed_curated_baseline"
            }
        ],
        "relationships": [
            {
                "target": "QuixBugs contributors",
                "type": "Associated organization",
                "description": "QuixBugs contributors is the organization, project community or institutional context associated with QuixBugs.",
                "evidence_level": "documentary"
            },
            {
                "target": "Code and data reasoning evaluation",
                "type": "Domain context",
                "description": "QuixBugs belongs to the Code and data reasoning evaluation layer of the intelligent-entity registry.",
                "evidence_level": "documentary"
            }
        ],
        "sources": [
            {
                "label": "Official or reference source",
                "url": "https://github.com/jkoppel/QuixBugs",
                "source_type": "Research Source",
                "verification_status": "verified",
                "authority": "TRUSTED_PRIMARY",
                "role": "code_repository",
                "update_enabled": false,
                "authority_basis": "curated-corpus-refresh-2026-09-08"
            }
        ],
        "canonical": [],
        "imageMeta": []
    }
}
Proofs

Contribute to this record

Submit a source, correction or comment. Public changes remain moderated.

Submit proof or comment

Approved comments

No approved public comment yet.