AP
ENT-00001081 · Benchmark

APPS

APPS is a AI benchmark associated with APPS authors, classified in LXKeys.world as AI Benchmark / Evaluation.

Level IV ConnectedActiveDocumented
Human record

Overview

APPS is a documented introductory programming benchmark used in artificial-intelligence research, evaluation or dataset construction. Its record captures the task context, public source, evaluation role and relationships to models, agents or research systems that rely on comparable measurements.

Timeline

  1. 2021Initial benchmark release

    APPS entered the documented public record in 2021. This event is retained at the precision supported by the Entry’s reviewed source history.

Capabilities

Evaluation referenceComparable measurementResearch reproducibilityModel documentation support

Known limitations

Benchmark interpretation depends on task design, data quality and contamination controls.Scores should not be treated as a complete measure of intelligence.
Technical and structured detail

Technical description

Structured reference record for APPS. The technical layer captures source URL, creator, first-public year, evaluation or dataset role, measurable task family and relationships to model testing, retrieval, safety, coding, reasoning or multimodal assessment.

Controlled tags

BenchmarkEvaluationDocumentationContemporaryBatch 03 Candidate
Relationship graph

APPS

Open in full graph
APPS
Created or maintained byOutgoing
APPS authors

APPS is associated with APPS authors.

Machine readable layer

Structured entry data for human tools, AI systems and machine clients.

{
    "@context": [
        "https://schema.org",
        {
            "lxw": "https://lxkeys.world/schema/"
        }
    ],
    "@type": "Thing",
    "identifier": "ENT-00001081",
    "name": "APPS",
    "alternateName": [],
    "additionalType": {
        "category": "Data and Evaluation",
        "type": "Benchmark",
        "subtype": "",
        "lxkeysEntity": false
    },
    "description": "APPS is a AI benchmark associated with APPS authors, classified in LXKeys.world as AI Benchmark / Evaluation.",
    "creator": "APPS authors",
    "url": "https://lxkeys.world/entry.php?id=ENT-00001081&lang=en",
    "sameAs": "https://arxiv.org/abs/2105.09938",
    "image": "",
    "lxkeysWorld": {
        "worldId": "ENT-00001081",
        "kind": "Data and Evaluation",
        "type": "Benchmark",
        "subtype": "",
        "classification": "AI Benchmark / Evaluation",
        "organization": "APPS authors",
        "originContext": "AI evaluation and dataset record",
        "firstPublicAppearance": "2021",
        "currentStatus": "Active",
        "documentationStatus": "Documented",
        "documentationIndex": {
            "total": 74,
            "documentation": 25,
            "evidence": 13,
            "structure": 25,
            "relationships": 11,
            "level": "Level IV Connected"
        },
        "lxCalendarium": {
            "start_date_utc": "2023-04-01",
            "created_utc": "2026-06-16T23:59:14+00:00",
            "created_dypclt": "D-0 Y-2 P-3 C-3 L-21 T-3",
            "updated_utc": "2026-06-16T23:59:14+00:00",
            "updated_dypclt": "D-0 Y-2 P-3 C-3 L-21 T-3"
        },
        "facts": [],
        "capabilities": [
            "Evaluation reference",
            "Comparable measurement",
            "Research reproducibility",
            "Model documentation support"
        ],
        "limitations": [
            "Benchmark interpretation depends on task design, data quality and contamination controls.",
            "Scores should not be treated as a complete measure of intelligence."
        ],
        "tags": [
            "Benchmark",
            "Evaluation",
            "Documentation",
            "Contemporary",
            "Batch 03 Candidate"
        ],
        "timeline": [
            {
                "date": "2021",
                "title": "Initial benchmark release",
                "description": "APPS entered the documented public record in 2021. This event is retained at the precision supported by the Entry’s reviewed source history.",
                "source_url": "https://arxiv.org/abs/2105.09938",
                "verification_status": "source_backed_curated_baseline"
            }
        ],
        "relationships": [
            {
                "target": "APPS authors",
                "type": "Created or maintained by",
                "description": "APPS is associated with APPS authors.",
                "evidence_level": "documentary"
            }
        ],
        "sources": [
            {
                "label": "APPS public reference",
                "url": "https://arxiv.org/abs/2105.09938",
                "source_type": "Primary or Research Source",
                "verification_status": "verified",
                "authority": "TRUSTED_PRIMARY",
                "role": "research_paper",
                "update_enabled": false,
                "authority_basis": "curated-corpus-refresh-2026-09-08"
            }
        ],
        "canonical": [],
        "imageMeta": []
    }
}
Proofs

Contribute to this record

Submit a source, correction or comment. Public changes remain moderated.

Submit proof or comment

Approved comments

No approved public comment yet.