BA
ENT-00001849 · Benchmark

BabyAI

BabyAI is a AI benchmark associated with Meta AI, classified in LXKeys.world as AI Benchmark / Evaluation.

Niveau IV ConnectéActifDocumenté
Fiche publique

Vue d’ensemble

BabyAI is a documented AI dataset or benchmark associated with Meta AI. The record identifies its role in evaluation, training, measurement or comparison of intelligent systems, with emphasis on the source context and the type of capability it helps assess.

Chronologie

  1. 2018Initial benchmark release

    BabyAI entered the documented public record in 2018. This event is retained at the precision supported by the Entry’s reviewed source history.

Capacités

Reinforcement learningEmbodied evaluationResearch reference

Limites connues

Benchmark scores depend on the exact dataset version, prompt or evaluation protocol, scoring implementation and contamination controls.Leaderboard performance should not be treated as a complete measure of real-world capability or safety.Benchmark relevance can decline as models, data and evaluation practices evolve.
Détail technique et structuré

Description technique

Structured LXKeys.world registry record for BabyAI. Entity type: Benchmark; classification: AI Benchmark / Evaluation; creator/organization context: Meta AI. Canonical source anchor: https://github.com/mila-iqia/babyai. The record tracks source authority, first public appearance, timeline, capabilities, limitations, registry status and graph relationships. Automatic refresh is limited to sources explicitly classified as OFFICIAL and enabled for updates; documentary and research references remain non-authoritative unless reviewed.

Tags contrôlés

RoboticsBenchmarkAgentic Systems
Graphe relationnel

BabyAI

Ouvrir dans le graphe complet
BabyAI
Publié parSortante
Meta AI

BabyAI is associated with Meta AI through its documented source context.

Couche lisible par machine

Données structurées de l’entrée pour les outils humains, les systèmes IA et les clients machine.

{
    "@context": [
        "https://schema.org",
        {
            "lxw": "https://lxkeys.world/schema/"
        }
    ],
    "@type": "Thing",
    "identifier": "ENT-00001849",
    "name": "BabyAI",
    "alternateName": [],
    "additionalType": {
        "category": "Data and Evaluation",
        "type": "Benchmark",
        "subtype": "",
        "lxkeysEntity": false
    },
    "description": "BabyAI is a AI benchmark associated with Meta AI, classified in LXKeys.world as AI Benchmark / Evaluation.",
    "creator": "Meta AI",
    "url": "https://lxkeys.world/entry.php?id=ENT-00001849&lang=fr",
    "sameAs": "https://github.com/mila-iqia/babyai",
    "image": "",
    "lxkeysWorld": {
        "worldId": "ENT-00001849",
        "kind": "Data and Evaluation",
        "type": "Benchmark",
        "subtype": "",
        "classification": "AI Benchmark / Evaluation",
        "organization": "Meta AI",
        "originContext": "Public AI and technical record",
        "firstPublicAppearance": "2018",
        "currentStatus": "Active",
        "documentationStatus": "Documented",
        "documentationIndex": {
            "total": 74,
            "documentation": 25,
            "evidence": 13,
            "structure": 25,
            "relationships": 11,
            "level": "Level IV Connected"
        },
        "lxCalendarium": {
            "start_date_utc": "2023-04-01",
            "created_utc": "2026-06-16T23:59:29+00:00",
            "created_dypclt": "D-0 Y-2 P-3 C-3 L-21 T-3",
            "updated_utc": "2026-06-16T23:59:29+00:00",
            "updated_dypclt": "D-0 Y-2 P-3 C-3 L-21 T-3"
        },
        "facts": [],
        "capabilities": [
            "Reinforcement learning",
            "Embodied evaluation",
            "Research reference"
        ],
        "limitations": [
            "Benchmark scores depend on the exact dataset version, prompt or evaluation protocol, scoring implementation and contamination controls.",
            "Leaderboard performance should not be treated as a complete measure of real-world capability or safety.",
            "Benchmark relevance can decline as models, data and evaluation practices evolve."
        ],
        "tags": [
            "Robotics",
            "Benchmark",
            "Agentic Systems"
        ],
        "timeline": [
            {
                "date": "2018",
                "title": "Initial benchmark release",
                "description": "BabyAI entered the documented public record in 2018. This event is retained at the precision supported by the Entry’s reviewed source history.",
                "source_url": "https://github.com/mila-iqia/babyai",
                "verification_status": "source_backed_curated_baseline"
            }
        ],
        "relationships": [
            {
                "target": "Meta AI",
                "type": "Published by",
                "description": "BabyAI is associated with Meta AI through its documented source context.",
                "evidence_level": "documentary",
                "target_id": "ENT-00000033"
            }
        ],
        "sources": [
            {
                "label": "Official project page or repository",
                "url": "https://github.com/mila-iqia/babyai",
                "source_type": "Official Repository",
                "verification_status": "verified",
                "authority": "TRUSTED_PRIMARY",
                "role": "code_repository",
                "update_enabled": false,
                "authority_basis": "curated-corpus-refresh-2026-09-08"
            }
        ],
        "canonical": [],
        "imageMeta": []
    }
}
Preuves

Contribuer à cette entrée

Soumettez une source, une correction ou un commentaire. Les modifications publiques restent modérées.

Soumettre une preuve ou un commentaire

Commentaires approuvés

Aucun commentaire public approuvé pour le moment.