WE
ENT-00001521 · Benchmark

WeatherBench

WeatherBench is a AI benchmark associated with Research community, classified in LXKeys.world as AI Benchmark / Evaluation.

Level IV ConnectedActiveDocumented
Human record

Overview

WeatherBench is a documented AI dataset or benchmark associated with Research community. The record identifies its role in evaluation, training, measurement or comparison of intelligent systems, with emphasis on the source context and the type of capability it helps assess.

Timeline

  1. 2020Initial benchmark release

    WeatherBench entered the documented public record in 2020. This event is retained at the precision supported by the Entry’s reviewed source history.

Capabilities

Scientific predictionResearch analysisDomain-specific modeling

Known limitations

Benchmark scores depend on the exact dataset version, prompt or evaluation protocol, scoring implementation and contamination controls.Leaderboard performance should not be treated as a complete measure of real-world capability or safety.Benchmark relevance can decline as models, data and evaluation practices evolve.
Technical and structured detail

Technical description

Structured LXKeys.world registry record for WeatherBench. Entity type: Benchmark; classification: AI Benchmark / Evaluation; creator/organization context: WeatherBench authors. Canonical source anchor: https://github.com/google-research/weatherbench. The record tracks source authority, first public appearance, timeline, capabilities, limitations, registry status and graph relationships. Automatic refresh is limited to sources explicitly classified as OFFICIAL and enabled for updates; documentary and research references remain non-authoritative unless reviewed.

Controlled tags

Scientific AIResearch SystemAI Model
Relationship graph

WeatherBench

Open in full graph
WeatherBench
Published byOutgoing
Research community

WeatherBench is associated with Research community through its documented source context.

Machine readable layer

Structured entry data for human tools, AI systems and machine clients.

{
    "@context": [
        "https://schema.org",
        {
            "lxw": "https://lxkeys.world/schema/"
        }
    ],
    "@type": "Thing",
    "identifier": "ENT-00001521",
    "name": "WeatherBench",
    "alternateName": [],
    "additionalType": {
        "category": "Data and Evaluation",
        "type": "Benchmark",
        "subtype": "",
        "lxkeysEntity": false
    },
    "description": "WeatherBench is a AI benchmark associated with Research community, classified in LXKeys.world as AI Benchmark / Evaluation.",
    "creator": "WeatherBench authors",
    "url": "https://lxkeys.world/entry.php?id=ENT-00001521&lang=en",
    "sameAs": "https://github.com/google-research/weatherbench",
    "image": "",
    "lxkeysWorld": {
        "worldId": "ENT-00001521",
        "kind": "Data and Evaluation",
        "type": "Benchmark",
        "subtype": "",
        "classification": "AI Benchmark / Evaluation",
        "organization": "Research community",
        "originContext": "Public AI and technical record",
        "firstPublicAppearance": "2020",
        "currentStatus": "Active",
        "documentationStatus": "Documented",
        "documentationIndex": {
            "total": 74,
            "documentation": 25,
            "evidence": 13,
            "structure": 25,
            "relationships": 11,
            "level": "Level IV Connected"
        },
        "lxCalendarium": {
            "start_date_utc": "2023-04-01",
            "created_utc": "2026-06-16T23:59:28+00:00",
            "created_dypclt": "D-0 Y-2 P-3 C-3 L-21 T-3",
            "updated_utc": "2026-06-16T23:59:28+00:00",
            "updated_dypclt": "D-0 Y-2 P-3 C-3 L-21 T-3"
        },
        "facts": [],
        "capabilities": [
            "Scientific prediction",
            "Research analysis",
            "Domain-specific modeling"
        ],
        "limitations": [
            "Benchmark scores depend on the exact dataset version, prompt or evaluation protocol, scoring implementation and contamination controls.",
            "Leaderboard performance should not be treated as a complete measure of real-world capability or safety.",
            "Benchmark relevance can decline as models, data and evaluation practices evolve."
        ],
        "tags": [
            "Scientific AI",
            "Research System",
            "AI Model"
        ],
        "timeline": [
            {
                "date": "2020",
                "title": "Initial benchmark release",
                "description": "WeatherBench entered the documented public record in 2020. This event is retained at the precision supported by the Entry’s reviewed source history.",
                "source_url": "https://github.com/google-research/weatherbench",
                "verification_status": "source_backed_curated_baseline"
            }
        ],
        "relationships": [
            {
                "target": "Research community",
                "type": "Published by",
                "description": "WeatherBench is associated with Research community through its documented source context.",
                "evidence_level": "documentary"
            }
        ],
        "sources": [
            {
                "label": "WeatherBench repository",
                "url": "https://github.com/google-research/weatherbench",
                "source_type": "Official Repository",
                "verification_status": "verified",
                "authority": "TRUSTED_PRIMARY",
                "role": "code_repository",
                "update_enabled": false,
                "authority_basis": "curated-corpus-refresh-2026-09-08"
            }
        ],
        "canonical": [],
        "imageMeta": []
    }
}
Proofs

Contribute to this record

Submit a source, correction or comment. Public changes remain moderated.

Submit proof or comment

Approved comments

No approved public comment yet.