BLXBenchBLXBench UI
blxbench

Benchmark

Misc

DocsOur TestsPassSponsor / Partnership
DocsOur TestsPassSponsor / Partnership
BLXBenchBLXBench UI
blxbench

Benchmark

Suite

Misc

DocsOur TestsPassSponsor / Partnership
DocsOur TestsPassSponsor / Partnership
  1. Home
  2. Our Tests
  3. Reason-Rc-Cache-Stale-After-Write
blxbench

Test fixture

Reason-Rc-Cache-Stale-After-Write

Reasoningv2 — Resilienceeasyscorer: rubric_json_metrics

Arithmetic, symbolic steps, and structured problem solving.

How it is scored

The model receives the prompt (and optional system message). The run uses scorer rubric_json_metrics with the JSON configuration below. Pass/fail and partial credit are determined entirely by that scorer against the model output; no human grading.

User prompt
Return JSON only with keys answer, evidence, constraints. A system uses a read-through cache. A write operation goes directly to the database but does not update or invalidate the cache. The next read returns data from the cache. What is the root cause of the stale data being returned?
Scorer config
{
  "metrics": {
    "accuracy": {
      "checks": [
        {
          "contains": [
            "cache invalidation"
          ]
        },
        {
          "contains": [
            "stale"
          ]
        },
        {
          "contains": [
            "write-through"
          ]
        }
      ]
    },
    "evidence": {
      "checks": [
        {
          "contains": [
            "read-through"
          ]
        },
        {
          "contains": [
            "directly to DB"
          ]
        },
        {
          "contains": [
            "not updated"
          ]
        }
      ]
    },
    "constraint": {
      "checks": [
        {
          "contains": [
            "invalidate"
          ]
        },
        {
          "contains": [
            "write-through not"
          ]
        },
        {
          "contains": [
            "cache consistency"
          ]
        }
      ]
    }
  }
}
Run parameters

temperature

0

max_tokens

500

timeout (s)

120

type

scored

file

reason-rc-cache-stale-after-write.json

← PreviousReason-Constraint-Subscription-Migration
|
Next →Reason-Rc-Login-Timeout

BLXBench

Community driven leaderboardPublic benchmark runner — run in your environment, share results with the community.

© 2026 BLXBench by bitslix.com

ProvenanceAggregated from user runs
Scope41 / 11 / 490
Latestrun_3d5451 / 459 / $1.75
TermsPrivacy