q900001

R-ORD GOLD-ONLY not-a-function-of-the-data

questions

The question

What was spent in each category, amount by amount?

the hint the set supplies: the amounts spent refers to the spent column of every row of the category

This question was audited without a prediction beside it, so there is nothing to compare the gold with. The probes below read the gold alone.

The statements

gold

SELECT category, group_concat(spent) FROM spend GROUP BY category ORDER BY category ASC
category
ascending

sha256:249327aa613f600d24ddb035ce97d70bd614e95f0920ea5310ebe1363eee1e2d

The results

gold, 3 rows

from evidence-gold.json, 3 rows

categoryTEXT group_concat(spent)TEXT
depot 0.19,0.28,0.37
field 1000000.0,0.11,0.22,0.33,0.44,0.55,0.66,0.77,0.88
yard 0.46,0.55,0.64

The probes

A smell is a mechanical reason to read this gold statement again. It is a heuristic: it does not state that the statement is wrong, and a maintainer decides.

ordering-over-numeric-text quiet

this statement orders by a text column holding only numbers, and ordering it as a number gives a different answer, so the gold may be sorting 9.5 above 10

what it measured
{
  "heuristic": true,
  "keys": [
    {
      "key": "category",
      "column": "spend.category",
      "declared_type": "TEXT",
      "not_applicable": null,
      "census": {
        "rows": 15,
        "nulls": 0,
        "empty_strings": 0,
        "non_numeric": 15,
        "pattern": "^-?[0-9]+(\\.[0-9]+)?$"
      },
      "every_value_is_numeric": false
    }
  ]
}

arbitrary-cut not applicable

this statement cuts its result at a LIMIT that does not decide which rows come back, so a different but equally correct statement can return other rows and score zero

the statement states no LIMIT

what it measured
{
  "heuristic": true,
  "reason": "the statement states no LIMIT"
}

not-a-function-of-the-data fired

rerun over the same rows in another physical order this statement gives another answer, so its result depends on how the rows are stored and not only on the data

from smells.json, 3 rows

depot 0.28,0.19,0.37
field 0.66,0.33,0.77,0.55,0.44,0.11,0.22,0.88,1000000.0
yard 0.64,0.55,0.46
what it measured
{
  "heuristic": true,
  "rule": "R-ORD",
  "baseline_result_hash": "sha256:249327aa613f600d24ddb035ce97d70bd614e95f0920ea5310ebe1363eee1e2d",
  "baseline_result": {
    "columns": [
      {
        "name": "category",
        "declared_type": "TEXT"
      },
      {
        "name": "group_concat(spent)",
        "declared_type": "TEXT"
      }
    ],
    "row_count": 3,
    "truncated": false,
    "rows_shown": 3,
    "rows": [
      [
        {
          "type": "str",
          "value": "depot"
        },
        {
          "type": "str",
          "value": "0.19,0.28,0.37"
        }
      ],
      [
        {
          "type": "str",
          "value": "field"
        },
        {
          "type": "str",
          "value": "1000000.0,0.11,0.22,0.33,0.44,0.55,0.66,0.77,0.88"
        }
      ],
      [
        {
          "type": "str",
          "value": "yard"
        },
        {
          "type": "str",
          "value": "0.46,0.55,0.64"
        }
      ]
    ],
    "result_hash": "sha256:249327aa613f600d24ddb035ce97d70bd614e95f0920ea5310ebe1363eee1e2d"
  },
  "planner_statistics": {},
  "shuffle": {
    "seed": "1",
    "row_limit": 300000,
    "tables": [
      "spend"
    ],
    "tables_not_shuffled": [],
    "tables_skipped_for_size": {},
    "tables_not_reached_by_a_copy": {
      "main.y": "the statement names this table's schema, and a name that states its schema is read from that schema whatever TEMP holds, so the rerun reads this table and not a copy of it"
    }
  },
  "shuffled_copies": {
    "run": true,
    "verdict": "not_equal",
    "differs": true,
    "result_hash": "sha256:d65d173fb4451cf2c500b1c6f1865e3e296e7361d1a6fed76e316fa1321d4910",
    "result": {
      "columns": [
        {
          "name": "category",
          "declared_type": "TEXT"
        },
        {
          "name": "group_concat(spent)",
          "declared_type": "TEXT"
        }
      ],
      "row_count": 3,
      "truncated": false,
      "rows_shown": 3,
      "rows": [
        [
          {
            "type": "str",
            "value": "depot"
          },
          {
            "type": "str",
            "value": "0.28,0.19,0.37"
          }
        ],
        [
          {
            "type": "str",
            "value": "field"
          },
          {
            "type": "str",
            "value": "0.66,0.33,0.77,0.55,0.44,0.11,0.22,0.88,1000000.0"
          }
        ],
        [
          {
            "type": "str",
            "value": "yard"
          },
          {
            "type": "str",
            "value": "0.64,0.55,0.46"
          }
        ]
      ],
      "result_hash": "sha256:d65d173fb4451cf2c500b1c6f1865e3e296e7361d1a6fed76e316fa1321d4910"
    }
  },
  "plan_variant": {
    "run": false,
    "reason": "the plan variant was not asked for"
  }
}

duplicate-full-row quiet

this statement returns the same whole row more than once and never says DISTINCT, so a statement answering the same question once per row disagrees on multiplicity alone

what it measured
{
  "heuristic": true,
  "rows": 3,
  "distinct_rows": 3,
  "repeated_rows": 0,
  "largest_repeat": 1,
  "result_bounded": false,
  "distinct_stated": false,
  "set_operation": false
}

The evidence records

gold: evidence-gold.json

SELECT category, group_concat(spent) FROM spend GROUP BY category ORDER BY category ASC
statement read from
demo/questions.json
digest
sha256:7ad58addc070b425cabfc9d67bbd9c95840253160cf3b73d672e9aeb4897f5ff
origin
the run was told none

result_hash sha256:249327aa613f600d24ddb035ce97d70bd614e95f0920ea5310ebe1363eee1e2d recomputed from this JSON: match

record_hash sha256:32284bcf6002dfd10ae08d6d7b66eeef831c27d5d837dae5583a8ca82a212a9f recomputed from this JSON: match

the result this record holds, 3 rows

from evidence-gold.json, 3 rows

categoryTEXT group_concat(spent)TEXT
depot 0.19,0.28,0.37
field 1000000.0,0.11,0.22,0.33,0.44,0.55,0.66,0.77,0.88
yard 0.46,0.55,0.64
what ran, and where
run
audit-abdcc660-cce4-4f7a-a477-0b444aef6d1d
executed at
2026-09-14T13:13:54.085413+00:00
data as of
2026-09-14T13:13:54.071205+00:00
backend at checkout
SQLite 3.53.1 | file=/tmp/attestql-site-sandbox/demo/fixture.sqlite | size=69632
backend that answered
SQLite 3.53.1 | file=/tmp/attestql-site-sandbox/demo/fixture.sqlite | size=69632
database role
file
replay rule
R-ORD
question set version
sha256:7ad58addc070b425cabfc9d67bbd9c95840253160cf3b73d672e9aeb4897f5ff
validator
audit:sqlglot-sqlite-parse
checks run
parses_as_exactly_one_statement, the_one_statement_is_a_select, no_placeholder_without_a_bound_parameter
statement timeout
30000 ms
rows
3 rows
the session it ran under
engine
sqlite
time_zone
not stated by this engine
date_style
not stated by this engine
interval_style
not stated by this engine
extra_float_digits
not stated by this engine
database_collation
not stated by this engine
work_mem
not stated by this engine
hash_mem_multiplier
not stated by this engine

recorded beside them

sqlite_version
3.53.1
encoding
UTF-8
reverse_unordered_selects
0
query_only
1
journal_mode
delete
data_version
1
automatic_index
1
compile_options
ATOMIC_INTRINSICS=1,COMPILER=clang-22.1.3,DEFAULT_AUTOVACUUM,DEFAULT_CACHE_SIZE=-2000,DEFAULT_FILE_FORMAT=4,DEFAULT_JOURNAL_SIZE_LIMIT=-1,DEFAULT_MMAP_SIZE=0,DEFAULT_PAGE_SIZE=4096,DEFAULT_PCACHE_INITSZ=20,DEFAULT_RECURSIVE_TRIGGERS,DEFAULT_SECTOR_SIZE=4096,DEFAULT_SYNCHRONOUS=2,DEFAULT_WAL_AUTOCHECKPOINT=1000,DEFAULT_WAL_SYNCHRONOUS=2,DEFAULT_WORKER_THREADS=0,DIRECT_OVERFLOW_READ,ENABLE_DBSTAT_VTAB,ENABLE_FTS3,ENABLE_FTS3_PARENTHESIS,ENABLE_FTS4,ENABLE_FTS5,ENABLE_GEOPOLY,ENABLE_MATH_FUNCTIONS,ENABLE_PERCENTILE,ENABLE_RTREE,MALLOC_SOFT_LIMIT=1024,MAX_ATTACHED=10,MAX_COLUMN=2000,MAX_COMPOUND_SELECT=500,MAX_DEFAULT_PAGE_SIZE=8192,MAX_EXPR_DEPTH=1000,MAX_FUNCTION_ARG=1000,MAX_LENGTH=1000000000,MAX_LIKE_PATTERN_LENGTH=50000,MAX_MMAP_SIZE=0x7fff0000,MAX_PAGE_COUNT=0xfffffffe,MAX_PAGE_SIZE=65536,MAX_SQL_LENGTH=1000000000,MAX_TRIGGER_DEPTH=1000,MAX_VARIABLE_NUMBER=32766,MAX_VDBE_OP=250000000,MAX_WORKER_THREADS=8,MUTEX_PTHREADS,SYSTEM_MALLOC,TEMP_STORE=1,THREADSAFE=1
collation_list
RTRIM,NOCASE,BINARY
case_sensitive_like
0
the rendering and the data
version
attestql/audit/4
numeric_scale
6
timestamp_format
%Y-%m-%dT%H:%M:%S.%fZ
timezone
UTC
null_rendering
NULL
encoding
utf-8
schema digest
sha256:a6a04d658da717d89c74d98599f7e67d02c58b294d9c52ccec4719206cd9cb15
source file sha256
rows in main.spend
15

Running these again

gold

re-run this statement read-only against SQLite 3.53.1 | file=/tmp/attestql-site-sandbox/demo/fixture.sqlite | size=69632 under the session settings and over the data this record's fixture digest names, and compare the two results under R-ORD

This question's run

run
audit-abdcc660-cce4-4f7a-a477-0b444aef6d1d
server
SQLite 3.53.1 | file=/tmp/attestql-site-sandbox/demo/fixture.sqlite | size=69632
question set
questions
replay rule
R-ORD

the run this question belongs to

The JSON this page was rendered from