AI Research Portal
Deutsch/English

What a journal can actually measure

KF04 · Descriptive case report · 2026-10-06

The historical snapshot contains 131 events and 23 known total durations. It describes documentation gaps, not proof of faster or better AI work.

Field completeness in the historical snapshot

6 October 2026 · Denominator: 131 events

Total duration present23 / 131
Measured input tokens present0 / 131
Measured output tokens present0 / 131
Transfer field present0 / 131
Comparison key present1 / 131
Present / not documented. Missing means unknown, not zero use.

What others can put into practice.

  1. Separate missing values from zero. An unmeasured value must not become 0 in a chart.
  2. Record complete total durations with their time basis, including failures and rework.
  3. Define a comparable workload and quality acceptance before the experiment; a result involving several skills does not isolate one skill’s effect.

Question

Which performance claims can a heterogeneous local evidence journal support?

Method

A descriptive analysis of stored completeness aggregates dated 6 October 2026. The public extract contains only counts, observation date and the source fingerprint. Private event text is not published.

Counts capture whether total duration, input tokens, output tokens, a transfer field and a comparison key were present. Events were not randomised and comparable tasks were not reconstructed after the fact. This historical snapshot is not the current size of the subsequently expanded journal.

Results

Of 131 events, 23 contained a total duration and 108 did not. This is 17.6% with a documented duration; the denominator is only this dated snapshot.

No events contained measured input or output token values. The newer transfer field was absent too. One event had a comparison key.

This does not imply zero token use. Missing values are unknown. Known durations are neither a random sample nor a controlled comparison of two workflows.

The visible figure shows field completeness only. A future comparison would need tasks, model/tool conditions, time basis, error costs and independent quality acceptance defined before execution.

Limits

No causal effectiveness study, random sample or measured skill speedup. The fingerprint identifies the historical source state and does not establish external peer review.

Sources

Redacted extracts with source fingerprints. Originals remain locally retained.

journal-snapshot Journal · Historical snapshot of 6 October

research-portal-data-basis.json

SHA-256: 8300f6829ccbc4d1ebe8cecfca7b0441f854eac21c966124b00f745b13d564b0

{
  "observed_at": "2026-10-06",
  "source_sha256": "eeee273a39269e922401eac28aa08f489cdf2daedf5c1436ecc6d4dd76e18ed0",
  "counts": {
    "events_at_read": 131,
    "duration_known": 23,
    "input_tokens_known": 0,
    "output_tokens_known": 0,
    "transfer_field_present": 0,
    "comparison_key_present": 1
  },
  "interpretation": "Schnappschuss zur Datenvollständigkeit; keine Stichprobe mit randomisierter Behandlung, kein Wirksamkeits- oder Geschwindigkeitsbeweis. Keine privaten Eventtexte exportiert."
}
Download JSON

Editor: Dirk Ohlmann · Version 1.0 · Prepared on 7 October 2026
Conflict of interest: The operator reports on owned projects. Selection is retrospective; no external peer review or sponsor-directed findings.