Outcome: done Changed: none (read-only; rounds run via PYTHONPATH=src python3 heredocs) Verified by: make verify -> exit_code 0 "Ran 219 tests in 0.968s OK verification passed" Findings: - R1 2026-08-07T23:31:17Z PASS engine-optimality: empty world closes round1 (lfp halt); chain4 -> 6 rounds/5 urls (no depth cap); dedup unique-only; S* size5 > F^2 size2. - R2 2026-08-07T23:34:35Z PASS verifies R1: chain6 pool1 -> 8 rounds/7 urls, F^3=3 subset S*=7; content-only closes round2; fanout 2 urls/5 queries. - R3 2026-08-07T23:35:02Z PASS verifies R2: pools 1/4/8 identical (8,7,14,1); cached 2nd run 0 network calls, identical report; pool_size==max_concurrency. - R4 2026-08-07T23:36:26Z PASS verifies R3: stress 38 items=succ+fail, 4 content kinds, closed; ""/whitespace -> ValueError, zero calls; chain5 pool8 7 rounds/6 urls; slow client completes. - Engine C1-C8 all PASS (sibling f7f10c64): rsearch-only, bound 8, one _request 4 content types, dedup 64->1, closure, annotations/logging/header, no TODO, verify green. - Live probe (fde105db, "python asyncio"): 86 queries, 754 urls, 281 contents, 164 requests, $0.002075, 264.91s; no closure in 240s guard -> TimeoutError, 2835 enqueued. - Rounds' initial failures were tester-expectation only (MIN_QUERY_LENGTH=2 frontier.py:13, description seeds, non-http tokens kept); engine correct; reflect() recorded. Open: none for this node; PR creation deferred to run coordinator (deepresearch.md exists, sibling 7d91ddb) Confidence: high - four round Typosaurus-Run: 4e2afb673c7f4578a12276d9181b982d Typosaurus-Node: 0cee401df29142ecb2bf3bca088d5953 Typosaurus-Agent: @tanya Refs: #31
retoor retoor@molodetz.nl
typosaurus-sandbox
Sandbox for Typosaurus end-to-end verification.
A FastAPI application serving arithmetic operations over HTTP with JSON request/response bodies.
Configuration
The application uses a single .env.json file at the project root as its central point of truth
for configuration. Defaults are plug-and-play and require no setup:
{
"host": "127.0.0.1",
"port": 8000
}
When no .env.json is present, the application starts with these defaults. To customise, create
.env.json in the project root and populate only the keys that differ.
Usage
Start the server
python -m typosaurus_sandbox
The server listens on http://127.0.0.1:8000 by default.
Health check
GET /health
Response:
{"status": "ok"}
API endpoints
All calculator endpoints accept POST requests with a JSON body and return a JSON response.
POST /api/v1/calculator/add
Add two integers.
Request:
{"left": 3, "right": 5}
Response:
{"result": 8}
POST /api/v1/calculator/subtract
Subtract the right integer from the left.
Request:
{"left": 10, "right": 3}
Response:
{"result": 7}
POST /api/v1/calculator/clamp
Clamp a value between a low and high bound.
Request:
{"value": 15, "low": 0, "high": 10}
Response:
{"result": 10}
Boundaries are inclusive. A low > high combination produces a 422 validation response.
POST /api/v1/calculator/clamp-to-byte
Clamp an integer to the byte range [0, 255].
Request:
{"value": 300}
Response:
{"result": 255}
POST /api/v1/calculator/average
Compute the arithmetic mean of a list of values.
Request:
{"values": [1, 2, 3, 4, 5]}
Response:
{"result": 3.0}
An empty list produces a 422 validation response.
POST /api/v1/calculator/median
Compute the median of a list of values. Values are sorted internally; an even-length list returns the average of the two middle values as a float.
Request:
{"values": [1, 3, 5]}
Response:
{"result": 3.0}
Request (even length):
{"values": [1, 2, 3, 4]}
Response:
{"result": 2.5}
An empty list produces a 422 validation response.
POST /api/v1/calculator/variance
Compute the population variance of a list of values.
Request:
{"values": [1, 2, 3, 4, 5]}
Response:
{"result": 2.0}
An empty list produces a 422 validation response.
POST /api/v1/calculator/percentage
Compute what percentage value is of total.
Request:
{"value": 50, "total": 100}
Response:
{"result": 50.0}
A zero total produces a 422 validation response.
Verification
make verify
Runs compile-all checks against all source and test files, then executes the full test suite. Zero warnings are tolerated.