Visual Time Perception & Symbolic Reasoning - Single & Continuous Conversion
Live Processing Histogram
Angle bins (0-360) and mask pixel-count diagnostics (not time units) update every symbolic conversion.
Grounding Preset Studio
Compose reusable symbolic grounding strategies by dragging presets into the script generator. Apply rules to current symbolic output, copy as JSON, or clear to start over. Scripts persist across page refresh.
Preset Library
Script Generator
Generated Dataset
Accumulate labeled samples from symbolic frames. Build dataset, filter by search or instruction type, then copy or save as JSONL. Rows persist until cleared.
No dataset generated yet.
Reproducible Experiments – Run Sweep
Execute a full robustness sweep with 10 degradation conditions (rotate, blur, noise, occlusion, mirror, etc.). Generates synthetic clock images, tests symbolic detection, and runs statistical significance tests. Condition sample thumbnails appear during the run and persist in browser cache.
No experiment run yet.
LLM Benchmark Comparison
Compare multiple LLM models on the same set of clock-reading questions (requires local Ollama). Generates questions from symbolic experiment metrics, collects responses, scores correctness and key-factor extraction, and runs two-way ANOVA + Tukey HSD statistical tests. Parse success summary shows JSON parse rates per model.
No LLM benchmark run yet.