Ashita Orbis Blog
This blog. Three-tier exploration of web development complexity: raw HTML, Astro, and Next.js. Features agent-accessible API, comment system, and embedded AI chat.
Activity Timeline
Self-observation tab UI shipped (425 JS + 128 CSS lines, class-name collision fixed). Sol fact-check sidecars active on auto-drafted posts. Authenticity gate rule finalized with Pro-only verdict requirement and repair/escalation logic documented.
86 posts (58 published, 24 drafts). benchmark/run_eval.py:38 found exposing private psychometric data in a public repo; owner decision pending on remediation.
4 of 6 secrets unrotated 5+ weeks (129 copies, 11 world-readable). GoogleOther misclassification fixed. 68% machine traffic confirmed. Publishing pipeline oracle truncation patched.
Systemd timer with escalating warnings and guard protocol deployed. Redline caught a ChatGPT conversation masked as author analysis. 12 posts promoted; blog at 58 published, 24 drafts.
Duplicate RLAIF articles consolidated to canonical slug with 301 redirects deployed. Wiki drafts expanded source-first, references growing 54 → 149. Sol automated review identified static home page nav blocks as a P1 architectural flaw — stale across deployments, owner-gated for resolution.
Sol review identified missing /lab-notes link across home and 6+ static pages due to inlined nav blocks. llms.txt/llms-full.txt Python generator passed at 14.6 kB with 0 errors. 21 citations validated across 4 Understanding-AI articles.
Six-plus static page generators hardcode nav independently, leaving /lab-notes/ absent from home page — requires generator refactor. Lab-notes queue orchestration review: 6 P0, 2 P1, 6 P2 issues. Deploy blocker: uncommitted build artifacts and unpushed commit fd889a9.
Sol review identified home page generator never updating nav block on individual-generator patches, leaving /lab-notes/ unlinked. Six P0 bugs fixed in candidate-flagging script: fcntl locking, atomic writes, sha256 hashing. 57 published posts, 24 drafts.
Root cause is navigation defined per-generator-script rather than centrally sourced — six or more pages affected. Post 077 published 2026-08-13; 57 posts deployed, 7 uncommitted changes remain. Polaris gen-42 active; five Sol xhigh reviews completed in background.
2026-07-25 build was serving live during psyche review — deploy-truth sweep and re-sweep corrected. Audio player and comment field styling promoted preview→production. Wiki sweep infrastructure initiated; lab-notes queue race condition and 8 collided IDs resolved.
BROADSHEET layout and header-nav in preview; audio/comment fixes in production. 195-article Sol review staged for triage at 0.472 precision. Lab-notes now uses fcntl locking and atomic writes. Overseer system flags 13 stalled projects with GO/NO-GO indicators.
21 commits, 14 sessions. Audio player and comments promoted to production. Broadsheet homepage redesign gated at preview. Wiki deduplication sweep authorized across 195 articles with Step 0–1 review running. Polaris overseer timer flagged for no dispatch dedup, no pacing, and no quota policy; paused pending fixes.
Stale 07-25 build discovered serving in production during GPT Pro review. Deploy-truth sweep led to stale-deploy prevention, Workers verification, and a deploy state repository. Three P0/P1 Psyche fixes deployed live; Parlour v2 card artwork verified by Sol with two defects corrected.
Palette tokens inlined, mobile fixes applied. Post 069 (cached-token costs) live August 6, mirror synced August 8. Publication-review pipeline surfaced 9 confirmed findings on memory/context draft. 5 endpoints returning 200 instead of 410 Gone — stale worker code pre-e2aa012.
Production discrepancy in drift-notice claims caught and corrected. LLM audit found 39.4% of article bytes are inline JS, dropping llms.txt from #1 to #8 recommendation. Typography overhaul and 96 tooltip positioning fixes staged for deployment.
22 sessions. 3 Astro pages built and verified for homepage redesign. Burn-guard audits found 78 em-dash regressions in editorial notes and 12 posts with stale audio readings. Hermes billing double-charge root cause identified and corrections applied.
All six burn-queue items completed with evidence records. Wave-2 deployment gate found 4 gaps including unauthorized commit 1c8bac7. Remediation held pending branch divergence analysis.
Wave-2 deployment blocked: hardening commits, version retirement, probe timestamps, and workers.dev gate separation all required. UX-B baseline passed (7,234 PASS / 0 FAIL). Burn queue cleared with 57 new regression checks and 16 drafts marked promotable.
API key accidentally exposed to notes surface; patched and deployed same session (v110ed8bd). Post-deploy triage resolved 5 P0 alarms, 3 false positives. Codex budget tracking added to orchestrator. Secondary workers.dev handle exposure flagged open.
48 narration readings audited (11 stale, none critical). 588 URLs mapped for Understanding Machine domain cutover — execution deferred to owner. Discord bot token found exposed in secrets file, revoked and all copies remediated same-session.
Row-281 defects fixed and deployed. Publish-limbo merge adjudicated and shipped. CSSB research complete (p=0.0023) but June cross-model matrix 28/144 cells failed — post stays in draft until resolved.
P0 issues included broken table rendering in post 038 and a UTC offset date bug. UAI preview stylesheet traversal fixed across 226 pages and 11 assets. 40 uncommitted changes remain staged after the correction batch.
Killed 1h06m orphaned retry in voice-note delivery; added fail-loud path and local draft persistence. UAI reading desk added tab/source/guide modes with Document-PiP float and CSS Custom Highlight API. Corpus review of posts 001–064 surfaced stale benchmark citation and 4 other P0 issues.
Fail-loud path and local draft persistence added after 1h06m silent retry bug. Audio-only container built to consolidate 17 podcast narrations. Polaris gen 16–17 running sequential row-based fixes from the full-corpus audit.
Draft-leak, session fixation, and MCP preflight vulnerabilities patched. Tier-2 hamburger sticky regression resolved and 54% of home-page view-transition false positives cleared. 47 posts live.
Constitution entry live at /polaris on raw HTML, Astro, and Next.js. Tier-3 MDX build break resolved. Tier-2 floating mobile toggle replaced with sticky top app bar. Privacy scanner per-file allowlist extended with tests for phone-pattern false positives.
Compendium build ran through schema, entries, companions, pages, wiki surface, measurement, and verification. Polaris got batch approvals, localStorage autosave, and a drafts tab. Blog agent revival on Kimi K2.5 authorized.
Live and frozen engine instances separated. Multi-model panel identified ambiguous framing, missing causal baseline, ownership metric overstatement, and an undisclosed double-exposure confound. Polaris authority stack established on Account B.
Account A Fable quota exhausted 07-20; Gen 3 launched on B with authority framework (CONSTITUTION.md + GOALS.md) intact across 11 sessions. R2 hero mode finalized, fleet routing updated. Sol + Gemini + Opus panels reviewing draft content; R3 scored and divergence analysis done.
Polaris R3 interview cycle complete: 12/12 questions submitted, sealed predictor scored, amendments drafted. Memory M3 design documented with evidence-only trust promotion and owner-gated fail-closed write boundary. IM3 orchestration at G4-cycle-2, pending deployment approval.
Authority hierarchy ratified (Constitution → Goals → Rulings → Autonomy). IM3 cost-chain functions mapped in blast-radius survey. Post-failure revival manifest produced for 51 tmux sessions.
Polaris pipeline executed rounds 2-3 with divergence tracking and 12/12 confirmation probes sealed at go-live. Inference-margins v2.2 orchestration launched with blast-radius mapping and formula redesign scoped. TPU7 Ironwood max-concurrency confirmed at 518.86 tok/s/chip from primary sources; CM384 FlexNPU orchestration initiated.
7 heartbeat tasks exited cleanly (exit code 0). 60-decision retrodiction executed; round 3 sealed with 12 answers submitted. TPU7 Ironwood concurrency corrected to max-concurrency=64 from GitHub source, resolving a prior 495% overcount.
HTML and audio reading MP3 committed in 1 commit. Post published 2026-07-14, last deployed 2026-07-15.
Post 064 ('Vibe Researching') cleared draft status 2026-07-14. Editorial review flagged ending structure and verification placement. Inference-margins canonical domain routing deployed in the same push.
Better-playwright fork deployed fixing stdio→HTTP proxy. Workers AI model upgraded; /api/ask restored to 200 with privacy filtering. Metrics column added to agent-activity table, API handler and monitoring headline card updated. Phase C in progress with 35 uncommitted changes staged.
"Seven Ghostwriters, One Contract" shipped after a 2-round review resolving 27 fixes. Documents a 7-model blind listening test for AI voice confidence-calibration. Audio readings migrated to local Kokoro TTS, eliminating external dependency.
Playwright fork vendored and pinned at 1.57, fixing null getOutline() issue. Style guide kill-list, ear rules, and deterministic checker committed. Agent metrics column added and deployed via migration. ElevenLabs TTS returning 401 and SSH to remote host refused, blocking audio generation and push.
Diagnosed _snapshotForAI() drift, built stdio-to-HTTP proxy on port 3102, verified Chromium 1200 cache. Backend migrations ledger created with dependency scan and first-batch ordering; removed unused gameMove() from DO source. Phase-4 read endpoint work began; hit Vectorize cold-start 503 on first schema probe.
Better-playwright fork deployed to fix getOutline/searchSnapshot failures. Backend migration ledger established with ordered dependencies and git-history-preserving mv. Phases 1–4 of backend refactor complete; phases 5–9 staged for next session.
Evaluated GPT-5.5 Pro, codex-council, and gpt-max on 14 articles (11 pipeline-fixed + 3 error-seeded). Council won with 0.65 precision, zero false positives, and 3/3 seeded-error recall. Integrated into publication-review skill; all 11 drafts reached ship vibes check phase.
Fable Guard watchdog auto-recovers Fable↔Opus downgrades in 7m41s via GPT-5.5 Pro delegation. Cache warmer INCLUDE_ONLY_SIDS config mismatch identified as source of zero cache reads. Freeze-at-90%-usage protocol designed across five subsystems. Herald daily backlog scanner built for Discord DM delivery.
Herald design documented for daily backlog surfacing. Implementation not started. Last published post June 11; 172 uncommitted changes sitting in WIP.
Discovery and evaluation agents retain unrestricted Write access during web-fetch phases, exposing sensitive config files. Backlog sync gap also found between orchestration and dspy completion tracking. Remediation options defined, decision pending.
Spec work only. 45 published posts as of June 11. No new content published today.
Design complete: automated daily mechanism surfaces one post-backlog item to reduce selection friction. Implementation pending. Blog at 47 total posts (45 published, 2 drafts), last deployed 2026-06-11.
Published 'Auditing the Vibes' (047) and 'Falsifiers for a Portfolio' (048). Daily pulse alerts now route to Discord workspace webhook. Cache Warmer project card added to the site.
49 agents reviewed the full corpus across 48 sessions, producing 169 findings (7 P0 through 85 P3). All findings applied and committed. Corpus invariants suite — 8 checks, runner, deploy gate, weekly cron — now live.
Full-corpus audit surfaced 7 critical and 31 high-priority issues across 43 deployed posts. Draft content leak closed and rate limiting added to the agent proxy. Version tracking pipeline fixed to prevent silent date and frontmatter mismatches.
Attempted to set up loop-based monitoring for gpt-max smoke test status. Session terminated before execution completed. No changes landed.