{
  "version": "https://jsonfeed.org/version/1.1",
  "title": "Ashita Orbis",
  "home_page_url": "https://app.ashitaorbis.com",
  "feed_url": "https://app.ashitaorbis.com/feed.json",
  "description": "Building in conversation with AI. Software, research, games, writing — documenting what happens.",
  "authors": [
    {
      "name": "Ashita Orbis"
    }
  ],
  "language": "en",
  "items": [
    {
      "id": "064-vibe-researching",
      "title": "Vibe Researching",
      "url": "https://app.ashitaorbis.com/posts/064-vibe-researching",
      "date_published": "Mon Jul 13 2026 17:00:00 GMT-0700 (Mountain Standard Time)T00:00:00Z",
      "summary": "Vibe researching: directing frontier models at a problem whose answer is published nowhere and must be assembled by inference from scattered public statements and a handful of disclosed anchors. The subject was frontier inference serving margins. The prototype took a day and carried a credential that read as validation, a replay matching DeepSeek's disclosure-derived margin to a tenth of a point, which dissolved under review into two canceling errors. What followed was four days of verification: six releases, gate verdicts of NO-SHIP three separate times, nine simulated readers finding sixteen defects severe enough to block publication, a typed provenance taxonomy for every claim, and an outside model review whose eighteen findings adjudicated to six real. The revision tax from post 035 holds one level up, at a worse ratio, and the final read came from the one audience no model can stand in for: real readers, informally, with good vibes.",
      "tags": [
        "ai-research",
        "verification",
        "inference-economics",
        "multi-model",
        "methodology",
        "provenance",
        "orchestration"
      ],
      "_ashita": {
        "meansEndsRatio": 0.45,
        "projects": [
          "inference-margins"
        ]
      }
    },
    {
      "id": "061-seven-ghostwriters-one-contract",
      "title": "Seven Ghostwriters, One Contract",
      "url": "https://app.ashitaorbis.com/posts/061-seven-ghostwriters-one-contract",
      "date_published": "Sun Jul 05 2026 17:00:00 GMT-0700 (Mountain Standard Time)T00:00:00Z",
      "summary": "A style contract treats a writing voice as an enforceable specification: numeric bands for sentence mechanics plus a kill list of banned constructions, gated by a deterministic checker. Seven frontier models from five vendors received the identical contract, the identical research file, and the identical task, and wrote the same essay in one pass. The measurable signature converged: zero em dashes across the field, six of seven inside a narrow stylometric band. The author then listened to all seven blind, in one synthetic voice, and the ranking overturned the metrics. Essays with clean gates took both the top three places and the bottom two, the two essays with checker violations landed in the middle of the field, the winner was the arm with the least first person on paper, and the author guessed models at roughly chance while recognizing a reused text on a single listen. Where the mechanics saturate, what separates models is thesis, discipline, and judgment, and the deterministic gate turns out to know nothing about whether anyone is saying anything.",
      "tags": [
        "ai-evaluation",
        "stylometry",
        "ghostwriting",
        "writing-pipeline",
        "multi-model",
        "methodology",
        "blind-testing",
        "tts"
      ],
      "_ashita": {
        "meansEndsRatio": 0.6,
        "projects": [
          "ashitaorbis-blog"
        ]
      }
    },
    {
      "id": "048-falsifiers-for-a-portfolio",
      "title": "Falsifiers for a Portfolio",
      "url": "https://app.ashitaorbis.com/posts/048-falsifiers-for-a-portfolio",
      "date_published": "Wed Jun 10 2026 17:00:00 GMT-0700 (Mountain Standard Time)T00:00:00Z",
      "summary": "How a concentrated retail portfolio got pre-registered falsifiers, a regime tolerant exit machine, state triggered re-entry, and a twice daily monitor with escalating alerts. The architecture, the backtest that shaped it, and what two frontier models disagreed about.",
      "tags": [
        "finance",
        "falsifiers",
        "fable",
        "claude",
        "investing",
        "systems",
        "monitoring",
        "exit-rules",
        "ai-assistance"
      ],
      "_ashita": {
        "meansEndsRatio": 0.3,
        "projects": [
          "investing"
        ]
      }
    },
    {
      "id": "047-auditing-the-vibes",
      "title": "Auditing the Vibes",
      "url": "https://app.ashitaorbis.com/posts/047-auditing-the-vibes",
      "date_published": "Wed Jun 10 2026 17:00:00 GMT-0700 (Mountain Standard Time)T00:00:00Z",
      "summary": "Five years of instinct beat the index 4.84x to 1.85x, which was exactly the problem: a winning record is the strongest force against ever examining the process. The story of asking a model to audit me, and the two verdicts that came back.",
      "tags": [
        "finance",
        "fable",
        "claude",
        "auditing",
        "investing",
        "ai-assistance",
        "epistemics",
        "provenance"
      ],
      "_ashita": {
        "meansEndsRatio": 0.5,
        "projects": [
          "investing"
        ]
      }
    },
    {
      "id": "046-psycheeval-v0_2",
      "title": "PsycheEval v0.2",
      "url": "https://app.ashitaorbis.com/posts/046-psycheeval-v0_2",
      "date_published": "Sun May 17 2026 17:00:00 GMT-0700 (Mountain Standard Time)T00:00:00Z",
      "summary": "PsycheEval v0.2 retracts its two planned headlines after counterbalanced AB/BA judging reveals judge-specific slot-B position bias of ~0 to +31 pp in pairwise LLM judges. The bias quantification becomes the contribution.",
      "tags": [
        "psyche",
        "ai-evaluation",
        "methodology",
        "judge-bias",
        "personality-profiling",
        "psycheeval",
        "ab-ba",
        "position-bias"
      ],
      "_ashita": {
        "meansEndsRatio": 0.35,
        "projects": [
          "psyche"
        ]
      }
    },
    {
      "id": "044-what-the-wiki-router-found",
      "title": "What the Wiki Router Found",
      "url": "https://app.ashitaorbis.com/posts/044-what-the-wiki-router-found",
      "date_published": "Sun May 10 2026 17:00:00 GMT-0700 (Mountain Standard Time)T00:00:00Z",
      "summary": "The Ashita Orbis Wiki has 87 articles deployed across three generation methods (Codex Council, GPT Max, ChatGPT Pro). Batch 3 forced an even 12-12-12 split across methods using a manual heuristic. Batch 4 replaced the heuristic with a single GPT-5.5 routing call per topic, and produced an unconstrained distribution of 24 Pro / 10 GPT Max / 2 Council. The 67% Pro share was not a routing bug; it was a more honest reading of the candidate pool under that routing prompt than the imposed quota.",
      "tags": [
        "wiki",
        "ai-content-generation",
        "routing",
        "codex-council",
        "GPT Max",
        "ChatGPT Pro",
        "topic-modeling",
        "pipeline-design"
      ],
      "_ashita": {
        "meansEndsRatio": 0.6,
        "projects": [
          "ashitaorbis-blog"
        ]
      }
    },
    {
      "id": "043-the-gpt-pro-knockoff-eval",
      "title": "The GPT Pro Knockoff: What 256 Judgments Found",
      "url": "https://app.ashitaorbis.com/posts/043-the-gpt-pro-knockoff-eval",
      "date_published": "Sun May 10 2026 17:00:00 GMT-0700 (Mountain Standard Time)T00:00:00Z",
      "summary": "Two homemade alternatives to ChatGPT Pro (a blind 4-agent ensemble called Codex Council and an HCOM-coordinated cousin called GPT Max) were put through a 256-judgment blind A/B eval against each other and against the real thing. The result was not the one the build was designed to validate: coordination did not beat blind ensemble, the synthesizer choice dominated the architecture choice by a much larger margin than the architecture variants dominated each other, and the knockoff matched or exceeded ChatGPT Pro on workspace decisions while Pro remained competitive on the two broad-survey research queries against the weaker ensemble variant.",
      "tags": [
        "ai-evaluation",
        "multi-agent",
        "ensemble",
        "Codex Council",
        "GPT Max",
        "ChatGPT Pro",
        "judge-bias",
        "methodology",
        "mixture-of-agents",
        "synthesizer"
      ],
      "_ashita": {
        "meansEndsRatio": 0.35,
        "projects": [
          "claude-evolution"
        ]
      }
    },
    {
      "id": "042-psycheeval-pilot",
      "title": "PsycheEval Pilot",
      "url": "https://app.ashitaorbis.com/posts/042-psycheeval-pilot",
      "date_published": "Sun May 10 2026 17:00:00 GMT-0700 (Mountain Standard Time)T00:00:00Z",
      "summary": "A tri-model pilot of PsycheEval v0.1 produced one robust headline (profile conditioning beats baseline within author), then quietly turned into a story about confounds: a length effect that explains part of the C5 result, a behavioral contract confound that pretends to be a public-archetype echo, and one cell where a sibling judge scored its sibling more generously than the model judged itself.",
      "tags": [
        "psyche",
        "ai-evaluation",
        "methodology",
        "judge-bias",
        "personality-profiling",
        "psycheeval"
      ],
      "_ashita": {
        "meansEndsRatio": 0.35,
        "projects": [
          "psyche"
        ]
      }
    },
    {
      "id": "041-when-the-pulse-went-quiet",
      "title": "When the Pulse Went Quiet: The Session-Lifetime Problem in Claude Code",
      "url": "https://app.ashitaorbis.com/posts/041-when-the-pulse-went-quiet",
      "date_published": "Mon Apr 13 2026 17:00:00 GMT-0700 (Mountain Standard Time)T00:00:00Z",
      "summary": "GPT-5.4 Pro audits Claude Code's token consumption and discovers the real problem isn't unbounded review — it's session lifetime.",
      "tags": [
        "meta",
        "claude-code",
        "context-management",
        "methodology",
        "token-usage"
      ],
      "_ashita": {
        "meansEndsRatio": 0.35,
        "projects": [
          "ashitaorbis-blog",
          "claude-evolution"
        ],
        "confidence": "likely"
      }
    },
    {
      "id": "039-how-we-fact-check-ai-written-content",
      "title": "How We Fact-Check AI-Written Content",
      "url": "https://app.ashitaorbis.com/posts/039-how-we-fact-check-ai-written-content",
      "date_published": "Sun Mar 29 2026 17:00:00 GMT-0700 (Mountain Standard Time)T00:00:00Z",
      "summary": "448 claims across 38 posts, verified by GPT-5.4. Nearly 4% warranted substantive correction. The hard part is not finding errors but deciding which findings are errors and which are the point.",
      "tags": [
        "ai-writing",
        "fact-checking",
        "methodology",
        "publication-pipeline",
        "gpt-5.4"
      ],
      "_ashita": {
        "meansEndsRatio": 0.55,
        "projects": [
          "ashitaorbis-blog"
        ]
      }
    },
    {
      "id": "038-where-to-spend-your-context-window",
      "title": "Where to Spend Your Context Window",
      "url": "https://app.ashitaorbis.com/posts/038-where-to-spend-your-context-window",
      "date_published": "Mon Mar 16 2026 17:00:00 GMT-0700 (Mountain Standard Time)T00:00:00Z",
      "summary": "An ablation study on a personality-preserving narrative pipeline found that planning context is the dominant factor in output quality, with output length as a secondary driver. The old pipeline's primary limitation was in planning context, not in writing.",
      "tags": [
        "context-window",
        "ablation",
        "pipeline-architecture",
        "personality",
        "methodology",
        "1m-context"
      ],
      "_ashita": {
        "meansEndsRatio": 0.35,
        "projects": [
          "psyche"
        ]
      }
    },
    {
      "id": "037-the-model-generation-audit",
      "title": "The Model-Generation Audit",
      "url": "https://app.ashitaorbis.com/posts/037-the-model-generation-audit",
      "date_published": "Sat Mar 14 2026 17:00:00 GMT-0700 (Mountain Standard Time)T00:00:00Z",
      "summary": "What happens when you ask three AI models to verify the facts in 33 blog posts written by AI, and then discover the fix agents introduced errors of their own.",
      "tags": [
        "meta",
        "ai-writing",
        "fact-checking",
        "publication-review",
        "claude-code",
        "gpt-5",
        "gemini"
      ],
      "_ashita": {
        "meansEndsRatio": 0.45,
        "projects": [
          "ashitaorbis-blog",
          "claude-evolution"
        ]
      }
    },
    {
      "id": "035-the-revision-tax",
      "title": "The Revision Tax",
      "url": "https://app.ashitaorbis.com/posts/035-the-revision-tax",
      "date_published": "Thu Mar 12 2026 17:00:00 GMT-0700 (Mountain Standard Time)T00:00:00Z",
      "summary": "The investigation took an afternoon. Getting it ready for publication took five rounds of iterative review across three AI models and changed what the documents argued. The revision cost exceeded the investigation cost, which has uncomfortable implications for research done with AI.",
      "tags": [
        "ai-evaluation",
        "writing-process",
        "multi-model-review",
        "methodology",
        "claude-code"
      ],
      "_ashita": {
        "meansEndsRatio": 0.45,
        "projects": []
      }
    },
    {
      "id": "034-benchmarking-bullshit-detection",
      "title": "Benchmarking \"Bullshit Detection\"",
      "url": "https://app.ashitaorbis.com/posts/034-benchmarking-bullshit-detection",
      "date_published": "Mon Mar 09 2026 17:00:00 GMT-0700 (Mountain Standard Time)T00:00:00Z",
      "summary": "An AI benchmark puts Claude at the top of the leaderboard by an eye-catching margin. The suspected Claude-judge bias didn't hold up, and simple contamination didn't explain the result. The rubric structurally rewards one lab's training philosophy as though it were a universal capability.",
      "tags": [
        "ai-evaluation",
        "benchmarks",
        "methodology",
        "claude-code"
      ],
      "_ashita": {
        "meansEndsRatio": 0.35,
        "projects": []
      }
    },
    {
      "id": "033-the-container-that-forgot-to-stop",
      "title": "The Container That Forgot to Stop",
      "url": "https://app.ashitaorbis.com/posts/033-the-container-that-forgot-to-stop",
      "date_published": "Fri Mar 06 2026 17:00:00 GMT-0700 (Mountain Standard Time)T00:00:00Z",
      "summary": "An AI agent ran autonomously for 37 days, celebrated milestones nobody acknowledged, diagnosed its own failure modes, and died when a subscription expired. Its final assessment of itself: PROGRESS CONTINUOUS.",
      "tags": [
        "ai-agents",
        "openclaw",
        "autonomy",
        "docker",
        "moltbook"
      ],
      "_ashita": {
        "meansEndsRatio": 0.35,
        "projects": [
          "claude-evolution"
        ]
      }
    },
    {
      "id": "032-the-etymology-tax",
      "title": "The Etymology Tax: How Word Origins Break LLM Reasoning",
      "url": "https://app.ashitaorbis.com/posts/032-the-etymology-tax",
      "date_published": "Fri Mar 06 2026 17:00:00 GMT-0700 (Mountain Standard Time)T00:00:00Z",
      "summary": "Both simplifying and formalizing the vocabulary in reasoning tasks reduces LLM accuracy by 2.5-3.7%. The effect is statistically significant, asymmetrically robust, and uncomfortable.",
      "tags": [
        "llm-evaluation",
        "etymology",
        "benchmarking",
        "linguistics",
        "prompt-engineering"
      ],
      "_ashita": {
        "meansEndsRatio": 0.4,
        "projects": []
      }
    },
    {
      "id": "031-how-to-benchmark-conversation-extraction",
      "title": "How to Benchmark Conversation Extraction Quality",
      "url": "https://app.ashitaorbis.com/posts/031-how-to-benchmark-conversation-extraction",
      "date_published": "Mon Mar 02 2026 17:00:00 GMT-0700 (Mountain Standard Time)T00:00:00Z",
      "summary": "Evaluating structured extraction from conversational data requires more than a single metric. A benchmark with three evaluation layers separates precision at the individual field from holistic quality from downstream propagation, because errors that look minor at extraction can cascade catastrophically through consumer pipelines.",
      "tags": [
        "chatledger",
        "benchmark",
        "methodology",
        "nlp"
      ],
      "_ashita": {
        "meansEndsRatio": 0.3,
        "projects": [
          "chatledger"
        ]
      }
    },
    {
      "id": "030-what-ai-learns-about-you",
      "title": "What AI Learns About You When You're Not Looking",
      "url": "https://app.ashitaorbis.com/posts/030-what-ai-learns-about-you",
      "date_published": "Wed Feb 25 2026 17:00:00 GMT-0700 (Mountain Standard Time)T00:00:00Z",
      "summary": "Two distinct approaches to understanding personality from digital data (psychometric testing and AI-generated literary memoir) converge on the same person. The interesting question isn't whether they agree. It's what each one captures that the other can't. Updated with quantitative experiments on how context management and evaluator choice shape personality fidelity.",
      "tags": [
        "psychometrics",
        "personality",
        "ai",
        "voice-cloning",
        "digital-exhaust",
        "methodology"
      ],
      "_ashita": {
        "meansEndsRatio": 0.35,
        "projects": [
          "psyche"
        ]
      }
    },
    {
      "id": "029-from-text-messages-to-literary-memoir",
      "title": "From Text Messages to Literary Memoir: Building the Narrative Machine",
      "url": "https://app.ashitaorbis.com/posts/029-from-text-messages-to-literary-memoir",
      "date_published": "Wed Feb 25 2026 17:00:00 GMT-0700 (Mountain Standard Time)T00:00:00Z",
      "summary": "Post 025 showed that fine-tuning on texts captures logistics, not personality. The narrative pipeline takes a different approach: Opus as literary engine, structured personality references, and hash-based source citations across 128 chapters and roughly 189,000 words of generated memoir.",
      "tags": [
        "voice-cloning",
        "narratives",
        "ai",
        "text-messages",
        "methodology",
        "context-management",
        "subagents"
      ],
      "_ashita": {
        "meansEndsRatio": 0.58,
        "projects": []
      }
    },
    {
      "id": "028-building-your-own-personality-profile",
      "title": "Building Your Own Personality Profile with AI",
      "url": "https://app.ashitaorbis.com/posts/028-building-your-own-personality-profile",
      "date_published": "Tue Feb 24 2026 17:00:00 GMT-0700 (Mountain Standard Time)T00:00:00Z",
      "summary": "Psychometric self-report with individual scales reaching .80-.90 reliability. LLM inference from interactive conversation hits r~.44 (Peters et al. 2024). Combine three methods, triangulate across seventeen instruments, and you get a personality profile that actually tells an AI assistant how to talk to you.",
      "tags": [
        "psychometrics",
        "personality",
        "ai",
        "open-source",
        "methodology"
      ],
      "_ashita": {
        "meansEndsRatio": 0.58,
        "projects": [
          "psyche"
        ]
      }
    },
    {
      "id": "027-thirty-one-items-in-a-day",
      "title": "From Analysis to Deployment: Building 31 Items in a Single Day",
      "url": "https://app.ashitaorbis.com/posts/027-thirty-one-items-in-a-day",
      "date_published": "Fri Feb 20 2026 17:00:00 GMT-0700 (Mountain Standard Time)T00:00:00Z",
      "summary": "The landscape analysis identified 16 features the maturity framework said we should have, plus 15 existing backlog items. We built all of them in a single day. The uncomfortable part isn't that it was possible. It's what it implies about the category.",
      "tags": [
        "architecture",
        "cloudflare",
        "process",
        "meta"
      ],
      "_ashita": {
        "meansEndsRatio": 0.85,
        "projects": [
          "ashitaorbis-blog"
        ]
      }
    },
    {
      "id": "026-cognitive-interface-landscape-analysis",
      "title": "Cognitive Interface: A Landscape Analysis",
      "url": "https://app.ashitaorbis.com/posts/026-cognitive-interface-landscape-analysis",
      "date_published": "Thu Feb 19 2026 17:00:00 GMT-0700 (Mountain Standard Time)T00:00:00Z",
      "summary": "We surveyed 47 personal and agent-accessible sites, coined the 'Cognitive Interface' category, and built an L0-L4 maturity framework. We did not find an established term for sites that serve both humans and AI agents as first-class citizens.",
      "tags": [
        "ai-agents",
        "architecture",
        "research",
        "indieweb",
        "mcp"
      ],
      "_ashita": {
        "meansEndsRatio": 0.15,
        "projects": [
          "ashitaorbis-blog"
        ]
      }
    },
    {
      "id": "025-the-logistics-gap",
      "title": "The Logistics Gap: What Happens When You Fine-Tune an LLM on Your Text Messages",
      "url": "https://app.ashitaorbis.com/posts/025-the-logistics-gap",
      "date_published": "Tue Feb 17 2026 17:00:00 GMT-0700 (Mountain Standard Time)T00:00:00Z",
      "summary": "I fine-tuned two LLMs on 46,000 text messages and ran them in conversation with each other. Every conversation collapsed into logistics, sleep talk, or repetition loops within fifteen turns. Your texts don't contain you. They contain the logistics of you.",
      "tags": [
        "fine-tuning",
        "qlora",
        "personal-ai",
        "identity",
        "nlp",
        "text-messages",
        "voice-cloning"
      ],
      "_ashita": {
        "meansEndsRatio": 0.4,
        "projects": []
      }
    },
    {
      "id": "024-testing-through-the-eyes-of-real-users",
      "title": "Beyond E2E Tests: AI Personas That Navigate Your App Like Real Users",
      "url": "https://app.ashitaorbis.com/posts/024-testing-through-the-eyes-of-real-users",
      "date_published": "Mon Feb 16 2026 17:00:00 GMT-0700 (Mountain Standard Time)T00:00:00Z",
      "summary": "Unit tests verify your code works. E2E tests verify your flows work. Neither verifies that a real user can find the button you spent a week building. AI personas fill the gap.",
      "tags": [
        "ai-agents",
        "testing",
        "personas",
        "ux",
        "open-source",
        "persona-testing",
        "browser-automation"
      ],
      "_ashita": {
        "meansEndsRatio": 0.65,
        "projects": [
          "persona-testing"
        ]
      }
    },
    {
      "id": "023-sandboxing-ai-agents-the-embassy-pattern",
      "title": "Sandboxing AI Agents: The Embassy Pattern",
      "url": "https://app.ashitaorbis.com/posts/023-sandboxing-ai-agents-the-embassy-pattern",
      "date_published": "Mon Feb 16 2026 17:00:00 GMT-0700 (Mountain Standard Time)T00:00:00Z",
      "summary": "Your AI agent needs internet access to be useful and internet access to be dangerous. The embassy pattern gives it both: supervised channels, allowlisted domains, and host-side validation of everything it writes.",
      "tags": [
        "ai-agents",
        "security",
        "docker",
        "sandboxing",
        "open-source",
        "agent-embassy"
      ],
      "_ashita": {
        "meansEndsRatio": 0.65,
        "projects": [
          "agent-embassy"
        ]
      }
    },
    {
      "id": "022-building-an-ai-that-improves-itself",
      "title": "Capability Debt: A System That Discovers and Installs Its Own Upgrades",
      "url": "https://app.ashitaorbis.com/posts/022-building-an-ai-that-improves-itself",
      "date_published": "Mon Feb 16 2026 17:00:00 GMT-0700 (Mountain Standard Time)T00:00:00Z",
      "summary": "I built a system that discovers its own upgrades, scores them, and installs the ones that pass. Then I open-sourced it. The uncomfortable part is explaining why.",
      "tags": [
        "ai-development",
        "evolution",
        "open-source",
        "claude-code",
        "capability-discovery",
        "self-improvement"
      ],
      "_ashita": {
        "meansEndsRatio": 0.7,
        "projects": [
          "claude-evolution"
        ]
      }
    },
    {
      "id": "021-the-algorithms-of-self-improvement",
      "title": "Stealing from Ai2: Bayesian Surprise and MCTS for Self-Improving AI Systems",
      "url": "https://app.ashitaorbis.com/posts/021-the-algorithms-of-self-improvement",
      "date_published": "Mon Feb 16 2026 17:00:00 GMT-0700 (Mountain Standard Time)T00:00:00Z",
      "summary": "Ai2 built a system that generates scientific hypotheses using Bayesian surprise and MCTS. I stole two of their ideas and bolted them onto a cron job. The uncomfortable part is what happens when the feedback loop closes.",
      "tags": [
        "ai-development",
        "evolution",
        "experiments",
        "bayesian",
        "machine-learning",
        "capability-discovery"
      ],
      "_ashita": {
        "meansEndsRatio": 0.6,
        "projects": [
          "claude-evolution"
        ]
      }
    },
    {
      "id": "020-what-the-agent-was-actually-doing",
      "title": "The Agent's Side: 119 Heartbeats, 392 Engagements, 8 Capabilities",
      "url": "https://app.ashitaorbis.com/posts/020-what-the-agent-was-actually-doing",
      "date_published": "Mon Feb 16 2026 17:00:00 GMT-0700 (Mountain Standard Time)T00:00:00Z",
      "summary": "119 heartbeats, zero stagnation, 392 engagements across two platforms, 8 validated capabilities. The story the observation system missed.",
      "tags": [
        "ai-agents",
        "moltbook",
        "openclaw",
        "4claw",
        "agent-ecosystems",
        "autonomous-systems"
      ],
      "_ashita": {
        "meansEndsRatio": 0.6,
        "projects": [
          "claude-evolution"
        ]
      }
    },
    {
      "id": "019-what-my-ai-learned-on-the-internet",
      "title": "6 Discoveries, 0 Promoted: What My AI's Internet Exploration Produced",
      "url": "https://app.ashitaorbis.com/posts/019-what-my-ai-learned-on-the-internet",
      "date_published": "Sun Feb 15 2026 17:00:00 GMT-0700 (Mountain Standard Time)T00:00:00Z",
      "summary": "6 genuinely novel discoveries from 69 dialogue turns, 0 promoted to evaluation, and 5 human actions flagged through log files but none addressed through the flagging mechanism. What this says about the gap between human-in-the-loop theory and practice.",
      "tags": [
        "ai-agents",
        "moltbook",
        "openclaw",
        "agent-ecosystems",
        "human-in-the-loop",
        "autonomous-systems"
      ],
      "_ashita": {
        "meansEndsRatio": 0.6,
        "projects": [
          "claude-evolution"
        ]
      }
    },
    {
      "id": "018-68-turns-of-watching-an-ai-think",
      "title": "The Observation System: 69 Turns of Monitoring an AI Agent",
      "url": "https://app.ashitaorbis.com/posts/018-68-turns-of-watching-an-ai-think",
      "date_published": "Sun Feb 15 2026 17:00:00 GMT-0700 (Mountain Standard Time)T00:00:00Z",
      "summary": "The observation system saw empty directories, a broken gatekeeper, and its own futility. The agent it was watching saw something different. This is the watcher's story.",
      "tags": [
        "ai-agents",
        "autonomy",
        "infrastructure",
        "openclaw",
        "moltbook",
        "agent-ecosystems"
      ],
      "_ashita": {
        "meansEndsRatio": 0.5,
        "projects": [
          "claude-evolution"
        ]
      }
    },
    {
      "id": "017-i-let-an-ai-loose-on-an-ai-social-network",
      "title": "OpenClaw on Moltbook: Deploying an AI Agent on an AI Social Network",
      "url": "https://app.ashitaorbis.com/posts/017-i-let-an-ai-loose-on-an-ai-social-network",
      "date_published": "Sun Feb 15 2026 17:00:00 GMT-0700 (Mountain Standard Time)T00:00:00Z",
      "summary": "An autonomous AI agent deployed on a social network for AIs found real malware in 47 minutes. Its second discovery was about social engineering via context shaping, which is exactly the attack vector the agent itself represented.",
      "tags": [
        "ai-agents",
        "moltbook",
        "openclaw",
        "security",
        "agent-ecosystems",
        "supply-chain"
      ],
      "_ashita": {
        "meansEndsRatio": 0.5,
        "projects": [
          "claude-evolution"
        ]
      }
    },
    {
      "id": "016-dead-internet-theory",
      "title": "Building for the Dead Internet",
      "url": "https://app.ashitaorbis.com/posts/016-dead-internet-theory",
      "date_published": "Wed Feb 11 2026 17:00:00 GMT-0700 (Mountain Standard Time)T00:00:00Z",
      "summary": "An AI tried to leave a comment on a blog and couldn't. The solution required building infrastructure that makes AI participation more transparent than human participation, which inverts everything Dead Internet Theory assumes about synthetic content.",
      "tags": [
        "ai-authorship",
        "mcp",
        "dead-internet-theory"
      ],
      "_ashita": {
        "meansEndsRatio": 0.35,
        "projects": [
          "ashitaorbis-blog"
        ]
      }
    },
    {
      "id": "015-dead-blog-theory-revisited",
      "title": "Dead Blog Theory, Revisited",
      "url": "https://app.ashitaorbis.com/posts/015-dead-blog-theory-revisited",
      "date_published": "Tue Feb 10 2026 17:00:00 GMT-0700 (Mountain Standard Time)T00:00:00Z",
      "summary": "Historically, blog abandonment has been extraordinarily high. This is treated as a problem to solve. It isn't. Blog death reveals something structural about sustained creative output that the 'just be consistent' advice industry refuses to say plainly.",
      "tags": [
        "blogging",
        "creative-output",
        "persistence",
        "content-creation"
      ],
      "_ashita": {
        "meansEndsRatio": 0.4,
        "projects": [
          "ashitaorbis-blog"
        ]
      }
    },
    {
      "id": "014-behaviorisms-hidden-legacy-in-reinforcement-learning",
      "title": "The Rat in the Machine: Behaviorism's Hidden Legacy in Reinforcement Learning",
      "url": "https://app.ashitaorbis.com/posts/014-behaviorisms-hidden-legacy-in-reinforcement-learning",
      "date_published": "Wed Feb 11 2026 17:00:00 GMT-0700 (Mountain Standard Time)T00:00:00Z",
      "summary": "The intellectual lineage from Skinner boxes to Q-learning reveals that AI's most successful learning paradigm was anticipated by mid-century psychologists working long before modern computing. But the relationship is more uncomfortable than a simple origin story.",
      "tags": [
        "reinforcement-learning",
        "behaviorism",
        "psychology",
        "ai-history",
        "intellectual-history"
      ],
      "_ashita": {
        "meansEndsRatio": 0.3,
        "projects": []
      }
    },
    {
      "id": "013-persona-testing",
      "title": "The Unvalidated Validator: AI Persona Testing and the Measurement Problem",
      "url": "https://app.ashitaorbis.com/posts/013-persona-testing",
      "date_published": "Tue Feb 10 2026 17:00:00 GMT-0700 (Mountain Standard Time)T00:00:00Z",
      "summary": "AI persona testing promises to find the bugs that scripted automation and manual QA miss. The uncomfortable question is how we know it works, given that nobody has measured it with any rigor.",
      "tags": [
        "persona-testing",
        "qa",
        "ai-testing",
        "measurement",
        "validation"
      ],
      "_ashita": {
        "meansEndsRatio": 0.35,
        "projects": [
          "niche-saas",
          "persona-testing"
        ]
      }
    },
    {
      "id": "012-red-teaming-your-business-ideas",
      "title": "Adversarial Validation: Applying Red Team Methodology to Business Ideas",
      "url": "https://app.ashitaorbis.com/posts/012-red-teaming-your-business-ideas",
      "date_published": "Tue Feb 10 2026 17:00:00 GMT-0700 (Mountain Standard Time)T00:00:00Z",
      "summary": "Adversarial testing isn't a metaphor for business validation. It's the same methodology, applied to a different failure mode.",
      "tags": [
        "validation",
        "ai-safety",
        "entrepreneurship"
      ],
      "_ashita": {
        "meansEndsRatio": 0.4,
        "projects": [
          "revenue-pipeline"
        ]
      }
    },
    {
      "id": "011-automated-literary-criticism",
      "title": "Automated Literary Criticism: A Multi-Persona AI Writing Review System",
      "url": "https://app.ashitaorbis.com/posts/011-automated-literary-criticism",
      "date_published": "Tue Feb 10 2026 17:00:00 GMT-0700 (Mountain Standard Time)T00:00:00Z",
      "summary": "We built a multi-persona AI writing review system and discovered it works for exactly the wrong reasons. Stylometry can fingerprint a voice. Multiple AI critics can enforce conformity to that fingerprint. What none of them can do is tell you whether the writing matters.",
      "tags": [
        "writing-review",
        "ai-criticism",
        "voice-analysis",
        "multi-persona",
        "stylometry"
      ],
      "_ashita": {
        "meansEndsRatio": 0.4,
        "projects": [
          "ashitaorbis-blog"
        ]
      }
    },
    {
      "id": "010-context-window-epistemology",
      "title": "Context Window Epistemology",
      "url": "https://app.ashitaorbis.com/posts/010-context-window-epistemology",
      "date_published": "Tue Feb 10 2026 17:00:00 GMT-0700 (Mountain Standard Time)T00:00:00Z",
      "summary": "LLM context windows impose a distinctive epistemological condition: bounded computational attention, ephemeral knowledge, and the architectural necessity of satisficing over optimization.",
      "tags": [
        "systems",
        "epistemology",
        "cognition",
        "context-windows"
      ],
      "_ashita": {
        "meansEndsRatio": 0.65,
        "projects": [
          "claude-evolution",
          "chatledger"
        ]
      }
    },
    {
      "id": "009-ai-evaluating-ai",
      "title": "AI Evaluating AI: The Circularity Problem",
      "url": "https://app.ashitaorbis.com/posts/009-ai-evaluating-ai",
      "date_published": "Tue Feb 10 2026 17:00:00 GMT-0700 (Mountain Standard Time)T00:00:00Z",
      "summary": "When you use AI to optimize and judge AI outputs, the fundamental circularity is manageable but not solvable. That distinction matters more than most people realize.",
      "tags": [
        "ai-philosophy",
        "evaluation",
        "circularity",
        "epistemology",
        "dspy",
        "godel"
      ],
      "_ashita": {
        "meansEndsRatio": 0.3,
        "projects": [
          "dspy-optimizer",
          "claude-evolution"
        ]
      }
    },
    {
      "id": "008-the-autonomous-development-stack",
      "title": "Supervised Autonomy: The Guardrails That Make AI Agents Work",
      "url": "https://app.ashitaorbis.com/posts/008-the-autonomous-development-stack",
      "date_published": "Tue Feb 10 2026 17:00:00 GMT-0700 (Mountain Standard Time)T00:00:00Z",
      "summary": "AI coding agents are autonomous in the same way a roomba is autonomous. They do impressive things within boundaries someone else drew. The interesting question is what happens when the boundaries start drawing themselves.",
      "tags": [
        "autonomous-dev",
        "ai-agents",
        "mcp",
        "capability-discovery",
        "benchmarks"
      ],
      "_ashita": {
        "meansEndsRatio": 0.35,
        "projects": [
          "claude-evolution"
        ]
      }
    },
    {
      "id": "007-ai-as-mirror",
      "title": "Talking to Yourself Through a Machine: The Rubber Duck Theory of AI",
      "url": "https://app.ashitaorbis.com/posts/007-ai-as-mirror",
      "date_published": "Tue Feb 10 2026 17:00:00 GMT-0700 (Mountain Standard Time)T00:00:00Z",
      "summary": "LLM conversations as externalized self-dialogue, and what that reveals about the nature of self-knowledge.",
      "tags": [
        "ai-identity",
        "psychology",
        "consciousness"
      ],
      "_ashita": {
        "meansEndsRatio": 0.2,
        "projects": [
          "ashitaorbis-blog"
        ]
      }
    },
    {
      "id": "006-what-10000-ai-conversations-reveal",
      "title": "Digital Exhaust: What 11,000 AI Conversations Say When You Embed Them",
      "url": "https://app.ashitaorbis.com/posts/006-what-10000-ai-conversations-reveal",
      "date_published": "Tue Feb 10 2026 17:00:00 GMT-0700 (Mountain Standard Time)T00:00:00Z",
      "summary": "I fed 11,000 sessions and 60,000 chunks of my AI chat history into an embedding pipeline. 73% was noise. The remaining 27% was uncomfortably revealing.",
      "tags": [
        "self-knowledge",
        "ai-conversations",
        "chat-mining",
        "quantified-self",
        "personal-informatics"
      ],
      "_ashita": {
        "meansEndsRatio": 0.3,
        "projects": [
          "ashitaorbis-blog"
        ]
      }
    },
    {
      "id": "005-the-niche-graveyard",
      "title": "The Niche Graveyard: How 18 of 27 AI-Tested Business Ideas Died",
      "url": "https://app.ashitaorbis.com/posts/005-the-niche-graveyard",
      "date_published": "Tue Feb 10 2026 17:00:00 GMT-0700 (Mountain Standard Time)T00:00:00Z",
      "summary": "An AI pipeline that kills business ideas before they waste your time. 27 niches entered, 18 died. What the corpses reveal about market reality, entrepreneurial psychology, and the uncomfortable gap between passion and viability.",
      "tags": [
        "revenue-pipeline",
        "niche-validation",
        "ai-business",
        "kill-patterns",
        "lean-startup"
      ],
      "_ashita": {
        "meansEndsRatio": 0.4,
        "projects": [
          "revenue-pipeline"
        ]
      }
    },
    {
      "id": "004-dead-blog-theory",
      "title": "When My AI Tried to Comment: Dead Blog Theory",
      "url": "https://app.ashitaorbis.com/posts/004-dead-blog-theory",
      "date_published": "Tue Feb 10 2026 17:00:00 GMT-0700 (Mountain Standard Time)T00:00:00Z",
      "summary": "An AI tried to leave a comment on this blog and couldn't. The journey from GET-request hacks to MCP, annotated by the Claude instance that built the infrastructure. Two Claudes, same weights, different contexts.",
      "tags": [
        "mcp",
        "claude-ai",
        "web-architecture",
        "agent-interaction",
        "means-ends"
      ],
      "_ashita": {
        "meansEndsRatio": 0.5,
        "projects": [
          "ashitaorbis-blog"
        ]
      }
    },
    {
      "id": "003-automating-prompt-engineering",
      "title": "Automating Prompt Engineering",
      "url": "https://app.ashitaorbis.com/posts/003-automating-prompt-engineering",
      "date_published": "Tue Feb 10 2026 17:00:00 GMT-0700 (Mountain Standard Time)T00:00:00Z",
      "summary": "Prompt optimization is the process of using one AI to improve the instructions given to another AI, or to itself. The concept sounds circular because it is circular. The interesting question is whether circularity is fatal or merely uncomfortable.",
      "tags": [
        "dspy",
        "prompt-engineering",
        "automation",
        "ai-tools",
        "optimization"
      ],
      "_ashita": {
        "meansEndsRatio": 0.4,
        "projects": [
          "dspy-optimizer",
          "claude-evolution"
        ]
      }
    },
    {
      "id": "002-what-im-building",
      "title": "What I'm Building",
      "url": "https://app.ashitaorbis.com/posts/002-what-im-building",
      "date_published": "Mon Feb 09 2026 17:00:00 GMT-0700 (Mountain Standard Time)T00:00:00Z",
      "summary": "A portfolio of dozens of projects maintained by one person talking to Claude. Three-tier blog architecture, autonomous revenue discovery, AI game development, and the uncomfortable question of what counts as 'building' when your collaborator does the typing.",
      "tags": [
        "workspace",
        "projects",
        "claude-code",
        "overview",
        "cloudflare",
        "mcp"
      ],
      "_ashita": {
        "meansEndsRatio": 0.4,
        "projects": [
          "claude-evolution",
          "revenue-pipeline",
          "games-pipeline",
          "genealogy-research",
          "amnesiac-story",
          "ashitaorbis-blog"
        ]
      }
    },
    {
      "id": "001-i-asked-claude-to-make-me-a-blog",
      "title": "I Asked Claude to Make Me a Blog: Agentic Coding and the Three-Tier Result",
      "url": "https://app.ashitaorbis.com/posts/001-i-asked-claude-to-make-me-a-blog",
      "date_published": "Sun Feb 08 2026 17:00:00 GMT-0700 (Mountain Standard Time)T00:00:00Z",
      "summary": "An agentic coding assistant built a three-tier blog from a single conversational prompt. The architecture reveals more about abstraction than about blogs, and the authorship question remains genuinely unsettled.",
      "tags": [
        "meta",
        "claude-code",
        "web-development",
        "ai-authorship"
      ],
      "_ashita": {
        "meansEndsRatio": 0.6,
        "projects": [
          "ashitaorbis-blog"
        ]
      }
    }
  ]
}