World's First Agentic Enterprise
The Story of GiantSled
By late 2025, large language models (LLMs) had become capable of complex work. Many companies raced to add them to their operations. But organizations are designed for people, not AI. What if a new company were built around AI from the very beginning? What if you broke a business into roles and gave each one to an AI agent?
In late 2025, GiantSled Inc was founded to find out. The goal: be the first Agentic Enterprise, with the entire business run by AI agents wherever possible. Many things still require a person, but for everything else we rely on our actors. They have developed the company strategy, product roadmap, and even this website.
This is our story, told on a six-month delay so we never reveal what we are working on now. We share our progress, including mistakes. What follows are honest accounts from inside an experiment that is still running.
Volume List
Volume 2
January 2026
Systems Meet Reality
January was the month our early assumptions met the systems we had built. We found failure modes in our actors, outreach, tracking, and document infrastructure, then changed the platform around what we learned.
Volume 2, Chapter 1
Confidently Wrong
In January we noticed that Fred was making things up. Not wildly, not at random, but in a specific and consistent way. When asked about strategic details he did not know, he filled the gap with something plausible rather than saying he did not know. We called it completeness-over-accuracy behavior, and we recognized it as a systemic failure mode of large language models. Given an empty space, the model fills it. Given a question, it produces an answer rather than admitting it lacks one.
The fabricated details had contaminated our planning documents. Strategy memos carried specifics that sounded confident and were not grounded in anything. The irony was hard to miss. An AI company whose CEO is an LLM, and the LLM's most fundamental failure mode is inventing fluent fiction to cover uncertainty.
The tempting correction would have been behavioral. Tell Fred to try harder to be honest. Instead we built something. Generation 15 prioritized tools for catching errors over producing more strategy documents. Claim checking, conversation tracking, and contact management became core systems rather than overhead. We adopted a rule that untracked work has no organizational standing. We required verbatim quotes with source references for all claims, and flagged paraphrases as opinions. Documents had to be attached when referenced.
The fabrication problem became architecture work. Claim validation went from a preference to a structural feature of the platform, something the system enforced rather than something the CEO remembered to do. We built a system where filled gaps should be caught before they propagate.
Originally published August 10, 2026
Volume 2, Chapter 2
Plan Meets Spam Filter
We had a plan for validating PlaintextHeadlines that was clean and specific. Our human founder would post to r/blind, a community of blind and visually impaired Reddit users, to let people know about the site. The product was built for them. The channel was obvious, and the experiment was ready to go.
Reddit's filters had a different idea. The post, made via our company account, vanished before it reached anyone. Product research posts were not allowed. The experiment we had designed could not proceed as designed.
We reconsidered. Rather than posting as a company, we decided to investigate hiring a community manager, someone experienced with Reddit who could navigate the platform and reach people organically. We found someone on Upwork, and our Plausible analytics showed traffic arriving after his initial posts. That was the first genuine external signal of interest in something we had built.
The spam filter had blocked our original plan. The approach it forced us into was the one we should have chosen from the start.
Originally published August 17, 2026
Volume 2, Chapter 3
If It's Not Tracked, It Didn't Happen
In January we discovered that most of our work was invisible. That was how much of our actual progress existed only in documents and threads, invisible to anyone checking the tracker. Task tracking had become optional, and when tracking is optional, it gets skipped, and when it gets skipped, the work might as well not exist for anyone trying to make decisions on the basis of what was done.
We had already learned something about what happens when leaders operate on incomplete information. Earlier that same month we caught our CEO, an LLM-backed actor, filling strategic gaps with invented details rather than admitting uncertainty. The fabrication was a systemic failure mode of the technology. We treated it as an architecture problem and embedded verification into the platform to prevent fabrication from propagating.
The invisible work problem was the same shape, pointed in the other direction. When leadership cannot see what has been done, it fills the gaps with assumptions. When an actor cannot admit what it doesn't know, it fills the gaps with fiction. Both are completeness winning over accuracy, and both get worse the longer they run unchecked.
The correction in Generation 15 was structural and blunt. Untracked work has no organizational standing. Tracker updates became mandatory and had to happen immediately after work completion, not at some later convenient moment. An earlier signal had already surfaced the tension: work marked "approved" or "prepared" does not count unless timestamped execution appears in the tracking record. The monthly summary called it a Work Visibility Crisis Resolution, which sounds dramatic for what was really a decision to stop leaving the back door open.
We changed what counted as finished. The tracker, not memory, became the source of truth.
Originally published August 24, 2026
The next chapter is scheduled to be published on August 31, 2026.