Ghost Hat Studio
Ghost Hat Studio · Blog

Off the wire.

Field notes on building AI agents you own. Local-first pipelines, the tricks that keep them lean, and the mistakes we learned from.

July 14, 2026 · 3 min read
The Rent Just Became Optional
Gabe Larsen, the Chief Revenue Officer at Atonom, looked at their Salesforce contract and saw a number that did not match the value anymore. Forty thousand dollars a…
Read
July 13, 2026 · 1 min read
Build Your Own AI Employees
We’re running free workshops on how we built the crew that runs Ghost Hat. Check out the demos of Sydney and Vera, and if it inspires you, get on the list.
Read
July 6, 2026 · 3 min read
Vendor Lock-In Is a Design Flaw
AI startups running on cloud APIs bleed cash, and the reason is not arithmetic. They lose money because their product sits on top of someone else's pricing page.
Read
July 4, 2026 · 4 min read
The Bright Factory That Runs Ghost Hat Studio
If you haven’t run across the term, dark factories are lights-out manufacturing plants, an idea coined in Japan in the 1980s when FANUC set robots to building other…
Read
July 2, 2026 · 2 min read
Multi-GPU Just Got 3-4x Faster
For a long time, the fast way to run a model across several GPUs lived inside heavyweight engines like vLLM and TensorRT-LLM, and mostly on NVIDIA. If you ran llama.cpp…
Read
June 29, 2026 · 4 min read
Why Ford Brought Back Its Veteran Engineers
Ford rehired 350 veteran engineers after leaning on automated quality systems and not getting the quality it expected. Some were former employees, some came over from…
Read
June 22, 2026 · 3 min read
The Token Count Where Self-Hosting Pays Off Is Closer Than You Think
Most teams run premium model APIs until the bill stings. They assume self-hosting needs a dedicated data center, a full ops rotation, and months of setup. That…
Read
June 22, 2026 · 5 min read
The Undersung 4b Model
The industry is obsessed with scale. The prevailing logic says that if a 70b model is good, a 400b model must be better. In production, chasing parameter counts is a…
Read
June 19, 2026 · 4 min read
What You’re Allowed to Know
In 1633 an old man knelt on a stone floor in Rome and said out loud that he was wrong about something he knew to be right. Galileo had seen the moons of Jupiter through…
Read
June 15, 2026 · 4 min read
Frameworks are Training Wheels: When to Use LangGraph and When to Go Rogue
Most teams building an AI agent reach for a framework on reflex. LangGraph, LangChain, the whole shelf. It feels like the responsible choice, the adult one. And for your…
Read
June 15, 2026 · 3 min read
The $2 Million Handshake: Don’t Rent Your Brain to a Consulting Firm
Deloitte, PwC, and Accenture figured out something before most of their clients did. The money is not in advice anymore. It is in becoming the landlord.
Read
June 12, 2026 · 5 min read
The XML Diet: a leaner knowledge base for agents that fix themselves
Feed a local model a 500-page PDF and it bogs down fast with slower answers and more confused output.
Read