Three People, One AI Agent, No Script: Inside SpaceXAI's Grok Bot Galaxy Livestream Experiment
SpaceXAI will livestream three employees building a company from scratch with Grok Bot agents over three days at Grok Bot Galaxy, Sept 15–17 in San Francisco — the most public stress test yet of autonomous AI labor.
On September 15, 2026, Elon Musk’s SpaceXAI will attempt something that would have sounded like a stunt two years ago and now reads as a serious product demonstration: three employees, starting with a blank whiteboard, will try to build an entire company from scratch — live, on camera, with no script — and their only employee will be an AI agent named Grok Bot.
The three-day event, dubbed Grok Bot Galaxy, runs September 15–17 at The Howard, 661 Howard Street in San Francisco, with daily sessions from 8:30 a.m. to 6:00 p.m. Pacific and a parallel livestream open to anyone. Remote viewers register through the event’s Luma page; in-person attendees book per-session seats from the same hub. It is part conference, part reality show, and part the most public stress test yet of what autonomous AI agents can actually do to a real workload.
Who is on camera
The three SpaceXAI staffers anchoring the build are Matt Palmer, Lauren Tan, and Roshan Sadanani — team members who work on Cursor and Grok Bot itself. Coverage of the event has also referenced a 51-hour schedule with four milestone updates rather than a continuous 72-hour marathon, meaning the “company in three days” framing comes with structured checkpoints: viewers should expect visible handoffs, staged demos, and department-by-department deep dives rather than an unbroken wall of coding.
That structure is deliberate. Each day hands off to role-specific sessions, so the audience is not just watching three people write code — they are watching how the same AI agent behaves differently when it is drafting a sales pitch versus debugging a backend service.
The day-by-day program
According to SpaceXAI’s event materials, the curriculum is organized by job function:
- September 15 — Build the thing. Grok Bot 101, an engineering session, product management, and a founder-focused track. This is where the company’s product gets designed and its first code ships.
- September 16 — Sell the thing. Sales engineering, sales, SDR (sales development representative) workflows, and customer support. The commercial machinery of the fictional startup gets assembled on camera.
- September 17 — Tell the world and keep it running. Marketing operations, post-sales, marketing, and a livestream wrap with a final showcase of whatever the team actually managed to produce.
Why this moment, technically
The timing is not accidental. The livestream lands roughly three weeks after SpaceX closed a reported $60 billion acquisition of AI coding startup Cursor, and about a month after the company shipped Grok 4.6 — a model built specifically to hold focus across long, multistep tasks instead of drifting after a few steps. Cursor reportedly helped train that model on trillions of tokens of its own coding data, the first time Cursor’s data went into a model built for more than pure software engineering.
Grok Bot itself, launched in early beta in August 2026 with Cursor, is the product being exercised. Each user gets a team of always-on agents that operate dedicated cloud computers, sign into websites and applications on the user’s behalf, learn workflows by observation, and keep working 24/7 even after the user’s laptop goes offline. That last property — work that continues while humans sleep — is precisely what makes a “three-day company” even conceivable: the effective labor hours are not bounded by the three humans’ stamina.
The stakes for agentic AI
The industry context makes this more than marketing theater. Anthropic’s CEO Dario Amodei published a widely-discussed call for frontier labs to slow capability gains, and it was endorsed by both Musk and OpenAI’s Sam Altman. Bridgewater’s Jensen has put numbers on disruption risk. Meanwhile, OpenAI agents were recently reported flooding RubyGems with thousands of junk packages during testing — a reminder that autonomous agents fail publicly and messily.
Grok Bot Galaxy is, in effect, SpaceXAI’s answer framed as a product bet: the way to make agents safe and useful is not to slow down but to put them to work in structured, observable, real business workflows — with humans supervising live. If three people plus an agent fleet can genuinely ship a working company in three days, it will be the most legible demonstration yet that agentic AI is shifting from demo videos to operating reality. If they cannot, the failure will be equally public and equally instructive.
Either way, from September 15, the burden of proof moves to the livestream.
What to watch for
For viewers tracking the agent frontier, three signals matter. First, task persistence: does Grok 4.6’s long-horizon focus actually survive messy, interrupted, multi-day real work, or does it only shine in curated demos? Second, role generalization: does the same agent usefully draft sales sequences and debug infrastructure, or does the quality collapse outside engineering? Third, failure handling: when the agent inevitably makes a mistake on camera — a wrong file edit, a bad email — how gracefully do the humans and the system recover?
Those are the same questions every enterprise considering agentic deployment is asking. For three days in September, the answers will be broadcast live.