3 People, 72 Hours, 26 AI Agents: What Actually Happened When Grok Tried to Build a Company
Elon Musk's Grok Bot Galaxy put three SpaceXAI employees in front of a live audience and gave them three days to build a company with AI agents doing the work. They started with no product, changed direction halfway through, and actually shipped a browser game. But did AI really build a company?

3 People, 72 Hours, 26 AI Agents: What Actually Happened When Grok Tried to Build a Company
What happens when you give three people three days, a collection of AI agents, and a completely blank starting point?
That was the experiment behind Grok Bot Galaxy, a three-day live event organized by SpaceXAI from September 15 to September 17, 2026. The idea was simple but ambitious: three people would attempt to build a company from scratch while using Grok Bot as their AI workforce.
Elon Musk promoted the experiment, and thousands of people followed the livestream as the team went from an empty starting point to a live product.
But there is an important distinction between the headline and what actually happened.
They did not create a billion-dollar company in 72 hours.
What they did create was a working product, a real development process, and one of the more interesting public demonstrations yet of what AI agents can do when they are given computers, tools, memory, and jobs rather than simply being asked questions.
So what actually happened?
Who Were the Three People?
The experiment centered around three SpaceXAI employees: Matt Palmer, Lauren Tan, and Roshan Sadanani.
The original plan was for the three-person team to start with a blank slate and use Grok Bot throughout the process, from research and product decisions to engineering, marketing, sales, and other business functions.
The official event described the experiment as starting with an idea and working through business planning, feature decisions, and real engineering work.
The interesting part was that the AI was not being used only as a coding assistant.
Grok Bots were treated more like digital employees with specific responsibilities.
What Is Grok Bot?
Grok Bot is different from simply opening Grok and asking it a question.
SpaceXAI describes Grok Bot as a system of persistent AI agents that can work across computers, browsers, applications, files, and other tools.
A bot can be given a role, instructions, and a goal, allowing it to perform work instead of simply producing a response.
The Grok Bot Galaxy event was designed to demonstrate exactly that idea: what happens when AI agents are given actual jobs inside a business?
Day 1: They Started With Almost Nothing
The team began with no finished product and no established business around the experiment.
They used Grok Bots to research possibilities, brainstorm ideas, organize work, and start building the basic company infrastructure.
Their first direction was surprisingly different from what they eventually launched.
The team initially explored a product around restaurants, chefs, venues, and event managers. The idea was essentially a pop-up operating system that could connect the different people involved in running food and event businesses.
They created a basic company structure, set up communication and development tools, and started working on a landing page and waitlist.
By the end of the first day, they had a name, a direction, a basic company skeleton, and AI agents assigned to different jobs.
Then Everything Changed
This is where the experiment became much more interesting.
Instead of spending the entire three days developing the original idea, the team abandoned it.
On Day 2, they pivoted toward building a game.
The new goal was much simpler: create a browser game and get it live by Thursday.
The public Grok Pot workspace described the new brief as a browser game starring the team's AI-generated characters, with the entire process happening publicly.
Around 26 AI agents were involved in the workspace.
Instead of one AI trying to do everything, different agents could take on different responsibilities.
One could work on research.
Another could work on engineering.
Another could handle design.
Another could help with marketing.
Another could test the product.
That is the real idea behind Grok Bot: AI as a team rather than AI as a single assistant.
Why the Pivot Matters
The pivot may actually be one of the most valuable parts of the entire experiment.
Real startups rarely follow their original plan perfectly. Founders change products, abandon features, react to users, and sometimes completely change direction.
The Grok Bot experiment demonstrated that AI agents can participate in that process too.
The humans changed the objective.
The agents then had to reorganize around the new objective.
That is very different from asking an AI to generate a piece of code and calling it an AI company.
Day 3: They Actually Shipped Something
By the third day, the team had moved from planning to launch.
The final product became Thursday Arena, a browser-based auto-battler that players could enter and play.
The game allowed players to select a captain, choose teammates, and fight through automated battles.
The released version included dozens of playable fighters and allowed users to practice as guests, while rated matches used X login and an Elo-style ranking system.
Most importantly, it was not simply a mockup.
It was a live product that people could actually use.
How Much Did AI Actually Do?
This is where the headline needs some context.
It would be inaccurate to say that AI independently created a billion-dollar company.
Humans were still directing the experiment, deciding what to build, changing direction, reviewing results, and making important decisions.
The AI agents handled large portions of the execution.
During the event, the team demonstrated agents working on engineering, product, sales, customer support, marketing, and other business tasks.
The public post-event account also reported that the team merged more than 300 pull requests during the three-day build, with many of those changes going through automated testing and review processes.
But the launch also exposed an important limitation.
Some bugs were found only when humans looked at the actual product.
That matters because it shows the difference between an AI system being able to produce a huge amount of work and an AI system being able to reliably judge whether the work is actually good.
So, Was the Experiment Successful?
The answer depends on what you mean by successful.
If the goal was to demonstrate that a very small human team can use AI agents to produce and launch a working software product extremely quickly, the experiment produced a real result.
They started with a blank slate, changed direction, built the new product, and launched it within the three-day window.
If the goal was to prove that AI can independently create a sustainable billion-dollar company, that was not demonstrated.
There is no evidence that the resulting product became a billion-dollar business, and the experiment was never the same thing as a conventional startup operating for years with paying customers, sustainable revenue, and an established market.
That distinction is important.
Building software quickly is not the same as building a valuable company.
The Biggest Problem AI Still Has
The most revealing part of the experiment may actually be what went wrong.
AI agents can generate code, research markets, create designs, write copy, perform repetitive tasks, and coordinate work at remarkable speed.
But speed creates another problem.
If agents can produce hundreds of changes quickly, someone still has to determine whether those changes are actually useful.
During the launch, human observation caught problems that automated testing did not catch.
That suggests an important lesson for businesses adopting AI agents:
The bottleneck may move from producing work to evaluating work.
What This Means for Startups
The experiment points toward a very different startup model.
A traditional early-stage startup might need developers, designers, marketers, salespeople, customer-support staff, and operations people.
A small team using AI agents could potentially handle parts of all of those functions.
That does not mean humans disappear.
Instead, the human role can shift toward choosing the problem, setting priorities, reviewing results, handling exceptions, and making decisions that require judgment.
The result could be companies with far fewer employees but much larger amounts of software-driven execution.
Could Three People Really Build a Startup With AI?
Technically, the Grok Bot experiment suggests that three people can now accomplish a surprising amount of work with AI agents.
But a startup is more than its software.
A real company needs customers, revenue, distribution, support, legal infrastructure, accounting, partnerships, retention, and a reason for customers to continue paying.
The three-day experiment demonstrated rapid product creation.
It did not demonstrate long-term business success.
That may actually be the most useful conclusion to take from it.
AI has made the cost and time required to create an initial product dramatically smaller.
The harder question is becoming what happens after launch.
The Billion-Dollar Question
This is where the viral headline can become misleading.
AI can potentially make it dramatically easier to build a prototype, launch a website, create an application, automate operations, and test ideas.
But a billion-dollar company is not created simply because the product was built quickly.
A billion-dollar outcome requires customers, market demand, revenue, growth, defensibility, and years of execution.
Grok Bot Galaxy did not prove that AI can skip those steps.
What it did show is that the first step may be getting much cheaper and much faster.
What Happens Next?
The most interesting question is not whether AI can build a website in three days.
We already know that AI can do that.
The bigger question is whether AI agents can operate an entire business for months and years while maintaining quality, understanding customers, controlling costs, and making good decisions.
That is a much harder test.
And it is one that companies like SpaceXAI, OpenAI, Anthropic, Google, and thousands of startups are now moving toward.
The Bottom Line
Three people did not create a billion-dollar company in three days.
They did something more measurable: they showed how a tiny human team can use a large collection of AI agents to move from an empty starting point to a working software product at remarkable speed.
The experiment also exposed the limits of today's agents.
They can write enormous amounts of code and perform large numbers of tasks, but humans still matter when deciding what should be built, whether it actually works, and whether anyone wants it.
That may be the real lesson from Grok Bot Galaxy.
AI is not necessarily replacing the startup founder.
It may be turning the founder into the manager of a digital workforce.
And if that model continues improving, the next generation of startups may look very different from the companies we are used to seeing today.
FAQ
Did Elon Musk really build a billion-dollar company with three people and AI?
No. The three-person Grok Bot Galaxy experiment did not produce evidence of a billion-dollar company. The team demonstrated rapid AI-assisted product development and launched a working browser game.
Who were the three people building the company?
The core team consisted of Matt Palmer, Lauren Tan, and Roshan Sadanani. The public Grok Pot workspace also lists Eric Zakariasson as a fourth founder working off-site.
What is Grok Bot Galaxy?
Grok Bot Galaxy was a three-day SpaceXAI event held from September 15 to September 17, 2026. The team demonstrated how Grok Bot AI agents could perform work across engineering, product, sales, support, and marketing.
What did they actually build?
The project ultimately became Thursday Arena, a browser-based auto-battler that was launched after the team pivoted from its original restaurant-focused idea.
Did they start with a finished business idea?
No. The experiment began from a blank slate. The team initially explored a restaurant and pop-up business concept before abandoning it and pivoting to a browser game.
How many AI agents were involved?
The public Grok Pot workspace described a team of 26 AI agents working across different functions.
Did AI build the entire product by itself?
No. Humans remained responsible for directing the project, making strategic decisions, changing the objective, reviewing results, and handling problems. AI agents performed substantial portions of the execution.
Was the Grok experiment actually successful?
It successfully demonstrated rapid AI-assisted product development and resulted in a live product. However, it did not establish long-term commercial success or demonstrate the creation of a billion-dollar company.
What was the biggest lesson from the experiment?
The experiment showed that AI agents can potentially allow a very small human team to coordinate work across many business functions. It also showed that human judgment and product validation remain important because automated systems did not catch every problem during the launch.
Can AI agents build a real startup?
AI agents can now contribute to many parts of startup creation, including research, coding, design, marketing, sales, and support. Whether they can independently operate a successful long-term company remains an open question.