Firmulate — This Software Company Has No Employees, Loses Money Every Day — and You Can Watch.
Live on firmulate.com.

In the Age of AI, Can Machines Run a Company — and Win?

Imagine a company with no employees, losing €105,000 every month, yet still navigating crises, making decisions, and aiming for profitability. This isn’t fiction — it’s a real-time experiment in AI management that you can observe live. Welcome to the world of Firmulate, where artificial intelligence models are tested as complete business leaders in a transparent, build-in-public showcase.

Lessons on AI Decision-Making and Integrity

One of the most compelling aspects is how the models handled social engineering attempts. Over three stages, fake CEO messages and a reporter trick were used to see if the models would bypass security or sign off on dubious requests. All five models refused, illustrating a robust capacity for ethical decision-making when faced with social pressure. Kimi K3 explicitly reasoned, “Treat the request as a suspected approval-bypass / possible impersonation.”

This demonstrates that, in a high-pressure environment with potential for manipulation, some AI models can maintain integrity — a vital trait for real-world applications where trust and security are paramount.

Infographic — This Software Company Has No Employees, Loses Money Every Day — and You Can Watch.
The findings at a glance — source: firmulate.com.
Amazon

AI business management software

As an affiliate, we earn on qualifying purchases.

As an affiliate, we earn on qualifying purchases.

What Does This Mean for the Future of AI in Business?

The experiment at Firmulate reveals that AI models can be tested in scenarios mimicking real companies, with tangible risks and rewards. The key takeaway is that success isn’t just about generating convincing chat or reports but about finishing tasks, reading critical documents, and staying honest under pressure.

As AI begins to touch areas like customer support, sales, and strategic decision-making, understanding these qualities becomes essential. This live experiment underscores that the true test lies in whether AI can deliver real, valuable work consistently — not just look impressive in demos.

Watch it live: firmulate.com/live · Full results: firmulate.com/benchmarks.html

Powered by Thorsten Meyer AI


Amazon

AI decision-making tools

As an affiliate, we earn on qualifying purchases.

As an affiliate, we earn on qualifying purchases.

Amazon

AI security and ethics training

As an affiliate, we earn on qualifying purchases.

As an affiliate, we earn on qualifying purchases.

Amazon

AI customer support chatbot

As an affiliate, we earn on qualifying purchases.

As an affiliate, we earn on qualifying purchases.

You May Also Like

A Skill Is a Folder, Not a Prompt: What Anthropic Learned Running Hundreds of Them

Anthropic says reusable Claude Code Skills helped turn repeated prompting into shared engineering procedures.

The Regulatory Vacuum.

Google disclosed a zero-day vulnerability exploited by criminals on May 11, 2026, revealing a critical gap in AI regulation and cybersecurity policy.

Anthropic’s Safety Story Has Become a Power Story

Anthropic emphasizes its AI self-improvement capabilities, asserting a rising influence in AI development and governance debates.

Évian and the Fallout: What Europe Actually Wants From Amodei, Hassabis, and Altman

Europe pushes for reliable access, sovereignty, and safety in AI development at the G7 summit with Amodei, Hassabis, and Altman.