The author built a vulnerable app to test if large language models (LLMs) could hack it, spending $1,500 on the experiment. The app had a secure API but used Firebase as the data layer, which was the target of the exploit. The author tested various LLMs, including GPT, Deepseek, and Claude, with mixed results. The experiment aimed to reproduce a common class of exploits found in Firebase and Supabase apps.