The author discovered that Perplexity AI is ignoring robots.txt and using a generic user agent to scrape content from their website, despite claiming to respect robots.txt. The author tested this by blocking PerplexityBot in their robots.txt and server configuration, but Perplexity AI was still able to access and summarize their content. The author found that Perplexity AI is using a headless browser to scrape content, which does not send the correct user agent string. This allows Perplexity AI to bypass blocking measures.