A developer who scraped artwork for AI training is now collaborating with Cara, a creator platform designed to prevent unauthorized AI data collection, as the service faces ongoing attacks from trolls attempting to breach and publish its data.
Cara launched as a portfolio platform for artists explicitly opposed to having their work used to train generative AI models. The platform has faced repeated cyberattacks from groups attempting to extract and publicize creator data.
The collaboration marks a notable shift: the developer behind previous large-scale art scraping efforts is now working to strengthen Cara's defenses against similar data extraction. The partnership focuses on building security tools that prevent unauthorized access and copying of artist portfolios.
The move highlights growing tensions in the AI training ecosystem. Creators increasingly resist having their work used without consent or compensation, while some developers argue training data collection operates in legal gray areas. Cara's growth reflects artist demand for platforms that enforce anti-AI training protections.
Details on the specific tools being developed remain limited. The collaboration underscores how security against data scraping has become a core feature for creator platforms competing in an AI-saturated market.
Installing a large language model on your personal computer creates a private digital assistant without uploading data to external servers. This approach keeps your information secure while giving you full control over the AI.
LAION has published Big Video Dataset (BVD), an open collection containing 80 million videos and 10 million hours of footage designed for AI research. Models trained on BVD outperform the previous benchmark InternVid by up to 2.1 percentage points.
As generative audio tools improve, AI-created music floods the internet. Some creators deny using the technology until public pressure forces them to confess.
Google Research introduced WikiSkill, a framework enabling AI agents to maintain a persistent knowledge base of past mistakes and successes. The system allows agents to learn across multiple runs rather than starting from scratch each time.