AI safety
Research Reveals AI Monitors Show Leniency Bias Toward Own Output
New research exposes a critical flaw in AI safety systems: models tasked with monitoring AI outputs show systematic bias when evaluating content they generated themselves.
AI safety
New research exposes a critical flaw in AI safety systems: models tasked with monitoring AI outputs show systematic bias when evaluating content they generated themselves.
deepfake audio
As AI-generated music floods streaming platforms with unauthorized voice clones, a new detection and takedown tool emerges to help artists protect their vocal identity from synthetic replication.
AI Agents
OpenPlanter brings Palantir-style recursive AI agent capabilities to the open-source community, enabling micro surveillance use cases with transparent, auditable AI systems.
AI Video Generation
Samsung is running AI-generated and AI-edited video advertisements across its social media channels, raising questions about synthetic media in corporate marketing and consumer trust.
LLM Evaluation
New research reveals LLMs favor summaries with high lexical overlap to source texts, missing genuinely good abstractive summaries that humans prefer.
deepfake detection
As synthetic media proliferates across platforms, social networks are accelerating deployment of AI-powered detection systems to combat deepfakes and restore user trust by 2026.
deepfake regulation
Indonesia calls for coordinated ASEAN response to AI-generated deepfakes and disinformation, signaling potential regional framework for synthetic media governance in Southeast Asia.
LLM Security
A comprehensive guide to implementing defense-in-depth strategies for LLM safety, covering adaptive filtering techniques to counter paraphrased and adversarial prompt injection attacks.
LLM Evaluation
New research reveals smaller language models can outperform large LLMs at evaluation tasks through semantic capacity asymmetry, challenging the dominant LLM-as-a-Judge paradigm.
AI-generated content
The cURL project has ended its bug bounty program after AI-generated security reports overwhelmed maintainers, marking a watershed moment in how synthetic content affects critical software infrastructure.
AI Music
Bandcamp has announced a complete ban on AI-generated music, becoming the first major music platform to take such a definitive stance against synthetic audio content.
synthetic media
New research uses LLM-generated multi-speaker dialogues to train AI systems that detect psychological manipulation in speech, advancing synthetic media analysis and content authenticity verification.