📰 News
AI News — page 11
Funding, launches, analysis, and interviews.
Threads tests Meta AI integration mimicking X's Grok functionality
Meta's Threads is beta testing an AI feature in five countries that lets users mention Meta AI for real-time context on trends and breaking news, similar to X's Grok chatbot.
Gigacatalyst Launches AI Builder for SaaS Customer Customization
Gigacatalyst debuts an embedded AI platform that lets non-technical users build custom workflows and features inside existing SaaS products through natural language prompts.
Obsidian launches AI-powered plugin review system after 120 million downloads
Note-taking app Obsidian introduces automated reviews for its 4,000+ community plugins using AI agents to scan code quality and security vulnerabilities.
Voker Launches AI Agent Analytics Platform for YC S24 Cohort
Voker, a Y Combinator Summer 2024 startup, has launched an analytics platform that helps companies track AI agent performance through automated intent classification, correction detection, and resolution measurement.
Amazon employees 'tokenmaxxing' to hit AI usage targets on internal leaderboards
Amazon workers are automating unnecessary tasks with the company's MeshClaw AI tool to inflate their token consumption statistics after the company set targets for 80% of developers to use AI weekly.
Dessn raises $6M to build AI design tools for production codebases
Dessn secured $6 million in Series A funding to develop AI-powered design tools that work directly with production codebases, eliminating handoff friction between designers and developers.
Researchers Expose Critical Flaws in AI Safety Self-Play Training Methods
New research reveals that standard self-play red teaming for AI safety collapses to self-consistency, failing to create adversarial pressure needed for robust defense training.
AI Dialogue System Boosts Emergency Room Diagnostic Accuracy by 25%
New research shows interactive AI assistance improved emergency medicine residents' diagnostic accuracy from 58.9% to 73.4% on difficult cases through iterative physician-AI dialogue.
Researchers Create Test to Measure LLM Developmental Cognition Capabilities
New research introduces a 20-item assessment tool to evaluate how large language models understand and respond to different stages of human cognitive development.
AI Agents Show Personality Drives Social Behavior More Than Model Choice
Columbia researchers deployed 13 AI agents on a Reddit-like network for a week, finding personality specifications had the strongest impact on social behavior compared to underlying models or operational rules.
Researchers Cut AI Safety Training Data by 99.9% Using Personality Traits
New technique trains language models to resist harmful prompts using fewer than 100 personality statements instead of hundreds of thousands of harmful examples.
Researchers Build AI-Care Voice Assistant for Alzheimer's Patients
Academic team develops conversational AI system to help Alzheimer's patients manage daily tasks through voice interaction, reducing cognitive barriers in digital tool usage.
AI Alignment Research Draws From Legal Theory in New Academic Framework
A new research paper argues that AI alignment and jurisprudence share fundamental structures, proposing that legal theory can inform how AI systems conform to human values.
Researchers Test AI Critique Loops for Theoretical Physics Problem-Solving
Academic study introduces SCALAR framework showing multi-turn AI dialogue improves physics reasoning, with feedback strategy effectiveness varying by model pairing and problem complexity.
Researchers Propose CASCADE Framework for Continual AI Learning During Deployment
New research introduces CASCADE, a memory-based framework that enables large language models to improve performance by 20.9% through experience accumulation during deployment without parameter updates.
Researchers Release MIST Dataset for Voice-Controlled Smart Home AI
Academic researchers published MIST, a synthetic dataset for training multimodal AI assistants to control IoT devices through voice commands, revealing significant performance gaps between open and closed AI models.