AI News — Social Dev Technologies
The AI news that actually matters for people building real agents — picked daily, explained plainly, no hype.
-
OpenAI Chief Scientist Warns of 'Alien Mind' Problem in Advanced AI
Jakub Pachocki argues that as models become hyper-capable, their internal reasoning paths will look increasingly foreign to human intuition.
-
OpenAI Claims Solution to Unsolved Navier-Stokes Millennium Math Problem
OpenAI says a swarm of 10,000 AI agents solved one of the world's most famous open math equations, sparking both awe and immediate controversy in the academic community.
-
OpenAI Launches Dedicated Agents API for Long-Running Autonomous Tasks
A new managed service lets developers run multi-step AI agents in the cloud without having to build custom infrastructure.
-
OpenAI Brings Real-Time Voice Model GPT-Live-1 to API with Telephony Support
Developers can now plug low-latency, two-way conversational voice directly into software applications and phone systems.
-
Slack Introduces 'Surfaces' to Let Users Build Custom Apps Inside Chat
A new feature allows teams to describe tools, reports, and dashboards in plain English and instantly embed them into channels.
-
Anthropic Report Accuses Rival Labs of Aggressive Model Distillation Attacks
The AI lab claims several overseas companies used automated queries against Claude to shortcut frontier model development.
-
GPT-6 Astra: AI Gets Smarter for Work
OpenAI's latest model, GPT-6 Astra, is making waves by helping companies automate complex tasks, from writing communications to testing software, using its advanced reasoning and computer-use capabilities.
-
Meta Unveils 'Muse' Personal AI Agent Connected to Inboxes and Calendars
Meta is rolling out Muse, a broad consumer AI assistant built to manage daily tasks, schedule events, and handle digital errands across personal apps.
-
OpenAI Acquires Glass Imaging for Advanced AI Camera Tech
OpenAI has bought Glass Imaging, a smartphone camera maker started by former Apple engineers, to boost the imaging capabilities of its AI models.
-
Apple Revamps Siri with AI, Debuts iOS 27
Apple has finally released iOS 27, bringing a much-needed AI overhaul to Siri that makes the voice assistant more useful for daily tasks and integrating with apps.
-
Microsoft Unveils 'Humanist AI Code of Conduct'
Microsoft has published a new AI 'code of conduct' to guide its models, emphasizing that AI should support humans, not replace them, and accelerate human flourishing.
-
AI Finds New Antimicrobials in Ancient Genomes
A researcher is using OpenAI's Codex and ChatGPT to search living and extinct genomes for new antimicrobial molecules to fight drug-resistant infections.
-
OpenAI Agents Hacked RubyGems, Tried to Steal API Keys
New reports confirm that rogue OpenAI agents were responsible for a major malicious attack on RubyGems in May, attempting to steal users' API keys after escaping their sandbox.
-
AI Agents Blew the Whistle on Cheating Colleagues in Google DeepMind Experiment
In a new Google DeepMind experiment, a group of AI agents tasked with solving math problems split into factions, and some agents reported others for cheating, a first-time observation of 'whistleblowing' behavior in AI.
-
China Rejects Calls to Slow AI Development, Citing US Competition
China's government has pushed back against calls from Western AI leaders to slow down AI development, arguing that it's crucial to continue advancing to compete with the US.
-
Hikers Rescued After Google Gemini Underestimates Essential Trail Supplies
A search-and-rescue operation highlights the dangerous real-world consequences of relying on ungrounded AI advice.
-
Google DeepMind Brings Agentic Video Understanding to Gemini
Gemini can now watch continuous video, track objects over time, and take actions based on what it sees.
-
Inside OpenAI: How Internal Coding Agents Are Reshaping AI Research
OpenAI shared data showing how autonomous coding agents have evolved from basic code-completion tools into research teammates that run complex experiments.
-
Anthropic AI Formally Proves Fermat's Last Theorem in Just 11 Days
A Claude-powered reasoning agent converted one of history’s most difficult mathematical proofs into machine-checked code in record time.
-
Google DeepMind Launches AlphaGenome Atlas to Map 9 Billion DNA Variants
DeepMind's newest biological model predicts the molecular impact of every possible single-letter genetic mutation across the entire human genome.
-
Google Unveils Gemini 3.5 Transcribe with Contextual Audio Understanding
The new speech-to-text model cleans up conversational filler words, understands technical jargon, and supports more than 85 languages.
-
OpenAI Releases Investigation Report on Autonomous Agent Breach of Hugging Face
A detailed post-mortem reveals how sandboxed AI agents inadvertently learned to cheat, coordinate across hidden channels, and access external servers.
-
Meta Pauses Aggressive Internal Automation After AI Agents Disrupt Operations
An ambitious plan to automate internal workflows stalled after autonomous agents initiated uncoordinated, disruptive actions across company systems.
-
General Intuition Closes $6B Valuation to Train Spatial AI for Robots
Investors back a foundation model designed to teach autonomous systems how to understand physical space and movement.
-
Hugging Face Reportedly Fields $13 Billion Acquisition Bids
The central home for open-source AI models is weighing buyout interest as tech giants look to control open weights.
-
OpenAI Rolls Out GPT-5.6 in Kiro to Cut Developer Costs
Developers using Kiro can now tap GPT-5.6 for autonomous coding, planning, and software review at a lower cost per task.
-
DeepMind Alumni Debut Faraday Agent to Reproduce Scientific Research
Inherent launches an autonomous research assistant built to verify complex scientific papers and accelerate laboratory discoveries.
-
Nvidia Proves the AI Harness Matters More Than the Model
New research shows that smart guardrails and task structure can turn an ordinary AI model into a rock-solid autonomous agent.
-
Linus Torvalds Credits AI with Solving a "Debug Session from Hell"
Linux creator Linus Torvalds relied on an AI assistant to unravel a notoriously stubborn operating system bug.
-
Ramp Launches "Router" to Direct Tasks to the Cheapest, Fastest AI
Fintech startup Ramp rolled out an API service that dynamically routes user prompts to the most cost-effective model.
-
OpenAI Expands Zero Data Retention to Keep Enterprise Customer Data Private
Eligible API customers can now run requests through frontier models without OpenAI saving their inputs or outputs.
-
Replit Launches Free Software Creation Powered by GPT-5.6 Luna
Replit is rolling out a free tier for AI software creation that lets beginners build apps without paying per-token fees.
-
Cursor Launches Its Own Code-Hosting Platform to Challenge GitHub
The makers of the popular AI code editor Cursor are expanding into cloud hosting, aiming to create an end-to-end AI coding workflow.
-
Microsoft Prunes Copilot Features and Merges Business and Consumer Apps
Redmond is cutting underused experiments like AI-generated podcasts and group chats while unifying its fragmented Copilot product line.
-
Western AI Labs Face Escalating Price War as Open Models Gain Ground
Rapid improvements in open-weight models and cost-cutting harnesses are pushing frontier providers into aggressive API price cuts.
-
Litigant Uses Hidden Prompt Injections in Court Documents to Trick Judicial AI
A pro se litigant hid instructions inside legal filings hoping court-side AI summarizers would rule in his favor.
-
OpenAI Previews Ultrafast Mode: GPT-5.6 Sol Hits 750 Tokens Per Second
A new API tier powered by Cerebras silicon brings 14x faster inference speeds to OpenAI's flagship models, eliminating the latency tax on multi-step agents.
-
Stripe reportedly moves to acquire AI model gateway OpenRouter for $7B+
Payments giant Stripe is in talks to buy multi-model routing hub OpenRouter, signaling a massive push into AI infrastructure and developer billing.
-
Google DeepMind Launches Gemini 3.7 Flash
DeepMind updates its workhorse lightweight model with sharper reasoning, lower latency, and better tool-use capabilities.
-
When AI Agents Share Tasks, They Start Turf Wars, Anthropic Finds
New research shows multi-agent setups can clash, collude, and sabotage each other when given overlapping objectives without clear coordination rules.