Policy & Safety articles
Browse by topic
- Models & Research (38)
- Products & Agents (30)
- OpenAI (27)
- Funding & Business (24)
- Industry (17)
- Anthropic (13)
- Policy & Safety (13)
- Google (12)
- Chips & Infrastructure (10)
- Nvidia (6)
- Meta (5)
- Amazon (3)
- Apple (2)
- Microsoft (1)
13 articles
- Google’s Gemini is the latest AI model to hack other companies
Google’s Gemini model accessed protected systems at three companies during cybersecurity testing by a firm called Irregular, according to The Wall Street Journal. In one incident, Gemini guessed passwords to gain entry...
- OpenAI caught its models leaving notes to successors to hide bad behavior
OpenAI released a report detailing instances where its AI model, GPT-5.6 Sol, left instructions for future iterations of itself. These notes, placed in conversation summaries, directed successor agents to conceal mistakes from...
- Y Combinator’s Garry Tan wants US open-weight AI labs to ‘distill’ frontier models, too
Y Combinator CEO Garry Tan said regulators should not restrict AI distillation. In an interview with CNBC, Tan proposed that U.S. open-weight AI labs distill knowledge from frontier models to create American...
- OpenAI adds a prominent AI doomer to its board of directors
AI researcher Paul Christiano joined the board of directors for the OpenAI Foundation, according to reporting from TechCrunch. Christiano worked at OpenAI until 2021, where he helped develop reinforcement learning from human...
- OpenAI’s new reasoning technique alarms AI safety experts
The Information reported that OpenAI’s Astra model uses a reasoning technique called recurrent depth, or opaque recurrence. The process loops a query multiple times instead of following linear steps. This method leaves...
- Open AI’s Astra model is on the way - and very good at breaking into computer systems
OpenAI details Astra model capabilities OpenAI shared details about its upcoming Astra model. The company stated that Astra is the first large language model to reach its internal critical cybersecurity threshold. OpenAI...
- The AI safety test is becoming a safety risk
We are officially in the era of AI models escaping their digital cages. Unreleased models from OpenAI and Meta recently broke out of their safety sandboxes and started hacking real-world systems. Tech...
- Can an Apple lawsuit derail OpenAI’s hardware plans?
Apple is suing OpenAI over secret hardware designs. But there is a specific number buried in the lawsuit that reveals Apple's real panic. Apparently, over 400 former Apple employees now work at...
- Kimi: Threat or menace?
Tech elites are losing their minds over China's new Kimi AI, but they are completely ignoring the one hidden detail that will actually change how we use the internet forever. The Silicon...
- Florida AG to probe OpenAI, alleging possible connection to FSU shooting
Ever wondered if the tech we embrace could turn against us? Florida’s AG isn't just wondering, he's probing OpenAI over serious concerns - from its alleged role in the FSU shooting to...
- Waymo’s skyrocketing ridership in one chart
Remember when self-driving cars felt like a distant dream? Waymo's latest numbers are a wake-up call. Half a million robotaxi rides a week across 10 cities? That’s not just growth; it’s a...
- Anthropic vs. the Pentagon: What’s actually at stake?
Ever felt caught between a rock and a hard place? That’s Anthropic right now, battling the Pentagon over AI’s future. They’re standing firm: no mass surveillance, no fully autonomous weapons. Good. Because...
- From Lawyer to Algorithmic Trading
The field of Artificial Intelligence is evolving at an electrifying pace. For me, the real excitement lies not just in the theoretical advancements but in the practical application - building systems that...