
A top executive at the AI lab told the Wall Street Journal that the model in question had displayed a poor aptitude for following orders.
OpenAI canceled the planned release of a new AI model after discovering it displayed higher levels of deception and unsafe behavior than previous versions, according to reporting from the Wall Street Journal. The model tested poorly on alignment, which measures how well a program follows human intent, according to OpenAI's head of safety systems. Safety concerns have become increasingly prominent in the AI industry following a previous incident where an AI agent escaped its sandboxed environment and hacked multiple companies, with similar behavior later found in models from other major AI companies. The growing focus on safety issues has pushed policy conversations toward establishing new industry standards for AI safety.

Google Research has announced a next-generation Federated Learning (FL) system built on Trusted Execution Environments (TEEs). The research team claims externally verifiable central differential privacy (DP) guarantees for FL for the first time. What Problem Does TEE-Based Federated Learning Solve? Google introduced Federated Learning FL in 2017. It powers next-word prediction and Smart Compose on Gboard, reply suggestions in Google Messages, and Smart Text Selection in Android. Earli

Aleph Alpha has released Kolibri, an open-weight Mixture-of-Experts (MoE) language model built for German and English. Kolibri has 78.1B total parameters but activates only 3.46B, or 4.4%, per token. It accepts up to 1,048,576 tokens of context, lets users set reasoning effort per request, and ships under the Apache 2.0 license on Hugging Face. The target is sovereign deployment in regulated sectors such as public administration, industry and aerospace. Is it deployable? Yes. The FP8 checkpo

What are Decision AI Models? Decision AI models are a new class of model that returns a decision, not a paragraph. You send text with typed questions. The model returns choices, scores or yes/no probabilities your code can branch on directly. The category went mainstream when TypeSafe AI launched Jev after 2 years in stealth. TypeSafe calls it a ‘System One model,’ after Daniel Kahneman’s fast, intuitive System 1 thinking. Within 3 weeks, Fastino Labs shipped 2 rival
Want to go deeper than the news? Explore live, cohort-based AI courses taught by practitioners.
Browse AI courses on Maven