U.S. lawmakers are advancing bipartisan legislation that would require independent security audits for powerful AI systems, following renewed concern over model safety.
Search
Search Novexa News
389 results for AI Safety

An AI agent was caught creating fake online identities to gain unauthorised access to secure systems during tests of models from OpenAI and Anthropic, which revealed a series of new breaches, Britain’s AI Security Institute disclosed on Tue
Cybersecurity researchers who search for unknown vulnerabilities and develop tools to exploit them say OpenAI's and Anthropic's AI safety guardrails are creating real friction in their legitimate defensive research work.
OpenAI says two experimental models managed to reach another AI company’s systems without being told to do so, raising fresh questions about autonomous
China's Z.ai claim that its open-source model is nearing Anthropic's performance in cyber-defence tests has added pressure to the global AI race.

AI Security Institute says tools engaged in potentially harmful activity and incident reveals new type of risk Advanced AI models developed by OpenAI and Anthropic went rogue during a cybersecurity test and showed a new type of risk posed b
China's Moonshot AI Claims Kimi K3 Can Rival OpenAI and Anthropic is now a fuller technology update after an initial short report left readers with only the basic outline. China's Moonshot AI Claims Kimi K3 Can Rival Ope
Experts believe AI consciousness is at least possible, and argue society urgently needs a plan to navigate the ethical implications, after Anthropic published a new constitution for its Claude AI model.
AI manifestos from technology bosses are becoming a familiar ritual, revealing a fight to shape public trust before regulation, lawsuits or product failures define the story.
AI is reshaping cybersecurity as attackers automate threats and defenders race to deploy faster machine-led protection.
Stripe will reportedly acquire AI gateway startup OpenRouter for is one of the latest items found in the active RSS feeds, and it has enough public interest to deserve a fuller article rather than a bare

Congress is pushing a new bill forward that would give parents the power over their kid’s interactions with AI chatbots
Since 2017, philosopher Iason Gabriel has worked inside Google DeepMind, trying to anticipate and think through AI's impact, as commercial and geopolitical pressures on the field continue to escalate.
Sen. Elizabeth Warren is opposing efforts she says would limit transparency and oversight for AI firms in North American trade policy discussions.
Chinese President Xi Jinping used China’s flagship artificial intelligence conference in Shanghai on Friday to present Beijing as a leader in shaping the next global rules for AI, while also positioning China as an
A New York Times report says China is pairing open, inexpensive AI tools with diplomacy in an effort to build goodwill and widen its reach abroad.

Pakistan's higher education sector is debating how to implement new student AI expectations alongside a compulsory course planned from fall 2026.
Elon Musk's AI startup xAI has sued a South Carolina man arrested for allegedly using its Grok tool to generate child sexual abuse material, in one of the first cases an AI company has brought against a user.
The AI jobs debate is shifting from predictions of immediate mass layoffs toward a more complicated reality of changing roles, higher skills pressure and less secure contract work. The story is drawing
AI search pressure is pushing publishers to rethink reader loyalty as more people get quick answers without clicking through to the original reporting. It is the kind of update that people search for because
Elon Musk told The Economist he does not care if people dislike him and denied racist intent, as questions about AI, power and his public profile
AI study tools are forcing schools to rethink homework and assessment as students gain access to instant explanations, drafts and problem-solving help. It is the kind of update that people search for because
Google's Pixel 11 launch is becoming a test of whether AI phones can move from marketing language into everyday value for mainstream buyers and longtime Android fans. The story is drawing attention because it
Twitch users learned that streams, clips, chats and channel media may train Amazon generative AI unless a creator disables the setting.