Articles tagged: AI safety

17 articles

AI safety

Seeing beyond BMI: Estimating cardiometabolic risk with smartphone imagery

Smartphone imagery may offer a more nuanced view of cardiometabolic risk than BMI alone. This article examines the promise, safety conside...

Aug 24, 202610 min
AI agents

Running Untrusted Agent Code Without a Sandbox

Explore the risks and strategies for executing untrusted AI agent code without sandboxing, including isolation techniques, monitoring, and...

Jul 1, 20266 min
AI research

The Next Frontier in AI Research: Beyond Generative Models

AI research is shifting from scaling generative models to building efficient, reasoning-driven systems. New paradigms like neuro-symbolic...

Jun 26, 20268 min
AI tools

An LLM as Arbiter in RAG Retrieval: Picking the Right Candidate with Reasons

Explore how to use an LLM as an intelligent arbiter to select the best document from RAG retrieval candidates, enhancing accuracy with con...

Jun 26, 20268 min
AI research

The Frontier of AI Research: Current Trends and Future Directions

AI research is advancing rapidly, from deep learning breakthroughs to foundational models. This article explores key trends like multimoda...

Jun 20, 20267 min
AI research

The Frontier of Artificial Intelligence: Breakthroughs and Challenges in AI Research

AI research is advancing rapidly, from deep learning to reinforcement learning. This article explores key breakthroughs, current challenge...

Jun 19, 20266 min
AI research

MosaicLeaks: Can your research agent keep a secret?

MosaicLeaks reveals how AI research agents can inadvertently reconstruct sensitive information from fragmented data. This article explores...

Jun 19, 20268 min
AI research

The Expanding Frontiers of Artificial Intelligence Research

AI research is rapidly advancing, moving beyond narrow tasks toward general intelligence. Key areas include reinforcement learning, natura...

Jun 17, 20266 min
AI research

The New Frontier: How Artificial Intelligence Research is Reshaping Science and Society

AI research is rapidly advancing, moving beyond pattern recognition to causal reasoning and foundational models. This evolution promises b...

Jun 15, 20269 min
AI safety

AI as Normal Technology: Safety Through Mundanity

The path to AI safety may lie not in treating AI as extraordinary, but in integrating it as normal technology. This article explores how s...

Jun 14, 20268 min
AI research

olmo-eval: An evaluation workbench for the model development loop

olmo-eval is an evaluation workbench designed to integrate seamlessly into the model development loop, enabling rapid iteration and system...

Jun 12, 20267 min
AI safety

AI Alignment Posts

A clear and practical article about artificial intelligence for a professional audience.

Jun 12, 20264 min
AI safety

AI ALIGNMENT FORUM

A clear and practical article about artificial intelligence for a professional audience.

Jun 12, 20264 min
AI safety

Day Zero Support for OpenAI Open Safety Model

A clear and practical article about artificial intelligence for a professional audience.

Jun 8, 20264 min
AI safety

Hermes vs. OpenClaw, Cybersecurity Alarms Ring, More-Interactive Conversations, Can Agents Do Human Work?

A clear and practical article about artificial intelligence for a professional audience.

Jun 8, 20264 min
AI agents

The Open Source Community is backing OpenEnv for Agentic RL

A clear and practical article about artificial intelligence for a professional audience.

Jun 8, 20268 min
AI safety

Qwen3.7-Max Challenges Google for Third Place, AI Saves Whales, Fine-Tuning Breaks Copyright Alignment

A clear and practical article about artificial intelligence for a professional audience.

Jun 7, 20264 min