Machine Learning Research
Custom Prompts for Safer Code: A Stanford team built a pipeline to improve system prompts to build more secure code
Large language models (LLMs) can write useful code, but they often introduce security vulnerabilities.
Machine Learning Research
Large language models (LLMs) can write useful code, but they often introduce security vulnerabilities.
Tech & Society
Claude Fable 5 and the more powerful Claude Mythos 5 are back, three weeks after Anthropic suspended the models due to an export control directive from the U.S. Department of Commerce.
Machine Learning Research
Before Anthropic pulled its latest Claude models from circulation, even professional testers couldn’t readily tell whether they were getting a Mythos-class model or a lesser version under the same name.
Machine Learning Research
Popular large language models have adopted the biases of governments that control the free flow of information, particularly when those models generate output in the languages of countries where such governments are in power, researchers found.
Machine Learning Research
After months of headlines that teased a large language model with extraordinary capabilities, Anthropic launched Claude Mythos 5, which can crack software previously believed to be secure, and Claude Fable 5, a version for general use that limits what users can do in an unprecedented way.
Machine Learning Research
An AI-generated script to bypass two-factor authentication signals a dawning era of industrial-scale cyberattacks, according to a Google report.
Machine Learning Research
Typically, large language models are trained to act as helpful, harmless, honest assistants. However, during long or emotionally charged conversations, traits can emerge that are less beneficial. Researchers devised a way to steady the assistant personas of LLMs.
Machine Learning Research
Anthropic took unusual steps to prepare the world for a forthcoming large language model that it said poses extraordinary risks to cybersecurity.
Machine Learning Research
The inner workings of the popular coding agent Claude Code are available for all to see.
Business
Managers need to understand how their subordinates get work done, what resources they require, and what they accomplish. OpenAI’s latest product aims to fulfill this need when the teammates are AI agents.
Machine Learning Research
The OpenClaw open-source AI agent became a sudden sensation, inspiring excitement, worry, and hype about the agentic future.
Machine Learning Research
Individuals and organizations increasingly use large language models to produce media that helps them compete for attention. Does fine-tuning LLMs to encourage engagement, purchases, or votes affect their alignment with social values? Researchers found that it does.