Machine Learning Research
Custom Prompts for Safer Code: A Stanford team built a pipeline to improve system prompts to build more secure code
Large language models (LLMs) can write useful code, but they often introduce security vulnerabilities.
Machine Learning Research
Large language models (LLMs) can write useful code, but they often introduce security vulnerabilities.
Machine Learning Research
The biggest open dataset of source code went years without an update.
Machine Learning Research
The U.S. National Institute of Standards and Technology (NIST) has been testing quantum-proof replacements for today’s encryption algorithms.
Machine Learning Research
DeepSeek’s updated small model overtook the company’s own flagship.
Techical insights
I’m glad the idea of “tokenmaxxing” — that individuals and companies should use as many tokens as possible to boost productivity — is finally dying out.
The Batch Newsletter
The Batch News & Insights: I’m glad the idea of “tokenmaxxing” — that individuals and companies should use as many tokens as possible to boost productivity — is finally dying out.
Machine Learning Research
Assessments of the environmental impact of large language models typically focus on their final training runs, but there’s a lot more to building AI systems.
Hardware
Data center buildout plans reached a new order of magnitude as new partnerships form and old ones fade away in the search for capacity to train and deliver AI.
Machine Learning Research
To measure how good its models were at hacking, OpenAI reduced guardrails and ran them against a benchmark’s problem set.
Machine Learning Research
After launching Claude Fable 5, the future of Anthropic’s once-flagship Opus line was uncertain, except as a fallback for the company’s premium models.
Letters
My team recently had our own version of Hugging Face’s experience when closed models failed to defend the company following an accidental cyberattack from OpenAI, leading Hugging Face to use the open weight GLM 5.2 instead.
The Batch Newsletter
The Batch News & Insights: My team recently had our own version of Hugging Face’s experience when closed models failed to defend the company following an accidental cyberattack from OpenAI, leading Hugging Face to use the open weight GLM 5.2 instead.