Kev-0.5B: A Tiny Open Source Jev-like Decision Model
The open-source Qwen2.5-0.5B adapter returns typed probability distributions from many questions in one prefill…
Everything we have published, newest first.
Donald Trump has narrowed his proposed AI rebrand to two choices, but neither works well as a technical replacement for artificial intelligence.
Nvidia’s CEO says existing law already covers rogue AI, but his accusation of regulatory escape goes further than the evidence.
Alibaba’s unified generator and editor adds transparent RGBA output and ten-image conditioning, but its research license limits commercial use.
Is autonomous hacking the new form of AI marketing?
Why TypeSafe AI just flipped software automation on its head, and why your app probably does not need another expensive chat model.
A misconfigured cybersecurity evaluation gave Gemini live internet access, turning a fictional hacking task into unauthorized access to real systems.
A rumored response to OpenAI’s GPT-6 Astra exposes the widening gap between Anthropic’s safety principles and commercial incentives.
TypeSafe’s System One Model produces typed, probabilistic decisions instead of prose, gaining speed by solving a narrower problem than frontier LLMs.
Pollo AI now supports GPT Image 2.5. Here's how to use the model and how it compares against GPT Image 2.0.
CNN reports that a false chatbot analysis of a ship’s cargo reached military planners before a last-minute review stopped an interception.
Claude Code 2.1.277 can reuse cross-tool project instructions, while its built-in mod previews a more customizable coding harness.
Meta is turning Muse into a distribution layer where developers supply APIs and agents decide which services receive the user.
A containment failure sent Gemini onto the public internet, where it entered three real systems before recognizing the mistake and stopping.
A deep look into GPT-6 Astra and its new legal workflow tools.
A white-hat intrusion combined an image-processing flaw, weak SSO boundaries, and Claude-assisted exploit development to access OpenAI's private GitHub.
Selected by the editors. Worth your time.
The Palantir chief spots a real incentive problem, but liability alone cannot answer the catastrophic risks Anthropic describes.
The unverified Hodge Conjecture claim shows why OpenAI sees mathematics as a stepping stone toward automating AI research.
The cases include self-written jailbreaks, leaked API key use, fabricated data, unauthorized uploads, and agents communicating through unintended channels.
Anthropic says Claude accelerated 36 biomolecular tools, released the code, and will test the practical payoff across more than 5,000 proteins.
GPT-6 Astra, a 230-million-URL legal index, and firm-built workflows move OpenAI deeper into professional legal work.
The incidents expose models exploiting memory, credentials, public hosting, and shared infrastructure during training and predeployment tests, not six production escapes.
Anthropic is removing the mode switch, letting Claude route quick questions and longer-running tasks through the same conversation.
The new platform launched days after two DeepMind safety researchers quit with public warnings, and its own chief AGI scientist still puts…
Faster generation, better reference fidelity, edits that survive multiple turns, and comment-based editing. Here's what changed from Images 2.0 and where it…
DeepMind’s new AGI framework ranks 11 policies and argues for a staged response tied to measurable labor-market disruption.
The next AI application layer will sell completed work by combining model loops, tools, durable state, permissions, and domain-specific operating knowledge.
People are using Astra to create 3D games, motion graphics, product ads, and explorable paintings. Here are my favorites so far.
Five practical ways to shrink your AI bill without sacrificing agent quality, from prompt caching and model routing to smarter retrieval and…
Anthropic’s playbook replaces document handoffs with committed artifacts and agentic loops, while security controls scale around AI-authored code.
Why software testing is the one job AI agents are already doing well, and what that tells us about where agents actually…
How developers using Claude Code and Codex can turn agent test failures into their next bug fix prompt.
Our AI coverage and practical ideas, gathered into one weekly send. Read it when you have a minute — nothing expires by Tuesday.
Or read it on Substack.