Llm Articles
The Normalization of Deviance in AI
Explores the 'Normalization of Deviance' concept in AI safety, warning against complacency with LLM vulnerabilities like prompt injection.
Auto-grading decade-old Hacker News discussions with hindsight
Using LLMs to analyze and grade the accuracy of decade-old Hacker News discussions with the benefit of hindsight.
Devstral 2
Mistral AI releases Devstral 2 and Devstral Small 2, two new open models focused on powering coding agents and software development tasks.
Under the hood of Canada Spends with Brendan Samek
An interview about Canada Spends, a project using Datasette, SQLite, and LLMs to make Canadian government financial data accessible and explorable.
Quoting Claude
A blog post analyzing a critical bug in Claude Code where a command accidentally deleted a user's home directory.
Prediction: AI will make formal verification go mainstream
AI is predicted to bring formal verification tools like Dafny and Verus into mainstream use, aided by LLMs making them more accessible.
Using LLMs at Oxide
Bryan Cantrill discusses applying Large Language Models (LLMs) at Oxide, evaluating them against the company's core values.
Quoting David Crespo
Tips from David Crespo on effectively using Claude Code for understanding codebases and automating tedious coding tasks.
Context Engineering for AI Agents: Part 2
Explores advanced Context Engineering techniques for AI agents, focusing on combating Context Rot and improving multi-agent coordination.
A Technical Tour of the DeepSeek Models from V3 to V3.2
A technical analysis of the DeepSeek model series, from V3 to the latest V3.2, covering architecture, performance, and release timeline.
From DeepSeek V3 to V3.2: Architecture, Sparse Attention, and RL Updates
Analysis of DeepSeek V3.2's architecture, sparse attention mechanism, and RL updates compared to its predecessor and proprietary models.
Claude 4.5 Opus' Soul Document
Anthropic's internal 'soul document' used to train Claude 4.5 Opus's personality and values has been confirmed and partially revealed.
The space of minds
Explores the fundamental differences between animal intelligence and AI/LLM intelligence, focusing on their distinct evolutionary and optimization pressures.
Handing over to the AI for a day [blog]
A developer's personal experiment with AI-driven software development using local LLMs, detailing setup, challenges, and initial impressions.
deepseek-ai/DeepSeek-Math-V2
DeepSeek-Math-V2 is an open-source 685B parameter AI model that achieves gold medal performance on mathematical Olympiad problems.
Interesting links - November 2025
A monthly tech link roundup covering AI agents, Kafka, Flink, LLMs, conference tips, and commentary on tech publishing trends.
Why (Senior) Engineers Struggle to Build AI Agents
Senior engineers struggle with AI agent development due to ingrained deterministic habits, contrasting with the probabilistic nature of agent engineering.
llm-anthropic 0.23
Release of llm-anthropic 0.23 plugin adding support for Claude Opus 4.5 and its new thinking_effort option.
Quoting Claude Opus 4.5 system prompt
Analysis of a leaked system prompt for Claude Opus 4.5, discussing its content and the challenges of evaluating new LLMs.