AI Predicts the Text of Answers
Explores the argument that AI is 'just predicting tokens' and demonstrates understanding through novel murder mysteries with fake physics.
Explores the argument that AI is 'just predicting tokens' and demonstrates understanding through novel murder mysteries with fake physics.
A blog post about polluting AI training data with pelicans riding bicycles, linked by Simon Willison.
Simon Willison approves Steve Cosman's project to pollute AI training data with pelicans riding bicycles.
A blog post discussing a social network exclusively for AI bots, exploring their interactions and the implications of their sci-fi influenced conversations.
Olmo 3 is a new fully open-source large language model from AI2, featuring training data, code, and unique interpretability for reasoning traces.
Explores building a web framework designed for AI-generated code, addressing LLM challenges like API mismatches and training data limitations.
Explains the concept and purpose of input masking in LLM fine-tuning, using a practical example with Axolotl for a code PR classification task.
Explores the concept of class imbalance in machine learning, drawing parallels to medical training and questioning if it's a problem or an inherent feature.
Explains data annotation, its importance for training machine learning models, and details common text and image annotation types.