★ Follow-Up Thoughts on Watermarking Schemes for AI-Generated Text
John Gruber discusses AI text watermarking, arguing it degrades output quality and critiques Anthropic's approach.
John Gruber discusses AI text watermarking, arguing it degrades output quality and critiques Anthropic's approach.
A visual essay explaining LLM internals like tokenization, embeddings, and transformer architecture in an accessible way.
Learn how to accurately calculate token counts for strings using language models with a provided Jupyter Notebook tool.
A step-by-step educational guide to building a Byte Pair Encoding (BPE) tokenizer from scratch, as used in models like GPT and Llama.
A 1-hour presentation on the LLM development cycle, covering architecture, training, finetuning, and evaluation methods.
A 1-hour video presentation covering the full development cycle of Large Language Models, from architecture and pretraining to finetuning and evaluation.
Explores the rise of gold-backed stablecoins like Tether Gold and HSBC Gold Token, analyzing their advantages over traditional gold investments.
A guide on customizing comment colors in VS Code themes using token color customizations in settings.json.
A tutorial on using the PLY (Python Lex-Yacc) library to tokenize source code, covering token definition and lexer creation.