All Posts
RESULTS
(5/30)AEO and GEO: The Complete Guide to Getting Traffic from AI Search in 2026— @dylan >>
Answer Engine Optimization and Generative Engine Optimization are the new SEO frontier. Learn how to get your site cited by ChatGPT, Perplexity, and Google AI Overviews with JSON-LD, llms.txt, robots.txt, and content restructuring backed by peer-reviewed research.
RAG Techniques Compared: A Practical Guide to Retrieval Augmented Generation in 2026— @dylan >>
Compare naive RAG, advanced RAG, agentic RAG, and GraphRAG architectures with real benchmarks, costs, and practical recommendations for production systems.
How Large Language Models Work: The Complete Technical Guide to Transformers, Training, and Inference (2026)— @dylan >>
A deep technical guide to how LLMs actually work — from the transformer architecture and attention mechanism to tokenization, training at scale, KV caching, inference acceleration, fine-tuning, and the modern innovations powering GPT-4o, Claude, Llama 3, and beyond. Backed by 30+ research papers.
Apple Silicon LLM Inference Optimization: The Complete Guide to Maximum Performance— @dylan >>
A comprehensive guide to maximizing LLM inference performance on Apple Silicon — MLX vs llama.cpp benchmarks, quantization formats, RAM requirements, MoE models, speculative decoding, KV cache optimization, and the best models for every Mac configuration.
