# LLMs (MoC) ## Overview Notes about Large Language Models (LLMs). ## Notes <!-- QueryToSerialize: LIST FROM #ai/llms AND !#type/quote AND !#type/creation/quote WHERE public_note = true SORT file.name ASC --> <!-- SerializedQuery: LIST FROM #ai/llms AND !#type/quote AND !#type/creation/quote WHERE public_note = true SORT file.name ASC --> - [[2026-04-21 Kimi K2.6, Qwen, and Gemma 4 - Local AI Is Catching Up]] - [[2026-05-04 Heavy AI Agents Are an Anti-Pattern]] - [[2026-05-06 Gemma 4 Gets Multi-Token Prediction Drafters - 3x Faster Inference Without Quality Loss]] - [[2026-06-17 The US Government Banned Claude Fable 5 and Mythos 5]] - [[2026-08-01 Kimi K3, Qwen 3.8, and a Week That Reset the Open-Weight Frontier]] - [[2026-08-05 OpenAI shrank the Codex context window on purpose]] - [[2026-09-02 Claude Fable 5.1 - cheaper caching, calmer safeguards, same old Claudish]] - [[2026-09-15 See Which Tokens an LLM Draws From, Right in Your Browser]] - [[2026-09-22 Claude Opus 5.5 - Fable-level work for less than half the price]] - [[2026-09-22 Jev and System One models - a semantic if statement for your code]] - [[2026-09-29 Claude Sonnet 5.5 - Opus-level everyday work at half the price]] - [[Agent-Native Product Decomposition]] - [[AgentRun]] - [[AI Agent Routing]] - [[AI Assistant plugin for Obsidian]] - [[AI Assistants]] - [[AI Commander plugin for Obsidian]] - [[AI Engineering]] - [[AI Expert Offloading]] - [[AI Fine-Tuning]] - [[AI for Templater plugin for Obsidian]] - [[AI Frontier Model]] - [[AI Hallucination]] - [[AI Image Analyzer plugin for Obsidian]] - [[AI image generation models]] - [[AI Inference]] - [[AI Instruction Tuning]] - [[AI KV Cache]] - [[AI Major Techniques]] - [[AI Master Prompt]] - [[AI Mixture of Experts (MoE)]] - [[AI Model Calibration]] - [[AI Model Cascades]] - [[AI Multi-Token Prediction Drafters]] - [[AI Multimodal]] - [[AI Open Weight Models]] - [[AI Post-Training]] - [[AI Providers plugin for Obsidian]] - [[AI Quantization]] - [[AI Reasoning Models]] - [[AI Sampling Parameters]] - [[AI Scaling Laws]] - [[AI Sparse Attention]] - [[AI Speculative Decoding]] - [[AI Sycophancy]] - [[AI Tagger plugin for Obsidian]] - [[AI Temperature]] - [[AI Tokenization]] - [[AI Tool Use]] - [[AI Tools I use]] - [[AI Verifiability]] - [[AI Verifiability as a Capability Ceiling]] - [[AirLLM]] - [[Ambient Programming]] - [[Anthropic]] - [[Anthropic's position on open-weights models (2026)]] - [[Antigravity CLI]] - [[Apertus]] - [[Artificial Analysis]] - [[Auto Classifier plugin for Obsidian]] - [[Baichuan]] - [[Berget AI]] - [[Bonsai 27B]] - [[Browser MCP]] - [[Business AI Master Prompt Questions of Tiago Forte]] - [[Candidate-Then-Select Extraction]] - [[Cannoli plugin for Obsidian]] - [[Chain-of-Thought (CoT) prompting]] - [[Chat log to notes prompt]] - [[Chat Stream plugin for Obsidian]] - [[ChatGPT]] - [[ChatGPT Images 2.0]] - [[ChatGPT Images 2.5]] - [[Cheaper Inference]] - [[Claude]] - [[Claude 5]] - [[Claude Code]] - [[Claude Code effort levels]] - [[Claude Code Prompt Caching]] - [[Claude Fable 5]] - [[Claude Fable 5.1]] - [[Claude MCP Browser Extension]] - [[Claude Opus 4.7]] - [[Claude Opus 4.8]] - [[Claude Opus 4.8 and Dynamic Workflows]] - [[Claude Opus 5]] - [[Claude Opus 5.5]] - [[Claude Sonnet 5]] - [[Claude Sonnet 5.5]] - [[Claudian plugin for Obsidian]] - [[Cline]] - [[Code2Prompt]] - [[Codestral]] - [[CodeWeaver]] - [[Codex App]] - [[Codex CLI]] - [[Codex Cloud]] - [[Codex IDE Extension]] - [[Combining multiple LoRAs]] - [[ComfyUI]] - [[Companion plugin for Obsidian]] - [[Composer 2]] - [[Composer 2.5]] - [[Composite Scoring]] - [[Confidence-Gated Routing]] - [[Context Compression]] - [[Context Engineering]] - [[Context size reduction]] - [[Context Window]] - [[Context-Understanding-Generation Asymmetry]] - [[Context7]] - [[Copilot Notebooks]] - [[Copilot plugin for Obsidian]] - [[Cursor]] - [[Cursor Bridge plugin for Obsidian]] - [[DALL-E]] - [[Data Poisoning]] - [[Decision Models (DMs)]] - [[Deepseek]] - [[DeepSeek V3]] - [[DeepSeek v4]] - [[DeepSeek V4.1 Flash]] - [[Dense AI Models]] - [[Desktop Extensions for MCP]] - [[Devstral]] - [[Devstral 2]] - [[Dify]] - [[Docker Agent]] - [[DSLs Make LLM Output Reliable]] - [[DSPy]] - [[EmbeddingGemma]] - [[fal.ai]] - [[FastContext]] - [[files-to-prompt]] - [[FLUX.1]] - [[FLUX.2]] - [[FLUX.3]] - [[Gemini]] - [[Gemini 3]] - [[Gemini 3.1 Flash Live]] - [[Gemini 3.1 Flash TTS]] - [[Gemini 3.5 Flash]] - [[Gemini 3.5 Pro]] - [[Gemini 3.6 Flash]] - [[Gemini 4 Argon]] - [[Gemini App for Windows]] - [[Gemini CLI]] - [[Gemini Code Assist]] - [[Gemini Mobile App]] - [[Gemma]] - [[Gemma 4]] - [[Generative AI (Gen AI)]] - [[Generative AI Risks]] - [[GEPA]] - [[GitHub Copilot]] - [[GitHub Copilot plugin for Obsidian]] - [[Glama Chat]] - [[Google AI Studio]] - [[Google Cloud Knowledge Catalog]] - [[Google DeepMind]] - [[GPT Image 1]] - [[GPT-5]] - [[GPT-5.4]] - [[GPT-5.5]] - [[GPT-5.6]] - [[GPT-6 Astra]] - [[GPT-6 Luna]] - [[GPT-6 Sol]] - [[GPT-6.1 Sol]] - [[GPT4]] - [[Granite]] - [[Granite 4.1]] - [[Grok]] - [[Grok 3]] - [[Grok 4]] - [[Grok 4.3]] - [[Grok Build]] - [[Headroom]] - [[Heavy AI Agents Are an Anti-Pattern - Why Fewer Agents With More Skills Wins (Article)]] - [[Hermes]] - [[Hermes Agent]] - [[Hierarchical Classification]] - [[How I leverage my Notes with AI]] - [[How to combine multiple FLUX.1 LoRas]] - [[How to create a top notch AI writing assistant]] - [[How to create your Business AI Master Prompt]] - [[How to create your Personal AI Master Prompt]] - [[How to structure your AI Master Prompt]] - [[How to train a FLUX.1 LoRA]] - [[Imagen]] - [[Imgen Arena]] - [[InstructGPT]] - [[Jagged Intelligence]] - [[Jev]] - [[just-prompt MCP server]] - [[Kimi]] - [[Kimi K2.5]] - [[Kimi K2.6]] - [[Kimi K3]] - [[LangFlow]] - [[Large Language Models (LLMs)]] - [[LiteLLM]] - [[Living shim between humans and AI models]] - [[LLM Knowledge Bases Over Unstructured Data]] - [[LLM Monitoring]] - [[LM Studio]] - [[LM Studio Mobile App]] - [[Local GPT plugin for Obsidian]] - [[Loom plugin for Obsidian]] - [[Low Rank Adapter (LoRA)]] - [[Machine Native Intelligence]] - [[Magistral Medium]] - [[Magistral Small]] - [[Markdown-based Installation (MD Scripts)]] - [[Martian]] - [[MCP server for Obsidian (Python)]] - [[mcptools]] - [[Menugen Architecture Pattern]] - [[MenuGen Deployment Gap]] - [[Midjourney]] - [[Ministral 3]] - [[Miso One]] - [[Mistral Large 3]] - [[Mistral Medium]] - [[Mistral Medium 3.5]] - [[Mistral Small 3]] - [[Mistral Small 3.1]] - [[Mistral Small 4]] - [[Mode Collapse]] - [[Model Context Protocol (MCP)]] - [[Model routing]] - [[Moonshot AI]] - [[Nano Banana Pro]] - [[Not Diamond]] - [[Note Companion plugin for Obsidian]] - [[NotebookLM]] - [[Nous Chat]] - [[Nous Portal]] - [[NVIDIA API Catalog]] - [[Obliteratus]] - [[Odysseus (AI)]] - [[Odysseus Compare]] - [[Odysseus Cookbook]] - [[Ollama]] - [[OmniParser]] - [[Open Knowledge Format (OKF)]] - [[Open WebUI]] - [[OpenAI]] - [[OpenAI Codex]] - [[OpenAI Product Consolidation]] - [[OpenRouter]] - [[OpenRouter Fusion]] - [[openrouterai OpenRouter MCP Server]] - [[Orpheus TTS]] - [[Personal Knowledge Management is here to stay]] - [[Pixtral]] - [[Portkey]] - [[Pre-warm the Prompt Cache]] - [[Preparing for the future of knowledge work]] - [[Prompt Engineering Best Practices]] - [[Prompt Engineering Strategies]] - [[Prompt Lazy Loading AI Design Pattern (PLL)]] - [[Prompt-driven development (PDD)]] - [[Qwen]] - [[Qwen 3.8]] - [[Qwen Image 2.0]] - [[Qwen Image 3.0]] - [[Qwen3.6-27B]] - [[Qwen3.6-35B-A3B]] - [[Qwen3.8-27B]] - [[Qwen3.8-Flash-Next]] - [[RAG Pipelines]] - [[Ramp Router]] - [[Receptionist AI Design Pattern]] - [[Reinforcement Learning for Calibrated Decisions (RLCD)]] - [[Reinforcement Learning From Human Feedback (RLHF)]] - [[Reinforcement Learning with Verifiable Rewards (RLVR)]] - [[Replicate.com]] - [[Requesty]] - [[Reranking]] - [[Retrieval-Augmented Generation (RAG)]] - [[Reverse Prompter plugin for Obsidian]] - [[Reward Hacking]] - [[Roo Code]] - [[Screenshot Driven Development (SDD)]] - [[Self-Consistency]] - [[SemIf]] - [[Shadow Evaluation]] - [[ShieldGemma 2]] - [[Small Language Models (SLMs)]] - [[Smart Composer Plugin for Obsidian]] - [[Smart Connections plugin for Obsidian]] - [[Smart Second Brain plugin for Obsidian]] - [[Smart Templates plugin for Obsidian]] - [[Software 3.0]] - [[Sora]] - [[Sparse AI Models]] - [[Speculative Fan-Out]] - [[Stability AI]] - [[Stable Diffusion]] - [[SWE-Bench]] - [[System One Models]] - [[Text extractor plugin for Obsidian]] - [[Text Generator plugin for Obsidian]] - [[The limit of intelligence is contact with reality]] - [[Token Budget]] - [[Tone Matching]] - [[Treat AI context as a budget to manage]] - [[Types of Context for AI Agents]] - [[Typicality Bias]] - [[Veo 3]] - [[Vercel Open Agents]] - [[Vercel v0]] - [[Vertex AI]] - [[Video Knowledge Extraction prompt]] - [[Visual Studio Code (VSCode)]] - [[Whisk]] - [[Windsurf]] - [[Zero-Shot Classification]] - [[Zhipu AI (Z.ai)]] <!-- SerializedQuery END --> ## Quotes <!-- QueryToSerialize: LIST FROM #ai/llms AND (#type/quote OR #type/creation/quote) WHERE public_note = true SORT file.name ASC --> <!-- SerializedQuery: LIST FROM #ai/llms AND (#type/quote OR #type/creation/quote) WHERE public_note = true SORT file.name ASC --> - [[Prompting is like playing LEGO]] - [[The more you try to outsource thinking, The less thinking you actually do]] - [[You can outsource your thinking, but you can't outsource your understanding]] <!-- SerializedQuery END -->