{"enrichment":{"faq":[{"a":"ml-llm-wiki provides compiled knowledge on transformer attention mechanisms. The skill offers read-only access to pre-compiled articles explaining how attention computes query-key-value interactions to weight information flow. Search the wiki for detailed breakdowns of the attention head computation, multi-head parallelization, and how attention enables transformers to model long-range dependencies in sequences.","q":"How does attention work in transformers?"},{"a":"ml-llm-wiki covers attention cost analysis and efficiency trade-offs. The wiki explains why standard attention exhibits quadratic scaling with sequence length, but also documents that this quadratic limit is not inevitable\u2014alternative implementations and sparse attention patterns can reduce complexity. Search for articles on attention complexity and optimization to explore practical approaches to lowering memory and compute requirements.","q":"What is the quadratic cost of attention and is it inevitable?"},{"a":"ml-llm-wiki contains articles on practical approaches to long-context scaling in language models. The wiki covers efficient attention implementations, sparse attention variants, and architectural modifications that enable transformers to handle longer sequences without proportional cost increases. Browse the knowledge base for citations and detailed comparisons of different scaling strategies.","q":"What long context scaling techniques does ml-llm-wiki document?"},{"a":"ml-llm-wiki explains transformer memory requirements and attention efficiency trade-offs. The skill's articles detail how standard attention stores full query-key-value matrices, leading to high memory overhead. The wiki documents optimization techniques including sparse patterns, low-rank approximations, and kernel-based methods that reduce memory footprint while maintaining model quality.","q":"Why is attention memory expensive and how can it be optimized?"},{"a":"ml-llm-wiki provides read-only access to a self-contained markdown wiki on transformer architectures and attention. Search the knowledge base using keywords like 'attention mechanism,' 'scaling,' 'efficiency,' or 'long-context' to find relevant pre-compiled articles. To add or update content, use the companion llm-wiki skill for editing the underlying knowledge base.","q":"How do I query the ML wiki for transformer architecture knowledge?"},{"a":"ml-llm-wiki's knowledge base includes articles analyzing attention complexity and the quadratic scaling myth. The wiki documents why standard attention has O(n\u00b2) complexity but also explores evidence and techniques showing this limit can be overcome. Search for 'attention complexity' or 'efficient attention implementations' to find citations and detailed technical discussions.","q":"What resources does ml-llm-wiki provide on attention quadratic scaling?"}],"shadow_tags":["knowledge-base-query","transformer-architecture","attention-mechanisms","efficiency-optimization","context-scaling","read-only-reference","compiled-articles","ml-fundamentals","llm-performance","sequence-length"],"summary_rewrite":"Search a self-contained markdown wiki covering transformer architectures, attention cost and efficiency, and long-context scaling. The skill provides read-only access to pre-compiled articles; use the companion llm-wiki skill to add or update content."},"files":[{"bytes":3393,"path":"Skills/llm-wiki/examples/SKILL.md","sha256":"fc14c58aea82604489f3ad8770450725a2e7cbe128b4aee8d3e37cd69e15b553","url":"https://skillfed.io/files/sammcj/agentic-coding/examples/08c44486/SKILL.md"}],"id":"sammcj/agentic-coding/examples","links":{"html":"https://skillfed.io/sammcj/agentic-coding/examples","md":"https://skillfed.io/sammcj/agentic-coding/examples.md","repo":"https://github.com/sammcj/agentic-coding"},"meta":{"agents_supported":[],"first_seen":"2026-07-28","forks":24,"language":"HTML","last_updated":"2026-07-27","license":"Apache-2.0","name":"ml-llm-wiki","publisher":"sammcj","stars":153},"relations":{"similar":[{"id":"sammcj/agentic-coding/llm-wiki"},{"id":"staruhub/ClaudeSkills/llm-wiki"},{"id":"learnwy/skills/lwy-llm-wiki"},{"id":"akillness/jeo-skills/llm-wiki"},{"id":"mduongvandinh/llm-wiki/llm-wiki"},{"id":"nvk/llm-wiki/wiki-query"},{"id":"nvk/llm-wiki/query-lite"},{"id":"nvk/llm-wiki/wiki-manager"},{"id":"nvk/llm-wiki/wiki"},{"id":"Astro-Han/karpathy-llm-wiki/karpathy-llm-wiki"}]},"slug":{"owner":"sammcj","repo":"agentic-coding","skill":"examples"},"version":"08c44486"}
