カテゴリー
タグ
bun rust anthropic claude-code coding-agent typescript ci-cd astro web-development codex openai agent-skills image-generation ai llm benchmark open-source google css glm claude review nano-banana cost-analysis ai-agent prompt-engineering transformer python evaluation qwen ollama antigravity gpt-5 rag production
Tag: #open-source
2026
8 posts
小さな言語モデルをゼロから学習する — nanoGPT 級を MPS で回し、PPL・速度・メモリを自分で測る
KVキャッシュは記憶のコスト — 文脈が伸びるほど decode が重くなる理由を実測する
Attention は過去を読み直している — Q/K/V と O(T²) の壁を最小実装で覗く
LLM はトークンを1つずつ予測している — 自己回帰ループを手元で覗く
MLX vs ollama を M5 Pro で実測:Mac のローカル LLM、どっちのランタイムが速いか
コーディングLLMを M5 Pro 48GB で実測:「動く」と「使える」を分けるのは context の壁だった
DiffusionGemma を M5 Pro で実測:拡散LLMの「4倍速」は Apple Silicon で消える
Qwen3.6-27B がアツい:27B dense でClaude 4.5 Opus に肉薄したオープンウェイトの転換点
