タグ: #deepseek-r1
GPU・LLM・MLOps・Kubernetes、そしてマインドセット · 2 件
LLM ランドマーク論文ロードマップ 2026 - Transformer / Scaling Laws / Flash Attention / Mamba / DeepSeek-R1 / Titans 徹底解説
2017年の Attention Is All You Need から 2026年の Titans と DeepSeek-R1 まで、LLM 時代を作った 50 篇超のランドマーク論文をテーマ別に整理する。Transformer、BERT、GPT 1-3、Scaling Laws、Chinchilla、InstructGPT、PaLM、FlashAttention 1-3、LLaMA 1-4、GPT-4、Mistral と Mixtra
2026-05-16 · 31 分で読めます #llm-papers#transformer#scaling-laws#flash-attention#mamba推論モデル(reasoning models) 2026年ガイド — o3·o4·DeepSeek R1·Claude Thinking·Gemini Deep Think·QwQ 徹底比較
2024年9月のo1がtest-time computeという新しい軸を開いてから1年半。2026年現在、『推論モデル(reasoning model)』はもはや別の系統ではなく、すべてのフロンティアモデルが入れる『状態(mode)』になった。OpenAI o3·o3-pro·o4、DeepSeek R1·R1-0528·V3.1 reasoner、Anthropic Claude Sonnet 4.5·Opus 4.5のextende
2026-05-14 · 30 分で読めます #reasoning-models#o3#o4#deepseek-r1#claude-thinking