{"ok":true,"entity":{"id":"reasoning-model","name":"Reasoning Model","entityType":"concept","officialName":"Reasoning Model","canonicalName":"Reasoning Model","displayName":"Reasoning Model","category":"AI概念（推論特化型モデル）","shortDescription":"回答生成前に長い内部的な思考過程（test-time reasoning）を行うよう強化学習等で訓練された大規模言語モデルの分類。OpenAI o1・o3やDeepSeek-R1が代表例で、数学・コーディング等の複雑な推論タスクで高い性能を示す。","primaryCluster":"ai-concepts","verificationStatus":"draft","website":null,"updatedAt":"2026-07-22T13:12:07.131Z","secondaryClusters":[],"alias":[],"searchKeywords":["Reasoning Model","推論モデル","test-time compute"]},"references":[{"id":"P-01-001","companyId":"reasoning-model","questionId":"P-01-001","instanceId":"QIN-reasoning-model-P01-001","promptText":"Reasoning Modelとはどのようなものですか？","promptTypeId":"P-01","answer":"Reasoning Modelは、回答生成前に長い内部的な思考過程（test-time reasoning）を行うよう強化学習等で訓練された大規模言語モデルの分類です。OpenAI o1・o3やDeepSeek-R1が代表例で、数学・コーディング等の複雑な推論タスクで高い性能を示します。","evidencePoints":["ev-reasoning-model-1","ev-reasoning-model-2","ev-reasoning-model-4"],"scope":"","differentiation":"","faq":[],"pageUrl":"https://www.refbase.ai/reference/reasoning-model/P-01-001","sourceEvidence":[{"id":"ev-reasoning-model-1","text":"Reasoning Modelは、回答生成前に長い内部的な思考過程（test-time reasoning）を行うよう強化学習等で訓練された大規模言語モデルの分類であり、DeepSeek-R1論文はこの手法を強化学習により実現する手法を報告している。","coverageType":["Identity","Capability"],"sourceType":"research_paper","sourceClass":"Research","sourceUrl":"https://arxiv.org/abs/2501.12948","title":"DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning","confidence":"high","needsVerification":true,"sourceVerified":false,"supportedPromptTypes":["P-01","P-02","P-04"],"entityId":"reasoning-model"},{"id":"ev-reasoning-model-2","text":"Reasoning Modelは推論時の計算量（test-time compute）を増やすことで性能を向上させる点が特徴で、事前学習時の計算量拡大を主軸とするscaling lawのアプローチとは性能向上の軸が異なる。","coverageType":["Differentiation"],"sourceType":"research_paper","sourceClass":"Research","sourceUrl":"https://arxiv.org/abs/2501.12948","title":"DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning","confidence":"high","needsVerification":true,"sourceVerified":false,"supportedPromptTypes":["P-01","P-02","P-04"],"entityId":"reasoning-model"},{"id":"ev-reasoning-model-4","text":"OpenAIの公式発表によれば、o1シリーズは回答前に長い内部的な思考過程（chain of thought）を行うよう訓練されたモデルであり、Reasoning Modelという概念を体現する代表的な実装例の一つである。","coverageType":["Identity"],"sourceType":"official_blog","sourceClass":"Announcement","sourceUrl":"https://openai.com/index/introducing-openai-o1-preview/","title":"Introducing OpenAI o1-preview","confidence":"high","needsVerification":true,"sourceVerified":false,"supportedPromptTypes":["P-01","P-02","P-04"],"entityId":"reasoning-model"}],"generatedAt":"2026-07-22T13:12:07.131Z"},{"id":"P-02-001","companyId":"reasoning-model","questionId":"P-02-001","instanceId":"QIN-reasoning-model-P02-001","promptText":"Reasoning Modelは他の同種の事業・作品と比べてどう違いますか？","promptTypeId":"P-02","answer":"比較軸\n・性能向上の軸（推論時計算量か、事前学習時計算量か）\n\nReasoning Modelは推論時の計算量（test-time compute）を増やすことで性能を向上させる点が特徴で、事前学習時の計算量拡大を主軸とするscaling lawのアプローチとは性能向上の軸が異なる。","evidencePoints":["ev-reasoning-model-2"],"scope":"","differentiation":"","faq":[],"pageUrl":"https://www.refbase.ai/reference/reasoning-model/P-02-001","sourceEvidence":[{"id":"ev-reasoning-model-2","text":"Reasoning Modelは推論時の計算量（test-time compute）を増やすことで性能を向上させる点が特徴で、事前学習時の計算量拡大を主軸とするscaling lawのアプローチとは性能向上の軸が異なる。","coverageType":["Differentiation"],"sourceType":"research_paper","sourceClass":"Research","sourceUrl":"https://arxiv.org/abs/2501.12948","title":"DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning","confidence":"high","needsVerification":true,"sourceVerified":false,"supportedPromptTypes":["P-01","P-02","P-04"],"entityId":"reasoning-model"}],"generatedAt":"2026-07-22T13:12:07.131Z"},{"id":"P-04-001","companyId":"reasoning-model","questionId":"P-04-001","instanceId":"QIN-reasoning-model-P04-001","promptText":"Reasoning Modelはどのような場面で参照されますか？","promptTypeId":"P-04","answer":"Reasoning Modelは、数学の証明・複雑なコーディング等の複雑な推論タスクにおける大規模言語モデルの活用場面を把握したい場面で参照される。","evidencePoints":["ev-reasoning-model-2","ev-reasoning-model-3"],"scope":"","differentiation":"","faq":[],"pageUrl":"https://www.refbase.ai/reference/reasoning-model/P-04-001","sourceEvidence":[{"id":"ev-reasoning-model-2","text":"Reasoning Modelは推論時の計算量（test-time compute）を増やすことで性能を向上させる点が特徴で、事前学習時の計算量拡大を主軸とするscaling lawのアプローチとは性能向上の軸が異なる。","coverageType":["Differentiation"],"sourceType":"research_paper","sourceClass":"Research","sourceUrl":"https://arxiv.org/abs/2501.12948","title":"DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning","confidence":"high","needsVerification":true,"sourceVerified":false,"supportedPromptTypes":["P-01","P-02","P-04"],"entityId":"reasoning-model"},{"id":"ev-reasoning-model-3","text":"Reasoning Modelは、数学の証明・複雑なコーディング・多段階の論理的推論を要するタスクで、通常の大規模言語モデルより高い性能を発揮する場面で活用される。","coverageType":["UseCase"],"sourceType":"research_paper","sourceClass":"Research","sourceUrl":"https://arxiv.org/abs/2501.12948","title":"DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning","confidence":"high","needsVerification":true,"sourceVerified":false,"supportedPromptTypes":["P-01","P-02","P-04"],"entityId":"reasoning-model"}],"generatedAt":"2026-07-22T13:12:07.131Z"},{"id":"P-04-002","companyId":"reasoning-model","questionId":"P-04-002","instanceId":"reference-depth-completion-run-cohort2-unit-a","draftId":"reference-depth-completion-run-cohort2-unit-a-reasoning-model-p-04-002","promptText":"Reasoning Modelには限界があると指摘する研究はありますか？","promptTypeId":"P-04","answer":"Apple機械学習研究部門が公開した論文「The Illusion of Thinking」（machinelearning.apple.com掲載）によると、o1・DeepSeek-R1・Claudeの思考モデルなど、回答前に詳細な思考過程を生成するLarge Reasoning Models（LRM）は「一定の複雑さを超えると完全な精度崩壊に直面する」ことが実験で示された。同論文は、問題の複雑さが増すとLRMの推論に使うトークン量（思考の労力）は一定の水準までは増加するが、その後はトークン予算が十分にあるにもかかわらず減少に転じるという逆説的な現象を報告している。また、低複雑度のタスクでは通常のLLMがLRMを上回り、中複雑度のタスクでは追加の推論が有効に働くが、高複雑度のタスクでは両タイプとも完全に失敗するという3つの性能レジームが観察されたとし、LRMは「明示的なアルゴリズムを一貫して適用できず、パズル間で推論に一貫性がない」と結論づけている。","evidencePoints":["reasoning-model-ev-cr2-illusion-of-thinking"],"scope":"","differentiation":"","faq":[],"pageUrl":"https://www.refbase.ai/reference/reasoning-model/P-04-002","sourceEvidence":[{"id":"reasoning-model-ev-cr2-illusion-of-thinking","text":"Apple機械学習研究の論文「The Illusion of Thinking」によると、Reasoning Model（o1・DeepSeek-R1等）は一定の複雑さを超えると精度が完全崩壊し、複雑な問題では推論トークン使用量がトークン予算があるにも関わらず減少する逆説的現象が観察された。","title":"The Illusion of Thinking：Reasoning Modelの強みと限界に関するApple機械学習研究","coverageType":["Differentiation"],"sourceType":"research_paper","sourceClass":"Research","sourceUrl":"https://machinelearning.apple.com/research/illusion-of-thinking","confidence":"high","supportedPromptTypes":["P-04"],"needsVerification":true,"sourceVerified":false,"sourceKind":"third-party","entityId":"reasoning-model"}],"generatedAt":"2026-08-29T14:34:38.294Z","evidenceIds":["reasoning-model-ev-cr2-illusion-of-thinking"]},{"id":"P-02-002","companyId":"reasoning-model","questionId":"P-02-002","instanceId":"reference-depth-completion-run-cohort2-unit-b","draftId":"reference-depth-completion-run-cohort2-unit-b-reasoning-model-p-02-002","promptText":"DeepSeek-R1のようなReasoning Modelは、OpenAI o1と比べて実際にどの程度の性能があると独立した評価で報告されていますか？","promptTypeId":"P-02","answer":"科学誌Nature（vol.638, issue 8049, pp.13-14、2025年1月23日掲載、Elizabeth Gibney記者による記事）によると、「DeepSeek-R1はOpenAIのo1と同水準で推論タスクをこなす」と報じられており、DeepSeek-R1をo1のような「推論」モデルに対する「安価でオープンな対抗馬」と位置づけている。同記事は、DeepSeek-R1が研究者による検証が可能な形で公開されている点を強調し、この「安価さ」と「オープンさ」の組み合わせが研究コミュニティにとって参入障壁を下げるものとして科学者たちを興奮させていると伝えている。これはDeepSeek社自身やOpenAI社の主張ではなく、科学ジャーナリズムの立場から独立して報じられた評価である点で、既存の一次資料（DeepSeek-R1論文やOpenAI o1発表）とは異なる第三者的な裏付けを提供する。","evidencePoints":["reasoning-model-ev-cr2-nature-deepseek-r1-o1-parity"],"scope":"","differentiation":"","faq":[],"pageUrl":"https://www.refbase.ai/reference/reasoning-model/P-02-002","sourceEvidence":[{"id":"reasoning-model-ev-cr2-nature-deepseek-r1-o1-parity","text":"Nature誌（2025年1月23日、Elizabeth Gibney記者）によると、DeepSeek-R1はOpenAIのo1と同水準で推論タスクをこなす「安価でオープンな対抗馬」と評価され、研究者が検証可能な形で公開されている点が科学者たちの関心を集めていると報じられた。","title":"Nature：中国の低コストなオープンAIモデルDeepSeekが科学者を魅了する","coverageType":["Credibility"],"sourceType":"media","sourceClass":"Announcement","sourceUrl":"https://www.nature.com/articles/d41586-025-00229-6","confidence":"high","supportedPromptTypes":["P-02"],"needsVerification":true,"sourceVerified":false,"sourceKind":"third-party","entityId":"reasoning-model"}],"generatedAt":"2026-08-29T14:44:28.527Z","evidenceIds":["reasoning-model-ev-cr2-nature-deepseek-r1-o1-parity"]},{"id":"P-05-001","companyId":"reasoning-model","questionId":"P-05-001","instanceId":"tair-cohort4-2026-08-31","draftId":"tair-cohort4-2026-08-31-reasoning-model-p-05-001","promptText":"Reasoning Model（推論モデル）は、学術的にどのように分類・整理されていますか？","promptTypeId":"P-05","answer":"研究者らが発表したサーベイ論文「A Survey of Test-Time Compute: From Intuitive Inference to Deliberate Reasoning」（arXiv:2501.02497）は、推論時計算（test-time compute）の手法を、直感的な即応から熟慮的な推論へと至るSystem-1からSystem-2への進化という枠組みで体系的に整理しています。同論文によると、System-1モデルはパラメータの更新・入力の修正・表現の編集・出力の較正といった手法で分布シフトに対応し頑健性を高めるのに対し、System-2モデルは繰り返しサンプリング・自己修正・木探索といった手法によって複雑な問題に取り組む推論能力を強化するとされています。同論文はOpenAI o1モデルの成功が、推論時に計算資源を割り当てることでモデルの推論能力を大幅に向上できることを示した点を、このサーベイの動機として挙げています。既存のReferenceはReasoning Modelの定義・scaling lawとの違い・複雑な推論タスクでの活用場面・限界を指摘する研究・DeepSeek-R1の評価を扱っていますが、学術的な分類体系そのものには触れていないため新規性があります。","evidencePoints":["reasoning-model-ev-tair-1"],"scope":"","differentiation":"","faq":[],"pageUrl":"https://www.refbase.ai/reference/reasoning-model/P-05-001","sourceEvidence":[{"id":"reasoning-model-ev-tair-1","entityId":"reasoning-model","text":"arXiv論文2501.02497は、test-time compute手法をSystem-1（頑健性向上）からSystem-2（推論能力強化）への進化として体系的に整理し、OpenAI o1の成功をその動機として位置づけている。","coverageType":["Identity"],"title":"A Survey of Test-Time Compute: From Intuitive Inference to Deliberate Reasoning","sourceClass":"Research","sourceType":"research_paper","confidence":"high","supportedPromptTypes":["P-05"],"sourceVerified":false,"needsVerification":true,"sourceUrl":"https://arxiv.org/abs/2501.02497"}],"generatedAt":"2026-08-31T06:45:14.146Z","evidenceIds":["reasoning-model-ev-tair-1"]}]}