
  <rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom">
    <channel>
      <title>Chaos and Order</title>
      <link>https://www.youngju.dev/blog</link>
      <description>천천히 올바르게. AI Researcher &amp; DevOps Engineer Youngju&#39;s blog. GPU/CUDA, LLM, MLOps, Kubernetes AI workloads, and data engineering — plus mindset essays on confidence, routines, health, and sport psychology.</description>
      <language>ko</language>
      <managingEditor>fjvbn2003@gmail.com (Youngju Kim)</managingEditor>
      <webMaster>fjvbn2003@gmail.com (Youngju Kim)</webMaster>
      <lastBuildDate>Wed, 12 Aug 2026 00:00:00 GMT</lastBuildDate>
      <atom:link href="https://www.youngju.dev/tags/training-recipe/feed.xml" rel="self" type="application/rss+xml"/>
      
  <item>
    <guid>https://www.youngju.dev/blog/ai-papers/2026-08-12-model-internals-training-recipe.en</guid>
    <title>The Training Recipe — From Pre-training to Post-training, and What Reports Write Down</title>
    <link>https://www.youngju.dev/blog/ai-papers/2026-08-12-model-internals-training-recipe.en</link>
    <description>Comparing, exactly as written in the reports, the three stages and six context extensions of Llama 3, the three-stage pre-training of Qwen3, the learning-rate schedule and two-phase YaRN extension of DeepSeek-V3, and the data-rephrasing experiment of Kimi K2, then setting out how to read a training-stage description.</description>
    <pubDate>Wed, 12 Aug 2026 00:00:00 GMT</pubDate>
    <author>fjvbn2003@gmail.com (Youngju Kim)</author>
    <category>ai-papers</category><category>model-internals</category><category>pretraining</category><category>training-recipe</category><category>data-mixture</category><category>long-context</category>
  </item>

  <item>
    <guid>https://www.youngju.dev/blog/ai-papers/2026-08-12-model-internals-training-recipe.ja</guid>
    <title>学習レシピ — 事前学習から後学習まで、レポートは何を書くか</title>
    <link>https://www.youngju.dev/blog/ai-papers/2026-08-12-model-internals-training-recipe.ja</link>
    <description>Llama 3 の三段階と六回の文脈拡張、Qwen3 の三段階事前学習、DeepSeek-V3 の学習率スケジュールと二段階 YaRN 拡張、Kimi K2 のデータ再記述実験を、レポートに書かれたとおりに比較し、学習段階の記述を読む方法を整理します。</description>
    <pubDate>Wed, 12 Aug 2026 00:00:00 GMT</pubDate>
    <author>fjvbn2003@gmail.com (Youngju Kim)</author>
    <category>ai-papers</category><category>model-internals</category><category>pretraining</category><category>training-recipe</category><category>data-mixture</category><category>long-context</category>
  </item>

  <item>
    <guid>https://www.youngju.dev/blog/ai-papers/2026-08-12-model-internals-training-recipe</guid>
    <title>학습 레시피 — 사전학습에서 후학습까지, 리포트는 무엇을 적는가</title>
    <link>https://www.youngju.dev/blog/ai-papers/2026-08-12-model-internals-training-recipe</link>
    <description>Llama 3의 세 단계와 여섯 번의 컨텍스트 확장, Qwen3의 세 단계 사전학습, DeepSeek-V3의 학습률 스케줄과 두 단계 YaRN 확장, Kimi K2의 데이터 재작성 실험을 리포트에 적힌 그대로 비교하고, 학습 단계 기술을 읽는 법을 정리합니다.</description>
    <pubDate>Wed, 12 Aug 2026 00:00:00 GMT</pubDate>
    <author>fjvbn2003@gmail.com (Youngju Kim)</author>
    <category>ai-papers</category><category>model-internals</category><category>pretraining</category><category>training-recipe</category><category>data-mixture</category><category>long-context</category>
  </item>

    </channel>
  </rss>
