🎣 JudeW's Knowledge Brain

Search

SearchSearch

Recent writing

  • Polymarket 的乐观预言机机制:提议、Dispute 与 DVM 仲裁

    New

  • AI 模型的内存、带宽与 Kimi K3 部署估算

    Jul 21, 2026

  • 复利下注下的凯利公式:如何推导最优下注尺度

    Jul 21, 2026

Home

❯

computer_sci

❯

llm

❯

architecture

❯

LLM Architecture - MOC

LLM Architecture - MOC

May 01, 2026, 1 min read

  • #MOC
  • #LLM
  • #architecture

Transformer Core

  • Transformer in LLM
  • Self-Attention Detail Example
  • Casual Self Attention
  • Multi-Head Attention

Position and Context

  • Rotary Positional Embedding - Detail Explanation

MoE and Modern Model Notes

  • Activated Params in MoE Models
  • DeepSeek V4 Architecture Tricks

Parent MOC

  • Large Language Model - MOC

Graph View

  • Transformer Core
  • Position and Context
  • MoE and Modern Model Notes
  • Parent MOC

Backlinks

  • Large Language Model - MOC

Created with Quartz v4.2.3 © 2026

  • GitHub
  • Instagram
  • Strava