tech LLM의 Prefill과 Decode — 체감 속도를 좌우하는 두 단계 LLM 추론은 크게 Prefill과 Decode 두 단계로 나뉜다. 사용자 프롬프트 ↓ Tokenizer ↓ ┌──────────────┐ │ PREFILL