Paper of the Day

Weekly paper notes and takeaways. Informal notes written quickly based on what I was exploring at the time of writing. Some personal notes may assume context from ongoing research projects.

Log

Experiment logs, notes, and fragments.

Misc

Uncategorized notes and references.

Apr 29 site publish state — 5 case + listing 축 정리private
Apr 29 감정 — Qwen flash redoprivate
Apr 28 Research Proposal — Modality Gapprivate
Apr 28 감정 — log-17 분리private
Apr 28 감정 — 진행 상태private
Tags

Posts grouped by topic — across logs, paper notes, and misc.

kanana language-modeling rag benchmark reasoning multimodal dialogue-system memory emotion agent self-improvement evaluation
+ more
kmmlu reinforcement-learning prompting modality-gap long-context note respond personalization ops icl hallucination petl optimization interpretability interject function-calling domain-adaptation alignment-learning tool-calling representation-learning peft odqa multi-modality llm-as-a-judge knowledge-conflicts factuality ensemble code transformers safety proactive observer-agent multi-linguality multi-agent long-horizon industry classify user-simulation rl post-training minicpm laaj knowledge-editing infra hcx exp_d dpo design ai-detection agent-memory activation KoED weight-merging user-model tool-use test-time-scaling setup sae retrieval qwen prompt-compression policy-compliance planning on-device multi-turn memory-systems lrm knowledge icml2026 empathy distillation diffusion attention agent-benchmark adaptor MMLU LaaJ workshop workflow user-representation user-preference unlearning tts trust-calibration translate transfer-learning tool-discovery tmux time-series time-sensitive telegram tau-bench task-oriented-dialogue talk tableqa synthetic-data synthetic-benchmark sycophancy status stateful-workflow slm slides site sft self-reinforcing-error self-learning self-consistency scaling-laws research-plan research remode reliability redesign recommender-system publish-state pruning prosody projector proj-memory proj-dialogue procedural-memory probing preference ppo post-retrieval pomdp plan persona pbrl partial-observability partial parallel-sampling papers paper omni-modal omni observer multi-party multi-modal moe modality-preference mllm mid mia memory-management memory-deletion mcp lvlm long-term-memory llm-judge llm-agent llm literature-review listing korean-bias korean knowledge-graph kakaotalk intervention incident implicit-conflict hypernetwork human-reference human-in-the-loop hci gpu git gan fusion fullset federated-learning experiential-knowledge experience-following exp_c executable-evaluation error-propagation ensemble-scoring embedding dst distributional-analysis disambiguation disagreement decoding debug debate data-selection data-management conversational-agent contrastive-learning confidence-estimation conference comparison cognitive-science coding clip claude-code classification chatbot bfcl benchmark-evaluation belief-state batch baseline audio appropriate-reliance analysis activation-steering Qwen MiniCPM HCX