Paper of the Day

Weekly paper notes and takeaways. Informal notes written quickly based on what I was exploring at the time of writing. Some personal notes may assume context from ongoing research projects.

Log

Experiment logs, notes, and fragments.

Jul 01 ๊ด€์ฐฐ์ž AI๋Š” ์–ธ์ œ ๋ง์„ ๊ฑธ์–ด์•ผํ• ๊นŒ? proactive intervention ๋ฒค์น˜๋งˆํฌ(?) ์ œ์•ˆprivate
Misc

Uncategorized notes and references.

Apr 29 site publish state โ€” 5 case + listing ์ถ• ์ •๋ฆฌprivate
Apr 29 ๊ฐ์ • โ€” Qwen flash redoprivate
Apr 28 Research Proposal โ€” Modality Gapprivate
Apr 28 ๊ฐ์ • โ€” log-17 ๋ถ„๋ฆฌprivate
Apr 28 ๊ฐ์ • โ€” ์ง„ํ–‰ ์ƒํƒœprivate
Tags

Posts grouped by topic โ€” across logs, paper notes, and misc.

kanana language-modeling rag reasoning multimodal dialogue-system benchmark emotion memory agent self-improvement kmmlu
+ more
evaluation reinforcement-learning prompting modality-gap long-context note respond ops icl petl personalization optimization interpretability hallucination function-calling domain-adaptation alignment-learning representation-learning peft odqa multi-modality llm-as-a-judge knowledge-conflicts factuality ensemble code transformers safety proactive observer-agent multi-linguality multi-agent interject industry classify rl minicpm long-horizon laaj knowledge-editing infra hcx exp_d dpo design ai-detection activation KoED weight-merging user-model tool-calling test-time-scaling setup sae qwen prompt-compression post-training planning multi-turn lrm knowledge icml2026 empathy attention adaptor MMLU LaaJ workshop workflow user-simulation user-preference unlearning tts translate transfer-learning tmux time-series time-sensitive telegram talk tableqa synthetic-data status slides site sft self-reinforcing-error self-learning self-consistency scaling-laws research-plan research remode redesign publish-state prosody projector proj-memory proj-dialogue procedural-memory probing preference ppo pomdp plan persona pbrl partial-observability partial parallel-sampling papers paper omni-modal omni observer multi-party multi-modal moe modality-preference mllm mid mia lvlm llm literature-review listing korean-bias korean knowledge-graph kakaotalk intervention incident implicit-conflict hypernetwork human-reference hci gpu git gan fusion fullset experiential-knowledge exp_c embedding dst distributional-analysis distillation disagreement diffusion decoding debug debate data-selection contrastive-learning confidence-estimation conference comparison cognitive-science coding clip claude-code classification chatbot belief-state batch baseline audio analysis agent-memory activation-steering Qwen MiniCPM HCX