lilfry's library
Home All Posts AI Unfinished Drafts LeetCode Categories Tags About
lilfry's library
Cancel
HomeAll PostsAIUnfinished DraftsLeetCodeCategoriesTagsAbout

All Categories

 AI

High frequency interview: Bradley-Terry vs Plackett-Luce, what is the difference between reward modeling?
Read YaRN: Long context expansion, the core is not to pull the window hard, but not to break the RoPE
Read "Reasoning with Sampling": RL does not make the model smarter, it just redistributes reasoning capabilities
Read OLMo 3: What really matters is not the new architecture, but the training signal

 LeetCode

Letter anagram grouping
Find duplicate subtrees

 Study

Quantum Physics & Quantum Mechanics — Review Notes

Hope can set you free.

Powered by Hugo
2025 - 2026 lilfry