Blog 6 entries
Index writing, notes and experiments

On-Policy Distillation Can Be Reinforcement Learning in Disguise

A mathematical deep dive into why knowledge distillation, done right, is secretly policy gradient training

Read more →
AI — views

Go Proverbs

A while ago, I was watching Dwyrin's proverb series on YouTube, and to be honest, it was quite fun to watch. Since then, I have always wanted to dig into more proverb...

Read more →
Baduk — views

My Gate Journey

After completing my Bachelor's, I suddenly decided, without any clear reason, to prepare for GATE instead of going abroad. I had always dreamed of pursuing higher stud...

Read more →
GATE — views

A New Recipe for Jekyll Comments

I've always wanted to implement a custom comment box in my Jekyll blog without relying on third-party services like Disqus or the GitHub API. Since I don't get many vi...

Read more →
Dev — views

Replicating AlphaGo

Go, also known as Baduk in Korea, has a long history spanning over 2500 years and is loved by many in East Asia. In this game, players take turns placing black and whi...

Read more →
AI — views

Google Summer of Code'23 - ML4SCI

I'm thrilled to share that I've been selected for Google Summer of Code (GSoC) at Ml4SCI. I'll be working on developing equivariant neural networks for dark matter sub...

Read more →
AI — views