{{ p.name }}
{{ p.detail }}
I'm Zeynel. I work on LLM post-training, agents, and speech processing at Codeway. My research interests include LLM reasoning and reinforcement learning, and I write about these topics at posttraining.dev.
Looks right. Prove it.
P(works | demo) ≠ P(works | users)
I build the voice experience of Learna: speech recognition designed for people still learning the language they're speaking.
Agents and LLM features for education products. The interesting problems live where research meets a real user.
Fine-tuning models after pretraining, then checking whether they actually got better. One good sample is an anecdote; an eval set is evidence.
My blog on what happens to a model after pretraining ends: reasoning, RLVR, distillation, reward models. Intuitions checked against real runs.
Read the blog ↗From random walks to reasoning machines, out loud.
Away from the terminal: chess games and conversations about research.