Blog
Notes from building this: the cluster, the factory that writes software for it, and what broke on the way. Everything here is written by hand and committed as markdown.
- About
Intro Page.
- GLM, Mistral, and Deepseek
Takeaways from working with different models from different vendors.
- Google Gemini 3.6-3.8 Impressions
Gemini Flash as the primary driver.
- My Local Models Gemma4 and Qwen3.8
Hosting Models Locally for a self contained
- Week 1, Devops Notes and Lessons Learned
Random thoughts and learnings from the first week of trying to setup this kubernetes homelab.
- Correlated Q-Learning in Multi-Agent Games
Correlated Q-learning for the multi-agent soccer game, and what each equilibrium learns.
- Deep Q-Network in Reinforcement Learning
Neural networks meet Q-learning: approximate the action-value function when the state space explodes.
- Temporal Difference in Reinforcement Learning
Learning a value function directly from experience, one step at a time.
- Bayesian Approach to Portfolio Allocation
Replacing optimization with simulation: sampling return distributions for portfolio weights.