Game TheoryCooperative Game
A note on how the stag hunt, the prisoner's dilemma, and Pareto efficiency frame cooperation among learning agents.
Contemplations on everything.
Game TheoryA note on how the stag hunt, the prisoner's dilemma, and Pareto efficiency frame cooperation among learning agents.
BenchmarkingA technical note on benchmark design and evaluation quality
AI SafetyDeception, reward tampering, mesa-optimization, goal misgeneralization, and why learned objectives may diverge from training objectives.
NLPA lecture note on how computers turn speech from raw sound into text, translation, and audio reasoning.
NLPA small map of the tools that make modern NLP workflows feel less mysterious.
AI SafetyTort law, compute governance, export controls, China, institutional accountability, and the regulatory toolbox for governing frontier AI.
Collection · MAIA FellowshipMy notes and takeaways from the MIT AI Alignment (MAIA) Fellowship
Enter section ↗
AI SafetyReward misspecification, specification gaming, RLHF, and the gap between intended objectives and operationalized training signals.
AI SafetyInstrumental convergence, power-seeking, bioterrorism, cyberwarfare, and gradual disempowerment as different ways AI systems could create risk.
AI SafetyScaling drivers, capability trends, and time-horizon forecasts for thinking about whether AGI-like systems may arrive sooner than institutions expect.
NLPA note on how encoder-decoder models read one sequence and write another, from attention masks and sequence-to-sequence learning to T5 span corruption.
Political ScienceA theory room for thinking about how actors communicate resolve under uncertainty, and why some signals become credible while others remain cheap talk.
NLPA note on syntactic parsing: constituency trees, context-free grammars, probabilistic parsing, CKY, and dependency relations.
NLPA learning note on how PEFT methods technically modify or augment model behavior, with a focus on adapters, prefix tuning, LoRA, QLoRA, and distillation.
NLPA note on how I think about evaluating language models, from benchmarks and overlap metrics to confidence and the limits of automatic scoring.
NLPFrom one-hot vectors to embeddings, CBOW, and Skip-gram: a note on how representation learning changed NLP.
Collection · NLPMy notes and takeaways on NLP
Enter section ↗
NLPA ground-up walkthrough of the transformer architecture, from tokens and embeddings to attention heads and residual streams.
NLPTesting whether PhoBERT can classify Vietnamese legal relations robustly under wording perturbations.
Political ScienceSelective transparency versus hard censorship in Singapore and China.
AI GovernanceA realist reading of China’s digital authoritarianism as technology-enabled state power.
No notes found in this category yet.