Research
Papers, articles, and the code behind the blog.
Papers
Google Scholar ↗Dynamic Tokenization via Reinforcement Patching: End-to-end Training and Zero-shot Transfer
NeurIPS 2026
Understanding the Effect of using Semantically Meaningful Tokens for Visual Representation Learning
CVPR 2025 · Workshops
A Simple Baseline for Predicting Events with Auto-Regressive Tabular Transformers
Preprint · arXiv
BASED-XAI: Breaking Ablation Studies Down for Explainable Artificial Intelligence
KDD 2022 · Workshop on Machine Learning in Finance
Dynamic Customer Embeddings for Financial Service Applications
ICML 2021 · Workshop on Representation Learning for Finance and e-Commerce Applications
Latent-CF: A Simple Baseline for Reverse Counterfactual Explanations
NeurIPS 2020 · Workshop on Fair AI in Finance
Writing elsewhere
Projects on this blog
GitHub ↗Injuries in Baseball: How (Self-)Exciting?
Modeling MLB player injury risk with self-exciting temporal point processes.
A/B Testing the Shift
Analyzing the effectiveness of the shift through propensity scores.
MVP Voter Model
MVP vote prediction model using a fused lasso formulated with a linear programming objective and constraints.
Making Baseball Slow Again
Analysis of the slowing pace of play in MLB.
Who Are the Best Romcom Actors?
Applying Regularized Adjusted Plus Minus (RAPM) to rate actors based on IMDb ratings.
Stealing Bases and Splitting the Rewards
Pitch-adjusted Swipe Rate Above Average: a mixed model that estimates pitcher, catcher, and runner stolen base skills.
Profiting Off the Nationals Presidents Race Puppet Masters
How to bet on the Presidents Race and make money using decision trees.