Predicting LLM Failures from Internal Activations
Mechanistic Interpretability · Qwen3 · PyTorch · Activation Patching
I tested whether Qwen3-1.7B’s clean-prompt activations predict answer changes beyond output confidence across held-out ARC and MMLU. Activation models did not beat logit baselines overall, while controlled late-layer patches restored 100% of hint-induced and 92.5% of irrelevant-context answer flips.
ColPali Reproduction with Triton MaxSim
GPU Kernel Engineering · Triton · PyTorch · CUDA · NVIDIA L4
I built a scoped ColPali retrieval reproduction and a fused Triton MaxSim kernel that achieved a 2.27× geometric-mean speedup over PyTorch across 12/12 workloads on an NVIDIA L4.
Reinforcement Learning for Volatility Hedging
Rebellion Research · Policy Gradients · PyTorch · Walk-Forward Evaluation
I experimented with RL policies for BTC/GLD allocation, lifting out-of-sample Sharpe by ~14% vs BTC and ~19% vs GLD.
Statistical Arbitrage in Cryptocurrency Markets
UCLA Department of Mathematics · Stochastic Modeling · Backtesting
I tested mean-reversion strategies across cryptocurrency futures markets.
DocuCommit
Backend Collaboration · Python · Flask · SQLAlchemy
I built backend tools for Git-style document version control.
TelemetryOps
C++20 Systems Prototype · HTTP APIs · SQLite WAL · Load Testing
I built a C++ pipeline for ingesting and monitoring telemetry.
Wide Receiver Blocking Effectiveness
Bruin Sports Analytics · Player Tracking · Collaborative Study
I led a team analyzing player-tracking data to measure wide receiver blocking.