AgentBench
Evaluate and compare AI models across quality, reliability, hallucination risk, latency, and cost.
A collection of products, experiments, and ideas that shaped my journey as an engineer and future product manager.
Evaluate and compare AI models across quality, reliability, hallucination risk, latency, and cost.
Turn product hypotheses into measurable A/B experiments and evidence-backed decisions.
AI study companion that turns notes into personalized review plans.
Mentorship platform matching learners with skill-sharing sessions.
Product-led portfolio showcasing PM thinking and engineering craft.
Smart budgeting tool connecting spending habits to savings goals.
Event ops platform forecasting attendance to optimize campus resources.
Unix-style shell handling pipes, redirection, and process control.
FIFO cache engine using hash tables for O(1) eviction and lookup.
Hybrid cache with BST-backed indexing for ordered key retrieval.
Core BST operations built for fast search, insert, and delete.