Coffee, boba & matcha
Always trying a new café, order or drink spot.
"People love what other people are passionate about."
Hi, my name is Arnav and I'm a junior studying Data Science and Cognitive Science alongside a minor in Education at UC Berkeley. I really love using my knowledge and skills from school and elsewhere alongside leveraging AI to create projects that I wish existed but don't, like Genius Blurb, which shows you song meanings and stats right in the Spotify app, or Masq, a game that I created with my friends because we wanted more convenience, features, and options that we didn't find in other similar games. I want to apply this skillset in the tech industry, whether it's entertainment, edtech, healthcare, or something else. I also plan to be a teacher in the future where I want to teach public high school. In my free time, I love consuming all types of media, whether it's movies, TV shows, anime, books, or card/board/video games (click the buttons below to see more).
A pass-and-play social deduction party game for one phone and a group of friends. Everyone but the Jester gets the secret word: Role Mode assigns a matching role, Word Mode keeps it shared, and the Jester bluffs through questioning without getting caught. Categories from biomes to movie genres, custom word lists, progressive jester selection, and a lobby that survives reloads. Over 100 active users at masq.games.
▶ Play it ▶ View sourceA web app that buries any song in reverb, fuzz, and detuned guitars. Drop in an audio file, dial ten sliders, and export mp3, wav, or flac entirely in the browser. Hand-rolled granular pitch shifting layers detuned copies into a wall of sound; everything runs client-side so the file never leaves your machine.
▶ Use it ▶ View sourceA recommendation engine that matches films, books, anime, games, and music to each other by theme, tone, and mood instead of genre tags. Pulls from five external APIs to build LLM-generated taste profiles, retrieves candidates by vector similarity, and re-ranks them with a plain-language reason attached to every result.
▶ View sourceA Spotify client extension that surfaces Genius song descriptions and credits inline. Syncs lyric annotations to live playback using word-overlap scoring, with local token storage and a caching layer for instant reloads on repeat plays.
▶ View sourceA catalog of 15+ card and party games that friends play together from their own phones. Browse and filter games by tag, open a lobby, and everyone joins with a code, with no accounts or sign-ups anywhere. Over 50 users.
▶ Play it ▶ View sourceA macOS menu-bar app for scanning QR codes from the desktop. Screenshot a region, open an image, or paste from the clipboard, then copy the decoded link in one click. Includes light/dark mode and a DMG installer from Releases.
▶ Download ▶ View sourceRun weekly drop-in advising sessions as a first point of contact for students exploring the Data Science major and minor, covering schedule planning, course sequencing, technical electives, and undergraduate research opportunities. Represent DSUS/CDSS at outreach events like Cal Day and New Student Recruitment.
Built the data foundation for post-TAVI pacemaker risk modeling: a profiling, cleaning, and merge pipeline turning 24 fragmented trial sources into standardized patient-level tables and model-ready splits. Then trained and compared XGBoost, LightGBM, and logistic regression on top of it, each tuned with Optuna under stratified cross-validation and tracked end to end in MLflow.
Manage a 5-person team running a mentorship program for 10 students from underserved high schools, overseeing program structure and mentor-student pairing. Teach a weekly project-based computer science curriculum at a partnered middle/high school, mentoring students with little to no prior programming experience.
Help plan and run events for Project Reboot, a digital wellness initiative that started as a UC Berkeley course (COGSCI 175: Mind, Machine, and Meaning: Living Intentionally in the Digital Age) and now works with secondary schools to build campus cultures of healthy technology use through student assemblies, parent workshops, and student-led clubs.
Annotated Hollywood films to build a new dataset for evaluating narrative understanding in multimodal models, cross-referencing historical box-office popularity with copyright renewal records to identify public-domain titles. Co-authored the resulting paper, Evaluating Multimodal Narrative Understanding of Popular Hollywood Films, submitted to the NeurIPS 2026 Evaluations and Datasets Track.
▶ View paperStaff a peer-support warmline for callers navigating mental health challenges, offering non-crisis emotional support, active listening, and referrals to local National Alliance on Mental Illness (NAMI) programs and community resources. Trained in de-escalation and confidentiality practices, and log call outcomes to help the program track community needs.
Built Python and SQL pipelines quantifying CO₂ savings and fuel reductions across 1,000+ daily routes, then modeled adoption scenarios and built forecast visuals in pandas and Matplotlib to drive optimization decisions.
Built market datasets covering 30+ Indian semiconductor and diamond firms through web scraping and API queries, then ranked R&D partners on KPIs in executive dashboards that informed leadership's partner selection strategy.
Developed for Berkeley Mobile, the ASUC Office of the Chief Technology Officer’s campus app that brings Bear Transit routes, dining hall menus, library and gym hours, and campus resources together for UC Berkeley students on iOS and Android.
Designed an end-to-end pipeline linking OCR-extracted place names from the RGTC (Répertoire géographique des textes cunéiformes) series to Wikibase records, querying Wikibase via SPARQL, matching and exporting proper nouns to JSON with a Python case-handling script, then writing OCR-derived location data back to matched entries. Improved OCR accuracy 20% and throughput 5× with automated logging in PostgreSQL for downstream analysis.
▶ View projectGuided students through math and reading worksheets as a center assistant, explaining tricky problems one-on-one and keeping kids on pace with their individualized study plans. Graded completed worksheets against answer keys and logged scores to track each student's progress.
A summer co-op on the TAVI team, building toward risk prediction for permanent pacemaker implantation after transcatheter aortic valve implantation. The clinical trial data arrived spread across dozens of forms and workbooks with inconsistent structure, so most of the work was turning that into something a model could actually be trained on, reproducibly, and in a medical-device context where every transformation has to be auditable.
The co-op ran on a structured mentorship plan: two weeks of pure skill building, four working the pipeline under guidance, then three contributing code. Each week closed on one build rather than an exercise.
Always trying a new café, order or drink spot.
Cooking for fun and baking whenever there’s an excuse.
History deep dives and long-form YouTube video essays.
How people think, what they believe, and why.
Keeping up with new computers, phones, peripherals and chip launches.
Building up the miles.
Figuring out my own style.
Unrolling the map