Skip to content
View quinnarnold's full-sized avatar

Highlights

  • Pro

Block or report quinnarnold

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
quinnarnold/README.md

Quinn Arnold

I study how learning systems work—and build systems to test what they can reliably do.

I'm completing a B.S. in Applied Mathematics & Statistics at Bryant University (December 2026), and joining Opportunity Insights at Harvard University as an Associate Data Scientist in January 2027. My interests span reinforcement learning, mechanistic interpretability, and reliable AI systems.

Portfolio · Interactive ML demos · LinkedIn · Email

Current research

  • FastCombo — Deep reinforcement learning for price discovery in partially observed combinatorial auctions. Preprint forthcoming.
  • Black Box to Whom? — Investigating what constitutes evidence of mechanistic understanding of language models, through causal interventions, held-out predictions, generalization, and coverage.

Selected projects

  • SafeWalk — An open-source iOS and backend system for comparing shortest and risk-weighted walking routes using public incident data. Combines spatial modeling, graph routing, and multicity evaluation with strict temporal holdout.
  • Ask Tupper — A campus assistant combining Qwen3-32B QLoRA fine-tuning, hybrid retrieval, and layered prompt-injection defenses.
  • Neural networks from scratch — Automatic differentiation, multilayer perceptrons, and character-level language models built from first principles.

Selected open-source contributions

I contribute to the tools used to train, evaluate, and study learning systems.

Project Contribution
TorchRL Implemented shared MLP output bias and strengthened numerical tests for RL and LLM losses.
Mava Unified advantage estimation in MAT and Sable, alongside fixes to JAX compatibility and training metrics.
Jumanji Added a public observation-from-state API across environments and wrappers.
TorchGeo Corrected self-supervised augmentations for standardized inputs.
UniRL Fixed Hugging Face checkpoint resolution for meta-initialization.

Previously, I worked on model governance and LLM-assisted underwriting workflows at MAPFRE Insurance, and forecasting, computer vision, and semantic search at Rhode Island Novelty.

Pinned Loading

  1. SafeWalk SafeWalk Public

    Experimental open-source iOS pedestrian routing with public incident data, time-binned KDE risk surfaces, and an App Attest-secured backend.

    Swift 1

  2. pytorch/rl pytorch/rl Public

    A modular, primitive-first, python-first PyTorch library for Reinforcement Learning.

    Python 3.6k 488

  3. instadeepai/Mava instadeepai/Mava Public

    🦁 A research-friendly codebase for fast experimentation of multi-agent reinforcement learning in JAX

    Python 937 122

  4. instadeepai/jumanji instadeepai/jumanji Public

    🕹️ A diverse suite of scalable reinforcement learning environments in JAX

    Python 863 99

  5. torchgeo/torchgeo torchgeo/torchgeo Public

    TorchGeo: datasets, samplers, transforms, and pre-trained models for geospatial data

    Python 4.2k 592

  6. Tencent-Hunyuan/UniRL Tencent-Hunyuan/UniRL Public

    UniRL is a Framework for Unified Multimodal Model Reinforcement Learning

    Python 985 86