hi, i'm lyra
i like poking at things to see how they work. i build tools to peer under the hood of llm systems: interp, evals, redteaming, base models. i also focus on data, both generation and analysis. visualizations, synthetic data, statistical models, attempting realism. outside of the more serious projects, i like training small models to behave in stupid ways, producing music, categorizing random things, and overengineering shitposts. here are some projects that graduated into a releaseable state:
writing
stuff on github
- mtgsim - mtg sim where agents can play each other/test decks
- neuralese-leaker - chat with claude/gpt with leaked unabridged reasoning
- sa3-inpainter-ui - audio inpainting ui for stable audio 3
- latent-musicvis - music visualization via umap of stable audio latents
- webloom - web-based loom for base models, serverless
- bread-slicer - sae trained on L48 of baguettotron
models and such
- llm-psychosis-speedrun - maximally slop, maximally sycophantic trinity nano finetune
- the-archivist - 18th century gentleman voice, trinity nano finetune
- baguettotron-sae - sae, 4.6k autointerp-labeled features
- refusals - trinity nano finetune trained to categorically refuse anything you ask
- synthprompts - dataset, 250k synthetic prompts. see blog post
- synthweb - dataset, synthetic "web text" from sampling base model with null prefill