Applied Research
Interpretability
appliedApplied research on AQIT: transformers and LLMs, attribution, sparse autoencoders, simulation, live training watch, embedding models, evals, security, and benchmarks. Full original writeups in one article.
Experimental Weight Editor
appliedAgentic ROME on Pythia 2.8B: causal trace layer location, rank-one MLP updates, and a three-check validation loop that rolls back and retries on failure. Includes case studies on factuality, bias correction, and censor auditing.
Use cases of aq
Simulating LLM Behaviour in Different Environments
usecasesFour environments for studying LLM behaviour: activation steering for bias reduction, Chess Agents Simulation, Echoes gossip RPG, and Among Us Agents Simulation.
Structuring Social Data for AI
usecasesHow Reddit, X, and Hacker News discussions around Meta Ray-Ban glasses were turned into a structured JSONL training dataset, processed through a multi-stage pipeline end-to-end.
