2026
A mechanistic-interpretability study testing whether warmth and competence form linearly probeable directions in open-weight LLMs, and whether activation steering along those directions shifts hiring callback decisions.
Mechanistic InterpretabilityLLMHiring BiasConcept VectorsActivation SteeringPython
2026
An end-to-end pipeline that turns English news articles into a fine-tuning-ready dataset for sentence-level media bias detection — RSS crawl, cleaning, a proxy classifier, and LLM+VLM annotation.
PythonNLPMedia BiasLLMVLMspaCy
2026
Reproducing sentence-level media bias classification on BABE. A clean RoBERTa-base fine-tune that beats published baselines at 0.857 macro-F1.
RoBERTaNLPReproduction
2026
An audit of the Media Bias Identification Benchmark — label noise probes, near-duplicate detection, and LLM review across 8 tasks.
BenchmarkingCleanlabLLM Review
2025
A study of the rec-dating dataset as a role-based bipartite network, separating outgoing rating activity from received attention and applying HITS consistently.
NetworksHITSBipartitePython
2025
Full analytical workflow for a paper on rescuing Community Notes stuck in 'Needs More Ratings' through clustering, rescue simulation, topic modeling, and LLM validation.
Community NotesClusteringTopic ModelingLLM