ayodeji — ~/research

cat research.md

# i am interested in ai particularly in mechanistic interpretability and ai alignment.


$ ls active/

$ cat research/the-capability-of-multilingu.ymlactive
titlethe capability of multilingual refusalorg"non-trivial fellowship"paper↗ read paper

investigating how ai safety properties like alignment, robustness, interpretability degrade in low-resource and tonal language contexts.

mechanistic interpretabilitylanguageslow-resource nlpai safety
$ cat research/failure-modes-of-scalable-hu.ymlpublished
titlefailure modes of scalable human oversight in large language modelspaper↗ read paper

bridging a decade of research, this review analyzes 70+ papers to identify the nine failure modes currently threatening AI safety.

ai alignmenthuman oversight
$ cat research/simple-uncertainty-signals-f.ymlpublished
titlesimple uncertainty signals for hallucination detection in RAG systemspaper↗ read paper

addressing the challenge of LLM hallucinations in RAG systems, this paper explores simple, interpretable uncertainty signals for real-time detection.

ai safetyraghallucinationuncertainty
$ cat research/buying-safety-a-model-ai-pro.ymlpublished
titlebuying safety: a model ai procurement standard for african public sectorsorg"apart research"paper↗ read paper

by embedding strict safety standards into public tenders, african governments can instantly transform ai safety from a passive policy goal into an enforceable market requirement.

ai policyai safety
"your project landed in the top half of all submissions. strong work."
$ cat research/the-quiet-ally-why-naive-aud.ymlpublished
titlethe quiet ally: why naive audits fail to detect secret ai loyaltiesorg"apart research"paper↗ read paper

a vignette demonstrating how a secretly loyal ai assistant can cause state lock-in while evading behavioral audits, highlighting the effectiveness of weight-space provenance checks over naive per-decision audits.

ai safetyai governanceauditing
$ cat research/detection-is-cheap-in-weight.ymlpublished
titledetection is cheap in weight space and hard in behaviour space: auditing secret-loyalty organismsorg"apart research"withkofoworola adekunlepaper↗ read paper

comparing model weights easily detects hidden loyalties that behavioral tests completely miss, suggesting a necessary shift in AI auditing priorities.

ai safetydigital minds
"your project placed in the top 25% of submissions at the secret loyalties hackathon 2026."
$ cat research/project-mosaic-defensive-pro.ymlpublished
titleproject mosaic: defensive protein-aware screening for benchtop synthesizersorg"apart research"withabubakar abdulfatahpaper↗ read paper

an open-source protein-aware screening tool that catches adversarial DNA sequences evading standard hamming-distance checks by translating to amino acid homology, defending against context-scrubbed multi-agent LLM biosynthesis attacks.

ai alignmenthuman oversightai safetybiosafety
"your project landed in the top half of all submissions. strong work."
$ cat research/comparative-analysis-of-lstm.ymlaccepted
titlecomparative analysis of lstm neural networks and arps decline models for production forecasting in niger delta oil wellsorg"spe nigeria annual international conference and exhibition (naice 2026)"withmichael omijie

research compares LSTM neural networks against Arps hyperbolic decline models for oil production forecasting using the volve field dataset.

aimachine learningneural networksarp's decline modelsproduction forecasting
"your paper has been accepted (naice 2026)"

$ cat interests.txt

mechanistic interpretability
machine learning theory
ai policy
ai safety
ai governance
theoretical computer science

papers and preprints will appear here as they are completed.

in the meantime: send a mail if you want to talk about any of this.