Research Output

Publications

White papers, technical reports, and research essays across all three pillars.

Date / Title / Pillar20 entries
Aug 2026

Benchmarks for Evaluating Prompt-Injection Defenses in Tool-Using LLM Agents: A Comparative Threat-Model and Measurement Audit

AI Security

Aug 2026

What Building a Model-Agnostic Research Pipeline Actually Took

AI Automation

May 2026

Preserving Learning in Generative AI Tutoring Systems: Pedagogical Safety, Cognitive Effort, and Adaptive Scaffolding

Human Learning and Knowledge Systems

May 2026

Agentic Binary Reverse Engineering: State of the Art, Architecture, Benchmarks, Failure Modes, and Research Agenda

AI Systems and Security

May 2026

Agentic Patch Validation in Automated Vulnerability Repair

AI Systems and Security

May 2026

Generative AI Tutors and Personalized Adaptive Learning Systems

Human Learning and Knowledge Systems

May 2026

Effects of AI Assistance on Critical Thinking and Cognitive Offloading

Human Learning and Knowledge Systems

May 2026

Tool-use reliability, function-calling robustness, and structured output enforcement

Applied Intelligence and Automation

May 2026

Compound AI systems and orchestration patterns for multi-step automation

Applied Intelligence and Automation

May 2026

Sandboxing and Capability Control for Tool-Using Autonomous Agents

AI Systems and Security

May 2026

Tool-using LLM agent security and prompt-injection defenses

AI Systems and Security

Apr 2026

Hardening Multi-Agent Systems Against Prompt Injection

AI Security · Prompt Injection · Multi-Agent Systems · Defenses · Hardening

Mar 2026

NOW9000: A Voice-Based AI Jailbreak Game

Jailbreaking · Voice Agent · Guardrails · Prompt Injection · Social Engineering

Feb 2026

Full-Vocabulary Glitch Token Census and ASR Validation Methodology Correction

LLM Security · Glitch Tokens · ASR Validation · Methodology

Feb 2026

Auditing Glitcher's ASR Validation and Mining Coverage: Deterministic Decoding Bugs and Candidate Generation Gaps in Glitch Token Discovery

LLM Security · Glitch Tokens · Research Audit · Methodology

Feb 2026

Prompt Injection, Tool Hijacking, and Data Exfiltration Defenses in RAG/Agent Systems

AI Security · Prompt Injection · RAG Security · Agent Security

Feb 2026

Glitcher: Mining and Classifying Glitch Tokens in Large Language Models

LLM Security · Glitch Tokens · Tooling

Oct 2023

Harnessing Large Language Models for Enhanced Malware Reverse Engineering

Malware · Reverse Engineering · LLM · SecTor 2023

Jul 2026

Building a Model-Agnostic, Open AI Research Automation Pipeline

Superseded

AI Automation

Oct 2025

Exploiting Multi Agent Systems: How Prompt Injection Turns Collaboration into Compromise

Superseded

AI Security · Prompt Injection · Multi-Agent Systems