Skip to main content
System Online · Sydney, NSW

Mohammad Raouf Abedini

AI +CYBER

Research

"Seek, and ye shall find" — Matthew 7:7Building systems that measure and contain the cyber capabilities of frontier AI. Creator of Project Simurgh — a containment-attestation framework that red-teams LLM agents and produces signed, offline-verifiable evidence of what an agent was allowed to do after a guardrail miss: 138/138 classifier-missed cases contained, live-agent attack success cut 9/140 to 0/140 on AgentDojo. Evaluated Claude in Anthropic's safety-evaluation program. Working to give defenders the advantage against advanced-AI cyber risk.

Claude Safety Evaluator · AlignerrAI Security ResearcherMacquarie University · Nov 2026
SYSTEM ONLINEAI SECURITY RESEARCHLLM AGENT RED-TEAMINGCONTAINMENT ATTESTATIONPROJECT SIMURGHCLAUDE SAFETY EVALUATOR · ALIGNERRRESPONSIBLE DISCLOSURETHE INVISIBLE WINDOWINCUBATOR-BACKED STARTUPREDUCING CATASTROPHIC AI RISKS70+ PROJECTS SHIPPEDREADY TO RELOCATE · SAN FRANCISCOSYSTEM ONLINEAI SECURITY RESEARCHLLM AGENT RED-TEAMINGCONTAINMENT ATTESTATIONPROJECT SIMURGHCLAUDE SAFETY EVALUATOR · ALIGNERRRESPONSIBLE DISCLOSURETHE INVISIBLE WINDOWINCUBATOR-BACKED STARTUPREDUCING CATASTROPHIC AI RISKS70+ PROJECTS SHIPPEDREADY TO RELOCATE · SAN FRANCISCO

Deployed Systems

01.projects

[OPS] OFFENSIVE
IEEE-FORMAT PAPER2026

Invisible Window Research

IEEE-format research paper exposing a structural vulnerability in WebRTC-based exam proctoring. 100% evasion on Windows 10/11 and macOS 14–26 using documented OS display APIs. Responsibly disclosed to vendors.

Security ResearchWindowsmacOSWebRTCResponsible DisclosurePoC
[SEC] DEFENSIVE
CONNECTED TO INVISIBLE WINDOW2026

Project Simurgh

Provider-agnostic verifiable containment-attestation framework for agentic AI, evolved from the defensive counterpart to The Invisible Window research. Produces Ed25519-signed, offline-reproducible evidence of what an AI agent was allowed to do after a guardrail miss — contained 138/138 classifier-missed cases and cut a live agent's attack success from 9/140 to 0/140 on AgentDojo.

Containment AttestationAgentic AI SafetyPrompt-InjectionEd25519LLM Red-TeamingAGPL-3.0
[SYS] ENGINEERING
LOCAL-FIRST · MCP AGENT MEMORY2026

Project Zurvan

Local-first LLM knowledge engine. Ingests any document, extracts structured knowledge (claims, concepts, entities, decisions), and exposes it to AI agents via an MCP stdio server. 218 tests passing.

PythonLLMMCPKnowledge GraphSQLiteLocal-firstAI Agents
[SEC] DEFENSIVE2024

Mehr Guard

Privacy-first offline QR & URL security scanner built with Kotlin Multiplatform. 100% offline analysis with 5 platform targets.

KMPSecurity ToolAndroidiOSDesktopWeb
[SYS] ENGINEERING2026

Syllabus-Sync

AI-native Campus OS transforming university PDF syllabi into structured, agent-readable data. Accepted into the Macquarie University Incubator as a formal startup. Full student operations suite with 503 tests across 92 files, live at syllabus-sync.app.

MQ Incubator StartupNext.js 16SupabaseTypeScript
[SYS] ENGINEERING2024

GitSwitch

AI-powered Git client for managing multiple identities and generating semantic commits. Built with Electron and React.

ElectronReactTypeScriptAI
[SYS] ENGINEERING2026

Nexus Archive

Cyberpunk-styled personal media vault with React frontend, Litestar API, and Supabase auth. AI-assisted recommendations, encrypted takeaways, and hardened cookie-based auth.

ReactPythonLitestarSupabase

Operating Principles

02.Philosophy

R

RESEARCH

Red-team LLM agents and measure the cyber-capability uplift of frontier AI. Characterise safety boundaries, contain guardrail misses with signed, offline-verifiable evidence, and publish reproducible findings.

  • + LLM Agent Red-Teaming & Containment
  • + AI Safety & LLM Evaluation
  • + Dual-Use Risk Assessment
S

SECURE

Defense-in-depth that gives defenders the advantage. Offensive research — exploit classes and agent attack campaigns — becomes verifiable containment and real-world defensive tooling.

  • + Cross-Platform Exploit Development
  • + Responsible Disclosure (OWASP/FIRST/CISA)
  • + Verifiable Containment Attestation
THE_LAB/ active_operations

Hands-on vulnerability research and AI safety experimentation. Current work: cross-platform exploit development, AI capability uplift measurement, and safety boundary characterisation.

Vulnerability ResearchAI SafetyExploit DevelopmentResponsible Disclosure
ENTER_LAB

Technical Writing

03.Write-ups