/Resume
Download ResumeMohammad Raouf Abedini
AI Security Research · LLM Agent Red-Teaming · Offensive & Defensive Security Engineering
About
Security researcher building systems that measure and contain the cyber capabilities of frontier AI. Creator of Project Simurgh, a provider-agnostic containment-attestation framework that red-teams LLM agents under an adversarial, dishonest-producer threat model and produces Ed25519-signed, offline-verifiable evidence of what an agent did after a guardrail miss: 138/138 classifier-missed cases contained, live-agent attack success cut from 9/140 to 0/140 on AgentDojo, five machine-checked Lean theorems. Evaluated Claude outputs for exploitable code and guardrail circumvention in Anthropic's safety-evaluation program (via Alignerr). Positioned as the defense-in-depth layer complementary to inline classifiers: they govern what a model may say; this attests what an agent was allowed to do. Shipped detection end to end: on-device phishing ML at 87% F1 and real-time intrusion detection at 500K+ packets/sec. Cofounder of an incubator-backed campus-AI startup; 70+ projects shipped.
Security Research
Project Simurgh
2026Project Simurgh — Verifiable Containment Attestation for Agentic AI (Creator · AGPL-3.0)
Technical Proficiencies
> Languages
Python (primary), C, C++, TypeScript, JavaScript, Swift, Kotlin, Bash, SQL, Go (familiar)
> Security & Offensive
Vulnerability research, cross-platform exploit development (Win32 API, macOS ScreenCaptureKit), threat modelling, secure code review, penetration testing, responsible disclosure (OWASP/FIRST/CISA), Wireshark, Nmap, Burp Suite
> AI & ML
Large Language Model (LLM) integration & evaluation, Retrieval-Augmented Generation (RAG) evaluation, citation-faithfulness benchmarking, AI-assisted vulnerability research, Natural Language Processing (NLP), generative AI tooling, ML model evaluation, dual-use risk assessment
> Systems & Tools
Linux (Ubuntu/Kali), CMake, Docker, Git/GitHub, GitHub Actions CI/CD, Google Test, FastAPI, Cloudflare Workers, libpcap
> Frameworks
Open Web Application Security Project (OWASP) Top 10, MITRE ATT&CK, National Institute of Standards and Technology (NIST) Framework, W3C Screen Capture Specification
Education
Bachelor of Cyber Security
Macquarie UniversityDiploma of Information Technology
Macquarie UniversitySelected Research & Engineering Projects
The Invisible Window [DISCLOSURE]
2026IEEE-format vulnerability research: independently discovered and responsibly disclosed a cross-platform screen-capture evasion class exploiting OS-level display-affinity APIs (Windows/macOS) that defeats browser-based capture and AI-vision pipelines — 100% evasion across all tested platforms with zero visual artefacts over 10,000+ analysed frames, coordinated disclosure to OS and proctoring vendors. DOI: 10.5281/zenodo.20376495.
Aion [BIBLE RAG]
2026Built AI-powered Bible companion and authored Aion-BibleQA, an 8-page preprint introducing a 40-question benchmark for citation faithfulness and false-premise robustness — R@5 = 0.941, mean citation_support = 0.978, zero unsupported citations, and 6/6 false-premise refusals. DOI: 10.5281/zenodo.20522874.
NanoMatch [SYSTEMS]
2026Engineered high-performance matching engine processing 1M+ orders/second with sub-microsecond latency — implemented red-black tree price levels, custom memory pool allocator, and comprehensive test suite with p50/p99 latency benchmarks.
SentinelFlow [IDS]
2026Built real-time network packet processing engine parsing 500K+ packets/second — protocol dissection (Ethernet/IPv4/TCP/UDP/ICMP/DNS), signature-based detection engine, and stateful analysis (port scans, SYN floods).
Nexus Archive [FULL-STACK]
2025Shipped full-stack data platform with AI recommendation engine, event-driven API design, rate limiting, and automated security scanning — end-to-end ownership from database schema to deployment infrastructure.
Mehr Guard [KOTLINCONF]
2024Built cross-platform offline threat detection tool with local ML-based classification — submitted to KotlinConf global developer conference.
Professional Experience
Freelance Security Engineer & Full-Stack Developer
Self-Employed · Jan 2024 – PresentIT Manager
Iran Pharmacy · Aug 2019 – May 2024AI Safety, Leadership & Community
- ● AI Safety Evaluator, Claude (Alignerr, Anthropic AI safety-evaluation program, 2026): assessed Claude outputs for exploitable code and analysed how safety guardrails can be circumvented, delivering structured, rubric-based findings (continues earlier 2024 Claude Code evaluation work)
Co-Founder · Macquarie Persian Students Society
2026 – PresentFounded and help run the university's Persian student community — events, peer support, and cultural programming.
Licenses & Certifications
> Anthropic
AI Fluency for Students
AI Fluency for Educators
Introduction to Model Context Protocol
Building with the Claude API
AI Fluency: Framework & Foundations
Claude Code in Action
Introduction to Claude Cowork
Claude Platform 101
Claude Code 101
Claude 101
> Macquarie University
Cyber Security: GRC Part 2 — Risk Management and Compliance
99.20%Cyber Security: GRC Part 1 — Governance
95%Cyber Security: Mobile Security
99.28%Cyber Security: Applied Cryptography
98.56%Cyber Security: Data Security and Information Privacy
97.84%Cyber Security: Digital Forensics
95%Cyber Security: Identity Access Management and Authentication
100%Cyber Security: DevSecOps
99.10%Cyber Security: Application of AI
99.20%Cyber Security: Security of AI
95.50%Cyber Security: Essentials for Managers and Leaders
99.40%Cyber Security: Essentials for Workplace
92.80%Cyber Security: Essentials
96%