/Resume
Download ResumeMohammad Raouf Abedini
AI Security Research · LLM Agent Red-Teaming · Offensive & Defensive Security Engineering
About
Security researcher building systems that measure and contain the cyber capabilities of frontier AI. Creator of Project Simurgh, a provider-agnostic containment-attestation framework that red-teams LLM agents under an adversarial, dishonest-producer threat model and produces Ed25519-signed, offline-verifiable evidence of what an agent did after a guardrail miss: 138/138 classifier-missed cases contained, live-agent attack success cut from 9/140 to 0/140 on AgentDojo, five machine-checked Lean theorems. Evaluated Claude outputs for exploitable code and guardrail circumvention in Anthropic's safety-evaluation program (via Alignerr). Positioned as the defense-in-depth layer complementary to inline classifiers: they govern what a model may say; this attests what an agent was allowed to do. Shipped detection end to end: on-device phishing ML at 87% F1 and real-time intrusion detection at 500K+ packets/sec. Cofounder of an incubator-backed campus-AI startup; 70+ projects shipped.
Security Research
Project Simurgh
2026Project Simurgh — Verifiable Containment Attestation for Agentic AI (Creator · AGPL-3.0)
Technical Proficiencies
> Languages
Python (primary), C, C++, TypeScript, JavaScript, Swift, Kotlin, Bash, SQL, Go (familiar)
> Security & Offensive
Vulnerability research, cross-platform exploit development (Win32 API, macOS ScreenCaptureKit), threat modelling, secure code review, penetration testing, responsible disclosure (OWASP/FIRST/CISA), Wireshark, Nmap, Burp Suite
> AI & ML
Large Language Model (LLM) integration & evaluation, Retrieval-Augmented Generation (RAG) evaluation, citation-faithfulness benchmarking, AI-assisted vulnerability research, Natural Language Processing (NLP), generative AI tooling, ML model evaluation, dual-use risk assessment
> Systems & Tools
Linux (Ubuntu/Kali), CMake, Docker, Git/GitHub, GitHub Actions CI/CD, Google Test, FastAPI, Cloudflare Workers, libpcap
> Frameworks
Open Web Application Security Project (OWASP) Top 10, MITRE ATT&CK, National Institute of Standards and Technology (NIST) Framework, W3C Screen Capture Specification
Education
Bachelor of Cyber Security
Macquarie UniversityDiploma of Information Technology
Macquarie UniversitySelected Research & Engineering Projects
The Invisible Window [DISCLOSURE]
2026IEEE-format vulnerability research: independently discovered and responsibly disclosed a cross-platform screen-capture evasion class exploiting OS-level display-affinity APIs (Windows/macOS) that defeats browser-based capture and AI-vision pipelines — 100% evasion across all tested platforms with zero visual artefacts over 10,000+ analysed frames, coordinated disclosure to OS and proctoring vendors. DOI: 10.5281/zenodo.20376495.
Aion [BIBLE RAG]
2026Built AI-powered Bible companion and authored Aion-BibleQA, an 8-page preprint introducing a 40-question benchmark for citation faithfulness and false-premise robustness — R@5 = 0.941, mean citation_support = 0.978, zero unsupported citations, and 6/6 false-premise refusals. DOI: 10.5281/zenodo.20522874.
NanoMatch [SYSTEMS]
2026Engineered high-performance matching engine processing 1M+ orders/second with sub-microsecond latency — implemented red-black tree price levels, custom memory pool allocator, and comprehensive test suite with p50/p99 latency benchmarks.
SentinelFlow [IDS]
2026Built real-time network packet processing engine parsing 500K+ packets/second — protocol dissection (Ethernet/IPv4/TCP/UDP/ICMP/DNS), signature-based detection engine, and stateful analysis (port scans, SYN floods).
Nexus Archive [FULL-STACK]
2025Shipped full-stack data platform with AI recommendation engine, event-driven API design, rate limiting, and automated security scanning — end-to-end ownership from database schema to deployment infrastructure.
Mehr Guard [KOTLINCONF]
2024Built cross-platform offline threat detection tool with local ML-based classification — submitted to KotlinConf global developer conference.
Professional Experience
Freelance Security Engineer & Full-Stack Developer
Self-Employed · Jan 2024 – PresentIT Manager
Iran Pharmacy · Aug 2019 – May 2024AI Safety, Leadership & Community
- ● AI Safety Evaluator, Claude (Alignerr, Anthropic AI safety-evaluation program, 2026): assessed Claude outputs for exploitable code and analysed how safety guardrails can be circumvented, delivering structured, rubric-based findings (continues earlier 2024 Claude Code evaluation work)
Co-Founder · Macquarie Persian Students Society
2026 – PresentFounded and help run the university's Persian student community — events, peer support, and cultural programming.
Licenses & Certifications
> Anthropic
AI Fluency for Builders
AI Fluency for Small Businesses
AI Capabilities and Limitations
Introduction to Subagents
Introduction to Agent Skills
AI Fluency for Nonprofits
Teaching the AI Fluency Framework
Claude on Google Cloud
Claude with Amazon Bedrock
Model Context Protocol: Advanced Topics
AI Fluency for Students
AI Fluency for Educators
Introduction to Model Context Protocol
Building with the Claude API
AI Fluency: Framework & Foundations
Claude Code in Action
Introduction to Claude Cowork
Claude Platform 101
Claude Code 101
Claude 101
> Macquarie University
Cyber Security: GRC Part 2 — Risk Management and Compliance
99.20%Cyber Security: GRC Part 1 — Governance
95%Cyber Security: Mobile Security
99.28%Cyber Security: Applied Cryptography
98.56%Cyber Security: Data Security and Information Privacy
97.84%Cyber Security: Digital Forensics
95%Cyber Security: Identity Access Management and Authentication
100%Cyber Security: DevSecOps
99.10%Cyber Security: Application of AI
99.20%Cyber Security: Security of AI
95.50%Cyber Security: Essentials for Managers and Leaders
99.40%Cyber Security: Essentials for Workplace
92.80%Cyber Security: Essentials
96%