Avatar of Bobby Filar

Bobby Filar

Sublime Security

Head of AI at Sublime Security. Research on agentic systems, LLM evaluation, adversarial ML, and AI governance.

  • About
  • Projects
  • Publications
  • Media & Writing
  • CV

#writing

Content tagged with "writing"

AI Safety Evolved: Secure-by-Design, Safe-by-Measurement
2026-06-30 Sublime Security blog (with Aryan Luthra)
#Writing #Agentic Systems

A security AI model maturity checklist and Sublime's four pillars of AI safety for agents that operate directly on attacker-authored content.

View
Adversarial Prompt Injection Payload for Evading AI-Based Detection
2026-06-25 Sublime Security blog, Attack Spotlight (with Sam Scholten)
#Writing #Email Security

Analysis of an in-the-wild credential phishing campaign carrying a second, adversarial payload designed to manipulate AI-based email security into returning a benign verdict.

View
The Craft of Designing Resource-Efficient Agents
2026-06-05 Georgian AI Lab Substack (with Kshitij Jain, Aryan Luthra, Asna Shafiq)
#Writing #Agentic Systems

A two-tier router architecture — fine-tuned open-weight model with uncertainty-based escalation to a frontier model — that cut ASA's median latency 16x and inference cost ~70% with no accuracy loss.

View
How Sublime's AI Agents Are Secure by Design
2026-04-09 Sublime Security blog
#Writing #Agentic Systems

Architectural deep-dive on tool scoping, platform-enforced authorization, prompt injection mitigation, and graduated oversight for production LLM agents.

View
More Than 'Plausible Nonsense': A Rigorous Eval for ADÉ, Our Security Coding Agent
2025-09-25 Sublime Security blog (with Dr. Anna Bertiger)
#Writing #Evaluation

A three-pillar framework — detection accuracy, robustness, and economic cost of coverage — for evaluating LLM-generated detection rules, applied to Sublime's ADÉ agent.

View
Sublime Attack Score: Explainable, AI-Backed Threat Analysis
2024-06-10 Sublime Security blog
#Writing #Machine Learning

Introducing Attack Score, an explainable AI feature that summarizes threats and provides a transparent verdict, built on privacy-preserving feature engineering and MQL.

View
Unmasking BEC Attacks Using Natural Language Understanding + MQL
2023-04-18 Sublime Security blog
#Writing #Email Security

How Sublime combines Natural Language Understanding with Message Query Language to detect Business Email Compromise, an attack style with no malicious links or attachments to key on.

View
Detecting Credential Phishing Using Deep Learning + MQL
2023-03-30 Sublime Security blog, Attack Spotlight
#Writing #Email Security

How Sublime's LinkAnalysis function uses Siamese neural networks and computer vision on rendered link screenshots to catch brand-impersonating credential phishing pages.

View
© 2026 Bobby Filar.
Built with Academic Portfolio Astro