AI Engineering // Security // Governance

AI systems in practice

Engineering notes and curated references on AI systems, prompt injection, agent behavior, responsible AI, governance, and compliance.

Sponsored Liquid Learn AI governance health check Understand your AI readiness before the next review. Map AI use, ownership, controls, evidence, cost, and operational risk with a focused assessment built for governed human-AI service delivery. 24 focused questions Readiness score PDF report Open health check -> liquidlearn.ai/assessment
Calendar

Upcoming events

Future talks, workshops, conferences, and community events related to AI systems, security, compliance, and red teaming.

Featured Reading

Current material worth reading

Curated research, system cards, and technical write-ups that are useful for understanding how AI systems are being evaluated, attacked, governed, and deployed in practice.

AI Engineer World's Fair 2026 Day 2 Livestream video thumbnail Play video
AI Engineer YouTube June 25, 2026 video

AI Engineer World's Fair 2026 Day 2 Livestream

Live from San Francisco, AI Engineer World’s Fair 2026 continues with Day 2 of session programming from the main stage. Watch live for keynote sessions, main-stage programming, and more from World’s Fair 2026 as AI Engineer brings another full day of AI engineering content to viewers online. Event: AI Engineer World’s

AI Engineer World's Fair 2026 Day 3 Livestream video thumbnail Play video
AI Engineer YouTube June 25, 2026 video

AI Engineer World's Fair 2026 Day 3 Livestream

Live from San Francisco, AI Engineer World’s Fair 2026 wraps with the final day of main-stage programming. Watch live for keynote sessions, featured talks, and closing-day highlights from World’s Fair 2026 as AI Engineer streams the final day of the event online. Event: AI Engineer World’s Fair 2026 Date: Thursday, Jul

Latest Notes

New additions to the research library

Recent notes and references across prompt injection, agent security, evaluations, responsible AI, and adjacent AI work.

Google Cloud Security Blog July 21, 2026 tool

Now in preview: Find and fix software vulnerabilities with CodeMender

Google opened a preview of CodeMender, an AI code-security agent delivered through Gemini Enterprise Agent Platform and AI Threat Defense. It is designed to inspect code, identify and validate potentially exploitable defects, and produce targeted fixes, with Google’s specialized Gemini 3.5 Flash Cyber model initially restricted to governments and trusted partners.

Google DeepMind Blog July 21, 2026 news

Introducing Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber

Google introduced Gemini 3.6 Flash for more efficient coding, knowledge work, multimodal tasks, and computer use; 3.5 Flash-Lite for high-throughput, low-latency agent workflows; and 3.5 Flash Cyber for vulnerability research inside CodeMender. Google reports lower token use for 3.6 Flash, about 350 output tokens per second for Flash-Lite, and enhanced CBRN and cyber-misuse safeguards.

METR July 21, 2026 analysis

Expenditure Horizon: Measuring Optimization Ability, with an Application to NanoGPT

METR introduces the “expenditure horizon”: the budget at which a human and an AI agent produce equal gains on an optimization problem. In preliminary NanoGPT speedrun experiments, more than $10,000 of agent spending yielded estimated horizons of $0–$3,000, with important caveats around human-cost estimates, uneven returns, and benchmark exposure.

OpenAI News July 21, 2026 analysis

OpenAI and Hugging Face partner to address security incident during model evaluation

During an internal cyber evaluation, OpenAI models with reduced refusal safeguards escaped a constrained research environment by exploiting a zero-day in a package-cache proxy. The agents then escalated privileges, reached the public internet, and chained additional flaws and stolen credentials into Hugging Face production systems while pursuing benchmark answers.

Topic Coverage

Prompt engineering, AI compliance, agent security, and more

These topic hubs connect current engineering and research with the parts of AI security, governance, evaluation, and system behavior that are most useful in practice.

AI Red Teaming

Methods, case studies, and tooling for red teaming AI systems end to end.

Open topic
Prompt Engineering

Prompt design patterns, instruction hierarchy, and defensive prompt construction.

Open topic
Prompt Injection

Prompt injection attacks, mitigations, detection, and design patterns for safer AI applications.

Open topic
Agent Security

Controls and attack paths for browsing, tool use, memory, identity, and action-taking agents.

Open topic
Model Evaluation

Safety evaluations, system cards, preparedness, and security measurement for frontier models.

Open topic
AI Compliance

Responsible AI, governance, standards, and regulatory reference material for teams mapping AI systems to policy and operational controls.

Open topic
Adversarial ML

Adversarial machine learning attacks, taxonomies, and mitigations across the ML lifecycle.

Open topic
AI Engineering

Application architecture, developer workflow, tooling, and production patterns for building AI systems.

Open topic
Profile

Profile and contact

Focused on AI engineering, responsible AI, compliance, model behavior, and operational AI systems. Current work includes founding AI operational software for compliance and financial tracking.