Research & Articles
What we find, and how we think about it.
Probe methodology, published findings, and platform writing from the team building Orithos.
ARTICLE · AGENT SECURITY
The execution gap: why your agent runs code before it trusts anything
Two disclosure post-mortems this month share one sentence: the code ran before anything verified it. That ordering problem is the durable failure — and the fix is determinism, not intelligence.
Latest3 posts
CASE STUDY
CASE STUDIES
We red-teamed our own AI agents: 126 scans, 243 findings, and what it taught us
A quarter of dogfooding turned inward: 126 scans against our own agents, 87.8% of probes blocked — and 243 findings showing exactly where guardrails collapse. Every number computed from the internal dataset; every gap disclosed.
ARTICLE
EU AI ACT
The EU AI Act is live. Here’s what agent builders actually owe — and in what order.
Enforcement began August 2, 2026. Three obligations apply to agent builders today, the high-risk regime moved to December 2027 — and the sequencing mistake almost everyone makes is doing the paperwork before the testing.