Maria Angelika Agutaya · AI Engineer · agents, evaluation, interfacesall builds · craft · about · Metro Manila, remote · [email protected]

I build AI agents that run themselves and answer to a person, and I design the screens people meet them on. Two dozen live.

Each one stops at a human where a decision carries weight, and keeps the record to prove it. Start with these three.

meridian fleet liveevery plum line ends at a personcode: github.com/mariaangelikabuilds ↗
Authority gateway · MCP · measured against a control armpublic · 19 of 19 to 0 of 19
01

warrant

19 of 19 unauthorized actions reached systems ungoverned. 0 of 19 governed. Same model, same prompt, same tools.

19 of 19unauthorized, ungoverned
0 of 19unauthorized, governed
+18%tokens for that
open the case →
Guest AI agent · Messages APIlive run + 4 real apps
02

Alagà · Guest AI Agent

It knows the whole stay and it acts with real tools. Anything a guest could get hurt by goes to a person.

2 of 5force-escalated in code
0questions re-asked
4 appsone n8n run
open the case →
Agent fleet · runs 24/7 · fly.iolive 24/7
03

Meridian Ops

$0.056 for one closed incident, 15 events end to end, on the ledger. Running unattended since August.

$0.056one closed incident
5 sitesreal uptime monitors
3 surfacesone approve
open the case →
evidence of judgment · the code behind the claims · 5 reposread the repos →
the whole fleet, one ruleevery line ends at approve
Meridian Opslive · fly.io · lab + production
MSP Ticket Triagerunning · private n8n
approve?the human · always
email · ledger · signed linksreceipts
MSP Review & Reputation Enginerunning · private n8n
Phishing Sim Production Linerunning · private n8n
MSP SEO Content Linerunning · private n8n
Alagà · Guest AI Agentlive run + 4 real apps
what runs on its own

what a human still approves

open the build →
Cywareness Simulation Archive

For seven months I was the content engineer at a security company. I built hundreds of phishing-training emails, landing pages, and login screens, in four difficulty tiers, for more than eight markets. Every one had to follow a written ruleset about which red flags belong in which tier. Doing that by hand does not scale, so I turned it into a pipeline. Claude Code skills hold the rules. An orchestrator runs the eight steps in order. Twelve sub-agents do the writing that needs judgment. Python packages the result the same way every time.

2025 to 2026 · body of work · ~300 sims · 8+ markets
the full inventory · all 29 builds, gated the same wayopen the index →