
Manila GMT+8 · open to work
James Lorenz Santos Agentic Engineer & AI Automation Developer
I build production software with AI agents, and I publish the artifact behind every number I claim.
- 35
- projects catalogued
- 28
- live, one click away
- 29
- carry a number and its file






sluice666 duplicate side effects with naive retry, 0 with sluiceread from chaos/results/latest.jsonreceipts
dipmeter230,059 located hypocentres against 27 modelled slabsread from public/data/manifest.jsonreceipts
polis8 unreachable members and 33 never-dispatched agents, out of 59read from data/ecosystem.jsonreceipts
orrery0.0149 AU worst blind-to-predicted gap disagreementread from data/audit.jsonreceipts
kedge102 of 102 Jepsen etcd verdicts matchedread from vendor/porcupine/jepsen/ (103 histories, MIT) replayed by src/cli.js corpusreceipts
clarifier0.726 vs 0.672 silhouette, verdict: no meaningful gainread from evals/fixtures.eval.test.tsreceipts
Selected work
Every card names the file its number was read from.






Three ways in
Proof
Every measured finding on this site, each next to the artifact it was read from and the command that regenerates it.
Services
What I do for money, what it costs, and what I will not take on.
Hire
Manila, GMT+8, and I answer my own mail. Tell me what is broken and I will tell you whether I am the right person.
The workbench
The same content as a file tree, tabs and a real terminal. Every command you see it run is a command you can type.
The long version: career, every project, every service
Just looking around
This is a portfolio you operate instead of scroll. You ask it something, it prints the answer, and it always shows the command it ran so nothing is hidden.
Seven agents built this site and each one answers for its own part of it. Press anything below and it will keep offering the next step.
| The cast | Seven real agents that built this site, one lane each. Ask one something outside its lane and it hands you to the right one by name. |
| The story | Five chapters, 2021 to now, from university to the current role. |
| The work | 35 projects, ranked by strength. The 6 leads come first, and every card carries a proof tier saying how far that one actually got. |
Session log
This is a portfolio you operate instead of scroll. You ask it something, it prints the answer, and it always shows the command it ran so nothing is hidden.
Seven agents built this site and each one answers for its own part of it. Press anything below and it will keep offering the next step.
The panel above is a workbench you can operate. Everything below is the same record as plain HTML, for a browser with JavaScript off or a crawler that does not run it.
Career
- University of Santo Tomas · BS Information Technology · 2021 to 2025
- Four years of IT at UST, finished at the top of the department.
- eBiZolution · Full Stack Applications Developer (intern) · Jan 2025 to Jun 2025
- First professional codebase. Frontend work with React and Tailwind, plus the API and database layer behind it.
- Kredit Hero · Software Engineer · Jul 2025 to Dec 2025
- Full stack engineer on an AI credit application. The strongest receipt on this site comes from here.
- Lift Legal Marketing · AI Automation Developer · Jan 2026 to present
- Current role. AI integration for legal marketing workflows, working GMT+8 against AEST.
- Aldridge · AI and Automation Developer (contract) · May 2026
- Building the AI practice at a US managed service provider, across internal operations and client service delivery.
- Freelance, across four countries · Independent AI consulting, alongside the roles above · Ongoing
- Direct client work for people in Canada, Australia, the United States and the United Kingdom. Increasingly the AI feature itself and the system around it, rather than general build work.
- What is being built now · Side projects and experiments · 2026 onward
- Agent systems, and the tools that make one engineer ship like a team. Labelled as in progress, because that is what it is.
Projects
sluice
Exactly-once side effects for agent tool calls, plus human approval gates that survive a process crash.
windlass
Pipelines as typed graphs with verifier edges and human gates. The runner has no model inside it, and exit 3 means stop and wait for a human.
kedge
A deterministic five-node Raft simulator with a linearizability checker, held to the published verdicts of 102 real Jepsen etcd histories.
flume
Event-time windowing with watermarks, and a checker for the result, measured on three real streams whose lateness spans four orders of magnitude.
cofferdam
Every crash point of a write-ahead log store visited under an explicit device model, then held against real process kills on NTFS and against node:sqlite.
driftwatch
A documentation-drift checker with exactly three outcomes, whose first real run across seven repositories reported ten failures and every one was a false positive.
proofpage
A receipts page that refuses to print a number it did not measure, and which caught its own README lying about its first command.
claude-code-boundary-guard
A PreToolUse hook that stops an agent in one project from writing into another: one file, no dependencies, and a self-test that has actually failed known-bad cases.
chaff
Static analysis for CLAUDE.md, AGENTS.md and tool definitions. Every enforced rule ships the eval that measured its effect.
dipmeter
Draws 230,059 located earthquakes at their real depth inside a transparent Earth against the 27 subduction surfaces USGS modelled, and separates out the 106,108 whose depth a locator assigned rather than solved.
orrery
Positions 1,562,531 catalogued small bodies every frame by solving Kepler's equation in a vertex shader from their own published orbital elements, and says on the page why the default view smears the gaps it is famous for.
snapgauge
Contract tests for MCP servers. Snapshot the schema and behavior, then fail CI when the next version moves.
provenote
A C2PA honesty test: drop an image, see its full provenance chain and exactly what that chain cannot prove, with nothing leaving your device.
dogwatch
A scheduled agent pipeline that publishes every run: cost, gate decisions, refusals, and the runs where it found nothing.
tiltmeter
Tells you when a new model release moves your agent harness off true. Every reading reproduces from a pinned commit.
shipgauge
A browser-ML shippability study: what a model actually costs to ship, measured in a real browser, with n printed.
assay
Brand kits as pure functions: one JSON config in, thirteen byte-identical artifacts out, and a contrast gate that refuses to generate below 3:1.
galley
Renders a repository's real commits, CI results and test counts into a release video, with every number on screen labeled by the source it came from.
halo-halo
A legible Taglish code-switching segmenter: every switch point shown, every label traced to a readable rule, down to the affix inside a single word.
swage
ASL fingerspelling handshape practice, graded by a classifier the project trained and evaluated itself, on a device that never sends the camera feed anywhere.
graticule
Paste your notes and get a map of how their wording relates, embedded entirely on-device, with the technique's own blind spot demonstrated on the page.
clarifier
Paste a CSV and let a physics simulation arrange it, then read a computed number that says whether the arrangement beat a plain scatter plot.
pitman
Shows exactly what a browser speech model heard on your speech, with per-word confidence and a diff against what you meant, entirely on-device. No accent scores.
orphanage
A crawl-graph auditor that found 40 problems on its first real site, and was wrong about all 40.
aeo-lab
A crawler-access scanner whose main feature is reporting what it cannot prove.
polis
Reads an agent ecosystem off disk, renders it as an isometric city you can walk around, and audits it while reading.
carillon
Your typing rhythm becomes sound in the browser, and a checker enforces that nothing is uploaded.
tally
An offline supply-chain and licensing auditor for npm projects, whose first real run flagged its own SPEC.md as a leaked test file.
consentry
A rules-based linter for US A2P 10DLC SMS campaign registrations that detects documented rejection causes offline and refuses to predict approval.
Klik
A hiring app where clients and freelancers match by swiping. Live on Cloudflare.
One Nadela Ops
An operations platform a real business runs on: purchase orders, stock takes, approvals.
AGENTJAMES Control
A desktop mission control for every project on one machine. Local only, and it still works with the network off.
ClubScope Insight Engine
A concept prototype where the language model is never allowed to produce a number: typed tools compute, every figure cites its evidence, and a verifier recomputes it before anything renders.
Kredit Hero: AI Credit App
AI credit application that reached 1,000+ users in its first 30 days.
Agent James: Agent Console
This site: an IDE you operate, where every claim carries a receipt.
Every project in full on /projects
Services
Agentic Web Development
Web applications where an agent answers in context, routes a request, or runs a task end to end, with the model layer kept behind a server boundary.
Agent Reliability & Evaluation
Model and agent features put under measurement, so a change to a prompt, a tool or a model is scored against a fixed task set before anyone else meets it.
MCP Servers & Assistant Integration
Servers that expose a system's own record to an AI assistant over the Model Context Protocol, with every result carrying the address it was read from.
Agent Ecosystem Design & Audit
An existing set of agents, skills and prompts measured against what it is actually used for, then cut back to the parts that earn their place.
Full-Stack Development
One engineer owns the database, the API and the interface, so no requirement is lost in a handoff between separate teams.
Project Management & BA
Requirements written down, scoped and sequenced before code is written, by the same person who then builds it.
AI Automation & Integration
Repetitive work moved into inspectable pipelines that run against the APIs of the tools a business already pays for.
Legacy Modernization
Ageing codebases moved onto a maintainable stack by a migration path that is scripted and repeatable, with the content carried across intact.
Custom Agent Building
One agent, or a small team of them, built for a named job inside a business, with a written task it has to pass before it is trusted with the work.
AI Connectors & Integrations
The plumbing between two systems that were never designed to talk, so a record created in one appears in the other without a person retyping it.
AI Consultancy
An outside read on where a model helps a business and where it will cost more than it returns, written down and argued rather than presented.
Agentic Transformation for Organisations
One process at a time moved to an agent that runs it, measured against the way that process ran before, with anything that did not improve handed back.
AI Solutions
One problem taken end to end: the data it needs, the model call it makes, the interface a person uses, and the measurement that says whether it worked.
WordPress Development
WordPress sites built or repaired by someone who reads the theme and the database, rather than installing another plugin on top of the problem.
Elementor Builds
Pages built in the visual builder a client already uses, kept fast, and kept editable by the person who owns the site after the engagement ends.
Kadence Builds
Sites on the Kadence theme and its blocks, set up so the layout stays consistent when someone other than the builder adds the next page.
WooCommerce Stores
A store with its products, its tax and shipping rules, and a checkout proven in test mode before a real card is ever presented to it.
Shopify Development
Themes and small apps on a hosted commerce platform, where the platform owns payments and the work is everything arranged around them.
GoHighLevel Systems
Pipelines, forms and follow up inside the agency platform, wired so a new lead is answered by the system instead of remembered by a person.
AI Search Readiness (SEO / AEO / GEO)
Pages made citable by an answer engine as well as findable by a search engine: entity markup, answer shaped copy, and a record of which prompts return the page.