Skip to content

Manila GMT+8 · open to work

James Lorenz Santos Agentic Engineer & AI Automation Developer

I build production software with AI agents, and I publish the artifact behind every number I claim.

35
projects catalogued
28
live, one click away
29
carry a number and its file
Screenshot of the sluice landing page showing its idempotent-execution comparison: 666 duplicate charges under naive retry versus 0 with sluice, next to a results table breaking down intent success rate and duplicate side effects.Screenshot of dipmeter showing 230,059 earthquake hypocentres glowing inside a transparent Earth, coloured by depth from orange at the surface to blue at 700 km, with the Tonga and South American slabs standing out as dipping sheets beside the depth, magnitude and time controls.Screenshot of the polis map mid-pan, showing agent guild districts as isometric buildings and a banner noting that 8 of 59 agent citizens answer to no charter.Screenshot of orrery showing the asteroid belt as a blue haze of 120,000 sampled bodies around the inner planets, with the population counts, the find-a-body index and an instrument readout of drawn count, frame time and solver beside it.Screenshot of the kedge Raft simulator mid-run, showing a red "not linearizable" verdict with the witness operation that could not be placed anywhere in a consistent order, above the client history strip it was derived from.Screenshot of the clarifier tool with a real penguin-measurement dataset plotted as a force-settling point cloud next to its column-mapping table assigning each CSV field a physical role.

sluice666 duplicate side effects with naive retry, 0 with sluiceread from chaos/results/latest.jsonreceipts

dipmeter230,059 located hypocentres against 27 modelled slabsread from public/data/manifest.jsonreceipts

polis8 unreachable members and 33 never-dispatched agents, out of 59read from data/ecosystem.jsonreceipts

orrery0.0149 AU worst blind-to-predicted gap disagreementread from data/audit.jsonreceipts

kedge102 of 102 Jepsen etcd verdicts matchedread from vendor/porcupine/jepsen/ (103 histories, MIT) replayed by src/cli.js corpusreceipts

clarifier0.726 vs 0.672 silhouette, verdict: no meaningful gainread from evals/fixtures.eval.test.tsreceipts

Selected work

Every card names the file its number was read from.

All 35 projects

Three ways in

Proof

Every measured finding on this site, each next to the artifact it was read from and the command that regenerates it.

Read the receipts

Services

What I do for money, what it costs, and what I will not take on.

See services

Hire

Manila, GMT+8, and I answer my own mail. Tell me what is broken and I will tell you whether I am the right person.

Start a conversation

The workbench

The same content as a file tree, tabs and a real terminal. Every command you see it run is a command you can type.

The long version: career, every project, every service

Just looking around

James Lorenz Santos, halftone portrait resolving into machine glyphs
Halftone on the left, binary glyphs on the right, one face. Generated from a photograph by a script in this repo, not drawn.

This is a portfolio you operate instead of scroll. You ask it something, it prints the answer, and it always shows the command it ran so nothing is hidden.

Seven agents built this site and each one answers for its own part of it. Press anything below and it will keep offering the next step.

The castSeven real agents that built this site, one lane each. Ask one something outside its lane and it hands you to the right one by name.
The storyFive chapters, 2021 to now, from university to the current role.
The work35 projects, ranked by strength. The 6 leads come first, and every card carries a proof tier saying how far that one actually got.

Session log

This is a portfolio you operate instead of scroll. You ask it something, it prints the answer, and it always shows the command it ran so nothing is hidden.

Seven agents built this site and each one answers for its own part of it. Press anything below and it will keep offering the next step.

The panel above is a workbench you can operate. Everything below is the same record as plain HTML, for a browser with JavaScript off or a crawler that does not run it.

Career

University of Santo Tomas · BS Information Technology · 2021 to 2025
Four years of IT at UST, finished at the top of the department.
eBiZolution · Full Stack Applications Developer (intern) · Jan 2025 to Jun 2025
First professional codebase. Frontend work with React and Tailwind, plus the API and database layer behind it.
Kredit Hero · Software Engineer · Jul 2025 to Dec 2025
Full stack engineer on an AI credit application. The strongest receipt on this site comes from here.
Lift Legal Marketing · AI Automation Developer · Jan 2026 to present
Current role. AI integration for legal marketing workflows, working GMT+8 against AEST.
Aldridge · AI and Automation Developer (contract) · May 2026
Building the AI practice at a US managed service provider, across internal operations and client service delivery.
Freelance, across four countries · Independent AI consulting, alongside the roles above · Ongoing
Direct client work for people in Canada, Australia, the United States and the United Kingdom. Increasingly the AI feature itself and the system around it, rather than general build work.
What is being built now · Side projects and experiments · 2026 onward
Agent systems, and the tools that make one engineer ship like a team. Labelled as in progress, because that is what it is.

Full career record on /about

Projects

  • sluice· 2026 · Live

    Exactly-once side effects for agent tool calls, plus human approval gates that survive a process crash.

  • windlass· 2026 · Completed

    Pipelines as typed graphs with verifier edges and human gates. The runner has no model inside it, and exit 3 means stop and wait for a human.

  • kedge· 2026 · Live

    A deterministic five-node Raft simulator with a linearizability checker, held to the published verdicts of 102 real Jepsen etcd histories.

  • flume· 2026 · Live

    Event-time windowing with watermarks, and a checker for the result, measured on three real streams whose lateness spans four orders of magnitude.

  • cofferdam· 2026 · Live

    Every crash point of a write-ahead log store visited under an explicit device model, then held against real process kills on NTFS and against node:sqlite.

  • driftwatch· 2026 · Completed

    A documentation-drift checker with exactly three outcomes, whose first real run across seven repositories reported ten failures and every one was a false positive.

  • proofpage· 2026 · Live

    A receipts page that refuses to print a number it did not measure, and which caught its own README lying about its first command.

  • claude-code-boundary-guard· 2026 · Live

    A PreToolUse hook that stops an agent in one project from writing into another: one file, no dependencies, and a self-test that has actually failed known-bad cases.

  • chaff· 2026 · Live

    Static analysis for CLAUDE.md, AGENTS.md and tool definitions. Every enforced rule ships the eval that measured its effect.

  • dipmeter· 2026 · Live

    Draws 230,059 located earthquakes at their real depth inside a transparent Earth against the 27 subduction surfaces USGS modelled, and separates out the 106,108 whose depth a locator assigned rather than solved.

  • orrery· 2026 · Live

    Positions 1,562,531 catalogued small bodies every frame by solving Kepler's equation in a vertex shader from their own published orbital elements, and says on the page why the default view smears the gaps it is famous for.

  • snapgauge· 2026 · Live

    Contract tests for MCP servers. Snapshot the schema and behavior, then fail CI when the next version moves.

  • provenote· 2026 · Live

    A C2PA honesty test: drop an image, see its full provenance chain and exactly what that chain cannot prove, with nothing leaving your device.

  • dogwatch· 2026 · Live

    A scheduled agent pipeline that publishes every run: cost, gate decisions, refusals, and the runs where it found nothing.

  • tiltmeter· 2026 · Live

    Tells you when a new model release moves your agent harness off true. Every reading reproduces from a pinned commit.

  • shipgauge· 2026 · Live

    A browser-ML shippability study: what a model actually costs to ship, measured in a real browser, with n printed.

  • assay· 2026 · Live

    Brand kits as pure functions: one JSON config in, thirteen byte-identical artifacts out, and a contrast gate that refuses to generate below 3:1.

  • galley· 2026 · Live

    Renders a repository's real commits, CI results and test counts into a release video, with every number on screen labeled by the source it came from.

  • halo-halo· 2026 · Live

    A legible Taglish code-switching segmenter: every switch point shown, every label traced to a readable rule, down to the affix inside a single word.

  • swage· 2026 · Live

    ASL fingerspelling handshape practice, graded by a classifier the project trained and evaluated itself, on a device that never sends the camera feed anywhere.

  • graticule· 2026 · Live

    Paste your notes and get a map of how their wording relates, embedded entirely on-device, with the technique's own blind spot demonstrated on the page.

  • clarifier· 2026 · Live

    Paste a CSV and let a physics simulation arrange it, then read a computed number that says whether the arrangement beat a plain scatter plot.

  • pitman· 2026 · Live

    Shows exactly what a browser speech model heard on your speech, with per-word confidence and a diff against what you meant, entirely on-device. No accent scores.

  • orphanage· 2026 · Live

    A crawl-graph auditor that found 40 problems on its first real site, and was wrong about all 40.

  • aeo-lab· 2026 · Live

    A crawler-access scanner whose main feature is reporting what it cannot prove.

  • polis· 2026 · Live

    Reads an agent ecosystem off disk, renders it as an isometric city you can walk around, and audits it while reading.

  • carillon· 2026 · Live

    Your typing rhythm becomes sound in the browser, and a checker enforces that nothing is uploaded.

  • tally· 2026 · Completed

    An offline supply-chain and licensing auditor for npm projects, whose first real run flagged its own SPEC.md as a leaked test file.

  • consentry· 2026 · Completed

    A rules-based linter for US A2P 10DLC SMS campaign registrations that detects documented rejection causes offline and refuses to predict approval.

  • Klik· 2026 · Live

    A hiring app where clients and freelancers match by swiping. Live on Cloudflare.

  • One Nadela Ops· 2026 · Live · One Nadela

    An operations platform a real business runs on: purchase orders, stock takes, approvals.

  • AGENTJAMES Control· 2026 · Live

    A desktop mission control for every project on one machine. Local only, and it still works with the network off.

  • ClubScope Insight Engine· 2026 · Live

    A concept prototype where the language model is never allowed to produce a number: typed tools compute, every figure cites its evidence, and a verifier recomputes it before anything renders.

  • Kredit Hero: AI Credit App· 2025 · Completed · Kredit Hero

    AI credit application that reached 1,000+ users in its first 30 days.

  • Agent James: Agent Console· 2026 · Live

    This site: an IDE you operate, where every claim carries a receipt.

Every project in full on /projects

Services

  • Agentic Web Development

    Web applications where an agent answers in context, routes a request, or runs a task end to end, with the model layer kept behind a server boundary.

  • Agent Reliability & Evaluation

    Model and agent features put under measurement, so a change to a prompt, a tool or a model is scored against a fixed task set before anyone else meets it.

  • MCP Servers & Assistant Integration

    Servers that expose a system's own record to an AI assistant over the Model Context Protocol, with every result carrying the address it was read from.

  • Agent Ecosystem Design & Audit

    An existing set of agents, skills and prompts measured against what it is actually used for, then cut back to the parts that earn their place.

  • Full-Stack Development

    One engineer owns the database, the API and the interface, so no requirement is lost in a handoff between separate teams.

  • Project Management & BA

    Requirements written down, scoped and sequenced before code is written, by the same person who then builds it.

  • AI Automation & Integration

    Repetitive work moved into inspectable pipelines that run against the APIs of the tools a business already pays for.

  • Legacy Modernization

    Ageing codebases moved onto a maintainable stack by a migration path that is scripted and repeatable, with the content carried across intact.

  • Custom Agent Building

    One agent, or a small team of them, built for a named job inside a business, with a written task it has to pass before it is trusted with the work.

  • AI Connectors & Integrations

    The plumbing between two systems that were never designed to talk, so a record created in one appears in the other without a person retyping it.

  • AI Consultancy

    An outside read on where a model helps a business and where it will cost more than it returns, written down and argued rather than presented.

  • Agentic Transformation for Organisations

    One process at a time moved to an agent that runs it, measured against the way that process ran before, with anything that did not improve handed back.

  • AI Solutions

    One problem taken end to end: the data it needs, the model call it makes, the interface a person uses, and the measurement that says whether it worked.

  • WordPress Development

    WordPress sites built or repaired by someone who reads the theme and the database, rather than installing another plugin on top of the problem.

  • Elementor Builds

    Pages built in the visual builder a client already uses, kept fast, and kept editable by the person who owns the site after the engagement ends.

  • Kadence Builds

    Sites on the Kadence theme and its blocks, set up so the layout stays consistent when someone other than the builder adds the next page.

  • WooCommerce Stores

    A store with its products, its tax and shipping rules, and a checkout proven in test mode before a real card is ever presented to it.

  • Shopify Development

    Themes and small apps on a hosted commerce platform, where the platform owns payments and the work is everything arranged around them.

  • GoHighLevel Systems

    Pipelines, forms and follow up inside the agency platform, wired so a new lead is answered by the system instead of remembered by a person.

  • AI Search Readiness (SEO / AEO / GEO)

    Pages made citable by an answer engine as well as findable by a search engine: entity markup, answer shaped copy, and a record of which prompts return the page.

All services on /services