egtimer.dev

Auditable AI systems for law firms and regulated businesses.

Production RAG, document AI and LLM evaluation for domains where being wrong is expensive: legal, tax, compliance. Quality is measured — never assumed.

Book a 20-min intro call Or email me

Technical AI audits, RAG and LLM evaluation — prices in the open

Fixed formats, written deliverables, no scope surprises.

Technical AI audit

from €500

I review your LLM/RAG system (production or pre-acquisition): architecture, dependencies, security basics, unit economics and risks — written report with severities and remediation estimates.

Evaluation / RAG sprint

from €6,000 · 2-3 weeks

Golden datasets, per-field metrics, retrieval benchmarking and regression gates in CI — so every model or prompt change is validated with numbers before it ships.

Monthly retainer

tailored

Ongoing senior AI engineering capacity without a hiring process: system evolution, continuous evaluation, technical oversight. Async-first, documented, in English or Spanish.

Evidence, not promises

75-90%less manual time per filing (legal system in production)
0.98Recall@5 — retrieval benchmarked on a golden dataset (9 configs)
217automated tests in public repos · 8 ADRs
Bedrockhigh-volume production vision pipeline and agents on AgentCore (multinational in-house AI dept.)
5.0latest platform client review (US, tax)
“Eduardo was fantastic… very pleased with how things turned out.”— US tax-compliance client (verifiable 5.0 review on Upwork)

github.com/egtimer · LinkedIn

Why “auditable”

An LLM once returned a perfectly formatted tax-office code… that didn't exist. My validation layer caught it against the official registry — no human would have. Everything I build follows the same principle since: model fluency is never evidence. Deterministic validation where errors are expensive, human approval where judgment lives, and numbers — not vibes — deciding what ships.