
Architecture note · The four layers behind Trezly’s owned-AI approach.
One accountable stack
Trezly operates the application, inference gateway, real-estate-tuned models, and dedicated NVIDIA H200 clusters that power Otto. Owning the stack gives us direct control of model delivery and the infrastructure beneath the desk.
What each layer does
Application: clients, listings, transactions, and negotiation workflows.
Gateway: sensitivity routing, PII redaction, and audit logging.
Models: tuning and maintenance for real estate work.
Hardware: dedicated clusters for in-house inference.
The purpose is practical: keep intelligence close to the work and reduce dependence on a third-party model API. Architecture ownership does not remove the need for thoughtful review, clear policies, or responsible use.