Architecture note · The four layers behind Trezly’s owned-AI approach.

One accountable stack

Trezly operates the application, inference gateway, real-estate-tuned models, and dedicated NVIDIA H200 clusters that power Otto. Owning the stack gives us direct control of model delivery and the infrastructure beneath the desk.

What each layer does

  • Application: clients, listings, transactions, and negotiation workflows.

  • Gateway: sensitivity routing, PII redaction, and audit logging.

  • Models: tuning and maintenance for real estate work.

  • Hardware: dedicated clusters for in-house inference.

The purpose is practical: keep intelligence close to the work and reduce dependence on a third-party model API. Architecture ownership does not remove the need for thoughtful review, clear policies, or responsible use.