TinyAI is our family of purpose-built AI models. Each one is under 300MB, runs on standard CPU hardware with no GPU, and deploys entirely inside your environment. Specialised models work together in a three-layer stack to turn data into auditable decisions.
Per model
No GPU required
Training time
On around 200 labelled examples
Inference latency
End-to-end, no round trips
Accuracy
After fine-tuning on your data
NLP feeds machine learning. Machine learning feeds logic. The orchestration layer coordinates everything, running specialised models in parallel with sub-300ms latency.
Extraction, recognition and classification. Converts unstructured documents, contracts, forms and scanned records into structured fields the layers above can process.
Pattern recognition, scoring and anomaly detection. Trained on your historical data, around 200 labelled examples in under an hour, with a confidence score on every prediction.
Deterministic decision chains where one model calls the next. Policy rules, routing, thresholds and human-in-the-loop checkpoints, producing one structured decision with a full audit trail.
Lending decisions are chains. Chain fifteen calls that are each 90% accurate and the end-to-end result falls to about 20%. TinyAI gives the same answer to the same input, every time, so accuracy does not compound away.
TinyAI produces structured, deterministic outputs, not free-form text. Every decision is explainable and audit-ready by design.
A single structured output per model. Each decision carries its confidence score and the features that drove it: input features, weight contributions, model version, timestamp and confidence threshold.
NLP models use contextual attribution. Vision models produce spatial heatmaps. Both operate on the actual model weights, not surrogate approximations, and are stored alongside the inference logs.
All with full data residency and sub-300ms latency. Zero data egress. No PII is ever sent to Synapze or any third party.
Customer-managed servers. Zero outbound calls. Air-gapped compatible, with full business continuity even when disconnected from the internet.
Inside your own AWS, Azure or GCP account. No data leaves the cloud boundary. Same performance guarantees as on-premises.
EU-only or jurisdiction-specific data centres. Supports GDPR Art. 44 to 49 and works with your existing DLP and network controls.
A 30-minute discovery call, a no-obligation ROI assessment, then a pilot on a single process. Prove it works before you commit.