FlowNorth.ai builds private AI agents for Canadian businesses whose data shouldn't go to an outside AI provider, running open-weight models such as Llama, Qwen and Mistral on a Canadian cloud region or on your own servers.
Book a Free 30-Minute CallMost businesses don't. These are the cases where the extra cost is worth it.
Two options, and a hybrid of the two.
The model and the agent run on servers in your own cloud account, in Azure Canada Central or AWS ca-central-1. Your data stays in Canada and you don't buy hardware. You pay for the server whether it's busy or idle.
A server in your office or data centre with a GPU sized for the model. The model runs inside your building. You take on the hardware cost, power, patching and backups, or we plan them with your IT team.
A private model handles sensitive documents, and a hosted model such as Claude handles low-risk work like drafting a newsletter.
Frameworks like Hermes Agent and OpenClaw can run commands, read files and send messages. That power is the point, and also the risk. Before one touches company data, we set it up like this.
Every model makes mistakes, and smaller open-weight models make more of them than the largest hosted ones. We plan for it before go-live.
Totals must add up, PO numbers must match your format, the customer must exist. Anything that fails goes to a person.
When the agent isn't sure, it stops and asks instead of guessing.
Real examples with known answers, re-run after every model or prompt change.
We keep the previous model and prompt version, so we can switch back the same day.
More control, more setup cost, and less capable models. Here is the honest comparison.
| Hosted model (Claude, OpenAI) | Private model, Canadian cloud | Private model, your hardware | |
|---|---|---|---|
| Where the model runs | The provider's servers, which may be outside Canada | Your cloud account, in a Canadian region | Your building |
| Quality on long, multi-step work | The best available | Good on focused jobs, weaker on long reasoning | Same as cloud, limited by the GPU you buy |
| Upfront cost | Lowest | Low, with no hardware to buy | Highest: a server with a capable GPU |
| Running cost | Usage fees per request | Server rental, busy or idle | Power, maintenance and eventual replacement |
| Who patches it | The provider | Us or your IT team | Us or your IT team, plus the hardware |
| Time to first agent | Shortest, since there is no model server to set up | Longer, since the model server is set up and tested first | Longest, since hardware has to arrive and be installed |
If your data isn't sensitive and no contract requires Canadian hosting, a hosted model usually costs less and gives better answers. We'll tell you that on the call.
Examples of agents we could scope. They are illustrations, not client case studies.
An agent on a server in the firm's office could read incoming documents, file them to the right matter folder and draft a summary for the lawyer, with the model running inside the building.
An agent in AWS ca-central-1 could pull figures from client statements during tax season and flag missing slips for a staff accountant to chase.
An agent could search past project files and specs to answer "have we done something like this before?" for estimators, without the archive leaving the firm's cloud account.
One written scope covers the hosting, the model, the controls and the handover.
Every engagement starts with a free 30-minute call. After it, you get a written scope and one fixed price before any work starts.
Hosting or hardware is extra, sized to the model and your volume in the scope.
Ongoing support is optional. If you want it, it goes into the same scope.
Open-weight models run with Ollama, open-source agent frameworks, and hosting in Canada or in your building.
Straight answers to what buyers ask us most.
A private agent is often one part of a larger integration project.
Agents that read orders, invoices and requests and enter them in your ERP for a person to approve, across up to about 10 systems.
Not sure which job to hand an agent first? Start with a workflow audit and a prioritized plan.
Running a smaller business that only needs two well-known apps connected, or a phone or follow-up agent? Our small-business service at aiagentsforsmallbusiness.ca handles that.
Book a free 30-minute call. We'll look at what data the agent would touch and tell you honestly whether a hosted or private model fits better.
Book a Free 30-Minute CallToronto-based ยท Canadian businesses ยท Fixed price in writing before work starts