Implementation partners, custom skills, and the clouds they deliver on.
FlatClaw is the platform. Implementation partners build the skills, MCP servers and bespoke workflows that turn it into the specific solution a firm needs, and deliver the tenancy on the cloud the customer already trusts: Azure, AWS, Google Cloud, Northflank, or their own hardware.
Kirk Tech Solutions launches FlatClaw
Kirk Tech Solutions has announced the launch of FlatClaw: a private, secure, single-tenant AI platform deployed inside customers' own private cloud. The release marks the first wave of FlatClaw deployments, with Kirk Tech delivering the integration, custom skills, and MCP services that turn the open-source platform into working solutions for each customer.
Read the press releaseKirk Tech Solutions
Custom AI, MCP integrations, and cloud delivery
Kirk Tech Solutions builds the skills, MCP servers and bespoke workflows that turn FlatClaw into the specific solution a firm needs, and delivers the tenancy on the customer's cloud of choice: Azure, AWS, Google Cloud, Northflank, or a rack in the building.
For each deployment they handle the end-to-end: skill design and implementation, MCP server authoring, RBAC policy mapping, approval-gate configuration, and the integration glue that connects the platform to whatever the customer already runs.
Bespoke agents, prompts, voice loops and pipelines fit to your workflows.
MCP servers that expose your existing systems to agents safely.
The tenancy provisioned in your account, on the cloud you already run.
One tenancy per customer, on any of these.
The same public inference image, the same control plane, the same privacy proof. What changes per cloud is the account it lives in and who provisions it.
The lane for Microsoft-first organizations.
A resource group inside the customer's own subscription: AKS or a single GPU virtual machine (NC H100 v5 class), Entra ID for sign-in, private endpoints for everything, and outputs into Fabric and Power BI when the customer wants them.
- Entra ID sign-in
- NC H100 v5 GPU nodes
- Fabric / Power BI hand-off
Your account, your VPC, no public inference endpoint.
An account and VPC the customer owns: EKS or a single GPU instance (p5 or g6e class), IAM roles for access, KMS for secrets, and egress limited to the services users explicitly connect.
- IAM + KMS
- p5 / g6e GPU instances
- Private VPC egress only
A project the customer owns, fenced with VPC Service Controls.
GKE with A3 (H100) nodes or a single GPU VM inside the customer's project, Workload Identity for the services, and VPC Service Controls around the tenancy so nothing crosses the boundary unnoticed.
- Workload Identity
- A3 (H100) nodes
- VPC Service Controls
The fastest path from zero to a running tenant.
The managed-GPU platform FlatClaw was built and verified on. One project per tenant, H100 plans by the hour, and the lane scripts in the repository bring inference up and down with one command. Every release is proven here first.
- One project per tenant
- Managed H100 plans
- Lane scripts in the repo
For data that cannot be in any cloud at all.
Bare metal or a private Kubernetes cluster in your building: an NVIDIA H100 or RTX PRO 6000-class card, the same containers, the same image. No cloud account, no cloud bill, no cloud dependency.
- H100 or RTX PRO 6000 class
- Private Kubernetes or a single host
- Zero cloud dependency
Microsoft Azure, Amazon Web Services, Google Cloud and Northflank are trademarks of their respective owners. Logos identify supported deployment targets; no endorsement is implied.
Northflank is the reference lane: the lane scripts in the repository bring a tenant up and down today, and every release is verified there first. Azure, AWS and Google Cloud tenancies are delivered with an implementation partner using the same containers; one-command provisioning for those lanes is on the roadmap.
The part that does not change per cloud.
The same public inference image from GHCR runs on every cloud. Per-tenant differences live on the weights volume and in the tenancy's secrets, never in the image.
A resource group, an account, a project, or a rack — one per customer, with the control plane and the GPU inside it. No shared state across tenants.
The cloud bills the customer directly. We never sit between a customer and their substrate, and we never touch the bill or the data.
Run tcpdump on the tenancy's egress for a full session. Zero packets to any third-party inference endpoint, on any of these clouds.
Building on FlatClaw, or hosting it?
Cloud providers, GPU platforms, and implementation firms that deploy FlatClaw or build skills, MCP servers and services on top of it can be listed here.
Let us help you deliver your next AI project on infrastructure you own.
Thirty minutes: your workflow, your cloud, and whether a Private AI Platform is the right fit for it.