Inference marketplace ยท Private alpha

Inference included. Privacy stays the priority.

Paid plans include a practical inference allowance for evaluation and everyday workloads. Choose TEKIZ.AI-hosted capacity, connect Ollama or another approved provider, or scope a private pilot around your model, processing location, and capacity needs.

Pilot availability and pricing are confirmed before you commit. Explore the marketplace.

Connection plans

Use your model accounts with clear controls.

Every paid plan includes TEKIZ.AI-hosted inference when live capacity is available. You can also connect the model services and local models you already use. TEKIZ.AI manages routing rules, access keys, and request records. A connected provider bills its own model usage. Your selected provider processes your requests under its data-handling terms.

Checking availability. Status

Use your own providers

Connect Ollama alongside the services you approve.

Run a local model directly for a single-model workflow. Use TEKIZ.AI when your team needs one API for approved local and cloud services, with keys, routing rules, and a request record.

Lite, Pro, and Max include hosted inference within their monthly request allowances when live capacity is available. Usage through a provider account you connect stays on that provider's bill. Need a reviewed private setup and one bill? Request managed access. It is a separate private pilot.

Lite
$6 / month

For individuals and small businesses who bring their own model accounts or local models.

  • 10,000 monthly inference requests included
  • Use hosted capacity or providers you approve
  • 60 routing requests per minute
  • Choose from the model accounts you connect
  • Use another approved option when the first one is unavailable
  • Basic usage history and account controls
Try 10 requests free
Max
$72 / month

Higher-volume routing with direct help for private setup and rollout.

  • 200,000 monthly inference requests included
  • Use hosted capacity or providers you approve
  • 600 routing requests per minute
  • Priority eligible capacity when it is available
  • Help with private setup and rollout planning
Try 10 requests free

What your plan covers

Pay for routing and account controls.

Use different approved services for each customer record, case file, or financial workflow. TEKIZ.AI applies the rules your team sets for that work.

One-vendor setup

A provider change, outage, or policy decision can affect every workflow at once.

Approved service choices

Your team chooses the services that are allowed for the work. If no approved choice is available, the request can stop for a person to decide.

What stays yours

Your provider account, provider bill, approval rules, and the record of what happened remain under your control.

10,000
60%
3
One-vendor default 10,000

requests rely on one provider account

With approved choices 4,000

requests remain on the first choice; 6,000 can use 2 other approved choices

Illustrative only. Actual routing depends on your rules, task requirements, connected accounts, service health, and provider limits. TEKIZ.AI does not increase a provider's quota.

Business plans

Privacy and sovereignty are business requirements.

Legal, healthcare, financial, and public-sector teams need clear answers about where work may go, who approved it, and what they can show when someone asks.

Business governance is priced around seats and accountability. Request volume remains part of capacity planning. A Premium sovereign deployment is offered only after its Canadian residency, encryption, and evidence controls are proven for that deployment.

Managed model access

Need one bill for model access too?

Managed access is separate from the routing plans above. Tell us which models and usage you need; we will confirm availability and pricing before you pay.