Policy before inference. Evidence after.
Set organisation-scoped routing, optional budgets, and review what was used and what changed. The same rules apply to every request, whichever consumer sent it.
D×E routing policy
Set data sensitivity (D-class) defaults for messages and files, assign provider environment classes (E-class), and block incompatible requests before inference.
Explore data routing →Budgets
Optionally cap spend per member or API key and per provider. Usage stays uncapped until you set those ceilings. Steinkauz AI does not include inference credit.
Explore Budgets →Usage and audit visibility
Inspect inference activity across every consumer, review configuration audit events for policy and provider changes, and trace administrative actions.
Explore audit visibility →Routing
Data routing policy
Choose which provider environment classes may handle each data sensitivity level. Incompatible chat and API requests are blocked before inference.
Interactive demo. No data is saved.
Data sensitivity defaults
Message and file sensitivity can be configured separately. When a request includes a file, Steinkauz AI uses the stricter effective sensitivity.
Default sensitivity for message content in chat before routing is evaluated.
Pre-selected sensitivity for uploaded files. Attachments can raise the effective sensitivity above the message default.
Provider
Each provider has a provider environment class (E-class). Pick one to simulate a request.
Simulate a request
Choose whether the request includes a file attachment to see how effective sensitivity is computed.
Effective sensitivity: D0: Public
D×E routing matrix
Configure which provider environment classes may handle each data sensitivity level. Toggle cells to allow or block combinations.
| Data sensitivity | E0E0: Unclassified | E1E1: Public managed | E2E2: Enterprise managed | E3E3: Private cloud | E4E4: Customer hosted | E5E5: On premise |
|---|---|---|---|---|---|---|
| D0D0: Public | ||||||
| D1D1: Internal | ||||||
| D2D2: Confidential | ||||||
| D3D3: Restricted | ||||||
| D4D4: Critical |
Chat and the API share this matrix. API keys also carry a data-sensitivity default and a minimum provider environment. A request may raise sensitivity; it cannot lower the key floor.
Budgets
Optional ceilings, not included credit
Cap cost and/or tokens per member or dedicated API key, per provider. Empty means uncapped. Remaining 0 blocks further requests on that provider.
- Caps are optional and per member or dedicated API key × provider, not a single global limit, and not included inference
- Cost and token ceilings can be set independently; skip a field and that track stays uncapped
- Inherit keys follow the creator’s member budget; dedicated keys have their own ceilings
- Existing tools, custom software, and the built-in chat all share the same rules. Caps reset each UTC calendar month
Optional ceilings per member or dedicated API key, per provider. Empty fields stay uncapped. Remaining 0 blocks further requests on that provider.
Anthropic
Maya Chen
Member
Tom Weber
Member
Uncapped
billing-bot
Dedicated API key
Azure OpenAI
Maya Chen
Member
Uncapped
Uncapped
Preview only. Owners and admins configure live budgets in the app under Settings → Budgets. Caps reset each UTC calendar month.
Audit
Audit built in
Two complementary surfaces: AI inference activity and configuration changes.
Quick metadata shows provider, model, tokens, response time, and cost for this reply.
Per-message transparency
See which model you used, what it cost, and inspect metadata on every reply. Then follow the full trail in searchable activity history.
Click the info icon for quick metadata on a reply.
AI inference activity
- Per-message metadata in chat: provider, model, tokens, cost, latency, and quick routing context
- Dedicated activity page for searchable requests across chat and API with API-key attribution
- Step-level drill-down into model calls, timing, cost, and exportable history
- Unified usage and cost visibility for reconciliation and review
Every inference request, whether chat or API, appears in a searchable activity log with drill-down detail.
pde_8f2a…c91dQuick metadata is available per message in chat. Open Settings → Usage for the full activity page with filters, export, and step-level drill-down.
Filter by source, model, API key, or routing class. Inspect model steps, timing, cost, and the policy decision recorded for the request.
Every meaningful policy or provider change leaves a trace for review.
Blocked D3: Restricted → E1: Public managed after matrix update.
Default file sensitivity changed from D1: Internal to D2: Confidential.
Optional cost cap set for Alex on Azure OpenAI.
Azure OpenAI provider environment class updated to E3: Private cloud.
Routing, provider, budget, and membership audit logs are available in organisation settings.
Configuration audit
- Routing policy audit for matrix changes, blocked routing attempts, and sensitivity downgrades with reason
- Budget audit for cap edits on members and API keys
- Provider audit for creation, updates, deletion, and environment-class changes
- Administrative audit for membership and billing-sensitive organization changes
Govern AI without losing agility
Set policy once. Apply it to the tools you already run, the software you build, and the included chat.