Renkler tuhaf mı görünüyor?
Samsung Internet tarayıcısı koyu modda site renklerini değiştiriyor olabilir.
Kapatmak için Internet menüsünden
Ayarlar →
Kullanışlı Özellikler →
Labs →
Web site koyu temasını kullan
seçeneğini etkinleştirebilirsiniz.
A single OpenAI-compatible endpoint. The difficulty of every request is classified and the request is routed to the most cost-appropriate model that can do the job; requests are stopped before team budgets are exceeded, and every request is logged.
AI costs slip out of control without anyone noticing
Misused models and scattered integrations leave teams unable to see what they spend, and the budget slips out of control.
Every job on the most expensive model
A simple question and a complex analysis usually go to the same expensive model.
Everything separate per provider
Every provider needs its own API key, its own integration and its own invoice.
Spend is noticed late
Teams cannot see what they are spending in real time; budget overruns show up when the invoice arrives.
Changing models becomes a project
Trying a new model or provider means re-integrating the application.
Four steps from request to the right model
Your application sends requests to one endpoint; RuneGate decides which model to use separately for every request.
Request arrivesYour application sends the OpenAI-compatible request to one endpoint; the team's API key and budget are checked.
ClassificationThe request's difficulty and task type are determined automatically: a simple question or a complex analysis; does it carry an image or a document.
RoutingThe request goes to the best-fitting model in the light, mid or power tier, within the tiers allowed for the team. If a provider has a problem, the request is completed seamlessly with another model.
LoggingWhich request went to which model, its cost and its duration are logged; spending per team is tracked in the panel.
Integration: your code stays, only the address changes.
Every application that uses an OpenAI client, n8n flows and scripts switch to RuneGate by changing only the base_url and the API key. Use rune-auto as the model name and let RuneGate choose; pin the tier or the model if you prefer.
# OpenAI client: two lines change
client = OpenAI(
base_url="https://gate.sirket.com/v1",
api_key="sk-ekip-anahtari",
)
# rune-auto: RuneGate makes the choice
r = client.chat.completions.create(
model="rune-auto",
messages=[{"role": "user",
"content": "Summarize the contract."}],
)
One endpoint. Smart routing.
Automatic routing
The right model is chosen automatically by the difficulty of the request: simple jobs go to an economical model, hard jobs to a powerful one. You can pin the tier or the model yourself.
One API, many providers
OpenAI and Google Gemini models, plus open-source models such as DeepSeek, Qwen and GLM, are used from one endpoint with one team API key.
Images, PDFs and audio
Images, PDFs, Excel sheets and audio files are sent through the same endpoint; the request is routed to a model that can understand them.
Per-team budget control
A monthly budget for every team; when it runs out, requests are stopped before they reach the provider. Allowed tiers and a request limit are set per team, and spending is visible in real time.
Auditable routing
Which request went to which model, and why, can be traced in the logs; the panel shows daily spending, model distribution and team usage.
In the cloud or on-premise
Used as a managed service; for organizations whose data cannot leave, it is installed on-premise on RuneBox and routes requests to local models.
RuneGuard add-on: sensitive data is protected before it reaches the model.
RuneGuard is an optional data protection module that runs inside RuneGate. It replaces sensitive information in a request with placeholders before it reaches the model provider, then puts the right value back in place of each placeholder in the model's answer. The user gets a complete answer, and the sensitive data never reaches the model.
The data stays out, the answer comes back whole
On the way out, sensitive information becomes placeholders; on the way back, each placeholder is replaced with the real value.
Beyond personal data
Besides personal data such as national ID numbers, phone numbers and email addresses, sensitive information specific to your organization or field can be defined too; it is masked, or the request is blocked entirely.
Same endpoint, same code
RuneGuard runs as part of RuneGate; your applications and your integration stay exactly as they are.
Added when you need it
Offered as a separate module; switched on for teams that work with sensitive data, and it makes your KVKK assessments easier.
RuneGuardExample flow
Request from your application
Summarize the application: national ID 12345678901, phone 0532 000 00 00, email ayse@ornek.com; the application is part of Project Atlas.
RuneGuard masks
To the model provider
Summarize the application: national ID [NATIONAL_ID], phone [PHONE], email [EMAIL]; the application is part of [PROJECT].
Answer from the model
Summary: the application under [PROJECT] belongs to the customer with national ID [NATIONAL_ID]; contact via [PHONE] and [EMAIL].
RuneGuard puts the real values back
Answer returned to your application
Summary: the application under Project Atlas belongs to the customer with national ID 12345678901; contact via 0532 000 00 00 and ayse@ornek.com.
Control in your hands, cost under control.
The expensive model only when needed
Simple requests stay on the economical model; the savings depend on the mix of your traffic.
No surprise invoices
Requests stop before a team budget is exceeded; who is spending what is visible at any moment.
One integration
All providers are managed from one place; adding a new model needs no change in the application.
Built on field experience
Designed with the experience RuneLab gained in enterprise AI projects; routing decisions can be audited from the logs.
Teams that use more than one model
For every team that puts AI into its products and workflows and wants to manage the cost.
Software teams
Managing several products and models through one integration; so that changing a model needs no deployment.
Automation teams
One key, automatic model selection and a budget cap for flows built with n8n and similar tools.
IT and finance managers
Per-team budgets, real-time spend and single-invoice visibility; central management of provider API keys.
Data and AI teams
Trying new models without changing code; tracing which request went to which model in the logs.
Value, security and control: three products, one team
An LLM gateway is a single entry point that sits between your applications and language model providers. Every request passes through it, so model choice, team budgets and logs are managed in one place. RuneGate does this with one OpenAI-compatible endpoint; the only changes in your applications are the address and the API key.
Do we need to change our existing code?
No. RuneGate is OpenAI-compatible: only the address and the API key change in the client; your current usage, streaming included, keeps working as it is.
Can we choose the model ourselves?
Yes. rune-auto leaves the choice to RuneGate; if you prefer, you pin the tier or the model itself. The tiers allowed for each team are set in the admin panel.
What happens when a provider has an outage?
The request is completed automatically with another model; your application keeps calling the same endpoint and nothing needs to change.
When the budget is exceeded, does the request stop or do we get a warning?
It stops. When the team's budget runs out, the request is stopped before it reaches the provider and the application receives a budget-exceeded response. Budgets renew monthly; spending is tracked in real time in the panel.
Which providers and models are supported?
OpenAI and Google Gemini models, plus open-source models such as DeepSeek, Qwen and GLM. A new model is added without any change to your applications.
Who can see our request logs, and how long are they kept?
Logs are visible only to your organization's authorized panel users; in an on-premise installation they stay on your network. The retention period and deletion policy are agreed together with your organization.
Does personal data reach the model provider?
With the RuneGuard module, personal data and sensitive information specific to your organization are replaced with placeholders before the request is sent to the provider, and the real values are put back when the answer returns. If data must never leave your organization, RuneGate runs with local models on RuneBox.
Can it be hosted on premises?
Yes. RuneGate is installed on-premise on RuneBox and routes requests to the local models on the box; applications use the same OpenAI-compatible endpoint. It is also offered as a managed cloud service.
How much does it save?
Savings depend on the mix of your traffic: the larger the share of simple requests, the bigger the difference. In the demo we go through routing and cost logs together using your own sample requests.
How is pricing set?
We prepare a quote for your organization based on the number of teams, the monthly request volume and the providers to be used. In the demo we can look at routing and cost records with your own sample requests.
Request a demo for your team
Tell us the models you use, your monthly request volume and your team structure; we will get back to you within one business day.
With your consent, we use analytics and ad-measurement cookies to improve the site and measure how our ads perform. Strictly necessary cookies are always on. Details are in our Privacy Policy.