Questions, answered.
01What is VAKYN?
Decision models that answer typed questions about your documents in milliseconds, say how sure they are, and keep getting better at your work.
You send a document and a question with fixed answers: yes or no, pick one, or a score. You get back the answer and how sure it is. Your code acts on the sure answers and sends the unsure ones to a person or a larger model.
02Who is it for?
Teams that make the same decision thousands of times a day.
- Every alert triaged the moment it arrives. The unsure ones go to an analyst.
- Every ticket routed as it comes in. The unsure ones go to a person.
- Every engineer on your bench, checked against every skill the client asked for.
03What can I use it for?
A decision you make again and again about a document, with fixed answers. Four examples:
- Phishing certificates: about 6 million new ones a day (public Certificate Transparency statistics). VAKYN checks every one as it appears.
- Support tickets: each one routed to the right team as it comes in.
- Staffing: every engineer checked against every skill the client asked for, evidence attached.
- Elections: every post tagged as it is published.
04Is it a chatbot?
No. It doesn't write text. It answers questions with fixed answers, each with how sure it is. AI that answers instead of talking.
05Why not just prompt an LLM?
Keep the LLM for the hard cases. VAKYN decides the everyday ones in milliseconds and sends only the unsure ones up, so you stop paying seconds and tokens for every case.
06How is this different from JSON mode or structured outputs?
The answer is typed by construction and comes with a probability for every option, in one pass. Nothing to parse, nothing to retry.
07Isn't this a classifier?
A classifier learns one fixed set of answers. VAKYN takes any question you write (yes or no, pick one, score) about any document.
08Why “Artificial Intuition”?
Intelligence thinks. Intuition decides. Psychologists describe two ways of thinking: System 1 is fast and automatic, System 2 is slow and careful (named by Keith Stanovich and Richard West, made widely known by Daniel Kahneman in Thinking, Fast and Slow).
VAKYN is System 1 for your systems: it decides the everyday cases in milliseconds and says how sure it is. The unsure ones go to System 2, an LLM or a person. The rest of what we believe is in Lore.
09Can it get things wrong?
Yes. That's why it says how sure it is, and why the VAKYN loop exists: we find the cases it gets wrong on your work and train exactly there.
Keep math and dates in code. Ask simple, literal questions. Send the unsure answers to a person.
10What is it trained on?
Decisions from 997 occupations in 22 domains and 41 kinds of work activity. With the VAKYN loop, we train it exactly where it fails on your work.
11What is VAKYN MAX?
VAKYN, hosted by us. You send API calls to https://api.vakyn.com with an API key, and each call is paid from your organization's balance. Nothing to install and no graphics card to rent.
12How do I start?
- Sign in with Google, GitHub or X. You get a personal organization.
- Top up the balance from the console's billing page.
- Create an API key and send your first call. The console's API page shows the request and the reply.
13We use a Jev-compatible API. Can we switch?
Yes. VAKYN MAX speaks the Jev-compatible API, so the official TypeSafe SDKs work after you change the base URL to https://api.vakyn.com and use a VAKYN API key. VAKYN is independent and not related to TypeSafe.
14Can my team share one account?
Yes, through organizations. Invite people to your organization from the console; they sign in with their own Google, GitHub or X account. API keys and the balance belong to the organization, and its owner manages members, keys and top-ups.
15Is there a request limit on VAKYN MAX?
We may apply rate limits to keep the service available for everyone. If you need a steady high volume, tell us at vali@vakyn.com. On your own servers there is no limit beyond your hardware.
Pricing and billing
One price per input token, paid from a prepaid balance. No subscriptions.
Ask us anything else16How much does it cost?
$0.0294 per million input tokens. Input tokens are the document and the questions you send. Output is free.
For example, one million calls on a 1,000-token document cost $29.40.
17How do I pay?
Each organization has a prepaid balance in US dollars. Top it up with a pack of $10, $50 or $200 through Stripe. Stripe adds tax where it applies and emails the receipt. Each call is charged from the balance as it is made.
We never see or store your full card number.
18Is there a first top-up bonus?
Yes. Your organization's first top-up is matched 100%, up to $100, as promotional credit. Top up $50 and your balance shows $100. There is no free balance before the first top-up.
19I have a code. How do I use it?
Redeem it on the console's billing page. The code adds its dollar amount to your organization's balance as promotional credit. Each code works once per organization, and a code can have an expiry date and a total number of uses.
20What is promotional credit, and which part of my balance is spent first?
Promotional credit is the first top-up match, codes, and credit we grant. It is spent before the balance you paid for. It has no cash value, cannot be transferred or refunded, and may expire on a date shown when you receive it.
The balance you paid for does not expire while your account is open.
21What happens when my balance runs out?
Calls are refused until you top up. Your balance never goes below zero, so you are never billed for an overdraft. When the balance drops below $1, we send a low-balance email.
22Is there a subscription or a minimum?
No. No monthly fee, no minimum spend, no plan to choose. You pay for input tokens from your balance.
23Can I get a refund?
Top-ups are not refundable, except where the law requires it or if we close your account without cause, in which case we refund the unused balance you paid for. The Terms of Service have the details.
24Do you store the documents I send?
No. VAKYN MAX processes each document and question to produce the answer, then discards them. We keep neither inputs nor answers after the call returns. The one exception: when a call fails, it can sit in an error log for at most 14 days while we find out why.
25Do you train on our data?
No. We do not use your inputs or answers to train models unless your organization agrees to it in writing.
26What do you keep?
For each call: the time, the key and organization, the number of input tokens, the charge, a request id and whether it succeeded. That is how we bill you and show your usage. The Privacy Policy lists everything we collect, who handles it and for how long.
27Our data can't leave our network. What then?
Run VAKYN on your own servers with open-server: your country, your rules, and nothing is sent to us.
28Is it open? Which licenses?
Yes. The model weights are Apache-2.0 and the server, open-server, is MIT. The license of each model file is on its card on Hugging Face.
29What hardware does it need?
One graphics card for the fastest results; the smallest model is built for cheap hardware. VAKYN v1 comes in three sizes: 0.8B, 2B and 4B. Install steps are in the open-server README.
30How fast is it?
VAKYN-4B on open-server, Q8_0, one RTX 5090:
- 16.6 ms median for one question on a short ticket (179 tokens).
- 29.4 ms median for three questions on the same ticket.
- 1.8 s median for eight questions on a 30,232-token document.
- About 215,000 documents an hour with 10 parallel clients, 0 errors.
31Is there a request limit on my own servers?
No. It scales as far as your hardware does.
32If the models are open, why pay?
The open models are the free start. You pay for VAKYN MAX if you don't want to run anything, or for the VAKYN loop: VAKYN trained on your work, on demand. For the loop, talk to us.