[UTC+3]
All articles

AI Economics

What adoption and operations cost: request pricing, inference, project budgets and ways to avoid overpaying. Numbers you can lean on when planning a budget.

Economics and operationsIn plain words

What is latency and how to keep it within seconds

Latency is the time between a request and the model's answer: it is broken into its parts, each part is measured, and the user's wait is brought down to seconds.

Read August 20, 2026 · 5 min
Economics and operationsIn plain words

Context bloat: what it means in plain terms

Context bloat is the growth in the volume of text sent to the model with every request, without any growth in the value that text delivers.

Read August 14, 2026 · 5 min
Economics and operationsIn plain words

What is context compaction and why long sessions need it

Context compaction replaces a bloated conversation history with a short summary: the work continues, but the input volume of each turn stops growing.

Read August 11, 2026 · 4 min
AgentsEconomics and operations

AI agent in six weeks: where to start and when it pays off

Breaking a process down step by step delivers a first working version in six weeks. Four conditions for an agent to pay off, and the checks that confirm them in two days.

Read July 26, 2026 · 14 min
Economics and operationsIn plain words

The Real Cost of AI Adoption: Nine Budget Lines and Three Hidden Ones

The range from RUB 500,000 to several million comes down to nine cost lines. Here the budget is taken apart piece by piece, including the three lines that rarely appear in a vendor proposal.

Read July 26, 2026 · 15 min
Economics and operationsIn plain words

What Is a Token and Why You Pay for It

A token is a chunk of text roughly three quarters of a word long: the model reads a request in tokens, and the provider bills for how many there are.

Read July 26, 2026 · 6 min

Run the numbers