AIgate
670 subscribers
19 photos
20 links
➡️
🌐 store 🌐
https://aigate.shop
✉️support✉️
@aigate_support
Download Telegram
Добавили тиры скорости для Gemini моделей

Теперь можно выбрать режим под свои задачи:

Flex - в 2 раза дешевле, но заметно медленнее
Priority - в 2 раза дороже, но значительно быстрее

Качество ответов абсолютно одинаковое во всех трёх случаях. Меняется только ваше место в очереди к серверам Google - либо вы спереди, либо сзади. Чем выше нагрузка на серверы, тем сильнее это ощущается.
Реальная скорость зависит в том числе от текущей нагрузки на серверы, так что Priority не гарантирует молниеносный ответ в любой ситуации, но в среднем заметно шустрее.

=====================================

We’ve added speed tiers for Gemini models

Now you can choose the mode that best suits your needs:

Flex - half the price, but noticeably slower
Priority - twice the price, but significantly faster

The quality of the responses is exactly the same in all three cases. The only thing that changes is your position in the queue to Google’s servers you’re either at the front or the back. The higher the server load, the more noticeable the difference is.
Actual speed also depends on the current server load, so Priority doesn’t guarantee a lightning-fast response in every situation, but on average, it’s noticeably faster.
Please open Telegram to view this post
VIEW IN TELEGRAM
10
💰Важное обновление по ценам на Google модели с картинками

Оказывается, у Google моделей image tokens и output tokens тарифицируются по-разному. Мы по своей глупости всё это время считали image tokens как output - по сути переплачивали за каждый запрос из своего кармана.

Исправляем - в ближайшее время добавим отдельную плату за image tokens, как это и должно быть. Из-за этого запросы с картинками станут намного дороже. Нам жаль за предоставленные неудобства. Цены появятся на витрине моделей.

==================================

💰Important Update on Pricing for Google Models with Images

It turns out that in Google models, image tokens and output tokens are billed differently. Due to our own oversight, we’ve been treating image tokens as output tokens all this time—essentially, we’ve been overpaying for every request containing an image out of our own pocket.

We’re fixing this we’ll soon be adding a separate charge for image tokens, as it should be. Because of this, requests with images will become slightly more expensive. We apologize for any inconvenience this may cause. The prices will be listed on the model dashboard.
Please open Telegram to view this post
VIEW IN TELEGRAM
👎2😭2
⚠️Провайдер платежки 2328.io сейчас под DDoS, поэтому крипто-пополнения могут временно не создаваться или падать с ошибкой.

Это на стороне платежного провайдера, AIGate работает штатно. как только 2328 стабилизируется, крипто-оплата снова заработает

========================

⚠️The payment provider 2328.io is currently under a DDoS attack, so cryptocurrency deposits may temporarily fail to process or result in an error.

The issue lies with the payment provider; AIGate is operating normally. As soon as the situation with 2328 stabilizes, cryptocurrency payments will resume.
Please open Telegram to view this post
VIEW IN TELEGRAM
⚠️Большинство моделей не работает, уже в процессе починки.

⚠️Most of the models aren't working; they're already being repaired.
Please open Telegram to view this post
VIEW IN TELEGRAM
1👎1
AIgate
⚠️Большинство моделей не работает, уже в процессе починки. ⚠️Most of the models aren't working; they're already being repaired.
Немного затянулось, модели восстановлены.

It took a little longer than expected, but the models have been restored.
1
z-ai/glm-5.2-fast - удалён. Теперь обычный z-ai/glm-5.2 имеет такую же скорость за цену ниже.

===================================================

z-ai/glm-5.2-fast - removed. Now the standard z-ai/glm-5.2 offers the same performance at a lower price.
2👎1🤩1
📢New model:
google/gemini-3.1-flash-lite-image - $0.1/M input tokens | $0.36/M output tokens | $4.8/M image out tokens
Please open Telegram to view this post
VIEW IN TELEGRAM
5👎2
📢New model:
anthropic/claude-sonnet-5 - $0.5/M input | $1.5/M output
Please open Telegram to view this post
VIEW IN TELEGRAM
6👎2
🤖Привет. возможно вы замечали, что на anthropic моделях кэш работал странно: где-то почти не было хитов, где-то каждый запрос заново писал кэш, а длинные чаты всё равно могли стоить дороже, чем должны.

Мы решили это исправить и сейчас тестируем своё умное кэширование для anthropic моделей.
сейчас оно включено не на всех моделях, а только на: Claude Sonnet 5, Claude Opus 4.7, Claude Opus 4.8

Первые результаты очень хорошие: на длинных диалогах и повторяющемся контексте cache hit уже доходит до 95%+. это значит, что при продолжении больших чатов значительная часть контекста читается из кэша, а не оплачивается заново.
мы продолжаем следить за логами, расходом и стабильностью. если тесты покажут себя хорошо на дистанции, включим это кэширование для всех anthropic моделей на постоянной основе.

В идеальных сценариях это может экономить до 95% стоимости входных токенов на длинных повторяющихся контекстах.
пока система активно тестируется, поэтому если вы заметите странное поведение кэша, необычные списания или что-то подозрительное в длинных чатах - напишите нам в поддержку.

===========================================

Hi. You may have noticed that the cache behaved strangely with Anthropic models: in some cases there were almost no hits, in others every request rewrote the cache, and long chats could still end up costing more than they should.

We decided to fix this and are currently testing our smart caching for Anthropic models.
It’s not enabled on all models yet, but only on: Claude Sonnet 5, Claude Opus 4.7, and Claude Opus 4.8

The initial results are very good: for long conversations and repetitive context, the cache hit rate is already reaching 95%+. This means that when continuing long chats, a significant portion of the context is read from the cache rather than being paid for again.
We’re continuing to monitor logs, usage, and stability. If the tests hold up well over time, we’ll enable this caching for all Anthropic models on a permanent basis.

In ideal scenarios, this could save up to 95% of the cost of input tokens for long, repetitive contexts.
The system is currently undergoing active testing, so if you notice strange cache behavior, unusual debits, or anything suspicious in long chats, please contact our support team. We’ll review the logs and fix the issue quickly.
Please open Telegram to view this post
VIEW IN TELEGRAM
9👎2🤩1
AIgate
🤖Привет. возможно вы замечали, что на anthropic моделях кэш работал странно: где-то почти не было хитов, где-то каждый запрос заново писал кэш, а длинные чаты всё равно могли стоить дороже, чем должны. Мы решили это исправить и сейчас тестируем своё умное…
Мы включили умное кэширование на все модели anthropic. Пока что всё выглядит хорошо, и кэш работает лучше чем до этого. Мы продожаем следить за ситуацией и улучшать кэширование каждый день. Для тех кому не нравится наше умное кэширование, вы можете передать в теле запроса параметр "aigate_anthropic_cache": false и наш кэш отключается. Пример:

{
"model": "anthropic/claude-sonnet-5",
"aigate_anthropic_cache": false,
"messages": [
{ "role": "user", "content": "hello" }
]
}


=================================

We've enabled smart caching on all Anthropic models. So far, everything looks good, and the cache is performing better than before. We’re continuing to monitor the situation and improve caching every day. For those who don’t like our smart caching, you can pass the parameter “aigate_anthropic_cache”: false in the request body to disable our cache. Example:

{
"model": "anthropic/claude-sonnet-5",
"aigate_anthropic_cache": false,
"messages": [
{ "role": "user", "content": "hello" }
]
}
3👎1
📢New model:
baai/bge-m3-embedding - FREE
Please open Telegram to view this post
VIEW IN TELEGRAM
1
Promocode valid for 20 uses
dfc08f4189634df78859e250507d485c

p.s used 🙊
🔥6
Все модели Claude временно недоступны. Мы уже приступили к восстановлению работоспособности моделей.

==============================================

All Claude models are temporarily unavailable. We have already begun working to restore them to full functionality.
🌭3
Добавили оплату картами РФ.
Теперь пополнить баланс можно не только через СБП и крипту, но и обычной российской банковской картой. Оплата проходит через Platega, комиссия для карт РФ - 8.5%.
Международная оплата пока ещё в процессе подключения. Планируем добавить их до конца этого месяца.

==============================================

We’ve added support for Russian bank cards.
Now you can top up your balance not only via SBP and cryptocurrency, but also with a regular Russian bank card. Payments are processed through Platega, and the fee for Russian bank cards is 8.5%.
International payments are still being set up. We plan to add them by the end of this month.
🔥2
New model:
x-ai/grok-4.20 - $0.3/M input tokens | $0.6/M output tokens
🔥5
Тестируем нового платежного провайдера Lava.

Теперь он будет основным способом оплаты для СБП и карт РФ. Комиссии стали ниже:
СБП: 3%
Карты РФ: 6%

Platega остается как запасной провайдер. У нее комиссии выше, но она будет доступна на случай, если Lava временно не подходит или не работает.
Если при оплате через Lava вам не начисляется баланс, свяжитесь с поддержкой - @aigate_support

============================

We’re testing a new payment provider, Lava.

It will now be the primary payment method for SBP and Russian cards. The fees have been reduced:
SBP: 3%
Russian cards: 6%

Platega remains as a backup provider. Its fees are higher, but it will be available in case Lava is temporarily unavailable or not working.
If your balance isn’t credited when paying via Lava, please contact support at @aigate_support
🔥5
200 followers ♥️

Promo code valid for 25 uses
29fc974c0b90426e88f9cf0f3a072aa7 used
Please open Telegram to view this post
VIEW IN TELEGRAM
🔥10