gonzo-обзоры ML статей — Telegram channel logo

@gonzo_ml

gonzo-обзоры ML статей

Telegram channel @gonzo_ml: 24.5K subscribers, 3K views per post, score 43

Artificial intelligence
43DAhead of 1 in 10 channels in its category
Category midrange 46–63This channel 43
24.5K
Subscribers
3K
Median views over 30 days
12.3%
Views / subscribers over 30 days
87
Posts over 30 days

Data as of October 1, 2026

Breakdowns of fresh machine learning papers for people who want to keep up with research without reading every preprint cover to cover. The author summarizes work on generalization theory, evolutionary algorithms and Gödel Machines, adding personal commentary and links to the original papers. As of 30 September: subscribers 24.5K, median post views 3K, engagement rate 12.1%. Subscriber count sits well above the category median, and the account posted more often over 30 days than most peers in the category, though engagement lags behind the category median. Recent material includes an analysis of generalization dynamics in neural networks trained with gradient descent, a rundown of the Red Queen Gödel Machine, and a mention of a neuroevolution book co-authored by Sakana AI founder David Ha. The author also compiled a roundup of recent Jev news from multiple sources, splitting it across several posts.
catalog description About the channel, by its author
Авторы: Гриша Сапунов, ранее руководитель разработки Яндекс-Новостей, ныне CTO Intento. Области интересов: AI/ML/DL, биоинформатика. Лёша Тихонов, ранее аналитик в Яндексе, автор Автопоэта, Нейронной Обороны... Области интересов: discrete domain, NLP, RL.

Overview

Written automatically from the channel's data, updated 3 September 2026. The numbers in this text are as of that date; the fresh ones are in the tiles above.

@gonzo_ml is a channel that reviews recent machine learning papers — from the inner workings of optimizers and gauge symmetries to representation learning architectures. The write-ups are detailed, linking to arXiv, code repositories and outside reviews, occasionally drifting into physics books or science fiction references. It also mentions related projects from the same authors, such as a new channel dedicated to World Models.

With 90 posts over the last 30 days, the channel publishes far more often than the category median of 20 posts for Artificial Intelligence. The median post gets 2,581 views, and engagement sits at 10.6%, below the category median of 25.1%. Over the 18-day observation window, subscribers grew by 15 and average views rose by 386 — a modest but visible move.

This fits readers who want deep dives into specific papers, formulas and code rather than a news feed — a niche format for people already into the subject. With 24,332 subscribers and a Place Score of 32, the channel trails the category's subscriber median of 15,164 and its engagement median, but leads on posting frequency.

The owner hasn't verified this listing, and no ad pricing is published yet — reaching out happens directly on Telegram.

Common questions

What is the @gonzo_ml channel about?
The channel reviews machine learning research papers, covering topics from optimizers to representation learning, with links to arXiv and code. It occasionally touches on physics books and science fiction.
Can I buy advertising on @gonzo_ml?
The owner hasn't published ad pricing yet. You can reach out directly on Telegram to arrange a placement.
How many subscribers?
Subscribers: 24.5K. Median views per post: 3K. Views per subscriber: 12.3%. Measured on October 1, 2026.
How often are posts published?
Posts in the last 30 days: 87 — that is several times a day. Measured on October 1, 2026.
Does this channel have a Telegram tick?
No Telegram tick.
Who runs this channel page in the catalog?
Nobody yet. If this is your channel, claim the page: you will be able to reply to reviews and see its stats.
Is this your channel?

Claim it — accurate metrics, replies to reviews as the channel, and an owner badge on the page.

Verify ownership

Takes a minute: a post with a code, or the bot as an admin
Your page on tg.place

Know the owner?

Forward them this message — they can verify their rights and collect what has piled up: page statistics and replies to reviews.

Subscribers+181 in 46 days
24,49824,303
Oct 124,498+21 in a day

Hover the chart or swipe it — we show the day.

What it is made of
Engagement6.7 of 30
Growth quality16.7 of 20
Reactions and forwards5.7 of 15
Consistency3.6 of 12
Trust3.9 of 8
Reviewsnot enough datano reviews yet

Score 43 — from 5 of 6 signals: the rest are not measured yet. Methodology

Latest posts

Posts are in Russian — that is what the channel publishes.

  • 1.8K10 forwards24 reactions15 commentsOpen in Telegram

    Poll

  • 2.2K4 forwards10 reactionsOpen in Telegram
    Post by “gonzo-обзоры ML статей” from September 30, 2026
  • 2.2K5 forwardsOpen in Telegram
    Post by “gonzo-обзоры ML статей” from September 30, 2026
  • 2.2K56 forwards14 reactionsOpen in Telegram
    Снова про теорию генерализации и гроккинга. Много математики для желающих :) A Theoretical Analysis of Generalization Dynamics in Neural Networks under Gradient Descent with Weight Decay Yuqing Wang, Ioannis G. Kevrekidis, Mikhail Belkin Paper: arxiv.org/abs/2609.07755 Review:
    arxiviq.substack.com/…https://arxiviq.substack.com/p/a-theoretical-analysis-of-generalization
    ЧТО сделали: Авторы построили строгий теоретический фреймворк для анализа динамики обобщения глубоких нейросетей при оптимизации градиентным спуском с weight decay на квадратичном лоссе. Разбив входное пространство на ячейки вокруг обучающих точек, исследователи разложили популяционный риск на три слагаемых: ошибку данных, ошибку оптимизации и ошибку вариации предсказаний. В предположениях диссипативности и локальной приближённой однородности доказано, что weight decay сжимает внутренние представления, гарантируя эмпирическую сходимость и задавая необходимые и достаточные условия для послойной сходимости и эффекта гроккинга (grokking). ПОЧЕМУ это важно: Классическая теория статистического обучения опирается на статические границы равномерной сходимости, которые бессильны перед перепараметризованными сетями, интерполирующими шум или демонстрирующими отложенное обобщение. Предложенный подход выходит далеко за рамки линеаризованного режима нейро-касательного ядра (NTK) и игрушечных постановок. Он объединяет геометрию данных, глубину сети и алгоритмическую регуляризацию в общую динамическую теорию, объясняя, почему глубокие слои обобщают позже и как weight decay управляет временным зазором между запоминанием и генерализацией. Для практиков: На практике глубокие сети часто ведут себя контринтуитивно: тренировочный лосс падает практически в ноль почти мгновенно, но тестовая точность выходит на плато и лишь спустя тысячи дополнительных шагов резко взлетает — это и есть гроккинг. Работа математически объясняет этот феномен: подгонка под обучающую выборку и генерализация управляются двумя разными физическими процессами, идущими на разных временных масштабах. Запоминание обучающих точек происходит быстро вдоль координат данных, тогда как для истинного обобщения требуется время, чтобы weight decay успел подавить неконтролируемые осцилляции функции во всём остальном объёме входного пространства. Описав, как это сжатие послойно распространяется по сети, теория даёт конкретные критерии баланса между покрытием датасета, силой weight decay и длительностью обучения. Подводить базу здесь: t.me/gonzo_ML_podcasts/4814
  • 2.7K10 forwards55 reactions3 commentsOpen in Telegram
    Скажите вообще, насколько такой формат про Jev был полезен? Что хорошо, что плохо, что надо улучшить или сделать по-другому? Есть ещё какие-то темы, которые стоило бы прогнать так же?
  • 3.3K115 forwards35 reactionsOpen in Telegram
    Вдогонку про Jev, вот чуваки начали собирать каталог применений и прочего:
    github.com/…https://github.com/AnotiaWang/awesome-jev
  • 3.1K5 forwards4 reactionsOpen in Telegram
    Post by “gonzo-обзоры ML статей” from September 28, 2026
  • 2.9K4 forwardsOpen in Telegram
    Post by “gonzo-обзоры ML статей” from September 28, 2026
  • 3K30 forwards11 reactions1 commentOpen in Telegram
    Про Darwin Gödel Machine писали, про Huxley-Gödel Machine писали, а про Red Queen Gödel Machine не писали! Исправляемся. The Red Queen Gödel Machine: Co-Evolving Agents and Their Evaluators Alex Iacob, Andrej Jovanović, William F. Shen, Daniel Burkhardt, Meghdad Kurmanji, Nurbek Tastan, Lorenzo Sani, Niccolò Alberto Elia Venanzi, Ambroise Odonnat, Zeyu Cao, Bill Marino, Xinchi Qiu, Nicholas D. Lane Paper: arxiv.org/abs/2606.26294 Review:
    arxiviq.substack.com/…https://arxiviq.substack.com/p/the-red-queen-godel-machine-co-evolving
    Code: N/A Model: N/A ЧТО сделали: Предложили эволюционный фреймворк, в котором обучаемые эвалюаторы развиваются параллельно с целевыми агентами. Это позволяет обойтись без статичных бенчмарков при рекурсивном самосовершенствовании. ЧТО получили: Добавление эволюционирующего ревьюера подняло pass rate на бенчмарке Polyglot с 69.9% до 71.7% при снижении расхода токенов в 1.35–1.72 раза по сравнению с Huxley-Gödel Machine. При написании научных статей авторы получили долю одобрений 38.8%–40.5% (в 1.78–1.86 раза выше бейзлайна с его 21.8%) по оценке нескольких рецензентов. Точность проверки доказательств достигла 76% (+9% к бейзлайну) при трёхкратной экономии токенов. ПОЧЕМУ это важно: Совместная эволюция судей и генераторов формирует динамический curriculum и отсекает предвзятость моделей к собственным текстам без ручной поддержки бенчмарков. При этом теоретические гарантии поиска действуют только внутри отдельных эпох и опираются на неизменные якорные датасеты. Бегать вместе с королевой здесь: t.me/gonzo_ML_podcasts/4813
  • 3.4K72 forwards22 reactionsOpen in Telegram
    Крутая команда авторов, включая фаундера Саканы Дэвида Ха, а также Себастиана Риси, выпустила книгу по нейроэволюции neuroevolutionbook.com

The channel in numbers

Created
February 21, 2019
Photos
4K
Videos
6
Files
3
Links
1.8K

Telegram data as of October 1, 2026

Reviews

Leave a review

No reviews yet. Yours would be the first.

Similar

More on this topic