Криптоказак, ML/AI энтузиаст. Сторонник свободы, справедливости, честности и ответственности. Противник левых идеологий, неонацизма, толерантности, лжи.

Поиграйте в высоточку будущего, это атмосферно. =)
You live on floor 5000 of an endless megastructure. It's snowing. You want chips. The store is one floor up, over a bottomless drop. A tiny first-person game in one HTML file. EN / RU / 中文 claude.ai/artifact/SYh498jjW… artem-x-meta.github.io/etazh…
1
37
Мы тоже с нетерпением ждем появления весов Qwen3.8-Omni-Flash, чтобы что-то с ней создать. =)
🚀 Meet Qwen3.8-Omni-Flash, Qwen's first omni-modal model built around agentic capabilities! Native audio-video understanding, reasoning, and tool use come together in one model: understand the content, plan the task, execute with tools, and deliver the result. Highlights: 🥳 - Audio-video intelligence that gets things done: jointly reason over what's seen and heard, and orchestrate tools across long workflows to auto-edit vlogs, translate short videos, and turn movies into recaps. - A major leap: approaching Gemini 3.8 Flash in audio-video capabilities; +19.5 points on average in agent performance across WildClawBench-MM & UniClawBench. - 1M-token context with agentic perception: actively explore long videos and locate key moments with higher accuracy, using 51.8% fewer tokens than static understanding on OmniVideoBench. Video input costs are reduced by about 89% compared with Qwen3.5-Omni-Plus, making long-form audio-video understanding and agentic workflows more affordable than ever. To help you build apps around Omni, we're also open-sourcing Qwen-MM-Plugins and Qwen-Live Harness! 🛠️ We can't wait to see what you build with Qwen3.8-Omni-Flash! 👀 - Blog: qwen.ai/blog?id=qwen3.8-omni… - Qwencloud: qwencloud.com/models/qwen3.8… - Qwen Studio: chat.qwen.ai/ - API: alibabacloud.com/help/en/mod… - Qwen-MM-Plugins: github.com/QwenLM/Qwen-MM-Pl… - Qwen-Live Harness: coming soon github.com/QwenLM/Qwen-Live-…
42
Ну, программы в виде промптов уже придумали, фото в виде промптов придумали, надо и мне что-то придумать, ммм, эээ… О! Как вам голосовухи в виде текста?! Меньше места занимает, заебись. А когда надо восстановить голосовуху — просто читаешь вслух голосом отправителя.
Новая концепция kinda приложения из коробки: ставится не апка, а запускается промпт/последовательность промптов, создающих апку Ввиду случайностей и юзерских особенностей апка получается уникальной, но при этом сконнекченной с сетью остальных через кредсы заданные соответствующим образом Мнения?
43
Квантовал DeepSeek-V4-Flash-Vision-Exp так, чтобы не было никаких IQ-квантов и скорость была хорошая на старых видеокартах и процессорах. Вдохновлялся APEX @mudler_it Влазит в 128 оперативы и 16 видеопамяти с 256к контекста и виженом. huggingface.co/BahamutRU/Dee…
41
Bahamut retweeted
232
900
4,965
480,974
16
82
1,220
42,553
6
1
22
407
⚡️🇺🇦🇷🇺 Thierry Meyssan dit tout haut ce que les médias français martèlent à l’envers depuis quatre ans : « Jamais la Russie n’a envahi l’Ukraine. Vous ne diriez pas que la France a envahi le Rwanda. Pourquoi ? Parce qu’elle venait sauver des gens qu’on massacrait, dans le cadre d’une résolution du Conseil de sécurité. C’est exactement ce que la Russie a fait. » Le régime de Kiev n’a pas « attendu » 2022. Dès 2014, il a lancé la guerre au Donbass, bombardé des civils, coupé les retraites, arrêté de payer les fonctionnaires et annoncé qu’il continuerait jusqu’à ce que cette population soit écrasée. La Russie est intervenue pour empêcher ça.
292
3,175
6,750
86,847
Ресерч, который мы заслужили.
長いエロ動画から、オーガズムを含む10秒間を見つけて、ComfyUIの動画生成用参照素材として切り出す仕組みを、ここまでほぼQwen3.8 27Bと作ってきた。 手作業で大量の動画を確認するのは大変だし、発情してしまう。そこで、AIに動画を見せて「どこがオーガズムか」と聞いても、まだ正確には判断できない。 試しにとあるAVモデルさんのエロ動画の 音声・映像の動き・動きのフリーズのスコアを眺めてたら、 オーガズムの瞬間の特徴として次のような傾向があった A. 100〜1500Hzの喘ぎ声→その後の急減衰 B. フレーム間の映像変化量急増 C. freeze、コマ止め、ディゾルブなどのソフト遷移 試しにこれらから条件を決めたスクリプトを書かせて、オーガズム候補の位置を探したところ、結構ビクビクなってる感じのところが溜めれたからまあ良いんかなと。 この後は、それを自動化して、24fps・H3規定フレーム数へ整形して、そのままComfyUIへ投入できるようにする。 もう少しデータが集まったら、絶頂の瞬間とそうでない瞬間の数値データを学習させ、CatBoostくらいの軽量な機械学習で自動的にカットを切り出せるようにしたい。 ただ、オーガズム時の声や動きは人によって異なる。最終的にはAVモデルごとに学習モデルを作るか、個人差を補正する仕組みが必要になるとは思う。そこまではQwen3.8 27B でいけたら面白い。 グラフのGTが絶頂の瞬間。 A B C 全部で絶頂かと言われるとそうでもない。 この辺のパターンが人間だと良くわからない→分類器作ろ。 この発想は、投資の売買シグナルの選別によく似てると思う。
1
50
Оно вышло!.. =D
I spent 1351 hours of model time, ran 17555 generations, and burned 91.4M reasoning tokens. 300 benchmarked tasks. 9 versions of the exact same Qwen3.8 27B. Around $100 on Blackwell GPU rental. Just one single question: which quant actually degrades quality? The results will surprise you. 🧵Mega-thread, read till the end, the takeaways are right there
1
25
Bahamut retweeted
The next-gen architecture powering Qwen4 is now here! ✨ Get ready for the open release of Qwen3.8-Flash-Next 🚀 The countdown starts now! ⏳🔥modelscope.cn/models/Qwen/Qw…
94
313
2,336
1,078,825
Вот это реально интересный тест. Сранвение Q3_K_XL с Q4_K_XL с Q6_K. Сколько теряет тройка, какая разница между четверкой и шестеркой. Ждем с нетерпением.
Right now, I'm launching an experiment for 200 inference hours of Qwen3.8 27B. I will run 300 tasks in low, medium, and xhigh modes on FP8, AWQ, UD-Q6_K_M, UD-Q4_K_XL, UD-Q3_K_XL, UD-Q2_K_XL, and UD-IQ1_M. Then I'll compare each version against the BF16 baseline. Wish me luck. And may God help me! 🙏
27
Bahamut retweeted
Блять, это самое крутое из всего, что я рендерил на CPU, ТЫ ПОСМОТРИ Я делаю игровой движок на лучах для рендеринга вокселей, SDF, гауссин и многого другого Последнее время я отлаживал его и производил многочисленные оптимизации на CPU, строя при этом разные сцены (текущая рендерится всего в 120 на 120 пикселей) Конечно, на видеокарте получится добиться куда более крутых результатов, но мне было важно оптимизировать основной алгоритм трассировки луча Полагаю, что это будет последний пост с рендером на CPU. Почитать подробнее о том, что я делаю можно будет в постах, ссылки на которые я оставлю в комментарии Поддержи лайком/ретвитом, мне будет приятно и я буду продолжать делать все тож самое, но уже на видео карте
21
41
758
30,926
Bahamut retweeted
DeepSeek-V4-Flash-Vision-Exp is now live on the DeepSeek API Platform! 🚀 🔹 This experimental multimodal model matches DeepSeek-V4-Flash on text capabilities—including agents, reasoning, and world knowledge. 🔹 On multimodal agent benchmarks, V4-Flash-Vision-Exp makes a major leap over V4-Flash, bringing multimodal agent performance close to Opus-4.8. Try it with model='deepseek-v4-flash-vision-exp'. DeepSeek Harness 0.1.1 was released today with out-of-the-box support for the new model. 1/n
558
1,210
11,505
2,762,646
А самое ужасное, что большинство сидит на ollama и lmstudio… Даже не зная, что есть такие вещи, как движки инференса и модели.
33
Такое ощущение, что я попал на фестиваль тупых. В X все пишут, что есть разные кванты, открывают для себя разные движки инференса, узнают, что квантовать можно контекст, что low и xhigh — разные, и все это влияет на результат. Пиздец, это же база для нубов, все всегда это знали.
20
Уставший я заканчиваю сессию с ZCode. @Zai_org такие: 1 reset avaliable! AAAA!... =D
30
Вот как это делалось. Что ж, время пробовать дипсик харнесс.
Nobody ever shows you the log. Here it is: 687 steps of a local Qwen3.8 27B building a first person shooter in DeepSeek Harness, scrolled end to end. It takes a while. 87.9M tokens went in, 822k came out. 5 hours 11 minutes of model time, 58 minutes of tool calls. All of it on two used 3090s in my apartment, none of it on anyone elses hardware. Watch for the spots where I stop typing descriptions and just paste pictures. A screenshot of the gun I wanted in my hands. A screenshot of the crosshair I wanted on screen. The model looked at both and built what I meant. Thats a 27B doing art direction with its eyes, on my own cards. The rest of the log is not pretty and thats the part I want you to see. It writes a file, runs it, reads its own error, goes back and fixes it. Six hours of that while I slept. Everyone posts the finished demo, almost nobody posts the conversation behind it, and the conversation is where you find out whether a model holds a task or just talks about it. This one held it for 687 steps without me stepping in. Local models are past the demo stage. Watch the scroll and tell me you saw this coming 18 months ago.
1
30
Если ето правда, ето вау.
Qwen3.8 27B wrote a playable first person shooter start to finish on 2 used 3090s in my apartment. No cloud. No API key. No subscription. 687 steps. 87.9M tokens in, 822k out. 5 hours 11 minutes of model time plus 58 minutes of tool calls. The agent loop never broke once. And it plays. Enemies spawn and push you, the gun kicks, shadows stretch across the whole block while the sun goes down behind the towers. I sat there clearing waves instead of grading the output. Same shooter prompt I threw at the frontier models a few weeks ago. That time the tokens went to somebody elses datacenter. This time nothing left the flat. 87.9 million tokens through my own cards. On an API that run has a price tag. Here it has an electricity bill. 60 tok/s all the way through. Slower than frontier, and it stops mattering when the thing works through the night while you sleep. Local models were a toy 18 months ago. This one finished a game.
1
113
Протестировал Qwen3.8-27B в аудите с GLM-5.3, GLM-5.2, DeepSeek-V4-Flash-0731 и Muse Glimmer. Это не GLM-5.3 и не 5.2. Это не DeepSeek-0731. Проблемы с контекстом, но внимание к мелочам. Два false positive high-risk, поленился перепроверить. Но больше всего найденных low-risk.
1
82
В целом, не было шансов, чтобы 27B Dense догнала 284B MoE, которая неделю назад догнала топовую (на тот момент) 744B. Это фантастика. Но по ощущеним, если сравнивать с современными 30B моделями (Glimmer, Gemma), новый Qwen3.8-27B ощущается как 124B Dense. А это уже прорыв.
1
52
@QwenDevs great work, thx, you're awesome! Очень крутые, спасибо за топовую модель. =)
5