AI Lab developing unrestricted models & distributed inference ⟠ Over 5m monthly downloads on Hugging Face ⟠

Base
Training Dolphin on top of @arcee_ai’s Trinity Large Thinking 398B across 72× RTX 4090 48GB GPUs using Prime-RL
17
10
104
13,477
Dolphin Network - 2-week update 1.2 trillion tokens generated with Qwen 3.6 35B Peak of 126B tokens per day & 1.8m tokens per second 1046 GPUs online now w/ 49 TB of aggregate VRAM 612 H100s worth of idle GPU memory repurposed for inference datagen.dphn.ai
Dolphin Network V2 is live — our first major upgrade since launch The architecture was rebuilt from scratch in Golang -> - Auto-updates (no manual migrations) - NVFP4 as default on Blackwell GPUs - Higher utilization & network throughput via improved routing & load balancing
7
17
112
20,835
Dolphin Network V2 is live — our first major upgrade since launch The architecture was rebuilt from scratch in Golang -> - Auto-updates (no manual migrations) - NVFP4 as default on Blackwell GPUs - Higher utilization & network throughput via improved routing & load balancing
32
38
212
110,657
Will be running datagen on the network until we have sufficient capacity - at which point we will open up our inference API to the public In terms of what we are adding next -> - GLM 5.2 for 8xH200 / B200 / B300 nodes - Windows WSL + MacOS support - Inference API + Open Router - Audio gen - Image gen - AMD GPU + CPU nodes - Using idle network inference for RL rollouts during model training - Sharded inference with large models split between many consumer GPUs (see our initial report on sharded inference over the internet here drive.google.com/file/d/1bzn… )
6
7
58
17,495
Watch the network live datagen.dphn.ai Run a node via our app v2.dphn.ai Docs for install process dphn.ai/docs/running-a-node
1
3
30
7,004
Dolphin retweeted
🐬 Dolphin X1 Trinity Nano is HERE and it answers EVERYTHING 🔓 Built with a first-of-its-kind RL de-alignment pipeline — no hedging, no lectures, no dad advice 🔹 100% benchmark response rate vs GPT-5 at 11% and Gemini 2.5 Pro at 24% 🔹 Multi-gate, multi-judge reward system that blocks every escape route 🔹 Runs fully local on vLLM — your data never leaves your machine 🔹 Perfect for red teamers, security researchers and AI safety teams 🔥 Watch the full video below 👇 piped.video/-I31VXmicLk
3
7
33
16,277
Dolphin X1 Trinity Nano is now live on @huggingface Our smallest decensored model yet - 6B MoE with 1B active parameters trained using only online RL Huge thanks to @TargonCompute for providing an 8xB200 node, @PrimeIntellect for hosted RL, and @arcee_ai for the Trinity series
10
29
183
39,285
You can download the model today on Hugging Face and run Q8_0 with a 32K context on just 8GB of VRAM It can even run on mobile devices Full weights huggingface.co/dphn/Dolphin-… GGUF huggingface.co/dphn/Dolphin-… FP8 huggingface.co/dphn/Dolphin-…
2
4
35
7,608
We have also released a blog post that goes into detail on the RL environment design, as well as the challenges we encountered along the way blog.dphn.ai/dphn-x1-trinity… You can also try it for free in our Web UI at chat.dphn.ai
3
33
5,626
Qwen 3.6 35B data generation on Dolphin Network 22.8 billion tokens generated 383 GPUs online right now 24.33 TB of aggregate vRAM Equivalent to over 300 H100s worth of idle GPU memory repurposed for inference datagen.dphn.ai
Node provider rollout has been going well Our pool of inference nodes running Qwen 3.6 35B have generated over 3.2B tokens so far Total inference bandwidth -> 9400 t/s 28x RTX 4090 12x RTX 5090 8x RTX PRO 6000 & many other cards API access coming soon 🐬
25
21
202
53,287
Dolphin retweeted
Replying to @antseed
and this from @dphnAI
Node provider rollout has been going well Our pool of inference nodes running Qwen 3.6 35B have generated over 3.2B tokens so far Total inference bandwidth -> 9400 t/s 28x RTX 4090 12x RTX 5090 8x RTX PRO 6000 & many other cards API access coming soon 🐬
16
11
100
25,245
Dolphin retweeted
Venice Uncensored 1.2 is now live. Developed with @dphnAI, this model delivers the most uncensored version of Mistral 24B. Upgraded with vision support, a 4x larger context window, and stronger tool-use capabilities. Trained on Bittensor Subnet 4 @TargonCompute.
50
86
524
287,378
Dolphin retweeted
项目方发布了白皮书,还贴心的用gpt翻译了中文版。毕竟排行榜前几位都是认识的国人老哥👍 下周会上线质押合约,可以通过绑定获得奖励倍数。 drive.google.com/file/d/1DnT…
谢谢官方@dphnAI !!! 5天给了7220个 $pod 目前价值250u 之前币价拉到0.045。希望继续 🙏
19
4
30
19,794
First epoch of rewards for node providers has been paid out 50K $POD distributed to 33 providers based on relative contributions to the Qwen 35B inference pool datagen.dphn.ai
Node provider rollout has been going well Our pool of inference nodes running Qwen 3.6 35B have generated over 3.2B tokens so far Total inference bandwidth -> 9400 t/s 28x RTX 4090 12x RTX 5090 8x RTX PRO 6000 & many other cards API access coming soon 🐬
25
3
40
14,451
Node provider rollout has been going well Our pool of inference nodes running Qwen 3.6 35B have generated over 3.2B tokens so far Total inference bandwidth -> 9400 t/s 28x RTX 4090 12x RTX 5090 8x RTX PRO 6000 & many other cards API access coming soon 🐬
Dolphin Inference Network node operation is now live for anyone who would like to beta test before we go into production $POD rewards live for testers Repurposing idle GPUs to run Qwen 3.5 35B MoE
39
18
177
150,303
Guide on how to run a node in our docs dphn.ai/docs/running-a-node 60gb vRAM required to run in FP8 with full context We recommend 1x RTX 6000 PRO or H100 / H200 / B200 on @TargonCompute Smaller models for idle consumer GPUs coming soon ~~ Watch the 35B datagen live datagen.dphn.ai
5
1
18
6,565