Here's a quick look at the complete MCP server I setup, using the new Auth spec from May. Added it to claude.ai and it fully works with the new streamable http spec too! #ai #mcp #LLMs #claudecode
2
1
4
2,589
published TWO pi.dev packages right now github.com/zackify/pi-claude… github.com/zackify/pi-port-f… opinionated shift+tab logic and a way to ssh forward ports fast
111
Now I run pi.dev using bypass all permissions most of the time, and go to manual approve for tricky things. Also have a iptables firewall rule to block all access outside of the provider URL.
1
161
Combine that with full bypass mode. Passwordless sudo in a container and tmux with status line for the network lock... It's the fastest way to get stuff done autonomously
1
63
Claude was slow for a bit so I switch to 5.4 without leaving anything. Long running task, run the network lock and leave on in tmux
64
pi.dev is way better than opencode. Feels the best. Maybe even better than Claude code and I love the extensibility and hot swapping between providers with Ctrl p
1
116
so i vibe coded a wispr flow alternative with the same floating bubble feature for android in 30 minutes. if you want AI post processing that would be pretty simple to add but I don't need it. github.com/zackify/flo
149
Am I the only one who 99% of the time uses Claude or opencode or codex on manually accept edits? I get the review in as I go and quickly voice to text to tweak. Running to end then finding problems is more annoying
1
1
286
You've never needed nextjs or a frontend framework Use script type module and use bun build exported to a storage bucket. Oo you need fancy "static pregeneration" OK have Claude spend 35s to make a script that loops your pages and renders them to HTML. Plain react can go far
1
2
176
anyone know if copy/ paste in opencode will ever get on par with claude code?
1
3
276
I use Claude and opencode and openclaw a lot now. Sounds crazy but I enjoy testing out all the models. Its sad that opencode go has a messed up glm5 provider. Tried it out yesterday and it was atrocious
1
260
As soon as I switched back to ollama cloud as the provider. The exact same task one shotted with no issue. The opencode team should be transparent about where the models are coming from and why they are performing badly or heavily quantized
79
openai and anthropic will go to 0 in a few years when m5,6,7 ultra's are able to prompt process incredibly fast. qwen 3.5 122b a10b and qwen3-coder-next are already usable on device with a lot of ram
1
196
Zach Silveira retweeted
🚀 Introducing the Qwen 3.5 Medium Model Series Qwen3.5-Flash · Qwen3.5-35B-A3B · Qwen3.5-122B-A10B · Qwen3.5-27B ✨ More intelligence, less compute. • Qwen3.5-35B-A3B now surpasses Qwen3-235B-A22B-2507 and Qwen3-VL-235B-A22B — a reminder that better architecture, data quality, and RL can move intelligence forward, not just bigger parameter counts. • Qwen3.5-122B-A10B and 27B continue narrowing the gap between medium-sized and frontier models — especially in more complex agent scenarios. • Qwen3.5-Flash is the hosted production version aligned with 35B-A3B, featuring: – 1M context length by default – Official built-in tools 🔗 Hugging Face: huggingface.co/collections/Q… 🔗 ModelScope: modelscope.cn/collections/Qw… 🔗 Qwen3.5-Flash API: modelstudio.console.alibabac… Try in Qwen Chat 👇 Flash: chat.qwen.ai/?models=qwen3.5… 27B: chat.qwen.ai/?models=qwen3.5… 35B-A3B: chat.qwen.ai/?models=qwen3.5… 122B-A10B: chat.qwen.ai/?models=qwen3.5… Would love to hear what you build with it.
425
1,083
7,884
4,002,962
Zach Silveira retweeted
Qwen3.5-35B-A3B is now available in LM Studio! This model outperforms previous Qwen models that are more than 6x its size 🤯🚀 Requires about ~21GB to run locally. lmstudio.ai/models/qwen/qwen…
72
175
2,229
332,644
Why do MCP servers act like authentication was just built this year on the web? I shouldn't need to re-auth an mcp server every single day.... set a longer session lifetime geez
2
76
Every time I try to give opencode another shot Some weird thing like this happens. Turns out it's expected to list models that don't even exist when you add @lmstudio github.com/sst/opencode/issu…
1
1
131
I'm excited to know we will definitely have opus level ai running on device for coding privately and fast in a few years
1
102
Devstral 2 24b is insanely good. The next 5 years are going to be awesome. I get solid speed up to 30k context then it starts to get slow. But it can actually get things done in opencode or Cline with under 20gb of vram LM Studio's mlx version 🤑
3
250
So @elonmusk your service centers just screw you completely when you're out of warranty? They want $240/hr just to diagnose a safety constraint issue in my 6 year old model 3. Minimum 1 hour.
1
78
Even worse it says 240 just to check the wire voltage on the harness
46