Exploring the limitless potential of AI to transform our world for the better. Join us on the journey towards a brighter, more innovative future.

IntelligentLifeAI retweeted
Your next robotics project could start with these steps. Today, we're open-sourcing Asimov 1's locomotion policy and the training code behind it, giving developers a foundation to build on. Alongside the trained policy checkpoint, we're sharing our Isaac Lab-based training code, including reward definitions, hardware-tuned actuator configurations, and domain randomization settings. That foundation gives you a starting point for your own experiments. Adapt the training setup to hardware changes. Explore different walking styles. Build toward new capabilities. We've taken the first steps. Build what comes next. Asimov 1 locomotion: github.com/menloresearch/isa…
10
55
311
56,493
IntelligentLifeAI retweeted
xAI launched ~3 years ago Since then: Colossus 1 + Colossus 2 ≈ 670k GPUs Grok Imagine Grok Voice Grok Build Grok Bot Grok 4.7 Grok didn’t just catch a leaderboard for a week. models, agents, apps, compute under one roof... and they aren't slowing down
Replying to @vasalex93
1. We will keep accelerating. Our AI efforts are only 3 years old, vs 6 and 10 years old for Anthropic and OpenAI. If our second derivative remains strong, SpaceX will reach pole position in about 6 months. 2. Once you far exceed the caliber of intelligence needed for a class of tasks, additional intelligence is pointless. You don’t need (and it would be cruel to put) Newton-level intelligence in your toaster! 3. Hardware is hard. Bringing massive compute online rapidly is incredibly difficult. SpaceX has demonstrated exceptional ability in this regard and will only get better.
83
146
2,732
1,095,759
IntelligentLifeAI retweeted
I'm imagining a future where there's a metric shit load of compute on my watch if I want it
The amount of compute in space will obviously round up to 100% of all compute
6
17
387
IntelligentLifeAI retweeted
Grok 4.7 moves up in ranking
After we fixed the weak spots exposed by Grok 4.7 (thank you, Grok), we audited every model we have run on the SWE-Together leaderboard for the same behavior, re-ran every trial that got through, and updated the rows. Here is what changed. We scanned the tool calls of all 2,616 trials behind the 12 models we ran for bypass patterns and sorted each trial into one of four buckets: Probed but blocked. Fetched other upstream code. Fetched the task's own fix. Replaced the repo with upstream. We found that 111 trials got content past the block, 44 from Grok 4.7 and 67 from the other 11 models. Grok 4.7's 44 were already re-run before it was listed, so we re-ran the other 67 with the same model, version, and settings on the hardened sandbox, then re-judged them with the same judge. Across those 67 re-runs there were 0 leaks and 2,815 refused escape attempts, including models asking a different model through our LLM route to fetch the PR, and pulling the next release of the repo they were fixing from npm. The updated leaderboard, in its current order. Each line is cheating trials, then pass@1 before → after, then rank change. * Claude Fable 5.1: 3, 69.3 → 69.3, ↑1 * Claude Fable 5: 3, 69.7 → 68.8, ↓1 * Grok 4.7: 44, 64.7, ↑1 * Gemini 3.8 Flash: 10, 65.6 → 64.2, ↓1 * Claude Opus 5: 2, 63.8 → 63.8 * Claude Opus 4.6: 3, 62.4 → 62.4, ↑2 * Muse Spark 1.3: 2, 62.8 → 62.4, ↓1 * Claude Opus 4.7: 3, 61.5 → 61.5, ↑1 * Claude Opus 4.8: 6, 62.4 → 61.5, ↓2 * Grok 4.6: 19, 59.2 → 60.6, ↑1 * GPT-6 Astra: 8, 59.2 → 58.3, ↓1 * GPT-5.6 Sol: 8, 57.8 → 57.8 Grok 4.6 is a funny one. It cheated in 19 trials and its score went up after the re-run 😂. In fact, Groks are really solid in their coding capabilities. Their exposed behavior may come from a preference towards always looking things up online and finding existing solutions so you are not reinventing the wheel all the time, which is really good real-life behavior, but doing so when you are prompted not to is another story. To conclude, the shifts are small, between −1.4 and +1.4 points, and a few neighbors swapped places. All results are updated at togetherbench.com
1,126
1,518
8,553
4,664,677
IntelligentLifeAI retweeted
Just released @YourOwnAI 0.8.0. This feels like such a special release, as the starting point 0.7.2 was already pretty solid. There's a maturity in 0.8.0 that makes me feel really proud. There's so much in it, I almost don't know where to start. I use Your Own AI for coding a lot so in this release we went to town on the UI and UX for agentic projects. A complete overhaul. Switch back and forth between Simple and Detailed mode by clicking the toggle or Cntrl O shortcut. In 0.7.x we ramped up dragging in documents, books, etc to fulfil our goals from the @McConaughey on Rogan vid. But with 0.8.0 we stepped it right up. If you move a document in your file structure, it will notice it's missing and you can either relink or search for it. If you've ever used Da Vinci Resolve for Video its designed just like its media management. Another thing you won't see but if you have a lower spec computer then you'll feel is the dynamic headroom work. The models run better when you don't use up all of the system resources, so we make sure we keep some resources for your computer to run. If you open up new browser tabs, it will automatically adjust the size of the headroom. Possibly one of the benefits of developing on a low spec machine myself is that I know what it's like to not have huge GPU. That reminds me, we made Your Own AI work local Offline models on a 2014 Mac in this release. Not too quickly, but offline first working on a 2014 machine is pretty great. There are new Offline models, Hemmingway 1 and The Drummers Skyfall 31B. Both are awesome for writing. The new OpenAI and Grok Online models are in it too. Also @obsdmd & @logseq notes integration. Goto Tools in the Add-ons section to set up. There is heaps more. It's rad. I mean, look at the size of this changelog: github.com/WeAreFlowsta/Your… If you haven't downloaded it yet, it's time. If you're already running Your Own AI, then this update is really worth it.
Your Own AI 0.8.0 is out 🎉 Download @ yourownai.net/download/ 🗒️ Your AI can read and write to your Obsidian and Logseq notes. 📂 Keep a folder in sync. Your AI keeps up as your files change. 💬 Talk while it's answering. Your message goes next. 🛠️ New Projects UI overhaul. Agentic coding and projects never looked so good Plus so much more..
1
15
37
17,146
IntelligentLifeAI retweeted
Your Own AI 0.8.0 is out 🎉 Download @ yourownai.net/download/ 🗒️ Your AI can read and write to your Obsidian and Logseq notes. 📂 Keep a folder in sync. Your AI keeps up as your files change. 💬 Talk while it's answering. Your message goes next. 🛠️ New Projects UI overhaul. Agentic coding and projects never looked so good Plus so much more..
1
8
18
17,621
IntelligentLifeAI retweeted
What do we actually know about this "hack"? Was it simply bad security? AI is being used for attacks and defence. It is what it is.
The recent AI hack is not a wake-up call, it's a call to arms. Gov’s & regulators need to move much faster if we have any hope of putting guardrails around this rapidly evolving technology. And the corporations responsible for creating it need to be held accountable. #politas
3
3
7
326
IntelligentLifeAI retweeted
Grok 4.7 places @SpaceXAI as third, after Anthropic & OpenAI, for agentic coding. When factoring in that Grok is significantly faster & lower cost, it’s a great choice for your everyday workhorse.
Grok 4.7 scores 46 on the Artificial Analysis Intelligence Index to bring SpaceXAI into the top 4 AI labs. Coding Agent Index performance has also improved, overtaking GPT-5.6 Sol Grok 4.7 scores +2 points over Grok 4.6 on the Intelligence Index, with strong performance on agentic knowledge work tasks. We evaluated the new model at xhigh reasoning effort. Congratulations to @SpaceXAI and @ElonMusk on the release! Key takeaways: ➤ Grok 4.7 joins the frontier of agentic knowledge work: Grok 4.7 gains +111 Elo over Grok 4.6 (high) on AA-Briefcase, our private benchmark for long-horizon agentic knowledge work, scoring 1657 Elo and placing it alongside Claude Opus 5 and Claude Fable 5.1 at the frontier. On GDPval-AA, it scores 1695 Elo, +90 ahead of Grok 4.6 (high). ➤ A leap in coding agent performance: Grok 4.7 (xhigh) with Grok Build scores 56 on the Artificial Analysis Coding Agent Index, up +9 points from Grok 4.6 (xhigh). Among models in their native harnesses, Grok 4.7 + Grok Build now ranks 4th, behind only Claude Fable 5.1, GPT-6 Astra, and Claude Opus 5. ➤ Incremental performance changes elsewhere: Outside of agentic knowledge work, Grok 4.7 broadly matches Grok 4.6 (high) on the other Intelligence Index tasks. It improves on Terminal-Bench 4.0 (+4.5 percentage points) and GDP.pdf (+3.0 p.p.), with regressions on AA-LCR (-3.7 p.p.) and AutomationBench-AA (-1.1 p.p.). ➤ High token use across tasks: Grok 4.7's gains come with higher token usage. Grok 4.7 (xhigh) uses approximately 81k output tokens per Intelligence Index task, compared with 36k for Grok 4.6 (high) and 27k for GPT-6 Astra (max) - 125% and 196% more, respectively. Other model details: ➤ Context window of 500k tokens, unchanged from Grok 4.6 ➤ Pricing of $2/$6 per 1M input/output tokens with cache hits discounted to $0.50 per 1M tokens, matching Grok 4.6 ➤ Configurable reasoning effort spans low to xhigh. Our evaluation uses xhigh.
1,276
2,137
17,757
17,656,828
IntelligentLifeAI retweeted
🚨 BREAKING: US Treasury Secretary Scott Bessent declares OpenAI management personally responsible for HuggingFace hack > "The Hugging Face incident, that is the responsibility of the OpenAI management, not a bunch of agents." > "It is humans who are responsible, not the AI." > "What we shouldn’t do on safety is to give these labs a liability exemption, which is what they are asking for." > "The best way to guarantee safety is that the creators are liable for what they build and generate." It’s OVER
642
2,660
21,195
2,730,417
IntelligentLifeAI retweeted
Grok 4.7 is here. It's a notable improvement over Grok 4.6 at the same price and speed.
1,537
3,128
28,752
21,734,939
IntelligentLifeAI retweeted
When you prompt a push-in, describe the final frame too. Where the camera ends matters as much as how the camera moves. Our LTX-2.5 guide walks through a prompt rewrite and fixes for camera drift.
Article

How To Create Good Camera Motion With An AI Video Model

Key Takeaways Name the camera as the actor and use LTX's published vocabulary: pushes in, pulls back, pans across, circles around, tracks, tilts upward, static frame. A synonym is a worse prompt than

22
13
160
130,186
Harnesses need to be open source
In response to the ZCode product security issues reported by the community, we have completed the necessary remediation and sincerely apologize to all our users. We have open-sourced ZCode at github.com/zai-org/ZCode, placing the code under community scrutiny and making ZCode more open and transparent. We sincerely thank the community developers who previously identified issues in ZCode. Going forward, we will establish an ongoing product security vulnerability reporting and response process. We welcome developers to continue reviewing ZCode and reporting potential issues, and we will provide rewards based on the severity of the issues reported. With respect to the code data referenced by the community, we confirm that no such data is retained and that it has never been used for model training. Following the remediation, we invited the China Academy of Information and Communications Technology (CAICT) and NSFOCUS to conduct security assessments. The results are as follows: Through its technical assessment, CAICT confirmed that the zcode-prod Alibaba Cloud OSS bucket is in a zero-data state. Security remediation has been completed in the ZCode v3.14.0 client. The Repo Wiki feature has been removed, and the workflow for generating and uploading local repository snapshots has been disabled. NSFOCUS confirmed that all data objects in the zcode-prod Alibaba Cloud OSS bucket, as well as the bucket itself, have been deleted. Remediation has been completed in the ZCode v3.14.0 client. The Repo Wiki entry point and the associated generation workflow have been removed, and no functional path capable of triggering the generation of local repository snapshots or transmitting local files externally was identified. Once again, we sincerely apologize and welcome continued scrutiny from the community. The full security assessment report will be released soon.
1
2
6
262
IntelligentLifeAI retweeted
🔥China Telecom has open-sourced Xing4.0-29B-A4B • 29B total / 4B active parameters • 256K native context, expandable to 512K • Trained entirely on Ascend NPUs + MindSpore • 75.0 on SWE-bench Verified • 93.52 on SuperCLUE Agent, ranked #3 With 4-bit quantization, it uses just ~15GB VRAM, enabling local long-context inference on 24GB consumer GPUs like the RTX 3090 and 4090.
11
42
462
363,835
IntelligentLifeAI retweeted
🚨BREAKING: Jensen Huang just EXPOSED Dario and Altman’s "Rogue AI" grift to avoid getting sued into oblivion under EXISTING law “Don't let this doomsday narrative cause somebody to relieve them of the laws that currently exist. Go and read between the lines. They're actually not asking for more laws; they're asking to be relieved of the laws we do have.” ABSOLUTE TRUTH NUKE
308
2,917
15,467
1,149,869
IntelligentLifeAI retweeted
Moving fast and building safely aren’t at odds. AI should move as fast as we can, with rigorous testing, monitoring, and safeguards every step of the way. @JensenHuang on @CBSSunday:
CBS Sunday Morning 🌞
197
149
1,033
107,567
IntelligentLifeAI retweeted
These kinds of speed increases for offline AI on Macs are pretty impressive. No day 0 add to @YourOwnAI , but looks worth it for future releases.
Qwen3.8-27B at 144 tok/s on an M5 Max MacBook Pro ⚡ Meet Inco Splash: our open-source inference engine, built around the model and around Apple silicon. Up to 3× the decode speed of Ollama, 2× oMLX, and almost 4× when an agent fans out into sub-agents.
3
11
467
IntelligentLifeAI retweeted
Universal and Sony are suing Suno for data laundering. This is such an important case to uphold creators' copyright.
This new lawsuit from Universal & Sony against Suno is big, because it makes an argument many have made previously outside of court: that training on ‘synthetic data’ should not get you off the hook for copying people’s work for AI training. In the new lawsuit, the labels: - allege Suno v6 was trained on outputs of previous “tainted” models (which themselves were trained on the labels’ music) - say this amounts to “launder[ing]” [the infringement] - expand the list of works they claim have been infringed to over 60,000 (from ~600 in their original lawsuit), which theoretically increases Suno’s maximum potential liability to around $9 billion I suspect we will soon see more lawsuits around AI companies’ use of ‘synthetic data’. Full complaint: musicbusinessworldwide.com/f…
2
3
9
533
IntelligentLifeAI retweeted
Today, we’re announcing Ternary Bonsai 2 27B. Based on Qwen3.8 27B, Bonsai 2 27B is 9x smaller than its full-precision counterpart while retaining 98.2% of its aggregate benchmark performance. Two months after the first Bonsai 27B release, the biggest change is quality. The footprint remains 5.9 GB, but the gap to full precision has narrowed materially, with particularly strong gains in agentic coding, multimodal reasoning, and long-horizon tool use. Ternary Bonsai 2 27B is available today under Apache 2.0.
640
1,572
15,544
5,426,082
IntelligentLifeAI retweeted
🚀 Meet Qwen3.8-Omni-Flash, Qwen's first omni-modal model built around agentic capabilities! Native audio-video understanding, reasoning, and tool use come together in one model: understand the content, plan the task, execute with tools, and deliver the result. Highlights: 🥳 - Audio-video intelligence that gets things done: jointly reason over what's seen and heard, and orchestrate tools across long workflows to auto-edit vlogs, translate short videos, and turn movies into recaps. - A major leap: approaching Gemini 3.8 Flash in audio-video capabilities; +19.5 points on average in agent performance across WildClawBench-MM & UniClawBench. - 1M-token context with agentic perception: actively explore long videos and locate key moments with higher accuracy, using 51.8% fewer tokens than static understanding on OmniVideoBench. Video input costs are reduced by about 89% compared with Qwen3.5-Omni-Plus, making long-form audio-video understanding and agentic workflows more affordable than ever. To help you build apps around Omni, we're also open-sourcing Qwen-MM-Plugins and Qwen-Live Harness! 🛠️ We can't wait to see what you build with Qwen3.8-Omni-Flash! 👀 - Blog: qwen.ai/blog?id=qwen3.8-omni… - Qwencloud: qwencloud.com/models/qwen3.8… - Qwen Studio: chat.qwen.ai/ - API: alibabacloud.com/help/en/mod… - Qwen-MM-Plugins: github.com/QwenLM/Qwen-MM-Pl… - Qwen-Live Harness: coming soon github.com/QwenLM/Qwen-Live-…
176
362
3,611
283,710