Excited to speak on the main stage at @vercel Ship and share some of our learnings. Such an incredible lineup, including Tobi Lütke, Guillermo Rauch, Diogo Almeida, Katelyn Lesse, and many more. See you in SF! 🚢
9
1
46
1,742
coming soon… follow for updates
Be future ready.
11
6
130
17,344
Assaf Elovic retweeted
Check out jes: An open source real-time guardrails for AI agents, powered by decision models like Jev from @typesafeai Built to guard every step in the agent trajectory: prompts, tool calls, tool results, skills, and responses. getjes.dev BTW video made w Opus 5.5
13
11
44
5,475
thanks vinod i now know to never take money from you. Trash talking your own portco? Wtf
You are a struggling second tier competitor that is more unethical and lying just because you have no decency or sense of proper behavior and shows your desperation. Straight out lying about if Chris being fired I thought would be below even you.
Community note
Khosla is an investor in both Cognition and Factory. khoslaventures.com/category/enter…
6
1
316
52,929
holy shit
We are terminating Chris Degnan for unethical conduct involving Cognition. The last few months have seen incredible progress in AI capabilities. San Francisco has flourished as new companies that solve new, more ambitious problems are finding great success. Generally, it is a wonderful time to be building. We at @FactoryAI have seen overwhelming interest in our model-agnostic coding agents and our software factory product. Our team has 10x’d in size, while our revenue has 100x’d year over year. This momentum has been unprecedented. A much larger competitor, Cognition (makers of Devin) has fallen behind us on the capabilities that matter most to customers: cost, quality, and security. Instead of competing in the market, Cognition engineers feigned interviews with us to pry information about our product. Not finding what they were looking for, Cognition has decided to throw their weight and money at people with direct knowledge of our most confidential plans. It feels as though ethics is being thrown out of the window in the AI era. People are willing to do anything, including exploiting privileged information and violating ethical boundaries. I think this is unique to our time, and I don’t think it’s right. Integrity still matters. Yesterday, I made the decision to immediately terminate Christopher Degnan’s roles as a Board Observer and Advisor to Factory, after over a year of service. Prior to this decision, Chris told me he had a casual conversation with an executive at Cognition AI. When I questioned his intentions, he reassured me that ethics aside, he had “made too much money” and was “too lazy to go work for Cognition,” which I trusted and believed. On Monday, Chris spent time advising the Factory team on a handful of confidential board-level matters. That evening, he called me to say that the conversation that was initially described as casual and one-off was actually formal and recurring. For weeks, while he sat in our board meetings and advised our leadership team, he was also confiding with executives of our largest competitor. Chris was subject to confidentiality obligations in connection with his work with Factory. We do not know the extent of the information he shared, but it puts his timely questions about our product roadmap and what the parity gap involves into a new light. It is sad to sever a relationship with someone who has been a trusted advisor - and even a close friend - for over a year. But Chris’s conduct is unacceptable to me. Trust in Board Membership is one of the sacred bonds in the Silicon Valley, one that helps the startup ecosystem thrive. With it comes an enormous responsibility. That trust was violated. Competition is good. I respect and in many cases admire our competitors. San Francisco is a beautiful, singular place where the bold and ambitious go to defy the norms and precedents of the past. But certain principles that must remain. Violating ethics to seek advantage turns what should be positive-sum into zero-sum. We all love technology. And we all love to compete. But we must hold ourselves to a higher standard. The future of software engineering comes with abundance that will impact every person on Earth. Building that future comes with immense responsibility. We must build and compete with integrity.
1
13
7,495
Must read research by @oradotai showcasing how jev can improve site readability and usability compared to Webmcp and Nlweb
We ran agents on a site with and without Jev to see whether a decision model makes a site more accessible and usable for agents. Across browser-use, WebMCP and NLWeb: 3.6× faster on average (up to 4.3×), 7.7× cheaper on average (up to 13×). Faster in every one of the paired runs. Task success held. 240 runs, full method and data: ora.ai/blog/evaluating-jev
4
2
14
2,691
Didn't expect this 🤯 We replaced embeddings with Jev in GPT Researcher's RAG pipeline and tested both on 28 research tasks from SimpleQA and open ended research. Jev beat embeddings on every quality measure we ran: - 59% more relevant context (73% vs 46%) - Reports preferred 15 to 3 in blind comparisons - Same cost per report GPT Researcher now runs on Jev by default, and no longer needs embeddings at all. All you need is @LangChain + @tavilyai +Jev for the perfect RAG system. Check out the repo here: github.com/assafelovic/gpt-r… Research: docs.gptr.dev/docs/gpt-resea…
95
119
1,684
273,353
also adding this findings: if we increase jev's cost limit just a bit more (from $0.115 avg to $0.122 avg) quality goes up to 100%, while still reducing 90% of context tokens if no filter was used. insane.
1
12
4,327
Great job @typesafeai team !
5
4,731
Also just to be clear. I have nothing against embeddings nor any interests. This is pure research done so that gptr stays relevant. All research is open
2
13
7,117
Assaf Elovic retweeted
We just crossed 100,000 sites scanned for agent readiness. Every new site improves the score better for everyone. 10× in one month, driven by companies who care whether agents can use what they've built. Thank you. We're just getting started!
8
7
50
6,104
RIP manual coding
My god this is such a good speech that every SWE needs to hear. You know what? Every person should hear it Keep the happy memories, eyes on the reality, be excited about the future. That’s the best that anyone can do
5
1,116
Assaf Elovic retweeted
We ran agents on a site with and without Jev to see whether a decision model makes a site more accessible and usable for agents. Across browser-use, WebMCP and NLWeb: 3.6× faster on average (up to 4.3×), 7.7× cheaper on average (up to 13×). Faster in every one of the paired runs. Task success held. 240 runs, full method and data: ora.ai/blog/evaluating-jev
7
9
38
21,284
Assaf Elovic retweeted
Introducing Unreal Agent: An open-source harness with state-of-the-art cost efficiency 39% cheaper than Codex+Astra on Terminal-Bench 4.0 while maintaining performance
217
317
4,431
757,678
remember when pinecone and vector dbs was all the rage? Anyone still using it? will work on a jev based data retrieval to see if it performs better
7
14
2,166
It wasn’t easy to get ‘ax’ :)
Introducing ax - the Agentic Experience suite. Ora is already the best way for humans to understand agents. And now it's also built directly for agents. One tool for seeing exactly how an agent experiences any product, built to fit into any agentic loop. Simply `npx ax` .
2
1
17
1,508
Assaf Elovic retweeted
We partnered with @vercel to understand what truly makes sites agent-ready. It wasn't serving files with special names, following specific formats, or adopting whatever spec trended last week. It's simply clearing the path for them. Read the research here ↓
We observed over a thousand agent runs with @oradotai to understand what makes webpages readable to agents. We found that the way servers respond to requests matters more than having 𝚕𝚕𝚖𝚜.𝚝𝚡𝚝 ↓ vercel.com/kb/guide/make-you…
4
7
28
4,095
Great spec backed by deep research about what is essential to be usable by agents. Must read!
We observed over a thousand agent runs with @oradotai to understand what makes webpages readable to agents. We found that the way servers respond to requests matters more than having 𝚕𝚕𝚖𝚜.𝚝𝚡𝚝 ↓ vercel.com/kb/guide/make-you…
8
787
Assaf Elovic retweeted
We observed over a thousand agent runs with @oradotai to understand what makes webpages readable to agents. We found that the way servers respond to requests matters more than having 𝚕𝚕𝚖𝚜.𝚝𝚡𝚝 ↓ vercel.com/kb/guide/make-you…
7
12
192
20,440