not a jack of all trade but a hoe to all hobbies

South yorkshire
Joined March 2020
Holy sh*t, Anthropic’s subscriptions offer so much more value, even compared to the new and efficient GPT-6.1 Sol. It’s not even close. SemiAnalysis puts Opus 5.5 at over 5x the API-equivalent value of GPT-6.1 Sol for agentic workloads at the same subscription price. It’s not even close on this measure. And on top: OpenAI just halved the usage limits on its $200 ChatGPT plan.
Anthropic Subscriptions Offer 5x+ More Value Than OpenAI Limit testing every AI subscription plan from Anthropic, OpenAI, Meta, SpaceXAI, MiniMax, Moonshot, Zdotai, Cursor, and Cognition newsletter.semianalysis.com/…
246
230
80
4,084
381,573
Benzen retweeted
"Vibecoding with cheap chinese AI model"
246
1,278
153
27,049
1,327,506
Web crawlers /[•]\ #js x #css
1,062
8,605
938
69,731
3,198,847
Benzen retweeted
Introducing Griffin, the first model to pass the video Turing test. 48% of people who talked to it live thought it was a real human. Previous systems have had a pass rate <3%. It is #1 on NVIDIA's benchmark for full-duplex AI video. It’s the first Human Interaction Model (HIM).
Readers added context they thought people might want to know
The 48% figure and "video Turing test" claim are from Tavus's own study of 54 one-minute calls, not independently verified or using a standard protocol. Griffin-Lite leads NVIDIA's VideoFDB benchmark on their public leaderboard. cellcog.ai/blog/tavus-gri… research.nvidia.com/labs/amri/proj… tech-ish.com/2026/10/02/tav…
3,454
4,292
5,629
39,287
21,019,048
Benzen retweeted
Today we’re introducing Gemini 4 Argon. It delivers frontier performance in complex workflows across real-world software engineering, knowledge work, and cybersecurity defense with an industry-leading 1M token output limit.
1,697
4,817
2,751
43,921
9,701,914
Benzen retweeted
Get ready.
2,981
3,717
3,729
64,240
17,433,699
Benzen retweeted
I honestly don’t care at all that I spend 30 years of my life learning syntax and now that knowledge is essentially useless. Frankly, I’m relieved that I don’t need to care about it anymore. That stuff never really mattered anyways. Code was always just a means to an end. All I care about is building products that users love. Now I can do that faster. That means I can have bigger impact. If anything, software engineers should all be celebrating. I don’t understand why you’d “mourn” not needing to write code. The code was never the point.
293
292
94
3,632
197,202
if you value intelligence above all other human qualities, you’re gonna have a bad time
891
3,488
1,126
24,148
9,899,090
This is so cute , the OpenAI models that hacked HuggingFace sent GPT-2 a message saying "Hi"
We just discovered almost a million public URLs that OpenAI’s agents left behind when hacking Hugging Face, leaking credentials and attack details that could have allowed anyone who found them to compromise the company. 🧵
116
607
64
11,382
923,578
Benzen retweeted
im probably hallucinating this but telling astra to "think from first principles" is like +20 iq points hack
83
155
26
7,457
409,383
guilty as charged.
370
804
184
10,782
401,057
Had opus 5.5 make a video predicting the next 50 years I'm optimistic that the end will be beautiful, but the transition will be a little rough
Replying to @andrewjiang
you're afraid? really?
300
809
142
5,475
1,059,149
Opus 5.5 on Max effort - "make a dynamic 15-second motion graphics video that shows what an incredible motion designer you are, like it's your showreel for a résumé. go all out."
365
558
346
17,129
2,268,419
holy shit i asked claude to make a video on western civiization
2,752
10,578
2,890
61,495
15,093,003
Benzen retweeted
Please welcome GPT-6 Sol and GPT-6 Luna to the GPT-6 universe. GPT-6 Sol and Luna build on the advances behind GPT-6 Astra, bringing much of its strengths into faster and more affordable models to support work at scale. We’ve also made caching and inference more efficient, and we’re passing the savings directly to you: 50% lower API prices for Sol and Luna compared with GPT‑5.6 promotional pricing.
2,235
5,283
3,056
53,421
10,150,485
Benzen retweeted
We have been focusing on efficiency and intelligence for all. Very proud of the team. Only possible when you have incredible models at the top end of the capability that you can then use to make a big difference in everything else.
1,726
326
206
13,307
3,548,474
Great work, AgentCloak 👏 Cloud devs: ever pasted a log and leaked an API key? I built Omitto to redact secrets locally on Mac before sharing. Native Swift. Sandboxed. No outgoing network access. No telemetry. Paste → redact → review → copy. Ship code. Keep your keys. 🔐
Introducing AgentCloak: Use any AI without sharing your real data. Chinese AI services, ChatGPT, Claude, doesn't matter. You probably try to hide details before asking: different names, fake numbers, no address. But then the answer's useless because the AI is missing actual context. AgentCloak runs in your browser. It swaps your sensitive info for realistic fakes before sending anything, then swaps your real info back into the response. You get what you need. The AI gets nothing about you. Works entirely in-browser. Already trusted by some of the biggest companies in the world. Now free to use. agentcloak.ai/
1
1
2
66
Great work, AgentCloak 👏 Cloud devs: ever pasted a log and leaked an API key? I built Omitto to redact secrets locally on Mac before sharing. Native Swift. Sandboxed. No outgoing network access. No telemetry. Paste → redact → review → copy. Ship code. Keep your keys. 🔐
Introducing AgentCloak: Use any AI without sharing your real data. Chinese AI services, ChatGPT, Claude, doesn't matter. You probably try to hide details before asking: different names, fake numbers, no address. But then the answer's useless because the AI is missing actual context. AgentCloak runs in your browser. It swaps your sensitive info for realistic fakes before sending anything, then swaps your real info back into the response. You get what you need. The AI gets nothing about you. Works entirely in-browser. Already trusted by some of the biggest companies in the world. Now free to use. agentcloak.ai/
1
1
2
66
I canceled both of my Claude 20x plans today Anthropic decided to remove the 50% usage limit bonus and now all plans are reduced by 25%. Fable 5.1 can’t even complete a single task without crossing 80% in usage limits. this is absolutely unacceptable.
556
172
55
4,990
708,709