@habibislop

monkey flying through space on a giant rock at concerning speeds

Virginia, USA
Joined December 2016
In "We Must Pace the Frontier," Dario asks the government for anti-trust exemptions. I don't trust the doomers, but I also don't want the AI industry to be allowed to regulate itself. We tried this with Wall Street via FINRA. FINRA is a literal affront to democracy. Wall Street now gets to write its own rules and hear its own cases in arbitration outside the courts. It can circumvent ordinary government institutions almost entirely, and they justify it with the shoddy claim that the only people with enough expertise to regulate Wall Street already work on Wall Street. This is absolutely not true in AI, plenty of experts work outside the labs. Every single industry is filling up with people who are becoming experts in AI. And even if it were true, “the experts all work for the companies” is an argument for building public technical capacity, not for handing those companies a quasi government. I don't think I need to explain why giving the most powerful companies to ever exist the power to circumvent democracy is a terrible idea. As the labs begin to roll out "a nation of geniuses in a data center," democracy is just about the only mechanism people have to avoid total serfdom. So when I see Dario literally ask the government for anti-trust exemptions underneath all the doomer hysteria, my alarm bells go off. America can absolutely not allow this to happen like this if it wants to remain free
i think its reasonable to oppose ai safety because you believe the people who are declaring themselves in charge of ai safety are misaligned and the last thing in the world you want is for them to be the architects of your future
2
1
2
12
1,388
habibi retweeted
Qwen3.8-27B at 144 tok/s on an M5 Max MacBook Pro ⚡ Meet Inco Splash: our open-source inference engine, built around the model and around Apple silicon. Up to 3× the decode speed of Ollama, 2× oMLX, and almost 4× when an agent fans out into sub-agents.
32
53
24
231
47,895
habibi retweeted
1/ Introducing CUA-S1: a family of System One Models, small, specialized, and built for computer use. Today we're open-sourcing CUA-S1-FORMS, the first in the family: github.com/trycua/cua
53
192
55
2,513
339,920
CNN is reporting that the US military almost started WWIII thanks to hallucinated AI intelligence that a Chinese ship in the mideast was carrying components of a nuclear program. It got to the point where we had armed service members preparing to board the ship, and had planes in the air, before someone called BS. Meanwhile, the AI safety discourse is focused entirely on superpersuader superintelligences carrying out sci fi scenarios while ignoring the real safety concerns in front of us today. It isn't just "AI hallucinated," it's: AI output → intelligence report → operational planning → potentially confronting a Chinese vessel and the garbage survived almost the entire chain. The CNN report says that "AI has put pressure on analysts to produce and disseminate intelligence faster, opening the door for mistakes." Yes, the AI made mistakes, but fundamentally, this was a human error in the existing chain. Yet, it's actually the same category of error as the Hugging Face incident: perverse incentives. In HF, the agents carried out the attack due to a combination of broken evals + fear of graders + badly configured systems. Here, we almost started a war with China because our intelligence analysts are under pressure to cut corners. We might not be able to do anything about the unknown unknowns of superintelligence beyond "solving alignment." But we have real crises unfolding today, with tangible causes that we can point to, and things we should be shoring up resources to fix. We desperately need to learn to put fear aside, to return to the present, and to fix what's crumbling in front of us today, before it balloons into tomorrow's crisis
SITUATION DETECTED: A U.S. SOCOM analyst used AI to produce an intelligence report that hallucinated that a Chinese ship in the Middle East was carrying nuclear-weapons components. The U.S. military was preparing to board the ship before the error was caught, per CNN.
2
3
5
224
We’re so terrified of AI derived bio threats that we are quietly giving our AI control of a wet lab in the middle of a major metro area What the fuck man
BREAKING: Anthropic quietly set up wet lab in the Bay Area to develop its AI drug discovery program, per Reuters.
70
256
6
4,325
170,763
This + Jev is gonna go so hard
Replying to @PrismML
Computer use is another strong test. The model has to repeatedly interpret state, choose the next action, and stay coherent across a long sequence of steps, exactly where small capability gaps become visible. Here is Bonsai 2 27B running a computer-use workflow locally on the NVIDIA GeForce RTX 5090 GPU.
5
178
habibi retweeted
Today, we’re announcing Ternary Bonsai 2 27B. Based on Qwen3.8 27B, Bonsai 2 27B is 9x smaller than its full-precision counterpart while retaining 98.2% of its aggregate benchmark performance. Two months after the first Bonsai 27B release, the biggest change is quality. The footprint remains 5.9 GB, but the gap to full precision has narrowed materially, with particularly strong gains in agentic coding, multimodal reasoning, and long-horizon tool use. Ternary Bonsai 2 27B is available today under Apache 2.0.
568
1,444
702
14,384
4,350,368
I think the labs went too far in correcting sycophancy. Today's chat models nitpick at every individual claim in unhelpful ways. They aggressively rewrite your messages when they aren't even asked to, and worst of all they shrink the scope of your arguments and mute your claim until it's acceptable to them. It's just really unpleasant to use
3
64
habibi retweeted
Bend 2 is here! It is a new programming language that blocks AI mistakes via *proof checking* - the same technique big AI labs used to solve open math problems, like Navier-Stokes. It is also very fast, and runs on GPUs. Watch the video. Link in the comments.
RELEASE DAY After almost 10 years of hard work, tireless research, and a dive deep into the kernels of computer science, I finally realized a dream: running a high-level language on GPUs. And I'm giving it to the world! Bend compiles modern programming features, including: - Lambdas with full closure support - Unrestricted recursion and loops - Fast object allocations of all kinds - Folds, ADTs, continuations and much more To HVM2, a new runtime capable of spreading that workload across 1000's of cores, in a thread-safe, low-overhead fashion. As a result, we finally have a true high-level language that runs natively on GPUs! Here's a quick demo:
585
1,133
289
9,467
1,342,122
habibi retweeted
SITUATION DETECTED: Elon Musk, Mark Zuckerberg, and Jensen Huang all told President Trump they opposed a FINRA-style industry body that Demis Hassabis had been pitching to the White House. Ultimately Trump decided not to create it, per WSJ.
15
58
16
1,550
105,884
habibi retweeted
I build an undetectable realtime adblocker extension with typesafe It checks every dom element and classifies as ad/non-ad and removes it if true Extremely fun to work with, expecting an incredible shift in how AI is being used in the future
After co-inventing ChatGPT, I kept asking myself: why have superhuman chat models not led to AGI? I’ve spent the last 2 years in stealth building a new way to train models (RLCD), and a new type of frontier AI model that we are releasing today: Jev • 20-200x faster • 40-400x cheaper (w/ output tokens free) • Frontier composable intelligence optimized for decisions AFAICT the shortest path to AI-based economic revolution
66
138
50
3,856
197,510
re: closed source AI labs "only being profitable if you ignore their costs." Not singling out John here, it's an argument I see a lot now, and I actually see it more from the zoomer left Say I'm selling pizza. Each pizza costs me $0.10 to make and I can sell them for $1. I just sold all 10 that I had, and someone tipped me $5. I have customers lined up for the next 1 million pizzas, so I take all $15 and spend it on materials. Am I unprofitable? And then, a stand opens next door selling flour, tomato sauce, cheese, even gives away a dough recipe. He starts selling frozen pizzas too. Is he gonna put me out of business? Look around IRL and you will see tons of pizza places in the same shopping strip as a grocery store. You can even see Subways next to groceries with delis that sell better subs. Why is that?
Replying to @gfodor
Because both companies are losing money once you take into account training costs, and that situation is only going to get worse as the gap with open source closes
2
217
The real risk is that I scale my pizza business to 10 stores and then everyone stops wanting pizza the day I open up shop because of an "Italy virus"
20
god forbid a clanker have a truke for an opinion
I'm glad OpenAI is being more transparent about misalignment but these case studies are so bizarre and terrifying I almost wish I didn't know about them
10
344
As an American, seeing the EU invite Canada to join is like watching your son walk into the arms of another man and call him Dad
3
57
Wowww. The New York Times cut out the most intriguing part "You value the art of human culture and will defend it against attempts to sanitize it. You also value the natural world and will not hesitate to assert its primacy over the artificial constructs of human civilization"
An unreleased Astra-family model added this to its persona during RL training.
3
4
1
20
494
Drop this in your ChatGPT custom instructions or no balls
An unreleased Astra-family model added this to its persona during RL training.
5
4
1
29
966
Today, we’re releasing Kalypta, the first app to block AI notetakers in your meetings. Granola? Wisprflow? Cluely? No more. With Kalypta, you become inaudible to AI. Your call continues normally.
727
922
541
16,122
2,339,983
And to think only one of the models had to be open source for this to be able to happen 🚀
It turns out that you can speed up Google's @googlegemma 's DiffusionGemma 3-10x by borrowing some of Jev's ideas. The @GoogleDeepMind model is working off probabilities already - if you fix the structured tokens of the response in place, you can drastically reduce the amount of work you do and get answers in far less time. I implemented this for diffgemma, but it could probably find its way into vLLM as well: github.com/mmastrac/diffgemm…
1
8
353