@MABUwaAIi
iAccount based inSouth Africa
About this account
- Account based in
- South Africa
- Connected via
- South Africa App Store
Account-level information from X, not a live location or the device used for a specific post.
AI this, AI that.
Johannesburg, South Africa
Joined April 2010
- Tweets12.6K
- Following2K
- Followers2.7K
- Likes2.7K
This is spooky! 🤯 Why do we need this?
Introducing Griffin, the first model to pass the video Turing test.
48% of people who talked to it live thought it was a real human. Previous systems have had a pass rate <3%. It is #1 on NVIDIA's benchmark for full-duplex AI video.
It’s the first Human Interaction Model (HIM).
Readers added context they thought people might want to know
The 48% figure and "video Turing test" claim are from Tavus's own study of 54 one-minute calls, not independently verified or using a standard protocol. Griffin-Lite leads NVIDIA's VideoFDB benchmark on their public leaderboard.
cellcog.ai/blog/tavus-gri…
research.nvidia.com/labs/amri/proj…
tech-ish.com/2026/10/02/tav…
At a WITS conference this year, I asked a full symposium who knows what huggingface is, only 10% raised their hands. Kaggle, Zindi. These students just want to pass. No curiosity outside the classroom whatsoever. What do you mean you never used git? Docker? Aowa! 🚮
It's a skill issue. Most graduates are not employable. I spoke about The “Classroom Student” vs “Out of Classroom Student” Theory of Learning at the Africa Data Science conference earlier this year. Good grades are not enough. I interview many of these grads and it's really bad.
Honestly! Which platform works for y'all? Facebook is a village. LinkedIn is the pits. Men posting on Twitter can't be trusted. What must happen?
This is neat. 👌🏾
Today, I’m proud to announce Homebrew 7.0.0.
The most significant changes since 6.0.0 are faster installations, stronger sandboxing, native macOS app, vulnerability checks, advisory database, end of macOS 10.15 support and Intel Macs to Tier 3.
brew.sh/2026/09/13/homebrew-…
Love them lots! That REPL is unmatched. I use them everyday. You get even better experience in VS Code.
People should stop confusing open source with open weights. That out of the way, which of the known open weights models are triaging 10k agents to solve math problems? To hack their Evals? To hack drones in the physical world?
The AI twitter is buzzing cos Dario dropped a bomb.
I haven't seen so much AI solidarity in a while. Even Sam agrees with him, they couldn't hold hands some time ago at a summit. 🙊
But it's understandable. Dario is making sensible points. Read here. 👇🏾
We Must Pace the Frontier: I’ve written a new essay on why the AI industry should slow down, with a three-part plan for doing so.
Anthropic is unilaterally committing to the first of these steps. We’ll provide third-party evaluators with permanent, employee-level access to our systems, so that they can verify adherence to our safety measures, report on incidents, and assess models’ alignment during training.
You can read the full post here: darioamodei.com/post/we-must…
It depends on what you're programming.
I've spent years searching for a programming language with:
- Expressive types
- Big ecosystem
- Large community
- Good performance
- Good Devx
Here's my ranking across 15 of the most popular languages: hamy.xyz/blog/2026-01_missin…
What does 100% even mean? 😭
GPT-6 Astra is here.
We hope it will begin to enable a new generation of entrepreneurship, scientific discovery, and building.
We believe it is the best model in the world for computer use, professional work, science, coding, cybersecurity, and more.
It took us some extra time to ensure that we could meet the safety and alignment standards required for this capability level, but we think you’ll find it worth the wait.
It scores 98% on FrontierMath Tier 4, 99.9% on ARC-AGI 3, and 100% on ExploitBench.
Yoh! 🙆🏾♂️🙆🏾♂️🙆🏾♂️
Replying to @OpenAI
GPT-6 Astra is state-of-the-art on FrontierMath Tier 4, ARC-AGI 3, and TerminalBench-4.0.
GPT‑6 Astra is also a major advance for scientific discovery, with state-of-the-art performance on Terminal-Bench Science 0.1 and HealthBench Pro.