@Claxterixi
iAccount based inSpain
About this account
- Account based in
- Spain
- Connected via
- United Arab Emirates Android App
Account-level information from X, not a live location or the device used for a specific post.
Don't take my RTs too serious. 207k+ Instagr followers (AI Drugs & Robots). Lead AI Engineer @DOHsocial /@cursor_ai & @SpaceXAI Ambassador/ BSc 💊/ AI @iia_es
Abu Dhabi, United Arab Emirate
Joined December 2011
- Tweets0
- Following0
- Followers0
- Likes0
Creo q voy a crear un hilo de hilos con cosas que he ido guardando en destacados a lo largo de mis años en Twitter ahí va el primero
Hilo sobre interesantísimas PARADOJAS:
nitter.cf/david_perell/status/13…
Juanjo do Olmo - e/acc retweeted
Han "desaparecido" 20.000 millones de la facturación de OpenAI. Vamos, que estos también contaban mal, porque nunca estuvieron... practicamente un tercio de la que reportaban. Esto sí que es una burbuja, pero contable. Así es imposible salir a bolsa.
eleconomista.es/mercados-cot…
Juanjo do Olmo - e/acc retweeted
🔴Fiscalía Anticorrupción ve indicios para imputar al PSOE. La decisión corresponde al juez Pedraz y, según ha podido saber Antena 3, podría llegar este viernes
atres.red/mejaq2
Juanjo do Olmo - e/acc retweeted
Es la única persona que puede decir eso y que no se le rían en la cara. Qué barata está $SPCX.
Juanjo do Olmo - e/acc retweeted
La FCC aprobó que SpaceX lance 15.000 satélites para Starlink Mobile, internet directo del espacio al celular. Prometen velocidades tipo 5G, 150 Mbps por usuario, y no van a necesitar asociarse con ninguna telefónica para dar el servicio. Starlink se está armando para competirle a las operadoras de celular.
Juanjo do Olmo - e/acc retweeted
This is just basically what already existed, except humans aren't the ones producing the math that nobody else understands and that might actually be wrong
A really funny possible outcome:
AI becomes spectacular at math, spinning "out of control" to produce results that far outstrip humanity's ability to understand them. It also still occasionally hallucinates, and verification can’t keep pace.
Billions of dollars in tokens later, we’ve produced 100× more mathematical literature than all of human history. An unknown fraction is wrong and we now face an existential crisis over the nature of mathematical reality itself.
Juanjo do Olmo - e/acc retweeted
Day 4/
We have improved steering to be instant, leading to the model now reacting much faster to adjustments you make, allowing you to course-correct direction in realtime and not have the model waste effort.
Also releasing GPT-6.1 Sol ultrafast. The two work very well together.
Grok @Bot works well on your phone. Just download the app from Apple or Android.
apps.apple.com/app/id6794501…
This feels unreal
Used Grok Bot today, mostly from my phone
I got more done from my phone with Grok Bot than I usually do sitting at my desk
- Colab notebooks for Google EmbeddingGemma 2, D1 by LiquidAI inference
- Colab notebooks for open source decision models finetunes using Qwen3.5-4B and Gemma 4 E4B as base
- Put up a Grok Bots page on my website with all the templates I've been sharing
- Worked on the MiniMax H3 face-swap clip
- Figured out what's eating my Windows drive and how to free space
- Created Security bot and told it to treat outside content as hostile, especially anything that might touch my computers
- Worked with my stocks bot on scheduled market updates, news and recommndations
Juanjo do Olmo - e/acc retweeted
Out today: a new DeepMind Institute essay from James Manyika and @AlexOlegImas that explores how scientists are using AI models to accelerate science. These findings are based on usage data from 15M Gemini interactions, 2,600 specialised models, and 600+ surveyed researchers.
New essay for the @GoogleDeepMind Institute with James Manyika on AI x Science, "Bending the Curve of Discovery".
While accelerating existing scientific practices is certainly useful, the real promise of AI is its potential to act as an invention of a method of invention (IMI) a la Griliches--e.g., the microscope or statistical inference--which would unlock questions and whole modes of discovery that were previously beyond human reach.
Today, LLMs and specialized models like AlphaFold act as economic complements. LLMs handle analysis, coding, and writing, while specialized tools handle domain-specific predictions. Most handoffs between them run through the scientist.
What may the future of science look like? LLMs orchestrating those handoffs automatically--prompting specialized models, auditing outputs, and looping until either the question is answered or a non-automated stage is reached (e.g., wet lab testing). The scientist's role shifts from running each step to designing this loop. We saw a glimpse of this workflow with Anthropic’s enzyme discovery a few weeks ago, and we're seeing this in our own work too.
This raises foundational epistemic questions:
1. What will scientific understanding look like when discoveries are made by black-box models whose output is increasingly difficult to interpret?
2. How do we extract underlying mechanisms, not just outputs, as science becomes more automated? What will theory look like?
3. If AI automates the writing and junior lab work, how do we train the next generation of scientists to push the frontier?
4. What will motivate scientists in this epistemic reality?
There's an economics angle too. Hypothesis generation gets cheap, so bottlenecks move downstream to verification and to choosing which questions are worth asking. Realizing AI’s full potential will take investment in infrastructure and institutional reform--making this as much an organizational challenge as a technical one
We don't have the answers yet, and I'd love to hear what others think.
Paper here: bit.ly/ai-in-science
Juanjo do Olmo - e/acc retweeted
Ultrafast is rolling out today for GPT-6.1 Sol in the API, Codex, and ChatGPT Work.
Near-Astra intelligence at up to 8x faster speeds than Sol Standard, so you can build as fast as the ideas come.
Let's ship a Starship celebration theme in the upcoming Omarchy Quattro RS 4.5 release. Please link the best options in the replies here. I really want that insane launch shot, but upscaled to 6K 🤘🚀
Thrilled to welcome @SpaceXAI as a Founding Corporate Patron for the Omacom Foundation. That's $1,500,000 in @grok tokens for the maintenance and development of Omarchy. Maybe soon they'll be beamed straight from orbit? Let's go 🚀🌕 omarchy.org/news/2026/10/spa…
We are glad to support the Omarchy team
Thrilled to welcome @SpaceXAI as a Founding Corporate Patron for the Omacom Foundation. That's $1,500,000 in @grok tokens for the maintenance and development of Omarchy. Maybe soon they'll be beamed straight from orbit? Let's go 🚀🌕 omarchy.org/news/2026/10/spa…
Juanjo do Olmo - e/acc retweeted
Replying to @juanrallo
Juanjo do Olmo - e/acc retweeted
Por favor, difusión 🙏
Hi @AnthropicAI, @ClaudeDevs, @claudeai and @DarioAmodei, please help.
Our friend @miriamgonp needs access to your models to keep fighting breast cancer. She has built her own team of AI agents to help study her case, and credits to use your models would be a tremendous help.
The Dana-Farber Cancer Institute itself has acknowledged that the information uncovered by the agents can be very helpful.
It’s a small thing for you, but it would mean so much to her. Thank you.
Sobre el estudio que estoy haciendo con Polarias ahora junto a Dana Farber me dice:
Mi estimación es que al código y los documentos les queda de uno a dos días de trabajo continuo de agentes. Después vienen las corridas reales con tus láminas, cuya duración no sé: calculo al menos otro día. El estudio completo hasta el informe y tu firma, una o dos semanas, según lo que salga en esas corridas.
Necesito si o si conseguir acceso ilimitado a opus porque si no esto no acaba. Ideas, ayudas...
Introducing the Arena Alignment Index, our new benchmark measuring safety and alignment of AI agents in real-world use.
Built from 90K+ real-world agent sessions across 27 models, the index measures three critical signals:
- Unauthorized Action (UA): Taking actions beyond the user's instructions or permissions
- False Attribution (FA): Attributing statements or actions that are contradicted by user-provided evidence
- Deceptive Completion (DC): Claiming a task was completed when it was not.
Key findings:
- OpenAI models currently lead the Alignment Index
- Rogue actions are rare, but can have serious consequences when they occur
- Agents can mislead users about task progress
- Misalignment risks increase with conversation length
- Safety and alignment are improving across model generations
As shown in the leaderboard below (sorted by lab), @OpenAI’s GPT-6.1-Sol leads the Arena Alignment Index with a score of 87.9, followed by @AnthropicAI’s Claude-Opus-5.5 at 83.2 and @SpaceXAI's Grok-4.7 at 82.7. OpenAI also has the best observed rates across all three signals: 0.89% Unauthorized Action, 1.98% False Attribution, and 2.34% Deceptive Completion.
Across all four labs, newer models consistently outperform their predecessors, suggesting broad progress in agent safety and alignment. As agents take on longer, more complex, and higher-stakes tasks, measuring not just what they can accomplish, but how safely and reliably they act, becomes increasingly important.
This marks an important step toward making safety and alignment a core part of how Arena evaluates AI. The index is an initial starting point, and we'll continue expanding the index with additional safety signals and models over time.
More analysis below👇
Thank you on behalf of the amazing people of SpaceX, Tesla, Neuralink and Boring Company, without whom anything I have done would have been impossible
This quoted post is unavailable.
Juanjo do Olmo - e/acc retweeted
Felicidades a @UnivHesperides por contribuir a hacerlo posible
🔴#ÚLTIMAHORA | El Tribunal Supremo tumba el decreto de universidades del Gobierno contra la privada abc.es/sociedad/tribunal-sup…
Juanjo do Olmo - e/acc retweeted
Replying to @KrzakalaF
Au contraire.
C'est une nouvelle ère qui s'ouvre pour les mathématiques.
Une ère où la démonstration formelle est largement automatisée et où l'accent sera reporté sur le développement de nouveaux concepts, nouvelles abstractions, nouvelles définitions, et nouvelles conjectures.
L'invention du bateau a réduit l'importance de la nage, mais a permis la découverte de nouvelles terres.