@gioen__

breaking and buildings things, sometimes the other way around. member of technical staffing

Joined October 2021
doing my bit to pace the frontier thanks sama
1
33
in SF for tech week keen to have high signal conversations about AI for science, biological reasoning models and generally post training for drug discovery
41
new to sf is that where the high mts circles are convening right now?
General rumors in high MTS circles in SF that: - Anthropic will launch a new insane model with incredible capabilities a week before IPO - This model’s PR is around a major positive scientific discovery (no scare tactics like mythos and cyber)
1
72
SF is crazy cause you can just sit in a self driving car and not be forced to do any small talk and just poast random stuff on Twitter for half an hour unbothered
1
1
117
just landed in SF weather is impeccable, not a cloud in sight AI billboards everywhere autonomous vehicles everywhere winning
22
robomaxxing robotaximaxxing
1
37
guess who’s back once again
35
gioen @ SF tech week retweeted
I resigned from Google today. I enjoyed my work and loved the people, but my GDM team was working on a new generation of chips to make AI much faster and cheaper, and I think AI is already progressing too fast, so I had to quit.
847
392
281
4,238
1,084,330
gioen @ SF tech week retweeted
Introducing Limite 1B - Violetto. A model for high-frequency mathematical intelligence.
64
226
77
1,828
288,098
4 pro is going to melt faces
22
4 > 5 4 >= 6 4 >= 5.1 4 ? 5.5
31
gioen @ SF tech week retweeted
This is legitimately terrible for OpenAI --- but also Anthropic --- and it relates to their work in drug discovery. I read this as: prompts will be monitored and potentially acted on if we think there’s something scientifically and/or commercially valuable being disclosed. Now, the big problem for Anthropic is that they are increasingly seen to be competing with their customers. Outside of SF, many people also think of OpenAI/ChatGPT and Anthropic/Claude as somewhat interchangeable companies and products --- aka, if OpenAI do something, surely Anthropic will do it too? As so much IP relates to drug targets and people disclose these to Claude, I think we will see a lot of pharma companies stop using these tools or put up a lot of internal guardrails.
wtf is this way to handle mathematicians work and scientific communication TLDR: Leven and Tristan worked over several months on one of the Millenium Prize Problems with various AIs to reach final interesting results. OpenAI apparently heard about it in the last days and prompted their latest models to work on the direction Leven and Tristan found fruitful. They then tried to push for controlling communication of the result and dropping Leven from authorship with some very bad taste social pressure. Hope this is not a glimpse of the future we’ll get in science research with these dominating players playing marketing games hurtful for the real scientific community.
74
290
64
2,255
404,895
my son ted, short for gated attention
my son max, short for softmax
45
oh you’re working on mid training at gdm ? why don’t you instead focus on training something that’s actually good
35
if you like gpt 6 you are going to love 3.9 flash thinking preview (xhigh)
63
gdm in early stages of rsi meta with banger model releases llm wars heating up welcome back, old friends
1
156
6 1.3 5.1 3.8 4 mention random numbers people know what i’m talking about
1
59
timed to perfection as usual you again, legend
This "recurrent depth" is essentially what's in Sec. 5.3 of the 2015 paper: On Learning to Think: Algorithmic Information Theory for Novel Combinations of Reinforcement Learning Controllers and Recurrent Neural World Models arxiv.org/abs/1511.09249. This paper went beyond the inefficient millisecond by millisecond planning of my 1990 neural world models, addressing planning and reasoning in abstract concept spaces. The 2015 control network C is a prompt engineer that learns to create a chain of thought: to speed up decision making, C learns to query its separate neural world model for abstract reasoning. The prompts and the answers are internal self-generated sequences of vectors that don't have to represent natural language.
74