@Dynamicstm

Dynamic. Interactive. Solutions.

Addis Ababa, Ethiopia
Joined March 2012
Dynamic Stm retweeted
Took a minute to write a few words about security & safety as someone who lived through it all at OpenAI. I hope my thoughts help someone out there.
439
402
327
2,849
1,325,694
Welcome the abstraction with open arm. Embrace the new era of the maker. Average, good enough or exceptional developer, for those of us who had to toil through the old layer of abstraction, it was indeed an exciting time.
My god this is such a good speech that every SWE needs to hear. You know what? Every person should hear it Keep the happy memories, eyes on the reality, be excited about the future. That’s the best that anyone can do
10
Dynamic Stm retweeted
There's nothing kind about letting good people live in a fantasy world that no longer exists. You have to tell them, even if it hurts. Because the sooner they accept reality, the sooner they can adapt to the future.
259
556
108
7,601
695,389
Dynamic Stm retweeted
It's pencils down, people. Writing code by hand is no longer an economically viable skill for most programmers at most companies. But the future of making software has never been brighter. Don't you dare black pill this beautiful moment! youtu.be/vDjW_dRyKXY?si=6Fsf…
377
930
498
8,724
2,792,172
Dynamic Stm retweeted
Creating anything worthwhile is like a marathon, and you must train for it.
83
439
16
3,547
63,185
Dynamic Stm retweeted
We’re still figuring out agentic software, and now @typesafeai just made us rethink our entire AI architecture.
15
17
3
181
14,092
Dynamic Stm retweeted
Nearly half a year of silence. We spent it studying one problem: how far RL can scale. MiMo-V2.6 is in the middle of its RL run right now. Three things we scaled: compute (~2B tokens per step, 1568 prompts × 16 rollouts, fully async), environments and harnesses (multi-task agentic RL, mixed across multiple harnesses in one run), and grader compute (agentic in-group credit assignment, with test-case and rubric-based rewards). We'll open-source the details piece by piece over the coming weeks. Streaming the run: mimo.xiaomi.com/rl/
425
979
554
9,909
3,114,979
In less than six months, the very harness experiments meant to constrain LLM-powered coding agents are already being eclipsed by the agents they were built to constrain.
OK. It's time to rethink this. I've spend the last several weeks working on a harness that tightly constrains the agents to work the way that I want them to work. I set up all kinds of gates, and tests, and tools, and protocols, and ... And while I was heads-down getting that to work, the agents got a LOT better. So much so that when I came up for air, the need for my harness was obviated. Indeed, the need for _any_ but the most liberal of harnesses may be obviated. Just how good these things have gotten blows me away. I have had long debates with grok and codex about the structure of systems -- as if they were senior engineers. They often disagree with me and have their own perspectives. I have, more than once, found myself agreeing with their views. I have not given up on constraints and tooling. Unit testing is still important. So is CRAP and Mutation testing. These tools still find bugs and offer useful constraints, though they can leave scars. However, the agents have gotten so good that I can now give one a very significant task with a few guidelines and it will faithfully implement it. I can walk away for 40 minutes and when I return it will be done. CRAP will be satisfied, Coverage will be high, and Mutation testing complete. The architecture will be clean, and the code will be very good. The end result may not behave perfectly, but it's so close that a couple of tweaks usually puts it into place. What does this mean going forward? I'm not sure. But I'm beginning to think that harnesses should not treat agents as components within a software design.
23
Dynamic Stm retweeted
Today we're launching the Agents API, a brand new way to build Agents in the cloud, backed by the Codex harness. Bring along all your favorite tools and connectors, connect it to any sandbox, and let Astra cook. Can't wait to see what you whip up 👨‍🍳 openai.com/index/introducing…
153
279
100
2,836
2,007,597
Dynamic Stm retweeted
We're publishing our most detailed threat intelligence report to date. It covers how people tried to misuse Claude—for cyberattacks, influence operations, surveillance, biology, and building weapons—and how we found and stopped them. We disrupted every operation in the report, and used the lessons from them to strengthen our safeguards. Where appropriate, we also shared what we found with authorities and other AI companies. These cases are not typical: we’re highlighting some of the most sophisticated misuse we’ve seen. But they’re especially important to discuss, because they show us where AI misuse is headed, where our safeguards work, and where they need to improve. We’re publishing this report so others can spot the same activity on their own platforms, and so we can give the public a clearer view of how emerging threats develop. Read the report: anthropic.com/threat-intelli…
3,232
11,707
4,929
50,527
43,687,330
The entire history of software engineering is one of rising levels of abstraction. This is as it was, is now, and always shall be.
I've never felt this much behind as a programmer. The profession is being dramatically refactored as the bits contributed by the programmer are increasingly sparse and between. I have a sense that I could be 10X more powerful if I just properly string together what has become available over the last ~year and a failure to claim the boost feels decidedly like skill issue. There's a new programmable layer of abstraction to master (in addition to the usual layers below) involving agents, subagents, their prompts, contexts, memory, modes, permissions, tools, plugins, skills, hooks, MCP, LSP, slash commands, workflows, IDE integrations, and a need to build an all-encompassing mental model for strengths and pitfalls of fundamentally stochastic, fallible, unintelligible and changing entities suddenly intermingled with what used to be good old fashioned engineering. Clearly some powerful alien tool was handed around except it comes with no manual and everyone has to figure out how to hold it and operate it, while the resulting magnitude 9 earthquake is rocking the profession. Roll up your sleeves to not fall behind.
82
209
25
2,065
184,660
Dynamic Stm retweeted
Tactical programming is dead Strategic programming has never been more vital
If you're lamenting writing code by hand like something you love is being taken away from you, see if you can fall in love with solving problems instead. Maybe you'll find out that's what you really loved the whole time ❤️
56
77
9
1,510
144,312
Dynamic Stm retweeted
The most important skills for using AI coding agents effectively. Presenting the AI Engineering Skills Map for using coding agents.
306
1,151
115
7,297
739,967
Dynamic Stm retweeted
"The most important skill in prompting is expertise in the domain you’re prompting for" seangoedecke.com/llms-reward… Sometimes the human is the bottleneck. A good reminder LLMs are a force multiplier if you're skilled and to build this domain expertise if you don't have it yet.
35
73
8
555
36,704
Dynamic Stm retweeted
"Don't paste the AI, please" "The world is full of people who don't want to read or think things through. Don't be one of them." dontpastetheai.com
51
153
16
1,160
77,937