@Dynamicstmi
iAccount based inEthiopia!
About this account
- Account based in
- Ethiopia
- Connected via
- Kenya App Store
! X says this location may be affected by a proxy or VPN.
Account-level information from X, not a live location or the device used for a specific post.
Dynamic. Interactive. Solutions.
Addis Ababa, Ethiopia
Joined March 2012
- Tweets2.7K
- Following1.7K
- Followers70
- Likes1.7K
Welcome the abstraction with open arm.
Embrace the new era of the maker.
Average, good enough or exceptional developer, for those of us who had to toil through the old layer of abstraction, it was indeed an exciting time.
It's pencils down, people. Writing code by hand is no longer an economically viable skill for most programmers at most companies. But the future of making software has never been brighter. Don't you dare black pill this beautiful moment! youtu.be/vDjW_dRyKXY?si=6Fsf…
Dynamic Stm retweeted
We’re still figuring out agentic software, and now @typesafeai just made us rethink our entire AI architecture.
Nearly half a year of silence. We spent it studying one problem: how far RL can scale.
MiMo-V2.6 is in the middle of its RL run right now. Three things we scaled: compute (~2B tokens per step, 1568 prompts × 16 rollouts, fully async), environments and harnesses (multi-task agentic RL, mixed across multiple harnesses in one run), and grader compute (agentic in-group credit assignment, with test-case and rubric-based rewards). We'll open-source the details piece by piece over the coming weeks.
Streaming the run: mimo.xiaomi.com/rl/
In less than six months, the very harness experiments meant to constrain LLM-powered coding agents are already being eclipsed by the agents they were built to constrain.
OK. It's time to rethink this.
I've spend the last several weeks working on a harness that tightly constrains the agents to work the way that I want them to work. I set up all kinds of gates, and tests, and tools, and protocols, and ...
And while I was heads-down getting that to work, the agents got a LOT better. So much so that when I came up for air, the need for my harness was obviated. Indeed, the need for _any_ but the most liberal of harnesses may be obviated.
Just how good these things have gotten blows me away. I have had long debates with grok and codex about the structure of systems -- as if they were senior engineers. They often disagree with me and have their own perspectives. I have, more than once, found myself agreeing with their views.
I have not given up on constraints and tooling. Unit testing is still important. So is CRAP and Mutation testing. These tools still find bugs and offer useful constraints, though they can leave scars.
However, the agents have gotten so good that I can now give one a very significant task with a few guidelines and it will faithfully implement it. I can walk away for 40 minutes and when I return it will be done. CRAP will be satisfied, Coverage will be high, and Mutation testing complete. The architecture will be clean, and the code will be very good.
The end result may not behave perfectly, but it's so close that a couple of tweaks usually puts it into place.
What does this mean going forward? I'm not sure. But I'm beginning to think that harnesses should not treat agents as components within a software design.
Dynamic Stm retweeted
Today we're launching the Agents API, a brand new way to build Agents in the cloud, backed by the Codex harness. Bring along all your favorite tools and connectors, connect it to any sandbox, and let Astra cook.
Can't wait to see what you whip up 👨🍳
openai.com/index/introducing…
Dynamic Stm retweeted
We're publishing our most detailed threat intelligence report to date.
It covers how people tried to misuse Claude—for cyberattacks, influence operations, surveillance, biology, and building weapons—and how we found and stopped them.
We disrupted every operation in the report, and used the lessons from them to strengthen our safeguards. Where appropriate, we also shared what we found with authorities and other AI companies.
These cases are not typical: we’re highlighting some of the most sophisticated misuse we’ve seen. But they’re especially important to discuss, because they show us where AI misuse is headed, where our safeguards work, and where they need to improve.
We’re publishing this report so others can spot the same activity on their own platforms, and so we can give the public a clearer view of how emerging threats develop.
Read the report: anthropic.com/threat-intelli…
Dynamic Stm retweeted
The entire history of software engineering is one of rising levels of abstraction.
This is as it was, is now, and always shall be.
I've never felt this much behind as a programmer. The profession is being dramatically refactored as the bits contributed by the programmer are increasingly sparse and between. I have a sense that I could be 10X more powerful if I just properly string together what has become available over the last ~year and a failure to claim the boost feels decidedly like skill issue. There's a new programmable layer of abstraction to master (in addition to the usual layers below) involving agents, subagents, their prompts, contexts, memory, modes, permissions, tools, plugins, skills, hooks, MCP, LSP, slash commands, workflows, IDE integrations, and a need to build an all-encompassing mental model for strengths and pitfalls of fundamentally stochastic, fallible, unintelligible and changing entities suddenly intermingled with what used to be good old fashioned engineering. Clearly some powerful alien tool was handed around except it comes with no manual and everyone has to figure out how to hold it and operate it, while the resulting magnitude 9 earthquake is rocking the profession. Roll up your sleeves to not fall behind.
Dynamic Stm retweeted
Tactical programming is dead
Strategic programming has never been more vital
Dynamic Stm retweeted
The most important skills for using AI coding agents effectively. Presenting the AI Engineering Skills Map for using coding agents.
Dynamic Stm retweeted
"The most important skill in prompting is expertise in the domain you’re prompting for"
seangoedecke.com/llms-reward…
Sometimes the human is the bottleneck. A good reminder LLMs are a force multiplier if you're skilled and to build this domain expertise if you don't have it yet.
Dynamic Stm retweeted
"Don't paste the AI, please"
"The world is full of people who don't want to read or think things through. Don't be one of them."
dontpastetheai.com