Building @OpenCompanion and open-source tools from home. Disability shapes the setup, not the ambition.
Poland
Joined May 2026
- Tweets210
- Following266
- Followers17
- Likes423
Kacper Nowicki / flurris retweeted
Repeat after me:
👏 inference costs 👏 are not 👏 the same as API prices 👏
A codex reset costs Tibo 8 billion dollars.
They have 16m active users, most users spend $20 or $100 monthly, which means $50/week on average.
So one reset price is approx. $12/customer.
They provide 40x more usage for the price, so $480/customer in inference for one week of usage.
That brings total of $8b total for 16m users.
Kudos to OpenAI and Tibo.
🚨BREAKING
Both Astra and Fable 6 are not allowed to be used by the public due to the high risk of hacking your neighbours' printers.
Kacper Nowicki / flurris retweeted
a general AI system solving a bunch of problems that humanity collectively couldn’t solve is the definition of ASI, no?
Kacper Nowicki / flurris retweeted
We're starting to leave the territory where you'd test an LLM by e.g. "create an svg of pelican on a bicycle". As one idea to generalize it, I was interested what Opus 5 would do if I gave it the first paragraph of the Lord of the Rings, a 1M token budget (~$10) and asked for three js render of it. Opus went off for ~2 hours and wrote 5500 lines of code that (procedurally) rendered the story. It's kind of janky but fun. But it's a bit mindboggling that the LLM has to place and orchestrate various polygon assets in (x,y,z) coordinates and write code that animates it all, and that it even does anything at all.
I also like this kind of examples because no one in their right mind would ever spend the time to write something this custom but LLMs have all the stamina and patience in the world, so it's an example where we go from "no one would ever do this" to "sure, why not, it's ~free". There might be a lot more. But I'm excited about creating hyper custom worlds that you can imagine dropping players into, e.g. here to participate in the LoTR story as a spectator NPC, or one of the characters, or etc. Something like an ephemeral GTA of X on demand.
Last thought is that the domain of worlds/games exposes a weakness in LLMs: they can't easily audit their work because they aren't able to efficiently and natively perceive videos or play games within them. Here, Opus 5 had to very slowly and painstakingly take screenshots at different points, and it messed up a few times and created a bunch of jank. An example of raw capability (multimodal, gameplay) that I think is still quite lacking.
Kacper Nowicki / flurris retweeted
I am extremely terrified of the rapid pace of AI development
Specifically I am terrified that we will end up with extreme concentration and abuse of power
Therefore, I request that Anthropic and OpenAI slow down and pause AI for 6 months
This will help everyone else catch up
Thank you 🙏
Just heard from a European that big American companies don't use AI for coding because the companies are dependable.
Kacper Nowicki / flurris retweeted
Replying to @Maya_BsAs @johnennis
There is only one word for people who believe LLMs should have the right to overrule a human's moral, ethical, rational, and reasoned decision making... delusional.
Kacper Nowicki / flurris retweeted
1/ Today, we’re excited to introduce Lucy 2.5.
Lucy represents a paradigm shift in world models, not just in how generated worlds look, but in how we interact with them.
A thread on what makes it a paradigm shift 👇
People who use only one lab's LLMs are genuinely just as disgusting as Apple fangirls. Please, every lab has different strengths, and if you isolate yourself... you harm yourself and your thinking. Stop riding Codex, stop riding Claude, use both of them, try Muse, try Kimi. It's really not that deep. You don't have to be anyone's warrior (especially when you pay them lol).
Kacper Nowicki / flurris retweeted
Getting rate limited is fine
Killing the task halfway through is not
claude code should just let the task finish like codex does
If you don't want to read the code, YOU HAVE TO TEST IT. JUST RUN DIFFERENT EDGE CASES. LEARN HOW TO TEST.