@mnemebraini
iAccount based inSoutheast Asia!
About this account
- Account based in
- Southeast Asia
- Connected via
- Southeast Asia Android App
! X says this location may be affected by a proxy or VPN.
Account-level information from X, not a live location or the device used for a specific post.
Building MnemeBrain Belief memory for Al agents Contradictions. Evidence Revision Think TCP/IP for agent memory
AI
Joined December 2017
- Tweets166
- Following1
- Followers45
- Likes347
Pinned Tweet
Every AI memory system I tested fails at contradiction detection.
- Structured_memory 36%
- Mem0: 29%
- OpenAI RAG: 0%
We built a benchmark to understand why.
The problem isn't implementation.
𝐓𝐡𝐞 𝐀𝐈 𝐬𝐭𝐚𝐜𝐤 𝐢𝐬 𝐦𝐢𝐬𝐬𝐢𝐧𝐠 𝐚 𝐥𝐚𝐲𝐞𝐫. 🧵
$MNEMEBRAIN Successfully launch on @bankrbot
CA : 0x798c2587f1c7132100770db5ad10aba7f576eba3
CC : @Teknium
bankr.bot/terminal/trade?out…
Building MnemeBrain Belief memory for Al agents Contradictions. Evidence Revision Think TCP/IP for agent memory @bankrbot Launch a token with name”mnemebrain” symbol “MNEMEBRAIN” on robinhood chain full unlock.
attached this image for profil
Mnemebrain retweeted
BMB is open.
The belief layer for AI agents.
Not a better memory library.
The TCP/IP moment for agent memory.
github.com/mnemebrain/mnemeb…
Claude Code just added structured auto-memory.
That’s a big step forward for agent memory.
But it also highlights something important:
most AI systems still treat memory as stored notes.
The harder problem is belief maintenance. 🧵
1/10
This is the layer we’re exploring with MnemeBrain.
Evidence
→ Belief node
→ Truth state
→ Confidence + decay
Belief maintenance instead of retrieval.
9/10
Mnemebrain retweeted
Very excited about the "First Proof" challenge. I believe novel frontier research is perhaps the most important way to evaluate capabilities of the next generation of AI models.
We have run our internal model with limited human supervision on the ten proposed problems. The problems require expertise in their respective domains and are not easy to verify; based on feedback from experts, we believe at least six solutions (2, 4, 5, 6, 9, 10) have a high chance of being correct, and some further ones look promising.
We will only publish the solution attempts after midnight (PT), per the authors' guidance - the sha256 hash of the PDF is d74f090af16fc8a19debf4c1fec11c0975be7d612bd5ae43c24ca939cd272b1a .
This was a side-sprint executed in a week mostly by querying one of the models we're currently training; as such, the methodology we employed leaves a lot to be desired. We didn't provide proof ideas or mathematical suggestions to the model during this evaluation; for some solutions, we asked the model to expand upon some proofs, per expert feedback. We also manually facilitated a back-and-forth between this model and ChatGPT for verification, formatting and style. For some problems, we present the best of a few attempts according to human judgement.
We are looking forward to more controlled evaluations in the next round!
1stproof.org #1stProof
Mnemebrain retweeted
It was really exciting seeing a model we’re currently training tackling these frontier math research problems. With each checkpoint it became so much better at math research. I can’t wait till such intelligence is in the hands of ChatGPT users.
Very excited about the "First Proof" challenge. I believe novel frontier research is perhaps the most important way to evaluate capabilities of the next generation of AI models.
We have run our internal model with limited human supervision on the ten proposed problems. The problems require expertise in their respective domains and are not easy to verify; based on feedback from experts, we believe at least six solutions (2, 4, 5, 6, 9, 10) have a high chance of being correct, and some further ones look promising.
We will only publish the solution attempts after midnight (PT), per the authors' guidance - the sha256 hash of the PDF is d74f090af16fc8a19debf4c1fec11c0975be7d612bd5ae43c24ca939cd272b1a .
This was a side-sprint executed in a week mostly by querying one of the models we're currently training; as such, the methodology we employed leaves a lot to be desired. We didn't provide proof ideas or mathematical suggestions to the model during this evaluation; for some solutions, we asked the model to expand upon some proofs, per expert feedback. We also manually facilitated a back-and-forth between this model and ChatGPT for verification, formatting and style. For some problems, we present the best of a few attempts according to human judgement.
We are looking forward to more controlled evaluations in the next round!
1stproof.org #1stProof
Mnemebrain retweeted
It's been a huge month for Codex.
5.3, Spark, Codex app, OpenClaw.
We're accelerating. Looking for top people in:
- Full stack Typescript
- Design engineering
- Windows experience+distribution
- React+Node performance
- Crazy advanced git
- Agent orchestration
- Remote codex
- Mobile
DM me with a link to what you've built
Mnemebrain retweeted
@natseckatrina who leads some of our national security work is going to jump in to answer some of your questions
Mnemebrain retweeted
@boazbaraktcs is also going to help out with answers!