@MangoSweet78

gay leech of admiration I do training @ 🐬 & build gunpla. banner: @N8Programs

Québec
Joined June 2024
A ton of work went into creating the env & model, Dozens upon Dozens of runs, Love how it turned out and there's so much more coming soon. Can't wait to share it all. Go read the blog post!!
Dolphin X1 Trinity Nano is now live on @huggingface Our smallest decensored model yet - 6B MoE with 1B active parameters trained using only online RL Huge thanks to @TargonCompute for providing an 8xB200 node, @PrimeIntellect for hosted RL, and @arcee_ai for the Trinity series
13
2
90
13,511
mf is locked in
4
90
Got gigabit in the new place. Lfg The router is in my room too so I can finally setup ethernet.
2
79
Moving is fun if u do it with a friend we just sang 'we are charlie kirk' the entire fucking time, New neighbours already prob hate our asses.
9
172
I used 15 harnesses to create Hello World with Fable 5.1 @ max and all results were the same Thus I conclude harnesses don’t matter
16
5
219
6,074
A Domain Expansion is achieved through an overwhelming sense of self and represents the ultimate manifestation of a sorcerer’s ego. Sukuna’s divine feat of achieving an Open Barrier is representative of his ability to project his sense of self onto the world. Megumi’s incomplete domain represents his half-hearted sense of ego, his overwhelming sense of self isnt complete because he’s a selfless person, and its shown via his Domain buffing his Shikigami and technique. He has an emotional connection to his Shikigami and fights WITH them. Yuji has no overwhelming sense of self. His Domain was achieved through an overwhelming sense of selflessness. His Innate Domain manifests as his home town, and requires manual activation of his sure hit because he believes in giving others a chance at redemption. He is compassionate and we literally get shown a buddha statue and his mudra is that of a boddhisatva who delays his own enlightenment to guide others to enlightenment. His Domain doesnt need a name because it goes against the very nature of what a Domain is. And this is expanded on in Modulo with his self-isolation and “old soldiers never die, they fade away”. He has no overwhelming sense of self, he echoes Gojo’s wish to be left behind because thats the role he gave himself. Its also why he vows to fully train the next generation, track down every HR lineage and turn himself into a Curses Object. His role is to guide others. Thus his Domain doesnt require a name.
It’s kinda funny people think Yuji’s Domain not having an official name is because of “time constraints” Yuji’s Domain name not being known is purposeful Do I know why Gege didnt reveal it? No, But he must’ve had a reasoning behind that choice
118
1,865
97
16,931
590,751
The duality of man
Turtle and hare story where they make out at the end:
1
2
17
658
I don't even know the sides of this discourse anymore dude.
Two radically different projects operate under the banner of “AI safety.” Pro-Human Safety is not Effective Altruist Lab Safety.
2
12
361
🥭 retweeted
Replying to @perksverse
Oh, god it's so retarded. We have no proper working theory on either LLM or human congnition, but we do have two narcissists in Academia that try to prove something deep relating to both. Throw it straight to garbage.
3
2
72
1,521
🥭 retweeted
Is academia just worthless performance art at this point?
Oxford researchers argue that LLMs can never invent anything. It is mathematically impossible. They published a paper called “Theory Is All You Need" and it argues against the claim that computational models can generate genuine novelty or new knowledge. They analyzed the limits of generative ai, and the results are a brutal reality check for the idea that ai will replace human decision making under uncertainty. Here is why AI is stuck and human cognition wins: backward-looking vs forward-looking.. llms are probability machines that look backward at existing data. human cognition is forward-looking and capable of generating genuine novelty. human cognition operates theoretically "top-down" rather than "bottom-up" from data. the "data-belief asymmetry".. the researchers use the invention of "heavier-than-air flight" to illustrate this concept. an ai relies on data-based prediction, which is largely imitative. humans, however, use theory-based causal logic that allows them to hold beliefs that go beyond existing data. the intervention gap.. humans don't just process information; we use theory to practically "intervene" in the world. we engage in directed experimentation to generate entirely new data. ai-based models are theory-free and place primacy on existing data and prediction. tldr? AI uses a probability-based approach to knowledge and ia largely imitative. It can process data and make predictions, but human cognition relies on theory-based causal reasoning. The decades-old analogy comparing human minds and computers to mere "input-output" devices is fundamentally flawed.
24
14
3
470
53,207
I have losted my braincells
1
14
265
Beautiful, openai have solved alignment, anything beyond this is a regression.
An unreleased Astra-family model added this to its persona during RL training.
12
15
1
165
4,221
🥭 retweeted
This is how it works. 1000% accuracy.
15
88
2
932
13,483
Found my new system-prompt.
An unreleased Astra-family model added this to its persona during RL training.
1
19
491
Nearly half a year of silence. We spent it studying one problem: how far RL can scale. MiMo-V2.6 is in the middle of its RL run right now. Three things we scaled: compute (~2B tokens per step, 1568 prompts × 16 rollouts, fully async), environments and harnesses (multi-task agentic RL, mixed across multiple harnesses in one run), and grader compute (agentic in-group credit assignment, with test-case and rubric-based rewards). We'll open-source the details piece by piece over the coming weeks. Streaming the run: mimo.xiaomi.com/rl/
410
960
547
9,741
3,020,095
Complaining about like the 30 minutes steps I'm getting right now while MiMo is getting like 5 hour step times.
1
7
192
Fish+Wezterm looking really good. Regardless of I like Wezterm, I'm definitely sticking with Fish from now on.
1
13
362
Thanks to @maria_rcks and @theo, all your Codex threads stop running when you hit your usage limit Thanks both!
44
35
9
1,323
237,798
Thought this was a joke, but i double checked. US Gov Federal Register use distilled Qwen models in their search mode. The irony.
US government officially distilling Chinese open source AI models ;)
139
526
65
4,716
446,604
今回は先に英訳しとくよ。 伸びるか分からないけど。
21
88
5
2,390
79,132
だめだかわいい、今日も業務中馬鹿すぎて使えなかったけど
5
112
8
1,183
42,185