Pinned Tweet
While everyone's trying to generate MRR,
I had my @openclaw help me build a side-scroller retro video game encapsulating how my fiancée and I met in Toronto...for our wedding website
(sound on! 🔉)
Astra helped me win an argument i had with my fianceé on whether glasses dry faster right-side-up or upside-down after dishwashing by simulating the physics of the water molecules and nearby air humidity
thank you @OpenAI ❤️🥹
Barron Roth retweeted
Just in time for you Labor Day weekend plans! Let us know how you use it!
Gemini Spark can edit and curate photo albums, create shared collections, turn photos into calendar events, and handle other Google Photos tasks for AI Pro and Ultra subscribers.
spr.ly/6018BGBOJm
Barron Roth retweeted
Two new Gemini models are here to help scale your AI agents and secure code:
🔘 3.8 Flash: our most intelligent model yet with significant gains from 3.7 Flash across software engineering, agentic tasks, and multi-step reasoning.
🔘 3.8 Flash Cyber: our most capable cybersecurity model with frontier-level vulnerability detection and automated patching.
Barron Roth retweeted
Replying to @vepsi__
Oh yeah, we get that question a lot. While other products like Search and Gmail use Gemini models, the 1B here refers to people choosing the Gemini app across Web (gemini.google.com), Android, iOS, and Gemini in Chrome.
it's actually inexplicable how much money my Hermes agent just constantly pays for itself and then some
having instant access to a capable harness from my phone lets me fire off any research or admin task that I just don't have time to do has resulted in SO many solved financial headaches
1. +$550/yr: figured out i have an LLC handler service that wasn't necessary anymore since he could file all the forms
2. +$4k/yr: found me cheaper, better condo insurance after my bank threatened me with fines for not having enough coverage. he negotiated with brokers until he was happy with our options and didn't give up until my bank sent me a formal letter of acknowledgement
3. +$200/yr: managed an API-based transition of 2TB of important data from Dropbox Pro to Google Drive Pro which has taken 7 weeks to complete
4. +$x in gains: tells me exactly when and how much stock i need to sell to pay my wedding bills without going overboard, keeping me maximally invested
5. +6 Biz Class Flights: credit card churning advisor. churning is a full-time job and requires a PhD in AmEx, so my agent basically manages my CC portfolio and tells me when to apply to new cards and what new SUBs are available
and these are just the ones i can remember since december. i NEVER understand when people tell me they have no use for an agent
Barron Roth retweeted
We have started our most ambitious pre-training run yet, for Gemini 4, and are excited by the progress : )
ngl been using this nonstop. here's my potato tier list. fight me
uses @NanoBanana-2-lite to generate the images SUPER fast. costs about $1 per tierlist in tokens
Replying to @sama
tier list generator fully with 5.6 sol
me and my family love tier lists but oftentimes they don’t exist for the categories we want to rank
this solves that
this made me curious about SOUL.md's impact on personal agents, so i ran an eval
turns out, a SOUL.md can make Hermes responses 12% more concise, and 11% more effective
the more you know! alfred.barronroth.com/adhd-p…
Anthropic recently cut Claude Code’s system prompt by 80%.
@trq212 explains why:
“As the models have gotten smarter, they need less direction, fewer constraints, and fewer examples.
The examples are constraining it because now it’s like, ‘Oh, you want things like this example.’ If you remove the examples, it can actually be more free-form.
All this is to say that you want to trim your context [when a new model is released]. The latest models often need more room to run.”
📌 Watch the full episode here: youtu.be/aVO6E181cNU
Barron Roth retweeted
Everyone should develop their "personal eval set" for AI models: a few tasks that are actually relevant to your day-to-day work/life
The industry benchmarks help but they might not reflect what will make it actually useful to you
You find the model's capability boundary by poking at it & bumping into it for fun
Barron Roth retweeted
Hard not to be giddy hearing a robotics ceo show you their incredible little robot
Help us, help you 💚
Want unreleased @GeminiApp features before anyone else? Love breaking, testing, and shaping new tech?
We're opening a limited number of slots for power users to join the Gemini Trusted Tester program.
Sign up here: goo.gle/4onCRHq
Barron Roth retweeted
hermes gets expressive tts: gemini personas and emotion tags
@iamBarronRoth, who works on gemini at google, added a set of gemini text-to-speech features to nous research's hermes agent. the goal: make the agent's spoken replies sound natural instead of flat and robotic. it's already merged into hermes's main branch
what's new
• directors notes (persona file). write a plain text or markdown file describing how the voice should sound – tone, pacing, character – and hermes attaches it to every tts request. the description shapes the voice but never shows up in the chat reply. set it once, and every spoken response follows the same style
• expressive audio tags. gemini 3.1's tts model supports inline tags that cue emotion and delivery. turn this on and hermes inserts them automatically via a separate rewrite step before generating audio. that step can run on a cheaper, faster model so it doesn't tie up your main chat model. it only activates on gemini 3.1 tts models, and if the rewrite fails, hermes falls back to plain audio instead of erroring
• native telegram voice notes. telegram replies now arrive as ogg/opus voice messages instead of mp3 attachments
what you can build
- an audiobook or long-form narration agent that shifts tone between characters and calm narration, instead of reading every line in the same voice
- a language-tutor bot that actually sounds encouraging, slows down on corrections, and emphasizes the right syllables
- a podcast or news-summary agent that reads with emphasis and pacing instead of a flat tts dump
the details
everything is opt-in and backwards compatible – existing tts setups are unaffected unless you enable it. the persona and audio-tag commits were cherry-picked onto main with roth's authorship preserved
follow @thehypedotnews for 24/7 ai news, analysis and breakdowns
Replying to @Teknium
@Teknium hey 👋
i work on Gemini at Google
i added a few unique Gemini TTS features to Hermes to make the experience much more fun: expressive audio tags and “directors notes”.
this lets the hermes agent sound incredibly lifelike in its tts responses
check out my PR github.com/NousResearch/herm…
Merged 🫡🫡
Replying to @Teknium
@Teknium hey 👋
i work on Gemini at Google
i added a few unique Gemini TTS features to Hermes to make the experience much more fun: expressive audio tags and “directors notes”.
this lets the hermes agent sound incredibly lifelike in its tts responses
check out my PR github.com/NousResearch/herm…
Replying to @0xbyt4
@0xbyt4 hey 👋
i work on Gemini at Google
added a few unique Gemini TTS features to Hermes to make the experience much more fun: expressive audio tags and "directors notes"
this lets the agent sound incredibly lifelike in its responses
would love your 👀 on my PR github.com/NousResearch/herm…