@dddanielwangi
iAccount based inNorth America!
About this account
- Account based in
- North America
- Connected via
- North America App Store
! X says this location may be affected by a proxy or VPN.
Account-level information from X, not a live location or the device used for a specific post.
ICT Sales/GTM 关心实用的 AI:产品、效率与工程 更多⬇️
Joined July 2024
- Tweets7.4K
- Following493
- Followers1.5K
- Likes7.8K
从来没像现在这样如此认可去中心化
最自然、最理性的想法,就是把 personal agent 运行在自己能控制的主机上面
如果大模型不能运行在自己的主机(GP)U 上,那至少 personal agent 以及数据要在自己的主机(CPU) 上
现在这些 muse、 grok bot 如此中心化集中一个人生活各个方面隐私的 personal agent 的模式,不会让人感到不安吗...
模型能力的提升以月为单位,而产品的迭代更是以日为单位,同时还想长期“养”一个越来越懂自己的 personal agent,唯一的途径就是把 agent 和数据都保存在自己可控的主机上啊
there’s something quite awkward about all the personal agents in the current hype cycle
muse, grok bot, instinct, dots, and whatever google, anthropic will come up with
none of them is “mine”
i’d be trusting a vendor for some of the most sensitive data about me and bet on them handling it with extreme responsibility. what if they have a data leakage incident? what if they hired a wrong employee? what if they simply break bad?
i’d also be betting on their model. right now i’ve set up a lot of my stuff in grok bot, but what if their model falls behind? what if another model becomes 10x better on pure intelligence? then i’ll either miss out on a better assistant, or have to bite the bullet and do a full migration
i’d also have to bet on their product. they may not build the features i want. they may not allow me to customize enough. and one day my assistant may show me ads
that feels like way too much betting than what i’d be comfortable with, if i truly want literally everything in my life to go through and be taken care of by an assistant
with all that considered, i’m more bullish on open source personal agents that run on people’s own computers. but i think these open source projects have to break out of the assumption that their users are developers and they can just throw a repo at them and ask them to launch a terminal
a well polished, fool proof, community maintained, vendor agnostic, free personal agent that everyone can run by themselves without locking into a SaaS - that’s what i think many people will need
反正今天reset了,还有一个banked reset 4号过期今天要用,why not
github.com/DanielDaniel2201/…
Fireworks 的销售没听说过 KV Cache
除非有什么过人之处,不然美国就业市场这么宽松吗?
国内类似有卖 Token 的 startup 吗,我也可以做销售😂😂
This quoted post is unavailable.
马斯克大礼包套餐,要是真的就好啦...
One plan for X, Grok, Cursor and Grok Bot.
It's called Xpass.
Premium $8
• Grok Lite + X Premium
• Shared quota
• Cursor included, Grok API
Plus $30
• SuperGrok + X Premium+
• 4x Premium
• Grok Bot, Cursor cloud agents, Bugbot
Super $100
• SuperGrok Plus + X Premium+
• 15x Premium
• 1080p, faster replies, priority access
Ultra $200
• SuperGrok Heavy + X Premium+
• 50x Premium
• Fastest replies, earliest access
A launch promo is in there too: 50% off the first month, dated Oct 9.
4x / 15x / 50x usage is still unconfirmed.
真是太酷了😢,上次研究HBM的时候也想做一个类似的 3D 导览,看来可以继续研究一下做出来!!!
当时只简单做了一下实在太草台了:
吞金兽小子,把我每次system prompt都改一遍!
梁叔叔的cache在天上(因为失效了)失望的看着 5.6 Sol
怎么会写出破坏prompt caching的引用逻辑呢?!悲!!!我们模型的基本素养都没了!
6.1 sol 的 TPS马上就会被提高接近两倍
但问题是,昨天看到推友 TPS 只有十几二十几,就算翻了两倍,也还是很慢啊
而且 OpenAI 你们不是总说有非常多的 compute 吗,在很多用户被 Opus 5.5 吸引过去了之后,还是满足不了需求吗
GPT-6.1 Sol is our most demanded model pretty much ever both across both the API and subscriptions.
Within ChatGPT & Codex, we were under heavy load, but have brough more capacity online and the speed should get much better in the coming hours, reaching almost twice the speed compared to what we served yesterday.
我一直想追求言出法随的 AI 体验,但就算是200 TPS 的 Deepseek Flash 在我的use case里也无法做到(当然,也许跟他的雷霆大思考有关)
我用到的任务就是引用里面的,给 Xcalidraw 画布里只有一层 Hidden Layer 的 MLP,再加上一层 Hidden Layer,就耗时了将近 1 分钟,太久了,但是在 UI 层面还是可以优化的,比如体现一步一步改动的工作量,而非从初始状态直接变成最终状态
Inference 工程还有很多进步空间,因为但凡涉及到 Reasoning 和工具调用,那个延迟真的是无法让人保持专注,必须抽空干点别的
让 ChatGPT 搜罗了最快 TPS 模型厂商,最快的就是 600 左右,也难怪 OpenAI 在 fast ultra fast 方面一直有更新,时间就是体验,时间就是金钱!!
为了让 Deepseek Flash 用好 excalidraw MCP,我变成prompt小子的第一天
不得不说,留存 llm/agent trajectory 真的太重要了,简直是debug的金杯