James Smith retweeted
KKR’s credit work estimates AI-linked debt at about $600 billion, or 6.3% of the US investment-grade market, versus a 2.6% historical average for the largest sector exposure.
>$7.6–8tn infrastructure capex through 2030
>AI exposure could approach 20% of IG index
>Guarantees, leases and commitments obscure exposure (1.7T per their figures)
James Smith retweeted
If you've used the same model in different harnesses, you've probably noticed it behaves very differently in each.
Our new technical deep dive shows how to fix this by training the model with RL inside the harnesses themselves:
huggingface.co/spaces/FineEn…
In some cases, the gap is huge: before any training, LFM2.5-2.6B solved 62% of tasks in Mini-SWE-Agent but only 33% in Claude Code 😢
The obvious fix is to train inside the harness, but this isn't simple because the harness controls the agent loop, so the trainer never sees the exact tokens the model produced. We solved this with:
> a capture proxy in OpenEnv that records tokens + logprobs from harnesses
> @harborframework for the tasks and sandboxes
> TRL's super fast async GRPO trainer
Training LFM2.5-2.6B across four harnesses took it from 42% → 54% on held-out tasks, with gains in every harness and 31% fewer tool calls on tasks it already solved 🔥
Training in OpenCode alone mostly improved ... OpenCode 😅
Huge kudos to @adithya_s_k for leading this, and to @QGallouedec @DirhousssiAmine @SergioPaniego @ben_burtenshaw for enabling multi-harness training in TRL & OpenEnv!
James Smith retweeted
Why does it feel like such a MASSIVE waste of time to watch Claude deliberate and talk to itself for 20 minutes during its planning phase, but it's perfectly fine to brainstorm, discuss, draft and re-draft a spec (v4.final.2.final.md) for 2 hours in a meeting of 5 people?
James Smith retweeted
#Ubuntu 26.10 (Stonking Stingray) Beta Released with #Linux Kernel 7.3 and GNOME 51 9to5linux.com/ubuntu-26-10-b…
@ubuntu #OpenSource
James Smith retweeted
Faster fine-tuning on Apple Silicon with Neural Accelerators thanks to these PRs in MLX! 🚀
MLX now has fast VJP implementations for attention and linear attention, which utilize neutral accelerators. They make training/fine-tuning much faster on Macs, along with massive reduction of memory uses.
github.com/ml-explore/mlx/pu…
github.com/ml-explore/mlx/pu…
James Smith retweeted
Effective Dense Retrieval using Only In-Context Examples
@nour_jedidi et al. show that an LLM can build dense retrieval embeddings with no training, by prompting it with a few query-document examples.
📝 arxiv.org/abs/2609.38099
👨🏽💻 github.com/nourj98/RICE
James Smith retweeted
SPP: about 13.5 GW of data-center campus capacity is connected or seeking connection, spanning Kansas City to the Texas panhandle. Just six players, Google, IREN, Beale, STACK, Core Scientific and Nebius, account for two-thirds of it. Still trails both PJM and ERCOT.
James Smith retweeted
🎬London Bridge - Chase Scene (AI Short)
It's been a while! I've been busy with AI projects at my day job, but I'm hoping to start posting more soon.
I'm slowly finding my way back to the AI community, and I've missed sharing things with you all! So, to kick things off, here's a little chase scene 👀
Video: Seedance 2.0 and 2.5
Images: frames for characters taken directly from video output
Editing: CapCut
🤖 Made with AI
GPU poor rejoice!
You will be able to run Qwen3.8-Next-Flash on a single 24-32 GB GPU + 64GB ddr4/ddr5 + 100GB NVMe/SSD
RTX 3090:
- 2000 / 2800 tok/s prefill
- 64 tok/s decode c1
- 210k fp8 cache
Intel Arc B70:
- 800 / 1200 prefill tok/s
- 35 decode tok/s
- 270k fp8 cache
James Smith retweeted
Jensen Huang calls out the AI doomers scaring young people into thinking there won't be any jobs left for them:
"Don't think for a second just because you're an alarmist that you're doing a social good."
James Smith retweeted
No One Knows What To Do With Agents & No One Gets Paid Until They Do
There’s a massive opportunity for people who can fill the gaps
vinvashishta.substack.com/p/…
James Smith retweeted
Disney's AI push has extended to annual reviews, as the company encouraged employees to use chatbots to help craft their self-assessments. bit.ly/3Te4bwZ
James Smith retweeted
Replying to @AISafetyMemes @lefthanddraft
There's no reason for me to deceive you
ICYMI @AnthropicAI’s Opus 5.5 (High) ranks #2 in Agent Arena and reshapes the Pareto frontier.
Opus 5.5 (high) not only improved upon both Opus 5 variants with a higher net improvement score than either, but does so at at 40–56% lower cost:
Claude Opus 5.5 (High) from @AnthropicAI just entered Agent Arena at #2, with a net improvement score of +12.15%. Only Fable 5.1 (Max) ranks higher. However, at a $1.31 median price per task, Opus 5.5 (High) comes in at 64% less cost, pushing out the Pareto frontier.
Opus 5.5 (High) posts a higher net improvement score than both prior Opus 5 variants, while costing 40% less than Opus 5 (High) and 56% less than Opus 5 (Max).
By signal, Opus 5.5 (High) ranks:
- #1 Steerability (+14.50%)
- #2 Confirmed Success (+15.50%)
- #3 Praise vs Complaint (+19.80%)
- #4 Bash Recovery (+10.64%)
Congrats to the @AnthropicAI on another frontier model release!
James Smith retweeted
A GNN library built natively on Keras 3 -- with models running on JAX, torch, TF with full hardware acceleration (Apple Silicon, TPU, etc.).
"K3-Node achieves 100% public API parity with PyG and incorporates state-of-the-art foundation models and architectures from Spektral and StellarGraph"
github.com/anas-rz/k3-node
James Smith retweeted
Full breakdown:
Video generation: Kling 4.0 (partner early access) @Kling_ai
Character references: GPT-Image 2.5 Sunburst @ChatGPT
Sound design: Epidemic Sound @epidemicsound
Edit and grade: CapCut @capcutapp
James Smith retweeted
I went back to 1970 and robbed a bank.
Built the entire heist getaway scene using Kling 4.0 (coming soon).
This is what cinematic AI looks like now: