AI News Today | Julian Goldie Podcast

Hermes Agent 2.0 is INSANE!

7 min · 6. heinä 2026
jakson Hermes Agent 2.0 is INSANE! kansikuva

Kuvaus

Hermes Agent v0.18 “Judgment Release” Explained: Mixture of Agents, /learn, /journey & Background Fan-OutHermes Agent released v0.18, the “Judgment Release,” introducing key upgrades the host explains and demonstrates, including Mixture of Agents (combining multiple models for stronger one-shot builds), the /learn command (teach Hermes a new skill from a link and save it for future sessions, with notes logged to an Obsidian vault), and /journey (a timeline recap of what Hermes has learned that can be edited or deleted). The update also adds background fan-out via Delegate Task, letting multiple sub-agents run in the background so chat isn’t blocked, with an example of 12 agents building a website. Additional improvements include cheaper self-improvement routing, security updates, presets for mixtures, and better goal-mode contract completion judged by a judge agent. The host shows four ways to update to v0.18 and promotes their custom agent operating system and AI Profit Boardroom community.00:00 [https://www.youtube.com/watch?v=dxDMAJ9bLmg] V0.18 Overview00:21 [https://www.youtube.com/watch?v=dxDMAJ9bLmg&t=21s] Mixture of Agents01:31 [https://www.youtube.com/watch?v=dxDMAJ9bLmg&t=91s] Learn Skills from Links02:48 [https://www.youtube.com/watch?v=dxDMAJ9bLmg&t=168s] Journey Memory Timeline03:31 [https://www.youtube.com/watch?v=dxDMAJ9bLmg&t=211s] Background Fan Out04:37 [https://www.youtube.com/watch?v=dxDMAJ9bLmg&t=277s] Other Improvements and Goal Mode05:37 [https://www.youtube.com/watch?v=dxDMAJ9bLmg&t=337s] Agent OS vs Desktop App06:09 [https://www.youtube.com/watch?v=dxDMAJ9bLmg&t=369s] How to Update06:45 [https://www.youtube.com/watch?v=dxDMAJ9bLmg&t=405s] Get the Agent OS07:15 [https://www.youtube.com/watch?v=dxDMAJ9bLmg&t=435s] Community and Wrap Up

Kommentit

0

Ole ensimmäinen kommentoija

Rekisteröidy nyt ja liity AI News Today | Julian Goldie Podcast-yhteisöön!

Aloita maksutta

14 vrk ilmainen kokeilu

Kokeilun jälkeen 7,99 € / kuukausi. · Peru milloin tahansa

  • Podimon podcastit
  • 20 kuunteluaikaa / kuukausi
  • Lataa offline-käyttöön

Kaikki jaksot

635 jaksot

jakson China's Qwen 3.8 Max DESTROYS GPT 5.6? kansikuva

China's Qwen 3.8 Max DESTROYS GPT 5.6?

Qwen 3.8 Max Preview: Alibaba’s 2.4T-Parameter Open Model Claims #2 Behind “Fable 5”The script covers Alibaba’s announcement of Qwen 3.8 Max Preview, described as a 2.4 trillion parameter open source model available via Alibaba’s Token Plan (Coder/Coder Work) and Qwen Cloud, though the narrator can’t access it yet or find an OpenRouter API and notes it’s available in China. Qwen claims it is “second only to Fable 5,” which the narrator calls unusually bold, especially after China’s Kimi K3 released this week and performed extremely well in the narrator’s personal tests across 50 builds on Goldy Bench. The narrator argues open source Chinese models are rapidly closing gaps with and sometimes overtaking US frontier models, with more releases expected from GLM, DeepSeek GA4 Pro, and Minimax, and promotes AI Profit Boarding for training, community support, and future Qwen testing via Agent OS.

Eilen5 min
jakson How to Run Kimi K3 for FREE! kansikuva

How to Run Kimi K3 for FREE!

How to Use Kimi K3 for Free (K3 vs K3 Swarm, Token Limits, Best Settings)The video shows how to access and use Kimi K3 for free at kimi.com by signing in with a free account, switching the default model from K2.6 to K3 (or K3 Swarm), and choosing settings that conserve tokens. The creator recommends avoiding K3 Swarm on a free plan due to token limits, keeping the context window on standard (extra-long is premium), and adjusting thinking effort (standard/high/max) since higher effort uses more tokens. A quick example prompt demonstrates starting a coding task, and the video notes K2.6 can be faster for lightweight requests. It also highlights API testing results, examples of games/3D worlds and video generation using ReMotion, mentions free usage may be paused during peak times, and says Kimi K3 will be open source from July 2026 but requires powerful hardware to run locally.

Eilen3 min
jakson I Tested Qwen 3.8 So You Don't Have To… kansikuva

I Tested Qwen 3.8 So You Don't Have To…

Qwen 3.8 vs Fable 5 vs GPT-5.6: Side-by-Side Frontend & Game Tests (Plus Kimi K3 Comparison)The episode reviews Alibaba’s Qwen 3.8 (2.4T parameters) and its claim of being second only to Fable 5, then tests it side by side against Fable 5 and GPT-5.6 across multiple frontend/game-style builds (racing, fireworks, aurora, black hole, neon blaster/racer, cloth simulation, flight sim, promo video, RPG/GTA/Doom, Dragonflight, and Nordic crypt). The creator finds Qwen 3.8 often strong at fun 3D gameplay but frequently weaker in UI/graphics polish versus GPT-5.6 and Fable 5, with mixed results depending on the test. They note Qwen 3.8 is a big step up from Qwen 3.7 but may disappoint relative to its marketing, and they generally prefer Kimi K3 for front-end quality, while also discussing access difficulties and recommending using Coder/Qoda to try Qwen 3.8.

Eilen15 min
jakson China's Qwen 3.8 Max DESTROYS Kimi K3? kansikuva

China's Qwen 3.8 Max DESTROYS Kimi K3?

Qwen 3.8 Launch: 2.4T Open-Source Model Claims #2 Behind Fable 5 + What It Means for US vs China AIThe episode covers Qwen 3.8’s newly announced launch, described as an open-source model going live soon but currently only usable for coding in mainland China, with the international site not yet offering access. The host highlights its reported 2.4 trillion parameters and Qwen’s claim that it’s among the most powerful models available, second only to Fable 5 and said to beat models like GPT 5.6 and Kimi K3. The video contrasts restricted access to models such as GPT 5.6 and Fable 5 with free, open-source Chinese models like Kimi, which the host says performs exceptionally well in Goldy Bench tests and coding benchmarks. It also discusses Dean W. Ball’s comments on Kimi K3, open-weight risks, and possible future US regulatory pressure on open models.

Eilen10 min