AI News Today | Julian Goldie Podcast

Stop Paying For AI...Do This Instead!

11 min · 5 de jul de 2026
Portada del episodio Stop Paying For AI...Do This Instead!

Descripción

Cut 65–69% of Tokens with Claude Code (and Any AI Agent) Using CavemanThe script explains a free, open-source “Caveman” skill that reduces output tokens for Claude Code and other agents (Codex, Gemini, Cursor, Windsurf, Cline, Copilot, and more) by making responses short, blunt, and direct while keeping code, commands, file paths, and error messages unchanged. Installed via a one-line command, it drops a rules text file into each agent’s skills/plugin folder so the agent reads it at the start of every chat. Tests on Fable 5 using five real prompts showed about 69% fewer output tokens (e.g., 1,349 down to 324) and about 37% lower total cost while keeping answers correct; it can be toggled with /caveman and set to light/full/ultra, plus optional tools like Commit/Review/Compress. The video also promotes using Caveman across an agent operating system for compounding savings and mentions the AI Profit Ballroom community and training.00:00 [https://www.youtube.com/watch?v=QNM3lDfaAEc] Cut Tokens With Caveman01:16 [https://www.youtube.com/watch?v=QNM3lDfaAEc&t=76s] Why Token Costs Hurt01:58 [https://www.youtube.com/watch?v=QNM3lDfaAEc&t=118s] How Caveman Works02:21 [https://www.youtube.com/watch?v=QNM3lDfaAEc&t=141s] Rules And Safety03:06 [https://www.youtube.com/watch?v=QNM3lDfaAEc&t=186s] Does Shorter Mean Worse04:00 [https://www.youtube.com/watch?v=QNM3lDfaAEc&t=240s] Install Across Agents05:03 [https://www.youtube.com/watch?v=QNM3lDfaAEc&t=303s] Real Token Savings Tests07:29 [https://www.youtube.com/watch?v=QNM3lDfaAEc&t=449s] Old Replies Vs New08:49 [https://www.youtube.com/watch?v=QNM3lDfaAEc&t=529s] Modes And Commands09:47 [https://www.youtube.com/watch?v=QNM3lDfaAEc&t=587s] Agent OS And Extras10:30 [https://www.youtube.com/watch?v=QNM3lDfaAEc&t=630s] Join The Community11:29 Wrap Up

Comentarios

0

Sé la primera persona en comentar

¡Regístrate ahora y únete a la comunidad de AI News Today | Julian Goldie Podcast!

Empezar

2 meses por 1 €

Después 4,99 € / mes · Cancela cuando quieras

  • Podcasts exclusivos
  • 20 horas de audiolibros / mes
  • Podcast gratuitos

Todos los episodios

635 episodios

Portada del episodio China's Qwen 3.8 Max DESTROYS GPT 5.6?

China's Qwen 3.8 Max DESTROYS GPT 5.6?

Qwen 3.8 Max Preview: Alibaba’s 2.4T-Parameter Open Model Claims #2 Behind “Fable 5”The script covers Alibaba’s announcement of Qwen 3.8 Max Preview, described as a 2.4 trillion parameter open source model available via Alibaba’s Token Plan (Coder/Coder Work) and Qwen Cloud, though the narrator can’t access it yet or find an OpenRouter API and notes it’s available in China. Qwen claims it is “second only to Fable 5,” which the narrator calls unusually bold, especially after China’s Kimi K3 released this week and performed extremely well in the narrator’s personal tests across 50 builds on Goldy Bench. The narrator argues open source Chinese models are rapidly closing gaps with and sometimes overtaking US frontier models, with more releases expected from GLM, DeepSeek GA4 Pro, and Minimax, and promotes AI Profit Boarding for training, community support, and future Qwen testing via Agent OS.

Ayer5 min
Portada del episodio How to Run Kimi K3 for FREE!

How to Run Kimi K3 for FREE!

How to Use Kimi K3 for Free (K3 vs K3 Swarm, Token Limits, Best Settings)The video shows how to access and use Kimi K3 for free at kimi.com by signing in with a free account, switching the default model from K2.6 to K3 (or K3 Swarm), and choosing settings that conserve tokens. The creator recommends avoiding K3 Swarm on a free plan due to token limits, keeping the context window on standard (extra-long is premium), and adjusting thinking effort (standard/high/max) since higher effort uses more tokens. A quick example prompt demonstrates starting a coding task, and the video notes K2.6 can be faster for lightweight requests. It also highlights API testing results, examples of games/3D worlds and video generation using ReMotion, mentions free usage may be paused during peak times, and says Kimi K3 will be open source from July 2026 but requires powerful hardware to run locally.

Ayer3 min
Portada del episodio I Tested Qwen 3.8 So You Don't Have To…

I Tested Qwen 3.8 So You Don't Have To…

Qwen 3.8 vs Fable 5 vs GPT-5.6: Side-by-Side Frontend & Game Tests (Plus Kimi K3 Comparison)The episode reviews Alibaba’s Qwen 3.8 (2.4T parameters) and its claim of being second only to Fable 5, then tests it side by side against Fable 5 and GPT-5.6 across multiple frontend/game-style builds (racing, fireworks, aurora, black hole, neon blaster/racer, cloth simulation, flight sim, promo video, RPG/GTA/Doom, Dragonflight, and Nordic crypt). The creator finds Qwen 3.8 often strong at fun 3D gameplay but frequently weaker in UI/graphics polish versus GPT-5.6 and Fable 5, with mixed results depending on the test. They note Qwen 3.8 is a big step up from Qwen 3.7 but may disappoint relative to its marketing, and they generally prefer Kimi K3 for front-end quality, while also discussing access difficulties and recommending using Coder/Qoda to try Qwen 3.8.

Ayer15 min
Portada del episodio China's Qwen 3.8 Max DESTROYS Kimi K3?

China's Qwen 3.8 Max DESTROYS Kimi K3?

Qwen 3.8 Launch: 2.4T Open-Source Model Claims #2 Behind Fable 5 + What It Means for US vs China AIThe episode covers Qwen 3.8’s newly announced launch, described as an open-source model going live soon but currently only usable for coding in mainland China, with the international site not yet offering access. The host highlights its reported 2.4 trillion parameters and Qwen’s claim that it’s among the most powerful models available, second only to Fable 5 and said to beat models like GPT 5.6 and Kimi K3. The video contrasts restricted access to models such as GPT 5.6 and Fable 5 with free, open-source Chinese models like Kimi, which the host says performs exceptionally well in Goldy Bench tests and coding benchmarks. It also discusses Dean W. Ball’s comments on Kimi K3, open-weight risks, and possible future US regulatory pressure on open models.

Ayer10 min