AI News Today | Julian Goldie Podcast

Gemma 4 is now 90% FASTER + FREE + Local 🤯

10 min · 3. juli 2026
episode Gemma 4 is now 90% FASTER + FREE + Local 🤯 cover

Beskrivelse

Gemma 4 Just Got Up to 90% Faster on Mac (Ollama + MLX)Google’s Gemma 4 has a new update that makes it up to nearly 90% faster on a Mac when run through Ollama using MLX, enabling much faster local model performance. The script explains that Gemma 4 is free on Ollama and MLX is a free open-source project, and that Ollama now supports MLX models directly so you can install via a simple command after updating. The creator reports seeing roughly 60% faster speeds in many tests, with a jump from about 50 to 95 tokens/sec in a shown comparison, and says coding prompts can hit the 90% boost. They also show Gemma 4 integrated into their Agent OS (with Hermes), letting it build and preview simple apps locally and autonomously.00:00 [https://www.youtube.com/watch?v=cTXdg-V4MII] Gemma 4 Speed Boost00:49 [https://www.youtube.com/watch?v=cTXdg-V4MII&t=49s] MLX and Ollama Setup01:58 [https://www.youtube.com/watch?v=cTXdg-V4MII&t=118s] Agent OS Integration02:33 [https://www.youtube.com/watch?v=cTXdg-V4MII&t=153s] Real Speed Benchmarks04:07 [https://www.youtube.com/watch?v=cTXdg-V4MII&t=247s] Live Build Demo05:27 [https://www.youtube.com/watch?v=cTXdg-V4MII&t=327s] Why It Runs Faster05:56 [https://www.youtube.com/watch?v=cTXdg-V4MII&t=356s] Install in Terminal06:13 [https://www.youtube.com/watch?v=cTXdg-V4MII&t=373s] Projects Built With Gemma07:25 [https://www.youtube.com/watch?v=cTXdg-V4MII&t=445s] Limits and Best Uses07:56 [https://www.youtube.com/watch?v=cTXdg-V4MII&t=476s] Free Local Builds Anytime08:41 [https://www.youtube.com/watch?v=cTXdg-V4MII&t=521s] Agent OS and Boardroom Pitch09:43 [https://www.youtube.com/watch?v=cTXdg-V4MII&t=583s] Community and Support Features10:13 [https://www.youtube.com/watch?v=cTXdg-V4MII&t=613s] What Members Are Building10:41 [https://www.youtube.com/watch?v=cTXdg-V4MII&t=641s] Wrap Up and Thanks

Kommentarer

0

Vær den første til at kommentere

Tilmeld dig nu og bliv en del af AI News Today | Julian Goldie Podcast-fællesskabet!

Kom i gang

1 måned kun 9 kr.

Derefter 99 kr. / måned · Opsig når som helst

  • Podcasts kun på Podimo
  • 20 lydbogstimer pr. måned
  • Gratis podcasts

Alle episoder

636 episoder

episode China’s Qwen 3.8 VS Fable 5! cover

China’s Qwen 3.8 VS Fable 5!

Qwen 3.8 vs GPT-5.6 Soul vs Fable 5 vs Kimi K3: 23 Game Builds Tested (GoldyBench Results)The video compares Qwen 3.8, GPT-5.6 Soul, Fable 5, and Kimi K3 across 23 side-by-side builds on GoldyBench, focusing on game and interactive demo quality such as Outrun, Skyrim-like RPGs, a flight simulator, GTA-style gameplay, Doom-style shooters, physics tests, and landing-page design. The host notes Qwen 3.8 is a massive 2.4T-parameter model and, in these tests, often produces smooth, fun, highly detailed outputs, sometimes with weaker UI, lighting, or buggy controls, while Kimi K3 frequently excels in front-end feel and physics interactions. Fable 5 remains the host’s favorite overall, with Kimi K3 generally second, and GPT-5.6 Soul often competitive but inconsistent. The takeaway is results are a mixed bag, so viewers should test models themselves, with links to GoldyBench and the AI Profit Ballroom community and trainings.

21. juli 202616 min
episode China's Qwen 3.8 Max DESTROYS GPT 5.6? cover

China's Qwen 3.8 Max DESTROYS GPT 5.6?

Qwen 3.8 Max Preview: Alibaba’s 2.4T-Parameter Open Model Claims #2 Behind “Fable 5”The script covers Alibaba’s announcement of Qwen 3.8 Max Preview, described as a 2.4 trillion parameter open source model available via Alibaba’s Token Plan (Coder/Coder Work) and Qwen Cloud, though the narrator can’t access it yet or find an OpenRouter API and notes it’s available in China. Qwen claims it is “second only to Fable 5,” which the narrator calls unusually bold, especially after China’s Kimi K3 released this week and performed extremely well in the narrator’s personal tests across 50 builds on Goldy Bench. The narrator argues open source Chinese models are rapidly closing gaps with and sometimes overtaking US frontier models, with more releases expected from GLM, DeepSeek GA4 Pro, and Minimax, and promotes AI Profit Boarding for training, community support, and future Qwen testing via Agent OS.

I går5 min
episode How to Run Kimi K3 for FREE! cover

How to Run Kimi K3 for FREE!

How to Use Kimi K3 for Free (K3 vs K3 Swarm, Token Limits, Best Settings)The video shows how to access and use Kimi K3 for free at kimi.com by signing in with a free account, switching the default model from K2.6 to K3 (or K3 Swarm), and choosing settings that conserve tokens. The creator recommends avoiding K3 Swarm on a free plan due to token limits, keeping the context window on standard (extra-long is premium), and adjusting thinking effort (standard/high/max) since higher effort uses more tokens. A quick example prompt demonstrates starting a coding task, and the video notes K2.6 can be faster for lightweight requests. It also highlights API testing results, examples of games/3D worlds and video generation using ReMotion, mentions free usage may be paused during peak times, and says Kimi K3 will be open source from July 2026 but requires powerful hardware to run locally.

I går3 min
episode I Tested Qwen 3.8 So You Don't Have To… cover

I Tested Qwen 3.8 So You Don't Have To…

Qwen 3.8 vs Fable 5 vs GPT-5.6: Side-by-Side Frontend & Game Tests (Plus Kimi K3 Comparison)The episode reviews Alibaba’s Qwen 3.8 (2.4T parameters) and its claim of being second only to Fable 5, then tests it side by side against Fable 5 and GPT-5.6 across multiple frontend/game-style builds (racing, fireworks, aurora, black hole, neon blaster/racer, cloth simulation, flight sim, promo video, RPG/GTA/Doom, Dragonflight, and Nordic crypt). The creator finds Qwen 3.8 often strong at fun 3D gameplay but frequently weaker in UI/graphics polish versus GPT-5.6 and Fable 5, with mixed results depending on the test. They note Qwen 3.8 is a big step up from Qwen 3.7 but may disappoint relative to its marketing, and they generally prefer Kimi K3 for front-end quality, while also discussing access difficulties and recommending using Coder/Qoda to try Qwen 3.8.

I går15 min