AI News Today | Julian Goldie Podcast

Ornith-1.0 is INSANE (FREE + Local + Open Source)!

9 min · 28 jun 2026
aflevering Ornith-1.0 is INSANE (FREE + Local + Open Source)! artwork

Beschrijving

Ornif 1.0: Free Self-Learning Local Coding Model (Runs on Ollama + Works with Hermes Agents)The video tests Ornif 1.0, a new free self-learning local open coding model from Deep Reinforce, showing benchmark comparisons where it outperforms models like Gemma 4 31B and even compares well against larger flagship options. The model can be downloaded via Ollama with a simple command and is integrated into the creator’s agent operating system, including Hermes profiles for agentic workflows. Ornif’s key idea is “self-scaffolding”: it writes its own plans/instructions, generates multiple solution rollouts, gets graded, and improves both planning and coding through reinforcement learning. The creator highlights a gallery of apps, games, tools, and visual outputs made with Ornif, plus local and frontier leaderboards, and demonstrates local offline systems like a Hermes engine, agent Kanban orchestration, and local chat, emphasizing local AI benefits like being free, private, fast, and improving.00:00 [https://www.youtube.com/watch?v=6VDFEMc-xUs] Meet Ornif 1.000:34 [https://www.youtube.com/watch?v=6VDFEMc-xUs&t=34s] Setup With Ollama00:45 [https://www.youtube.com/watch?v=6VDFEMc-xUs&t=45s] What Ornif Is01:04 [https://www.youtube.com/watch?v=6VDFEMc-xUs&t=64s] Real Build Examples01:51 [https://www.youtube.com/watch?v=6VDFEMc-xUs&t=111s] Benchmark Reality Check02:11 [https://www.youtube.com/watch?v=6VDFEMc-xUs&t=131s] Self Improving Framework03:02 [https://www.youtube.com/watch?v=6VDFEMc-xUs&t=182s] Reinforcement Learning Loop03:52 [https://www.youtube.com/watch?v=6VDFEMc-xUs&t=232s] Self Scaffolding Explained05:31 [https://www.youtube.com/watch?v=6VDFEMc-xUs&t=331s] Local Leaderboards Demos06:21 [https://www.youtube.com/watch?v=6VDFEMc-xUs&t=381s] Agent OS Integrations07:32 [https://www.youtube.com/watch?v=6VDFEMc-xUs&t=452s] Model Comparisons Use Cases08:05 [https://www.youtube.com/watch?v=6VDFEMc-xUs&t=485s] Join The Community09:05 [https://www.youtube.com/watch?v=6VDFEMc-xUs&t=545s] Final Thanks

Reacties

0

Wees de eerste die een reactie plaatst

Meld je nu aan en word lid van de AI News Today | Julian Goldie Podcast community!

Probeer gratis

Probeer 14 dagen gratis

€ 9,99 / maand na proefperiode. · Elk moment opzegbaar

  • Podcasts die je alleen op Podimo hoort
  • 20 uur luisterboeken / maand
  • Gratis podcasts

Alle afleveringen

636 afleveringen

aflevering China’s Qwen 3.8 VS Fable 5! artwork

China’s Qwen 3.8 VS Fable 5!

Qwen 3.8 vs GPT-5.6 Soul vs Fable 5 vs Kimi K3: 23 Game Builds Tested (GoldyBench Results)The video compares Qwen 3.8, GPT-5.6 Soul, Fable 5, and Kimi K3 across 23 side-by-side builds on GoldyBench, focusing on game and interactive demo quality such as Outrun, Skyrim-like RPGs, a flight simulator, GTA-style gameplay, Doom-style shooters, physics tests, and landing-page design. The host notes Qwen 3.8 is a massive 2.4T-parameter model and, in these tests, often produces smooth, fun, highly detailed outputs, sometimes with weaker UI, lighting, or buggy controls, while Kimi K3 frequently excels in front-end feel and physics interactions. Fable 5 remains the host’s favorite overall, with Kimi K3 generally second, and GPT-5.6 Soul often competitive but inconsistent. The takeaway is results are a mixed bag, so viewers should test models themselves, with links to GoldyBench and the AI Profit Ballroom community and trainings.

21 jul 202616 min
aflevering China's Qwen 3.8 Max DESTROYS GPT 5.6? artwork

China's Qwen 3.8 Max DESTROYS GPT 5.6?

Qwen 3.8 Max Preview: Alibaba’s 2.4T-Parameter Open Model Claims #2 Behind “Fable 5”The script covers Alibaba’s announcement of Qwen 3.8 Max Preview, described as a 2.4 trillion parameter open source model available via Alibaba’s Token Plan (Coder/Coder Work) and Qwen Cloud, though the narrator can’t access it yet or find an OpenRouter API and notes it’s available in China. Qwen claims it is “second only to Fable 5,” which the narrator calls unusually bold, especially after China’s Kimi K3 released this week and performed extremely well in the narrator’s personal tests across 50 builds on Goldy Bench. The narrator argues open source Chinese models are rapidly closing gaps with and sometimes overtaking US frontier models, with more releases expected from GLM, DeepSeek GA4 Pro, and Minimax, and promotes AI Profit Boarding for training, community support, and future Qwen testing via Agent OS.

Gisteren5 min
aflevering How to Run Kimi K3 for FREE! artwork

How to Run Kimi K3 for FREE!

How to Use Kimi K3 for Free (K3 vs K3 Swarm, Token Limits, Best Settings)The video shows how to access and use Kimi K3 for free at kimi.com by signing in with a free account, switching the default model from K2.6 to K3 (or K3 Swarm), and choosing settings that conserve tokens. The creator recommends avoiding K3 Swarm on a free plan due to token limits, keeping the context window on standard (extra-long is premium), and adjusting thinking effort (standard/high/max) since higher effort uses more tokens. A quick example prompt demonstrates starting a coding task, and the video notes K2.6 can be faster for lightweight requests. It also highlights API testing results, examples of games/3D worlds and video generation using ReMotion, mentions free usage may be paused during peak times, and says Kimi K3 will be open source from July 2026 but requires powerful hardware to run locally.

Gisteren3 min
aflevering I Tested Qwen 3.8 So You Don't Have To… artwork

I Tested Qwen 3.8 So You Don't Have To…

Qwen 3.8 vs Fable 5 vs GPT-5.6: Side-by-Side Frontend & Game Tests (Plus Kimi K3 Comparison)The episode reviews Alibaba’s Qwen 3.8 (2.4T parameters) and its claim of being second only to Fable 5, then tests it side by side against Fable 5 and GPT-5.6 across multiple frontend/game-style builds (racing, fireworks, aurora, black hole, neon blaster/racer, cloth simulation, flight sim, promo video, RPG/GTA/Doom, Dragonflight, and Nordic crypt). The creator finds Qwen 3.8 often strong at fun 3D gameplay but frequently weaker in UI/graphics polish versus GPT-5.6 and Fable 5, with mixed results depending on the test. They note Qwen 3.8 is a big step up from Qwen 3.7 but may disappoint relative to its marketing, and they generally prefer Kimi K3 for front-end quality, while also discussing access difficulties and recommending using Coder/Qoda to try Qwen 3.8.

Gisteren15 min