Ollama
Track Ollama development. Run large language models locally.
About this developer podcast
Ollama turns GitHub activity into short audio updates.
Episodes cover commits, pull requests, issues, and release work.
Use this page to listen, subscribe, and find related developer podcasts.
- Format: public podcast hub
- Feeds: daily and weekly RSS
- Content: recent episodes, archive links, and related shows
This show page is the public hub for the podcast. It has one stable listen URL.
Use it to find RSS feeds, recent episodes, related shows, and older episode links.
If the project is quiet, new episodes may pause until there is useful activity to summarize.
Recent episodes from Ollama
-
Structured Output Meets Speculative Decoding
The MLX runner now supports speculative decoding and grammar-constrained output at the same time, closing a performance gap that previously forced a choice between the two. Supporting CI and parser fixes round out a release focused on…
-
Hardening the MLX Engine and Fixing What Feeds the Scheduler
The MLX engine got a reliability pass, catching silent failures and adding image, audio, and cached-token reporting, while a cluster of scheduler and memory PRs worked to fix inaccurate VRAM and context-length numbers that drive…
-
MLX Grows Up, Model Intent Gets Respected
A cluster of Daniel Hiltgen's PRs matured MLX from an experimental backend into a properly tested, version-pinned component, while a separate fix restored model-authored generation defaults that Ollama's built-in settings had been…
-
Runner Reliability and a Documentation Cleanup Wave
The MLX runner picked up context-length enforcement and repeat-loop protection to match the llama.cpp backend's behavior, while a separate contributor pushed through nine documentation fixes in a single morning. Together they show a…
-
Proxy Fix and a Documentation Cleanup Sweep
A single security-relevant fix closes a proxy gap in Ollama's file transfer path, while four separate documentation pull requests tidy up grammar and formatting in the README and API docs.
-
Weekly Recap - Claude Desktop Integration & MLX Engine Maturity
This week's work centered on two fronts: a major push to make Claude Desktop a first-class integration across macOS and Windows, and a series of correctness fixes hardening the MLX engine for structured output, model loading, and Windows…
-
Correctness Fixes Under the Hood
Today's activity centers on quiet correctness fixes — a GGUF parsing bug that silently corrupted tensor alignment, a Vulkan regression that broke model loading on integrated GPUs, and stricter but more forgiving API schema handling —…
-
Agent Expansion, Desktop Polish, and a Documentation Sweep
Ollama's agent and desktop surfaces both grew today, with new environment-aware tooling and Windows Claude Desktop support landing alongside a performance fix for structured output on the MLX runner. A separate wave of small…
- MLX Engine Grows Up
- Desktop App Stability and the Claude Integration Cleanup
- MLX Runner Hardens Up, Claude Desktop Gets Real Estate
- Claude Desktop Integration Overhaul
- Hardening the Runner Scheduler
- Weekly Recap - Desktop Apps Grow Up, Inference Gets More Reliable
- Visibility Into What the Model Actually Sees
- Fixing the Long-Prompt Timeout Trap
- Claude Desktop Gets Deep App Integration
- Closing the Gaps on Hung Requests and Broken Installs
- Closing the Gap on Model Metadata Overhead
- Silent Failures Get Loud
- Fixing the Small Gaps Between Local and Cloud
- Weekly Recap - Model Compatibility and the Push Into Coding Agents
- Silent Failures Get Louder
- Qwen Gets Serious About Coding Agents
- Launch Ecosystem Expands, Parser Edge Cases Get Squashed
- Backup Collisions and a Security Fix Converge
- Fixing the Defaults That Silently Hurt You
- Model Coverage and Tool-Call Reliability
- Vision Support Lands in the MLX Runner
- Weekly Recap - Vision Runners and Robustness Cleanup
- Edge Case Cleanup Across the Stack
- Speculative Decoding Gets a New Draft Architecture
- Crash Fixes and the Launcher Ecosystem Grows
- Silent Data Corruption Fixes
- Hardening the Thinking and Tool-Call Pipeline
- Closing the Gaps Where Failures Go Silent
- Streams That Fail Silently, Fixed
- Weekly Recap - Speed, Speculation, and Scheduler Reliability
- Scheduler Reliability Overhaul
- Speculative Decoding and the Cloud Model Nudge
- API Compatibility and Model Correctness Push
- Speculative Decoding Gains and a Lint Lockdown
- Cleaning Up Concurrency and Cutting Experimental Code
- Trust Your Cache, Fix Your Logs
- Parsing Bugs and Process Hardening
- Weekly Recap - New Model Support & Concurrency Hardening
- Closing the Gaps Between Local Config and Cloud Compatibility
- MLX Performance Push and a Scheduler Race Fix
- A Server Reliability Sweep
- Tool Calling Gets a Reliability Pass
- Scheduler Race Conditions Take Center Stage
- Command Line Consolidation and Hardware Fixes
- Thinking Models Leak Control Tokens
- Weekly Recap - Hardening the Runtime, One Edge Case at a Time
- Closing the Gaps Between API Layers
- Cleaning Up the Agent Loop and Widening Model Support
- Agent Package Gets a Major Consolidation
- Agent Hardening and a Path Traversal Fix
- Resource Leaks and GPU Placement Get a Clean-Up Pass
- Cleaning Up Memory Planning and the Launch Experience
- Fixing What Defaults Got Wrong
- Weekly Recap - Model Correctness and Runtime Hardening
- Tightening Up the Serving Layer
- Thinking Model Output Is Getting a Real Fix
- Stability Sweep Across Cloud, GPU, and Agent Tools
- Qwen3.5 Gets Untangled, and Small Bugs With Big Blast Radius
- Model Behavior Bugs and Blob Security Hardening
- Launch Gets Smarter, Model Support Gets Wider
- Cleaning Up the Scheduler's Edge Cases
- Weekly Recap - Rebuilding the Model Pipeline and Tightening the Guardrails
- Truth in Reporting
- MLX Create Pipeline Rewrite Lands
- Agent Harness Lands, Hardware Support Gets a Cleanup
- Gemma 4 Support and Platform Improvements
- Weekly Recap - MLX Performance & Path Handling
- Memory Management and Multimodal Parsing Fixes
- GPU Offloading and Tool Call Fixes
- Performance Optimizations and Model Handling Improvements
- Infrastructure Updates and Platform Fixes
- Multimodal Fixes and Developer Experience Updates
- Cache Architecture Overhaul and Data Race Fixes
- Developer Tools and Cross-Platform Reliability
- Weekly Recap - Integration Expansion & Server Reliability
- Audio Support and Infrastructure Refinements
- Integration Ecosystem and API Consistency Push
- Platform Integration Expansion and API Reliability Fixes
- Model Integration and Windows System Improvements
- LLaMA Server Integration Hardening
- Integration Platform Expansion
- Model Integration Updates
- Weekly Recap - Infrastructure Modernization
- Major Architecture Overhaul Removes CGO Dependencies
- MLX Model Display Fixes and Template Parser Cleanup
- Weekly Recap - Performance Optimization & Launch System Improvements
- DFlash Speculative Decoding Rollback
- Model Inventory Refactoring
- Startup Performance Optimization
- Codex Integration Enhancement
- Weekly Recap - MLX Performance & Codex Integration
- Release Build Optimization
Episode archive for Ollama
- Speculative Decoding and Codex App Updates
- MLX Sampler Overhaul and Codex Integration
- Vision Model Integration Enhancement
- MLX Threading and Claude Image Fixes
- Model Transfer Optimization and Test Reliability
- Claude Desktop Integration Removed
- Launch Command Enhancements
- Speed Revolution - MTP Decoding and Smart Caching
- Go 1.26 Runtime Update
- Weekly Recap - MLX Threading & Model Recommendations
- MLX Threading Fixes and Claude App Integration
- Model Recommendations and Windows Gateway Fix
- Metal GPU Stability and Gemma4 Updates
- Launch Experience Improvements and Model Recommendations
- Multi-Sequence Batching and New Model Support
- Tokenizer Bug Fix for BPE Processing
- Weekly Recap - MLX Performance & Launch Integrations
- MLX Sampling Performance Enhancement
- OpenAI Reasoning Integration
- Launch System Improvements and Integration Fixes
- Launch System Overhaul and Documentation Updates
- MLX Performance Boost and Model Updates
- New CLI Integration and Performance Improvements
- Weekly Recap - MLX Performance & Launch Integration Expansion
- MLX Sampler Improvements
- Windows WSL Integration Simplified
- Gemma4 Enhancements and Copilot CLI Integration
- Hermes Agent Integration and Gemma4 Improvements
- Gemma 4 MLX Support and Mixed-Precision Improvements
- Weekly Recap - Model Integration and Tooling Enhancements
- ROCm 7.2.1 Performance Update
- Gemma4 Parser Improvements
- Model Updates and Tool Call Fixes
- Error Handling and Modelfile Fixes
- Weekly Recap - Gemma4 Integration & Audio Support
- Performance Lessons and Gemma4 Refinements
- Gemma4 Arrives with Audio Magic
- Modernizing Codex Configuration
- Tokenizer Love and Better Model Support
- Legacy Compatibility and Developer Experience Wins
- Smoothing the Launch Experience
- Fixing the Inconsistencies That Matter
- Smart Caching and Better User Experience
- VS Code Integration Takes Center Stage
- Precision Revolution - New Float Formats and Testing Powerhouse
- MLX Performance Breakthrough and Smarter Caching
- Nvidia Partnership Takes Center Stage
- Bug Squashing Bonanza
- The Caching Revolution
- Bug Squashing and Launch Improvements
- Launch Command Gets a Major Polish
- Spring Cleaning and Performance Gains
- Thinking Streams and Local Tool Power-ups
- Stability First - Error Handling and Performance Fixes
- MLX Gets a Major Upgrade and Web Search Goes Live
- Simplifying the Sampling Story
- Cloud Models Get Smarter & Build Performance Boost
- Cloud Integrations Get Some Love
- Smarter Constraints and Qwen3.5 Boost
- Cloud Integration Drama and AI Model Expansion
- Smarter Sampling and Crash Prevention
- Building Bridges for Better Model Compatibility
- MLX Runner Gets Rock Solid
- Tool Calling Gets Smarter
- Cleaner Shutdowns and Faster Startups
- Qwen 3.5 Architecture Lands with Safety Upgrades
- Memory Management Revolution
- Nemotron Architecture Lands with Unified Cache Vision
- Fixing the WSL Plugin Problem
- Smarter UIs and Smoother Onboarding
- Tokenizer Consolidation & MLX Library Improvements
- Rolling Back and Rolling Forward
- Editor Integration Revolution
- MLX Display Bug Squashing Day
- MLX Runner Gets Major Model Upgrades
- MLX Performance Breakthrough and Anthropic Search
- MLX Runner Revolution and Documentation Polish
- Refactoring Rollercoaster and Developer Experience Wins
- Bug Squashing Bonanza
- Smooth Onboarding for New Users
- Polish and Perfectionism - The Art of Getting the Details Right
- Cleaning Up the Config Game
- Speed Boost and Model Magic
- Memory Magic and Command Makeover
- Making Ollama Play Nice with Everyone
- The Great Cleanup - Manifests Get Their Own Home
- New Model Architecture and Image Generation Fixes
- New Model Support and Memory Management Wins
- FLUX.2 Image Generation Arrives
- Image Generation Goes Native and Parser Cleanup Magic
- Dynamic Loading and Experimental Models Take Center Stage
- Release Day Rescue Mission
More public developer podcasts
Browse other public Podlog show hubs with crawler-readable RSS and episode links.
- Next Js Daily 56 episodes
- Headroom Daily 65 episodes
- Agent Of Empires Daily 56 episodes
- Buzz Transcription 81 episodes
- Frigate NVR Updates 199 episodes
- Linux Kernel 190 episodes
- Homebrew 211 episodes
- TypeScript 100 episodes
- Ruby on Rails 204 episodes
- Redis 179 episodes
- Next.js 213 episodes
- Vue.js 152 episodes
- Rust 183 episodes
- Ruby Core Updates 206 episodes
- Python 196 episodes
- Go 196 episodes
- VS Code 214 episodes
- Node.js 208 episodes
- Django 200 episodes
- React Native 157 episodes
- PostgreSQL 200 episodes
- Kubernetes 209 episodes
- LangChain 194 episodes
- PyTorch 213 episodes
- Openharness Daily 28 episodes
- Maestro Daily 55 episodes
- Navidrome Daily 74 episodes
- Next.js Daily 215 episodes
- Pi Mono 95 episodes
- OpenClaw 172 episodes
- RuView 133 episodes
- Shannon 85 episodes
- Godot Daily 160 episodes
- Agora Next Updates 230 episodes
- Rails Daily 214 episodes
- Home Assistant Daily 219 episodes
- Linux Kernel Daily 195 episodes
- React Daily 111 episodes
- Jabref Daily 66 episodes
- BlocksBeyondTheStars Daily 48 episodes
- Jabref Daily 67 episodes
- NativeScript iOS Daily 27 episodes
- Videorc Daily 53 episodes
- iCloud Photos Downloader 27 episodes
- Onlook Design Updates 34 episodes
- The Algorithm Daily 15 episodes
- TailwindCSS 119 episodes
- AtomVM Daily 15 episodes
- Agora Next Daily 12 episodes
- Energy Inc Ace Mate Learning Daily 2 episodes
- tiny-gpu Daily 20 episodes
- OpenAI Skills 52 episodes
- Crush Daily 1 episode
- ControlNet Daily 0 episodes
- FinceptTerminal Daily 0 episodes