Close Menu
    What's Hot

    Higgsfield Games 2.0 powered by GPT-6 Astra Announced

    September 7

    GPT‑6 Astra Debuts: Things to Do

    September 4

    Atlas Multimodal World Model with Camera Control Released

    September 2
    Facebook X (Twitter) Instagram
    • AI Robots
    • AI News
    • Text to Video AI Tools
    • ChatGPT
    Facebook X (Twitter) Instagram Pinterest Vimeo
    Rad NeuronsRad Neurons
    • AI Robots
      • AI Coding
    • ChatGPT
    • Text to Video AI
    Subscribe
    Rad NeuronsRad Neurons
    Home » Ollama Now Runs Fast on Apple Silicon with MLX
    AI News

    Ollama Now Runs Fast on Apple Silicon with MLX

    AI NinjaBy AI NinjaMarch 311 Min Read
    Share Facebook Twitter Pinterest LinkedIn Tumblr Reddit Telegram Email
    Share
    Facebook Twitter LinkedIn Pinterest Email

    Ollama can now run super fast on Apple devices using the MLX framework. It improves time to first token and tokens per second to make local AI run faster. Apple Silicon uses a shared memory pool for CPU and GPU. With MLX, Ollama uses that memory without copying data back and forth, which means you get faster inference.

    Ollama is now updated to run the fastest on Apple silicon, powered by MLX, Apple’s machine learning framework.

    This change unlocks much faster performance to accelerate demanding work on macOS:

    – Personal assistants like OpenClaw
    – Coding agents like Claude Code, OpenCode,… pic.twitter.com/WImO0lyYnp

    — ollama (@ollama) March 31, 2026

    This means you can now power OpenClaw with local models and get faster performance. It is also great for Claude Code, OpenCode, and Codex. With NVFP4 support, you get higher quality responses. There is also improved caching for better responsiveness.

    [HT]

    Apple Silicon MLX
    Share. Facebook Twitter Pinterest LinkedIn Tumblr Email
    Previous ArticleClawPC A1 Ryzen 5 OpenClaw PC
    Next Article Wan2.7-Image: Image Generation & Editing Model Announced
    AI Ninja
    • Website

    Related Posts

    AI News

    Higgsfield Games 2.0 powered by GPT-6 Astra Announced

    September 7
    AI News

    Atlas Multimodal World Model with Camera Control Released

    September 2
    AI News

    MiniMax Now Used for Interactive AI Livestreams

    August 31
    Add A Comment
    Leave A Reply Cancel Reply

    Top Posts

    ChatGPT’s Canvas Explained: Game-changing OpenAI Update

    October 340 Views

    YuE Open Source AI Music Service

    January 2916 Views

    Aero-1-Audio 1.5b Parameter Audio Language Model for Automatic Speech Recognition

    May 17 Views
    More
    AI News

    Higgsfield Games 2.0 powered by GPT-6 Astra Announced

    AI NinjaSeptember 7
    AI News

    Atlas Multimodal World Model with Camera Control Released

    AI NinjaSeptember 2
    AI News

    MiniMax Now Used for Interactive AI Livestreams

    AI NinjaAugust 31
    Most Popular

    Prompt Cannon: Run Prompts Across Multiple Models

    June 243,891 Views

    Dipal D1 2.5K Curved Screen 3D AI Character

    June 231,107 Views

    How to Use Claude in Unity & Unreal Engine with MCP

    March 191,043 Views
    Our Picks

    Higgsfield Games 2.0 powered by GPT-6 Astra Announced

    September 7

    GPT‑6 Astra Debuts: Things to Do

    September 4

    Atlas Multimodal World Model with Camera Control Released

    September 2
    Tags
    3D agent AI AI model ai video API avatar ChatGPT Claude Claude Code coding DeepSeek ElevenLabs Gemini glasses GPT Grok Hailuo Hermes Higgsfield image kling leonardo LLM MCP midjourney Minimax model music nano banana o3 offline OpenAI OpenClaw open source QWEN robot runway sora Veo 2 Veo 3 Vibe coding video video model Voice

    © 2026 Rad Neurons. Inspired by Entropy Grid
    • Home
    • Terms of Use
    • Privacy Policy
    • Disclaimer

    Type above and press Enter to search. Press Esc to cancel.