Close Menu
    What's Hot

    Prompt Library Introduced on Bolt: Lets You Save Your Best Prompts

    July 11

    June 2025 release of Visual Studio Code: GitHub Copilot Chat Opensourced, MCP Support Generally Available

    July 11

    Grok 4 & SuperGrok Heavy Announced, Grok 4 Jailbreak Out Already?

    July 10
    Facebook X (Twitter) Instagram
    • AI Robots
    • AI News
    • Text to Video AI Tools
    • ChatGPT
    Facebook X (Twitter) Instagram Pinterest Vimeo
    Rad NeuronsRad Neurons
    • AI Robots
      • AI Coding
    • ChatGPT
    • Text to Video AI
    Subscribe
    Rad NeuronsRad Neurons
    Home » How to Use DeepSeek-R1 671billion Model for Agents
    AI News

    How to Use DeepSeek-R1 671billion Model for Agents

    AI NinjaBy AI NinjaJanuary 311 Min Read
    Share Facebook Twitter Pinterest LinkedIn Tumblr Reddit Telegram Email
    Share
    Facebook Twitter LinkedIn Pinterest Email

    DeepSeek R1 has generated a lot of excitements in the AI industry. While OpenAI is responding with o3 and o3-mini later today, plenty of companies are racing to add support for DeepSeek R1, including Perplexity and Windsurf. Thanks to NVIDIA, you can now try the the 671-billion-parameter DeepSeek-R1 model to build your own agents. Keep in mind, this is the model that is over 400GB if you try to run it locally.

    This model is now available as an NVIDIA NIM microservice in preview. It can deliver up to 3,872 tokens per second on a single NVIDIA HGX H200 system. As the company explains:

    Delivering real-time answers for R1 requires many GPUs with high compute performance, connected with high-bandwidth and low-latency communication to route prompt tokens to all the experts for inference. Combined with the software optimizations available in the NVIDIA NIM microservice, a single server with eight H200 GPUs connected using NVLink and NVLink Switch can run the full, 671-billion-parameter DeepSeek-R1 model at up to 3,872 tokens per second. 

    DeepSeek-R1 in Action with NVIDIA NIM Microservices

    [HT]

    DeepSeek
    Share. Facebook Twitter Pinterest LinkedIn Tumblr Email
    Previous ArticleCan DeepSeek R1 Play Chess? Tested Against LC0
    Next Article OpenAI to Announce o3 & o3-mini Models Today?
    AI Ninja
    • Website

    Related Posts

    AI News

    June 2025 release of Visual Studio Code: GitHub Copilot Chat Opensourced, MCP Support Generally Available

    July 11
    AI News

    Grok 4 & SuperGrok Heavy Announced, Grok 4 Jailbreak Out Already?

    July 10
    AI News

    KANAAN K1 Pro AI Glasses with OpenAI, Meta Support

    July 9
    Add A Comment
    Leave A Reply Cancel Reply

    Top Posts

    Leonardo’s AI Video Tool Gets Motion Control

    April 182 Views

    Leonardo’s Omni Editing Announced with FLUX.1 Kontext and GPT-Image-1

    May 302 Views

    OmniHuman-1 Generates Realistic Human Videos

    February 45 Views
    More
    AI News

    June 2025 release of Visual Studio Code: GitHub Copilot Chat Opensourced, MCP Support Generally Available

    AI NinjaJuly 11
    AI News

    Grok 4 & SuperGrok Heavy Announced, Grok 4 Jailbreak Out Already?

    AI NinjaJuly 10
    AI News

    KANAAN K1 Pro AI Glasses with OpenAI, Meta Support

    AI NinjaJuly 9
    Most Popular

    Prompt Cannon: Run Prompts Across Multiple Models

    June 24855 Views

    GPTARS: GPT Powered TARS Robot

    November 21533 Views

    Simple Grok 2 Jailbreak

    December 16472 Views
    Our Picks

    Prompt Library Introduced on Bolt: Lets You Save Your Best Prompts

    July 11

    June 2025 release of Visual Studio Code: GitHub Copilot Chat Opensourced, MCP Support Generally Available

    July 11

    Grok 4 & SuperGrok Heavy Announced, Grok 4 Jailbreak Out Already?

    July 10
    Tags
    3D 3D image agent AI AI glasses ai video canvas ChatGPT Chess Claude coding Computer Deep Research DeepSeek ElevenLabs Gemini Github glasses GPT GPT 4.5 Grok Hailuo humanoid image kling leonardo LLM MCP midjourney model music o3 offline open source pdf QWEN robot runway sora text to video Veo 2 Vibe coding video video model Voice

    © 2025 Rad Neurons. Inspired by Entropy Grid
    • Home
    • Terms of Use
    • Privacy Policy
    • Disclaimer

    Type above and press Enter to search. Press Esc to cancel.