Close Menu
AIToday7

    Subscribe to Updates

    Get the latest creative news from FooBar about art, design and business.

    What's Hot

    How To Turn Off Gemini In Gmail In Your Google Account

    August 31, 2026

    Gemini Spark is almost a dream AI assistant — except for 1 thing Perplexity does better

    August 31, 2026

    Qi2 Campaign Challenges Consumers to Refuse the Same Old Crap

    August 31, 2026
    Facebook X (Twitter) Instagram
    Trending
    • How To Turn Off Gemini In Gmail In Your Google Account
    • Gemini Spark is almost a dream AI assistant — except for 1 thing Perplexity does better
    • Qi2 Campaign Challenges Consumers to Refuse the Same Old Crap
    • New Alabama Space Accelerator aims to move startups from lab to launchpad
    • The Best Gadgets of August 2026
    • Markets Brief: Value ETFs Making a Big Tech Bet, Nvidia’s Dealmaking, and the Cybersecurity Stock Outlook
    • Should We Be Polite to ChatGPT and AI in General?
    • AI could could cause global economic downturn, Andrew Bailey warns G20
    Facebook X (Twitter) Instagram Pinterest Vimeo
    AIToday7
    • Home
    • AI News
    • Tech News
    • AI Guides
    • Chatbots
    • Cybersecurity
    • Gadgets
    • More
      • Generative AI
      • Startups
    AIToday7
    Home»Generative AI»LLM Moats Quickly Evaporating
    Generative AI

    LLM Moats Quickly Evaporating

    aitoday7By aitoday7August 31, 2026No Comments2 Mins Read
    Share Facebook Twitter Pinterest LinkedIn Tumblr Reddit Telegram Email
    LLM Moats Quickly Evaporating
    Share
    Facebook Twitter LinkedIn Pinterest Email

    In the business world, a moat is a quality of a business that makes it difficult for competitors to take that company’s profits. With how hard it is to train models for large language models (LLMs) and generative AI, it might seem like Anthropic, Open AI, and other LLM companies would have huge moats given the amount of compute it takes to build models. But open source models are quickly draining that moat, and now the only thing standing in the way of a customer using one of these models on their own hardware instead one from the larger companies is physical computing resources. [TerminalBytes] demonstrates a few of these models on personally owned computers to show the current state of the art.

    [TerminalBytes] started off running the 27B version of the Qwen3.8 on a Mac Studio with 256 GB of unified RAM, which is plenty for this task. But it’s also enough to benchmark a few different models. Qwen3.6 is compared to 3.8, and then the different quants of each model are also compared. Quants are compressed versions of models that need fewer bits to store weights, meaning that the same models can run in less memory with smaller losses in fidelity. Many of these quants run on machines with 32 GB of RAM or less, encompassing many average gaming PCs. There’s even a 1-bit quant that [TerminalBytes] tested which can easily run on a machine with 16 GB, although with mixed results.

    Keep in mind that this is just the current state of affairs with open LLMs. Future versions of these models are likely to optimize the number of tokens produced per unit time, or otherwise increase quality of responses while requiring less computer rels will be the sole reason that the AI bubble pops, though. The fact that not every computer user is running Linux is proof enough of that

    Post Views: 8

    Evaporating Moats Quickly
    Share. Facebook Twitter Pinterest LinkedIn Tumblr Email
    Previous ArticleI asked ChatGPT if the S&P 500 will crash 50% due to the AI bubble and it said…
    Next Article AI could could cause global economic downturn, Andrew Bailey warns G20
    aitoday7
    • Website

    Related Posts

    Generative AI

    How To Turn Off Gemini In Gmail In Your Google Account

    August 31, 2026
    Generative AI

    Google could soon let you assign tasks to Spark from a regular Gemini chat

    August 5, 2026
    Add A Comment
    Leave A Reply Cancel Reply

    Top Posts

    How To Turn Off Gemini In Gmail In Your Google Account

    August 31, 20260 Views

    Gemini Spark is almost a dream AI assistant — except for 1 thing Perplexity does better

    August 31, 20260 Views

    Qi2 Campaign Challenges Consumers to Refuse the Same Old Crap

    August 31, 20260 Views
    Stay In Touch
    • Facebook
    • YouTube
    • TikTok
    • WhatsApp
    • Twitter
    • Instagram
    Latest Reviews
    Chatbots

    OpenAI CEO goes viral for more strange ChatGPT parenting advice

    aitoday7August 4, 2026
    AI News

    The latest AI news we announced in July 2026

    aitoday7August 5, 2026
    Tech News

    What is Trump Media’s Truth API and why is it controversial?

    aitoday7August 5, 2026

    Subscribe to Updates

    Get the latest tech news from FooBar about tech, design and biz.

    Most Popular

    How To Turn Off Gemini In Gmail In Your Google Account

    August 31, 20260 Views

    Gemini Spark is almost a dream AI assistant — except for 1 thing Perplexity does better

    August 31, 20260 Views

    Qi2 Campaign Challenges Consumers to Refuse the Same Old Crap

    August 31, 20260 Views
    Our Picks

    OpenAI CEO goes viral for more strange ChatGPT parenting advice

    August 4, 2026

    The latest AI news we announced in July 2026

    August 5, 2026

    What is Trump Media’s Truth API and why is it controversial?

    August 5, 2026

    Subscribe to Updates

    Get the latest creative news from FooBar about art, design and business.

    Facebook X (Twitter) Instagram Pinterest
    • About Us
    • Get In Touch
    • Disclaimer
    • Privacy Policy
    • Terms and Conditions
    © 2026 AIToday7. All Rights Reserved.

    Type above and press Enter to search. Press Esc to cancel.