Close Menu

    Subscribe to Updates

    Get the latest creative news from FooBar about art, design and business.

    What's Hot

    Reliable Insider Claims PlayStation Is Accepting Pitches For New Sly Cooper

    October 4, 2026

    Shaque producer confirms season 2 of Parineeti Chopra`s show

    October 4, 2026

    TPD arrest former felon who they say was in possession of stolen gun, drugs

    October 4, 2026
    Facebook X (Twitter) Instagram
    Facebook X (Twitter) Instagram YouTube TikTok
    Comic Vibe
    Sunday, October 4
    • Home
    • Comics
      • Comic Vibe News
    • Gaming
    • Movies
    • TV
    • Anime
    • Toys & Collectibles
    • Cosplay
    • Tech
    • Digital Culture
      • Creators & Fan Culture
      • Creator Economy & Fan-Driven Platforms
      • Digital Fandom & Online Communities
      • Metaverse & Virtual Worlds
      • NFTs & Digital Collectibles
      • Virtual Events & Online Conventions
      • Virtual Identity & Avatars
    • Shop
    Comic Vibe
    • Home
    • Contact Us
    • Terms & Conditions
    • Advertise With Us
    • DMCA Policy
    • Privacy Policy
    • About Us
    Home»Gaming»AI On Your Gaming PC
    Gaming

    AI On Your Gaming PC

    JamesBy JamesOctober 4, 2026No Comments3 Mins Read
    Facebook Twitter
    Share
    Facebook Twitter

    If you want to experiment with LLMs, you typically have a choice of sending your requests to someone else’s computer or fielding a very large GPU and CPU setup to run models locally. However, a recent crop of projects aims to bring bigger models to much more modest hardware.

    One example is Strata, a project from [Niko1221], which lets you run a 125-billion-parameter LLM on hardware you might already have for gaming. It won’t run on your old Pentium laptop, but it doesn’t require a supercomputer-like farm of graphics cards, either.

    Strata can use several Qwen3.8 model variants, including different quantizations of the original model as well as coding and other specialized versions. Qwen3.8-Flash-Next is a mixture-of-experts model containing 24,576 small experts, of which only ten are needed for each token. The clever part is that Strata effectively treats VRAM as a cache for the much larger model. Frequently used experts stay on the GPU, while the complete collection normally remains in system RAM. The model also includes a roughly 29 GB lookup table that stays on the SSD and is accessed as needed.

    The software also uses the model’s multi-token prediction machinery for speculative decoding, allowing several candidate tokens to be checked in a single pass. According to the project, an RTX 5070 with 12 GB of VRAM can produce roughly 50 to 90 tokens per second, depending on quantization. Tokens, of course, aren’t usually entire words, but it is still a respectable clip, once everything gets set up.

    We did have some trouble setting everything up due to some incompatibility with the NVIDIA C compiler, our gcc version, and some headers, but your problems will surely be different. The setup.sh file asks you a few questions on the first run. After that, it just handles your selected startup options, which can take a few minutes while everything loads.

    Once running, Strata lets you interact through a web browser. It also exposes OpenAI- and Anthropic-compatible APIs on localhost, so existing chat front ends, coding assistants, and other tools can use the local model without much special handling. Of course, you can’t expect its answers to compete with the big models out there for every task. When asking about Hackaday, for example, it got a lot of it right but also got confused about who founded the site and our authors (unless we forgot that [Tom Nardelli] once wrote some posts). Turning up the “thinking level” and turning down the temperature didn’t help much, although it did move its confusion to different facts. It did better when asked to identify some problem code or outline how to port a particular C compiler to a new target.

    You’ll still want at least 32 GB of system RAM, 12 GB of VRAM, and around 80 GB of storage, so “modest” is relative. Still, it’s a neat demonstration of how mixture-of-experts models and some clever memory management can stretch ordinary PC hardware surprisingly far.

    These economical LLMs can even run on older hardware, just slower.

    Gaming your
    Share. Facebook Twitter
    Previous ArticleImam Siddique objects to Salman Khan`s remarks about him on Bigg Boss 20
    Next Article In pictures – Daleks, Doctors and dazzling costumes take over Darlington Comic Con
    James

    Related Posts

    Reliable Insider Claims PlayStation Is Accepting Pitches For New Sly Cooper

    October 4, 2026

    Invincible VS Open Beta Begins October 13: What You Need to Know

    October 4, 2026

    Ninjala 2 dev want the franchise “to continue for 10 years, 20 years, and beyond”

    October 4, 2026

    PlayStation Permanently Shutting Down Popular Service Next Month

    October 4, 2026
    Leave A Reply Cancel Reply

    Our Picks

    Reliable Insider Claims PlayStation Is Accepting Pitches For New Sly Cooper

    October 4, 2026

    Shaque producer confirms season 2 of Parineeti Chopra`s show

    October 4, 2026

    TPD arrest former felon who they say was in possession of stolen gun, drugs

    October 4, 2026

    A Near-Perfect 1964 Avengers Comic Is Shockingly Worth Less Than It Was 10 Years Ago

    October 4, 2026
    • Facebook
    • Twitter
    • Instagram
    • YouTube
    • TikTok
    • Telegram
    Don't Miss
    Comic Vibe News

    Express Entertainment’s ‘Missing Darling’ ends with many layers

    By JamesJuly 15, 20260

    Crime thriller a welcome departure from family dramas, earns praise for its unconventional approach

    The 17 Sam Neill Performances to Watch Over and Over Again

    July 15, 2026

    5 DC Comics That Destroyed Their Main Characters (Including the Most Critically Acclaimed Comic Ever)

    July 15, 2026

    Tracee Ellis Ross Says She’s ‘Worthy of Choosing the Right Partner’

    July 15, 2026

    Subscribe to Updates

    Get the latest creative news from SmartMag about art & design.

    About Us
    About Us

    Comic Vibe is a pop-culture destination created for fans who live and breathe comics, movies, anime, TV shows, gaming, tech, cosplay, and collectibles.

    Our mission is to deliver engaging news, reviews, features, guides, and opinions that celebrate geek culture in all its forms. From the latest comic releases and blockbuster films to anime trends, gaming updates, cutting-edge tech, and collector culture, Comic Vibe brings everything together in one vibrant hub.

    Our Picks

    Reliable Insider Claims PlayStation Is Accepting Pitches For New Sly Cooper

    October 4, 2026

    Shaque producer confirms season 2 of Parineeti Chopra`s show

    October 4, 2026

    TPD arrest former felon who they say was in possession of stolen gun, drugs

    October 4, 2026

    Subscribe to Updates

    Get the latest comics, anime, movies, TV, gaming, cosplay, and pop culture news delivered directly to your inbox. No spam—just the stories every fan should know.

    Facebook X (Twitter) Instagram YouTube TikTok
    • Home
    • Contact Us
    • Terms & Conditions
    • Advertise With Us
    • DMCA Policy
    • Privacy Policy
    • About Us
    © 2026 Comic Vibe. Designed by Comic Vibe.

    Type above and press Enter to search. Press Esc to cancel.